动态

GPT-5.5编码性能优于Claude Sonnet,DeepSWE评分70%

GPT-5.5编码性能优于Claude Sonnet,DeepSWE评分70%
Peter Steinberger 🦞
Many developers have suspected for months that GPT-5.5 outperforms Claude Sonnet for coding. But SWE-Bench reported near-parity, and it made people question what they’d been seeing in practice.

DeepSWE aligns more closely with that day-to-day experience: GPT-5.5 scores 70% https://t.co/bDmRTCgIJ5
动态Peter Steinberger 🦞2026-05-27原文

相关内容