GPT-5.5编码性能优于Claude Sonnet,DeepSWE评分70%
Many developers have suspected for months that GPT-5.5 outperforms Claude Sonnet for coding. But SWE-Bench reported near-parity, and it made people question what they’d been seeing in practice.
DeepSWE aligns more closely with that day-to-day experience: GPT-5.5 scores 70% https://t.co/bDmRTCgIJ5
DeepSWE aligns more closely with that day-to-day experience: GPT-5.5 scores 70% https://t.co/bDmRTCgIJ5