动态

5.2版本用更多RL可改进。

jietang
5.2 could be better with more RL ...
青龍聖者
Deepswe's benchmark results are my own experience.
I've used all models,
GLM 5.2 ≈ Claude Opus 4.6–4.7.
Kimi 2.7 code more like inference optimization.
Looking forward to K3.
Doubao-seed 2.1 Pro around 37% ≈ Gemini 3.5 Flash.
code are quite weak, but visual are strong. https://t.co/BDMuaC9n5T
动态jietang2026-06-24原文

相关内容