待验证50% 置信事实精确时间
Gemini 3.1 Pro High在Sonar评估中综合质量最优,代码简洁度较好,圈复杂度为234
1
来源数
50%
置信度
中期 (~90 天)
时效性
2026/7/2
首次发现
有效期至:2026/9/30
来源
Sonar实测53个大模型代码质量:Claude安全漏洞最多,GPT-5代码量暴增5倍
bilibiliAI_Express2026/6/1
相关事实
待验证Gemini 3.1 Pro High在SWE-bench测试中通过率为84.17%,排名第一,生成代码量为307,000行,Bug密度为614个/百万行代码,安全问题为210个/百万行代码72% 相似待验证Gemini Pro/Ultra系列在复杂推理、大规模代码分析方面表现更优71% 相似待验证Gemini 3.1 Pro High showed the most balanced performance and is described as the best overall quality choice among evaluated models68% 相似待验证用户认为Gemini 3.1 Pro的回复高度可读、内容组织清晰、逻辑层次分明68% 相似待验证Gemini 3.5 Pro在ARC-AGI-2通用智能评估中得分42分,远超同期竞品65% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/51219API
curl https://kongchang.com/api/v1/knowledge/claims/51219MCP
get_claim(id=51219)