- GPT0%
- MUSE0%
- MINI0%
- SPK0%
- OPUS0%
Perceive Beating AI News Flash: Shortly after the release of Gemini 3.8 Flash, Google achieved a score of 73.7% on DeepSWE v1.1. The DeepSWE official leaderboard currently displays it as 74%±1%, ranking above Claude Opus 5 and GPT-5.6 Sol. Several hours later, Meta announced Muse Spark 1.3, reaching 75.4% on the same DeepSWE suite, pushing the public score up by another 1.7 percentage points.
This test involved the Agent autonomously handling long-term software engineering tasks in 113 real code repositories. Both Google and Meta used the mini-swe-agent, with Google running on high inference mode and Meta on max. Meta's previous generation, Spark 1.2, scored only 55.0%, seeing a direct increase of 20.4 percentage points this time. However, Meta's 75.4% has not yet been included in the DeepSWE official leaderboard.
In the Artificial Analysis Coding Agent Index, the preview-limited Spark 1.3 max received a score of 68, ranking just below Claude Opus 5; the currently publicly available xhigh version scored 64.
면책 조항: 현재 콘텐츠는 제3자 관점에서 제공되거나 제3자 관점에서 AI가 직접 번역한 것입니다. CoinEx는 콘텐츠의 진위성, 정확성, 독창성을 보장하지 않으며 CoinEx의 투자 조언으로 간주하지 않습니다. 암호화폐 가격은 변동성이 크므로 잠재적인 위험에 유의하시기 바랍니다.
- 코인가격24시간 변동