شراء العملة
الأسواق
فوري
العقود
الأرباح
الأنشطة
المزيد
reward-centerمنطقة المبتدئين
الصفحة الرئيسية للتغذيةتفاصيل الفلاش
Terminal-Bench 4.0: GLM-5.3 Surges to Third Place, Surpassing GPT-5.6
  • GPT0%
  • OPUS0%
  • GLM0%

Dynamic Beating AI News Flash: Terminal-Bench has released version 4.0, recalibrating the time, CPU, and memory when the agent performs tasks, and fixing 19 tasks. It has removed 8 tasks with saturation, denial, openly disclosed solutions, or quality issues. The maximum execution time for all tasks has been standardized to 8 hours, mainly to reduce interference from timeouts and environmental issues on performance.

In the latest leaderboard, Opus 5 + Claude Code ranks first with 51.8%, followed by Fable 5 at 44.5%. GLM-5.3 + Claude Code scored 41.8%, claiming the third spot, surpassing GPT-5.6 Sol + Codex at 37.3%. Among the top three, GLM-5.3 is the only model not from Anthropic.

In Terminal-Bench 3.0, GLM-5.3 ranked fourth with 32.4%, trailing GPT-5.6 Sol at 34.6%; by version 4.0, GLM-5.3 has climbed to third place, leading Sol by 4.5 percentage points in turn.

المصدر:BlockBeats

إخلاء المسؤولية: يتم الحصول على المحتوى الحالي من وجهات نظر خارجية أو تتم ترجمته مباشرة بواسطة الذكاء الاصطناعي من وجهات نظر خارجية. لا تضمن CoinEx صحة المحتوى ودقته وأصالته، ولا تشكل أي نصيحة استثمارية من CoinEx. أسعار العملات المشفرة متقلبة للغاية، يرجى الانتباه إلى المخاطر المحتملة.

عمليات البحث الساخنة
  • العملات
    السعر
    التغيرات في 24 ساعة