- GLM0%
BlockBeats News, August 18, according to MaxForAI, today, the AI project J-Space Cognition Suite that went viral on the X platform has been questioned by the community, with its promoted DeepSeek V4 test results being unreproducible. The project previously claimed that V4 Flash combined with J-Space could match GLM-5.3, while V4 Pro could outperform Fable 5 on multiple Agent Benchmarks, with a 2.53x speed improvement and a 2.21x Token efficiency boost.
GitHub user GoForceX used 87 questions from Terminal Bench 2.1 for high-concurrency retesting and confirmed the loading of J-Space-related modules. The results showed that after integrating J-Space, the benchmark performance slightly decreased, and both Token usage and costs increased, contrary to the project's claims of performance, speed, and Token efficiency improvements.
Subsequently, the community requested the project to publicly release full evaluation configurations, per-question results, run logs, raw execution times, and Token consumption data. Currently, the project has mainly released summarized results and has not provided complete original experimental records substantial enough to verify the precise data mentioned above. Of note, in response to the skepticism, the project's authors admitted that the related data was indeed "exaggerated," stating that the actual improvements were roughly between 1.6x to 3x. As of now, the authors have not issued a formal response to the community's questions, and some relevant skeptical Issues have been deleted.
免責事項:現在のコンテンツは第三者の視点に基づくもの、または第三者の視点からAIが直接翻訳したものです。CoinExはコンテンツの信頼性、正確性、独創性を保証するものではなく、CoinExからの投資アドバイスを構成するものではありません。暗号資産の価格変動は急激に変動します。潜在的なリスクにご注意ください。
- コインリスト価格24時間価格変動