- GLM0%
- NVDAX0%
- GPU0%
- Token-54.39%
Dynamic Beating AI News Flash: Following the GLM-5.3 Flash release, Intellisense has confirmed a previously undisclosed detail: during the Ox Alpha anonymous testing phase, all traffic was powered by a domestically produced Chinese AI chip for inference computation.
Ox Alpha, launched on OpenRouter, processed 23.2 trillion Tokens in just 6 full calendar days, more than double the throughput of the concurrent DeepSeek-V4-Flash. OpenCode had even suggested earlier that there was the capability to provide a daily free quota of 100 trillion Tokens.
Intellisense stated that they have optimized end-to-end inference performance on the same Chinese domestic hardware to initially triple the efficiency, and now the hardware efficiency and per-Token cost have approached levels close to mainstream NVIDIA GPUs.
SemiAnalysis specifically pointed this out. Previously, observers speculated whether the supply capacity of 100 trillion Tokens per day could only be sustained by top-tier labs and NVIDIA GPUs. However, this time, the entire process was run using domestically produced Chinese chips.
免責聲明:當前內容均來自第三方觀點或由AI直接翻譯第三方觀點,CoinEx不保證內容的真實性、準確性和原創性,不構成CoinEx相關的任何投資建議。數字資產價格波動劇烈,請注意潛在風險。
- 幣種價格24H漲跌