- GLM0%
- NVDAX0%
- GPU0%
- Token-54.39%
Dynamic Beating AI News Flash: Following the GLM-5.3 Flash release, Intellisense has confirmed a previously undisclosed detail: during the Ox Alpha anonymous testing phase, all traffic was powered by a domestically produced Chinese AI chip for inference computation.
Ox Alpha, launched on OpenRouter, processed 23.2 trillion Tokens in just 6 full calendar days, more than double the throughput of the concurrent DeepSeek-V4-Flash. OpenCode had even suggested earlier that there was the capability to provide a daily free quota of 100 trillion Tokens.
Intellisense stated that they have optimized end-to-end inference performance on the same Chinese domestic hardware to initially triple the efficiency, and now the hardware efficiency and per-Token cost have approached levels close to mainstream NVIDIA GPUs.
SemiAnalysis specifically pointed this out. Previously, observers speculated whether the supply capacity of 100 trillion Tokens per day could only be sustained by top-tier labs and NVIDIA GPUs. However, this time, the entire process was run using domestically produced Chinese chips.
Disclaimer: The current content is sourced from third-party perspectives or directly translated by AI from third-party perspectives. CoinEx does not guarantee the authenticity, accuracy, and originality of the content, and it does not constitute any investment advice from CoinEx. The prices of cryptocurrencies are highly volatile, please be aware of the potential risks.
- CoinsPrice24H Change