BTC $77,401.35 -1.31%
ETH $2,410.36 -2.19%
BNB $686.07 -0.93%
XRP $1.35 -1.96%
SOL $99.98 -3.25%
TRX $0.3217 -3.07%
DOGE $0.0812 -1.92%
ADA $0.1976 -1.14%
BCH $247.82 +0.26%
LINK $11.21 -1.26%
HYPE $82.98 -0.99%
AAVE $128.58 +3.49%
SUI $0.7231 -0.86%
XLM $0.1758 -1.00%
ZEC $835.92 -2.00%
BTC $77,401.35 -1.31%
ETH $2,410.36 -2.19%
BNB $686.07 -0.93%
XRP $1.35 -1.96%
SOL $99.98 -3.25%
TRX $0.3217 -3.07%
DOGE $0.0812 -1.92%
ADA $0.1976 -1.14%
BCH $247.82 +0.26%
LINK $11.21 -1.26%
HYPE $82.98 -0.99%
AAVE $128.58 +3.49%
SUI $0.7231 -0.86%
XLM $0.1758 -1.00%
ZEC $835.92 -2.00%

x

All
Article
Flash

first_img Google claims that the cost of AI server memory has exceeded 75%, promoting a dual-track strategy for software and hardware

The SEMICON Taiwan 2026 Memory Summit took place on the 1st, where Nikhil Cherian, Senior Director of Supply Chain Infrastructure at Google Cloud under Alphabet, pointed out that with the popularity of multimodal and mixed expert architectures, AI computation has shifted from being power-limited to memory-limited, with high-performance memory accounting for over 75% of the cost of AI server hardware bill of materials. In the face of capacity, bandwidth, and power consumption bottlenecks, Google is breaking through the AI memory bottleneck through a dual-track strategy of hardware offloading for inference and training, and lossless quantization software algorithms.Google adopts an offloading strategy in hardware architecture, launching TPU 8i for low-latency inference and TPU 8t specialized for large-scale training. The TPU 8i is equipped with 288 GB of high-bandwidth memory, with SRAM capacity on the chip increased threefold to 384 MiB, placing dynamic conversation states and key-value caches on the chip itself to achieve zero chip-off latency. The TPU 8t forms a super-large computing cluster with 9600 chips, achieving a shared pool of HBM at a scale of 2 PB, eliminating chip-off data transfer bottlenecks, along with TPU Direct Storage technology.Google has developed the training-free TurboQuant lossless quantization algorithm, compressing the key-value cache of large models from 32 bits to 3 bits, reducing memory usage by six times without loss of accuracy, resulting in an eightfold acceleration in attention computation, and integrating old-generation DRAM technology to extend the lifecycle of components.

first_img The core ARR of China's open-source large models is approximately 6-8 billion USD

Robonomics author FD published the China Open-Source LLM Tracker on September 1, 2026, stating that the total ARR of China's open-source large models is approximately $10-15 billion, with core LLM revenue around $6-8 billion after excluding ByteDance / Seedance. Growth has not relied on price wars; after DeepSeek raised prices by about 3-12 times, usage still increased, and the average price of the Zhipu API rose by 101% while token usage exceeded 40 times within the year. The best estimate for overseas revenue is approximately $2 billion.Zhipu MaaS ARR increased from about $250 million in March to an annualized monthly rate of about $1.6 billion in August, with a weekly annualized rate of $2 billion. August revenue has already surpassed its API revenue for the first half of the year; the API gross margin is 24.6%, and inference costs per token decreased by 80%. MiniMax ARR rose from about $100 million in December 2025 to over $800 million in annualized weekly revenue by August 2026, with B2B accounting for about 80%. Kimi ARR increased from about $100 million in March to about $300 million by mid-June. DeepSeek's revenue from January to July was approximately $70 million, with the latest ARR estimated at about $1.2 billion.ByteDance's annualized AI revenue is approximately $4 billion, of which Seedance accounts for about $2 billion.

hot_img UBS: It will be difficult for China to develop EUV lithography equipment comparable to ASML's top products within the next ten years; DUV mass production may take two to five years

According to Bloomberg, UBS analysts expect that China's semiconductor manufacturing capabilities may struggle to make sufficient progress in the next decade to develop solutions that can truly replace ASML's advanced EUV lithography technology. UBS analysts wrote in their report, "The stage they are currently at seems comparable to ASML's in 2004." Based on patent application activities, especially in the fields of light sources and laser subsystems, China's current technological maturity is roughly equivalent to ASML's level 15 years before it began mass production of EUV equipment.UBS also expects that China is likely to achieve large-scale production capabilities for immersion DUV lithography machines within the next two to five years. While immersion DUV is not as advanced as EUV, it remains an indispensable tool for chip manufacturing. ASML's immersion DUV machines are priced at nearly $90 million, while EUV equipment costs over $200 million each. China is one of ASML's largest markets, but due to export restrictions, ASML is prohibited from selling EUV equipment to China. UBS pointed out that considering the yield and capacity gaps, as well as regulatory restrictions, Chinese lithography equipment is unlikely to be applied overseas. An ASML spokesperson did not respond to requests for comment. This report is based on UBS analysts' predictive views, and actual progress will depend on industry dynamics.

first_img Silhouette launches the xStocks inquiry trading system on Hyperliquid

According to The Block, Hyperliquid's block trading layer Silhouette launched its Request for Quote (RFQ) trading system on the mainnet on Tuesday, initially supporting tokenized stocks from Payward's tokenized stock framework xStocks. Silhouette stated that the system allows traders to initiate quote requests for any supported xStock and receive competitive quotes from connected market makers, with winning trades able to settle on-chain at any time and in any size.The RFQ system enables trading of supported xStocks without a dedicated order book, and tokens that generate sufficient trading activity can subsequently be upgraded to an independent HyperCore market. Silhouette founder Chandler De Kock mentioned that while tokenized stocks continue to go on-chain, they mostly lack trading venues, and market makers compete for each trade, with settlement completed on-chain, proving that assets with real liquidity will be upgraded to their own HyperCore market.xStocks was launched in June 2025 and is said to have processed over $40 billion in total trading volume, covering more than 200,000 holders, with nearly $20 billion settled on-chain. The platform issues publicly listed stock tokens pegged 1:1, and its business has expanded from U.S. stocks to European and Asian markets. RWA.xyz data shows that the existing tokenized stocks amount to approximately $2.53 billion, with xStocks accounting for about $620.1 million, ranking as the third-largest issuer, following Ondo (approximately $856.3 million) and bStocks (approximately $621.5 million).
app_icon
ChainCatcher Building the Web3 world with innovations.