- NVDAX0%
- UOS0%
BlockBeats News, August 30th, Citrini analyst Jukan posted that, according to sources, the HBM specification used by NVIDIA's Rubin Ultra may be reduced from 12-layer HBM4E to 8 layers. Customers such as OpenAI and Anthropic even requested a 4-layer product but were rejected by the memory manufacturer. The current downspec may be limited to 8 layers. The main reason for the downspec is believed to be yield and cost pressure: if 12-layer HBM4E were used with price increases taken into account, the memory cost could account for approximately 70% of Rubin Ultra's overall material cost.
Software optimizations such as model quantization, MLA, and compute task partitioning are transferring low-frequency access KV caches and model states to LPDDR, CXL, and NAND, while HBM mainly retains the working set required for current computations. Therefore, after meeting the minimum capacity, customer emphasis on HBM bandwidth is beginning to outweigh capacity.
It is believed that reducing the stack layers can increase packaging yield, increase HBM and AI accelerator shipments, and may actually expand total HBM demand; higher bandwidth requirements will also reduce the proportion of chips in the wafer sorted by speed grades, further consuming DRAM wafer capacity. In the long term, HBM will eventually be replaced by new architectures, and the ultimate direction may be the fusion of storage and logic chips. The next two years will be a crucial stage to see if memory manufacturers can expand into the logic domain.
면책 조항: 현재 콘텐츠는 제3자 관점에서 제공되거나 제3자 관점에서 AI가 직접 번역한 것입니다. CoinEx는 콘텐츠의 진위성, 정확성, 독창성을 보장하지 않으며 CoinEx의 투자 조언으로 간주하지 않습니다. 암호화폐 가격은 변동성이 크므로 잠재적인 위험에 유의하시기 바랍니다.
- 코인가격24시간 변동