The Register · 4h ago
The Register · 4小时前
Qualcomm is proposing a next-generation AI accelerator approach that places compute close to, or physically under, DRAM. The Register frames it as Qualcomm's attempt to attack the memory wall rather than compete only on raw GPU-style compute. If the architecture becomes product silicon, it would put memory bandwidth, packaging density, and thermal design at the center of Qualcomm's AI infrastructure bid.
Qualcomm提出一种下一代AI加速器思路,把计算单元放到DRAM附近,甚至放在DRAM下方。The Register将其描述为Qualcomm试图绕开只拼算力的路线,直接针对memory wall。若该架构进入量产硅片,内存带宽、封装密度和热设计将成为Qualcomm AI基础设施方案的核心。
Wccftech · 5h ago
Wccftech · 5小时前
Nvidia says software tuning on Blackwell systems has reduced DeepSeek v4 token costs by up to 5x about one month after launch. The number matters because cost per token is becoming the practical purchasing metric for inference clusters, not just peak FLOPS. The claim also shows how much of Blackwell's near-term value depends on compiler, kernel, and serving-stack work after hardware ships.
Nvidia称,通过Blackwell系统的软件调优,DeepSeek v4上线约一个月后token成本最高下降5倍。这个数字重要,是因为推理集群采购越来越看每个token的成本,而不只是峰值FLOPS。它也说明Blackwell的短期价值很大一部分来自硬件交付后的编译器、kernel和服务栈优化。
Wccftech · 3h ago
Wccftech · 3小时前
AMD is rolling out Versal Gen 2 with on-package LPDDR5X memory, positioning it as a response to constrained and expensive HBM supply. The candidate report cites a 60% smaller board footprint, 10x compute uplift versus Versal Gen 1, PCIe 6 support, and a 15-year lifecycle. The move is not a direct replacement for HBM accelerators, but it shows more vendors designing around memory availability as a first-order constraint.
AMD推出带封装上LPDDR5X内存的Versal Gen 2,把它定位为应对HBM供应紧张和价格高企的方案。候选报道提到,相比Versal Gen 1,新方案板级占用面积减少60%,算力提升10倍,支持PCIe 6,并提供15年生命周期。它并不是HBM加速器的直接替代品,但显示更多厂商开始把存储可得性当作一阶设计约束。