3 ms·
LPDDR6... that's not enough for speedy inference
by khalic 2mo ago
LPDDR6... that's not enough for speedy inference
- rjzzleep 2mo agoMaybe,but memory bandwidth is much better than on the DGX Spark at least.
- khalic 2mo agoI hate this RAMpocalypse so much
- fauigerzigerk 2mo agoOn the other hand, this year's RAMpocalypse could be 2028's RAMbundance ;P
- khalic 2mo agoRAMvana was right there!!
- kcb 2mo agoNo. its has a 96 bit bus.
- simlevesque 2mo agoNot every computing platform is expected to do fast inference.
- CharlesW 2mo agoIt depends on the number of channels. The Xring O3 appears to have 4×24-bit channels, so 113.8 GB/s. The iPhone 17 Pro's memory bandwidth is ~76.8 GB/s. Seems fine?
- rpdillon 2mo agoIf I recall correctly, Strix Halo gets about 450G/s. The M series processors get between 400 and 800G/s, and an actual NVIDIA card is like 5.5T/s. 113G/s is pretty slow for inference.
- trvz 2mo agoStrix Halo is 256 GB/s.
- rpdillon 2mo agoThank you! Got my wires crossed: Strix Halo at 256G/s, M1 at 450G/s. Apologies, and thanks for the correction!
- dust42 2mo agoThe CPU is for mobile phones. An Nvidia H200 has 4.8TB/s. An RTX 5090 has 1.8TB/s. Both use ~700W - not comparable with a phone.
- GeekyBear 2mo agoAccording to Mark Gurman, the base M6 is bumped up to 200 GB/s of memory bandwidth with a single core performance bump of 15%.
- hoherd 2mo agoOh thank heavens, computing hardware that the AI companies will not buy all of the supply of.