I find it just misleading by advertising 700GB RAM as AI headline. I could plug a 32GB GPU to my 1.5TB ram server and call it “AI station with 1.5TB+ RAM” just so that you find it actually useless compared to the headlone
I don't know about "useless" (it seems quite useful to me) but I do feel mislead. It's unified memory in the same way that my current dGPU has unified memory. I guess nvlink-c2c probably (?) doesn't introduce a bottleneck but it's still two distinct arenas with very different performance characteristics.
7.1TB/s of HBM is not "useless" -- that's 30 times the memory bandwidth of my DGX Spark -- and nobody is expecting such a machine to run "Opus 5" on its own. For such large models even datacentre GB300 NVL72 are multiple trays linked together via NVlink etc. This machine has QSFP ports and ConnectX for linking up for larger models.
It's a workstation, not a rack. It's for AI researchers. I'd love to have one on my desk.
https://www.gigabyte.com/in/Enterprise/Tower-Server/W775-V10...
https://www.msi.com/Landing/NVIDIA-DGX-STATION
I think you might need a bit more than that at this price point ...
https://www.guru3d.com/story/nvidia-dgx-spark-achieves-175-f...
So it has only 256GB of actual ”AI” memory making it “useless”/toy for actual real world AI workloads(I.e it can’t replace something like opus 5)
It's a workstation, not a rack. It's for AI researchers. I'd love to have one on my desk.
What even is this comment?
What is it with those stupid names?
No price, so of course this is not for the smelly working class.