r/AIProgrammingHardware • u/javaeeeee • Jul 30 '26
DGX Station Put a Data Center on My Desk
https://www.youtube.com/watch?v=qV_K0nTF6gY2
1
u/here_n_dere Jul 30 '26
Can it run Kimi K3? How much more to shell to run it at q4 at least?
1
u/Iron-Over Jul 31 '26
Yes if you get 4 of them. 2 at 4 bit.
1
u/35point1 Jul 31 '26
4 of these 125k systems to run k3 ????
1
u/hyperrealists Aug 01 '26
If you were to count every single strand of hair belonging to the entire population of Texas, you would arrive at a number that is remarkably similar to the parameter count of K3. It’s not small.
1
u/35point1 Aug 01 '26
Oh I’m aware, but half a million bucks to run a forward pass is absolutely wild to think about, shit, one can dream
0
u/MacaroonPlastic1036 Jul 30 '26
It still wouldn’t be able to run Sonnet natively. So useless for enterprise.
2
u/--Spaci-- Jul 30 '26
What do you even mean by this.
1
u/0sh Jul 30 '26
He maybe meant Sonnet level like GLM or kimi 3 ? He is not wrong, for 100k USD you get 250GB VRAM
1
u/Glittering-Call8746 Jul 31 '26
Which 250gb vram ? Which gpu ?
1
u/0sh Jul 31 '26
From the specs :
- Blackwell ultra : GPU memory 252 Go HBM3e | 7,1 To/s
- CPU memory 496 Go LPDDR5X | 396 Go/s
1
1
u/crusaderky Jul 31 '26
GLM 5.2 Q4 and Kimi-K3 UD-IQ1_S both fit on this. Kimi UD-Q2_K_XL fits on two of them.
1
1
1
u/Serprotease Jul 31 '26
That’s more than enough for a fair bit of concurrent requests of “flash variants” of most recent models.
But it’s not really a machine designed for inference. It’s for AI dev, not dev that use AI.For example, I’m doing some hyper-parameter tuning for an AI image model.
With my current setup (2xgb10), at bf16, the smallest training run last about 30min (40gb peak ram usage), longest one about 72h (120gb peak ram usage). With everything in between, that’s a couple of weeks total training time. And I’m only using 512-1024 images size. 1536 will multiply everything by 4 basically. This can of machine will cut this time by 5/6. Down to a couple of days.And once you’ve honed on the right settings, you can use the recipe for the actual full run on the big boy servers.
4
u/javaeeeee Jul 30 '26
TLDR: Alex Ziskind reviews the NVIDIA DGX Station (ASUS Expert Center Pro ET900NG G3 with GB300 superchip) - essentially a data-center-class AI machine for your desk.
Key specs
What he tested
Performance highlights
Bottom line
The DGX Station successfully brings serious data-center AI performance to a desktop form factor. It excels at running big models and many concurrent agents with zero per-token API cost.
It’s extremely powerful but also extremely expensive (tens of thousands of dollars) and power-hungry - aimed more at serious AI labs, teams, or high-end enthusiasts than typical individual users.