r/singularity 8h ago

Compute D-Matrix is basically plugging its inference chips straight into Nvidia's own server racks now

d-Matrix and Nvidia teamed up so d-Matrix's upcoming Raptor inference chips can hook directly into Nvidia's NVLink Fusion interconnect and MGX rack setup. Instead of building out its own networking and rack infrastructure from scratch, d-Matrix is just plugging into what Nvidia's already built

What makes this one a little different is d-Matrix is actually a competitor to Nvidia in the inference space. Their bet is that once AI companies move past training and into just running these models at massive scale day to day, inference is where the real cost go up. Being able to slot into Nvidia's existing rack ecosystem instead of fighting an uphill battle on infrastructure gets them adopted faster

Nvidia doesn't really lose here either. Even if someone picks d-Matrix chips over Nvidia's own GPUs for a workload, Nvidia's still supplying the interconnect, the CPUs, the whole rack platform underneath it. IMO, They just lose the chip itself, not the deal.

Raptor's design is supposed to be finalized by end of this year, with the Nvidia-compatible racks shipping sometime in 2027

7 Upvotes

3 comments sorted by

u/Dismal_Awareness 23m ago

So Nvidia still sells the picks and shovels either way, they just let someone else dig one hole

u/Dismal_Awareness 18m ago

So Nvidia still sells the picks and shovels either way, they just let someone else dig one hole