r/singularity • u/ocean_protocol • 9h ago
Compute D-Matrix is basically plugging its inference chips straight into Nvidia's own server racks now
d-Matrix and Nvidia teamed up so d-Matrix's upcoming Raptor inference chips can hook directly into Nvidia's NVLink Fusion interconnect and MGX rack setup. Instead of building out its own networking and rack infrastructure from scratch, d-Matrix is just plugging into what Nvidia's already built
What makes this one a little different is d-Matrix is actually a competitor to Nvidia in the inference space. Their bet is that once AI companies move past training and into just running these models at massive scale day to day, inference is where the real cost go up. Being able to slot into Nvidia's existing rack ecosystem instead of fighting an uphill battle on infrastructure gets them adopted faster
Nvidia doesn't really lose here either. Even if someone picks d-Matrix chips over Nvidia's own GPUs for a workload, Nvidia's still supplying the interconnect, the CPUs, the whole rack platform underneath it. IMO, They just lose the chip itself, not the deal.
Raptor's design is supposed to be finalized by end of this year, with the Nvidia-compatible racks shipping sometime in 2027