Photo: Matheus Bertelli / Pexels
D-Matrix targets Nvidia MGX rack integration for Raptor XPU by Q4 2027
The AI inference startup plans to slot its custom accelerators into Nvidia's rack infrastructure, promising a nearly 5x throughput advantage over conventional HBM designs
AI inference chip startup d-Matrix wants its upcoming Raptor XPU to live inside Nvidia’s MGX rack architecture by the end of 2027, and the company has a fairly concrete roadmap to get there. The Santa Clara-based firm expects its first Raptor tape-out to land before the close of 2026, with full rack-scale deployment targeting Q4 2027.
D-Matrix is integrating Raptor with Nvidia’s NVLink Fusion interconnect, which means up to 144 Raptor XPUs could operate inside a single NVLink fabric by the time 2027 wraps up.
What Raptor is actually built to do
Raptor debuted publicly at Hot Chips 2026 in August, and the technical specs are worth paying attention to. The platform combines a TSMC 4nm logic die with 3D-stacked DRAM, targeting over 100 TB/s of memory bandwidth with 32 GB of capacity per card.
D-Matrix claims Raptor delivers approximately 4.7 times higher throughput per card compared to HBM-based alternatives for generative inference tasks.
AI, tech, and the markets they move—in one daily briefing.
Daily. Free. Join 34,000+ readers across crypto, finance, and policy.
The Corsair foundation and the funding behind it
Raptor is d-Matrix’s next act, but the company is not a pre-revenue concept. Its current platform, Corsair, is already in production and uses digital in-memory compute built on SRAM.
The company has raised between $450 million and $500 million in total funding, including a $275 million Series C round that valued the company at roughly $2 billion. Partners in the broader ecosystem include Alchip and Andes, and the Raptor integration specifically involves Astera Labs, a connectivity chipmaker that sits inside Nvidia’s technology ecosystem.