DigestAI news desk
OpenAI board member warns company is not on track to prevent catastrophic AI loss of control OpenAI launches Agents API beta for long-running cloud agents OpenAI solves Navier-Stokes problem, sparking academic controversy over data use RTK Token Savings Debunked: Cost Benchmarks Disagree The Waymo effect: AI making research less collaborative Meta’s Muse AI Agent Seeks User Trust with Secure Architecture Thelio Mira AI Linux Workstation with 192 GB GPU Memory Claude No Longer Available to Minors
Hardware & Compute updated 2 min read

d-Matrix adopts NVIDIA NVLink Fusion to connect Raptor XPUs to MGX rack platform

d‑Matrix, a specialist in AI inference silicon, announced that its upcoming Raptor XPU family will be linked to NVIDIA’s AI infrastructure via NVLink Fusion. The move ties the custom XPUs into NVIDIA’s MGX rack architecture, Spectrum‑X networking, and the broader AI factory stack, giving customers a unified, liquid‑cooled platform for ultra‑low‑latency inference.

1 source primary source

Key points

  • d-Matrix will connect its Raptor XPU line to NVIDIA’s MGX rack architecture via NVLink Fusion.
  • The integration promises ultralow‑latency inference with shared power, cooling and supply‑chain infrastructure.
  • d‑Matrix will also use NVIDIA Vera CPUs, ConnectX‑9 NICs, BlueField‑4 DPUs and Spectrum‑X Ethernet in the same racks.

CEO Sid Sheth highlighted that soaring inference demand is constrained by capital, time and energy, and that the NVLink Fusion integration reduces risk and accelerates deployment. d‑Matrix will also bundle NVIDIA Vera CPUs, ConnectX‑9 SuperNICs, BlueField‑4 DPUs and Spectrum‑X Ethernet, allowing the same rack to host GPUs, CPUs and XPUs without separate designs. The partnership leverages NVIDIA’s proven supply chain, power and cooling solutions, aiming to lower cost per token and improve performance‑per‑watt for AI workloads.

The announcement underscores a growing trend toward heterogeneous compute in data centers, where custom silicon can plug into a common, vendor‑validated ecosystem rather than building bespoke infrastructure from scratch.

The story so far

7 episodes →
  1. d-Matrix adopts NVIDIA NVLink Fusion to connect Raptor XPUs to MGX rack platform this story
Full story from NVIDIA Blog · by Jesse Clayton primary source Open source ↗

d-Matrix Adopts NVIDIA NVLink Fusion for Rack-Scale XPU Deployment

NVIDIA Blog · 10 September 2026

AI inference chipmaker d-Matrix today announced it will use NVLink Fusion to connect its next-generation Raptor XPUs to NVIDIA’s AI infrastructure platform — joining a growing roster of partners building on the NVIDIA AI platform.

By connecting Raptor to NVIDIA NVLink scale-up and Spectrum-X scale-out networking, the NVIDIA MGX rack architecture and the broader NVIDIA AI platform, NVLink Fusion gives d-Matrix an accelerated, lower-risk path from custom silicon to large-scale deployment.

“Demand for inference is soaring, but capital, time and energy remain finite,” said Sid Sheth, cofounder and CEO of d-Matrix during a press briefing yesterday. “With NVLink Fusion and MGX, we can integrate our Raptor XPUs into a broadly deployed, liquid-cooled architecture, giving customers a faster, lower-risk path to deploy and scale ultralow-latency inference.”

The NVIDIA AI platform is vertically integrated and horizontally open. NVLink Fusion extends this openness to XPUs and CPUs, allowing silicon companies to focus on their processor innovations while using NVIDIA infrastructure to deploy them at AI factory scale.

From Custom Silicon to Rack-Scale Deployment

Building an XPU is only the first step. Deploying it at AI factory scale requires a complete platform spanning networking, rack architecture, power, cooling, software and a proven supply chain. Each of the steps — sourcing chips, integrating high-speed interfaces, validating a scale-up networking solution, designing and certifying rack architecture — adds time, costs and risks.

NVLink Fusion lets silicon innovators connect directly into NVIDIA’s proven platform. It’s the high-bandwidth, low-latency technology that connects custom XPUs and CPUs to the NVIDIA stack.

By adopting NVLink Fusion, d-Matrix can tap into the NVIDIA MGX ecosystem’s mature, validated rack designs, supply chain, power and cooling infrastructure. By standardizing on a common rack, data centers can be built once and support GPUs, CPUs and XPUs — without requiring a separate rack architecture for each processor type.

NVLink Fusion Integrates d-Matrix XPUs Into AI Factories

Using NVIDIA NVLink, d-Matrix plans to connect its XPUs in a single high-bandwidth, low-latency scale-up domain. Its racks can also work alongside NVIDIA GPU-based systems like NVIDIA Vera Rubin NVL72 for disaggregated inference.

d-Matrix also plans to integrate NVIDIA Vera CPUs, NVIDIA ConnectX-9 SuperNICs, NVIDIA BlueField-4 DPUs and NVIDIA Spectrum-X Ethernet networking. Together with NVLink and MGX, these technologies give d-Matrix a proven foundation to deploy specialized inference alongside NVIDIA systems within flexible, unified AI factories.

NVLink Fusion Opens NVIDIA AI Factories to Specialized XPU Architectures

The NVIDIA full-stack AI factory platform includes NVIDIA Vera Rubin NVL72, Groq 3 LPX, the Vera CPU rack, Vera BlueField-4 STX storage and Spectrum-6 SPX Ethernet networking. It’s designed to be completely fungible — running every AI workload, model and model architecture — with the best performance per watt and lowest cost per token.

NVLink Fusion gives customers the flexibility to match the right compute to each workload within a common AI factory platform, opening access to NVIDIA networking, systems, software and global supply chain. Silicon innovators like d-Matrix can use NVLink Fusion to increase performance, accelerate time to market and reduce the risk of deploying semi-custom AI factories.

This text was published by NVIDIA Blog and written by Jesse Clayton. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗

Topics · follow one to build your own front page

The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us.

Comments

via GitHub Discussions

Related stories