> ## Content Index
> Fetch the complete content index at: https://www.theleftshift.com/llms.txt
> Use this file to discover other available public pages before exploring further.

# d-Matrix Partners With NVIDIA to Bring Inference XPUs Into AI Factory Racks
- URL: https://www.theleftshift.com/d-matrix-partners-with-nvidia-to-bring-inference-xpus-into-ai-factory-racks/
- Published: 2026-09-11T09:52:51.000Z
- Updated: 2026-09-11T09:52:51.000Z
- Description: d-Matrix will integrate Raptor with NVIDIA’s latest rack architecture, which includes Vera CPUs, NVLink switches, BlueField-4 DPUs, ConnectX-9 SuperNICs and Spectrum-X Ethernet networking.
- Author: The Left Shift Bureau
- Tags: Partnerships, Chips

AI inference chipmaker d-Matrix has [announced](https://www.d-matrix.ai/announcements/d-matrix-rackscale-nvidia/?ref=theleftshift.com) a multi-year collaboration with NVIDIA that will integrate its next-generation inference XPUs into NVIDIA’s widely deployed AI factory infrastructure.

The collaboration will initially centre on d-Matrix’s Raptor XPU being integrated into an [NVIDIA MGX](https://www.nvidia.com/en-in/data-center/products/mgx/?ref=theleftshift.com) rack-scale system enabled by NVLink Fusion. The system is for AI labs, hyperscalers and neocloud providers running latency-sensitive AI services where faster responses can command a premium.

[Microsoft’s New AI Tool Targets the Pain of Moving Off SalesforceSupport for other customer relationship management (CRM) platforms and enterprise resource planning (ERP) migration scenarios is planned for later.![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/icon/assets_task_01jvq6mt6ee65vskk3794wghn2_1747756766_img_3-1-4726ec01-f0f8-4dff-a351-039a28e26a55.webp)The Left ShiftThe Left Shift Bureau![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/thumbnail/Microsoft---s-New-AI-Tool-Targets-the-Pain-of-Moving-Off-Salesforce-1ecb7b6f-93fd-4e36-b72f-c29d093153de.webp)](https://www.theleftshift.com/microsofts-new-ai-tool-targets-the-pain-of-moving-off-salesforce/)

As an NVIDIA NVLink Fusion partner, [d-Matrix](https://www.theleftshift.com/d-matrix-launches-jetstream-network-card-to-accelerate-ai-inference-in-data-centers/) will integrate Raptor with NVIDIA’s latest rack architecture, which includes Vera CPUs, NVLink switches, BlueField-4 DPUs, ConnectX-9 SuperNICs and Spectrum-X Ethernet networking. 

> *“Being integrated into NVIDIA’s latest MGX rack-scale infrastructure with NVLink Fusion means our customers can deploy our inference XPUs alongside the broadly available NVIDIA AI factory platform. That’s the future d-Matrix has been building toward—ultra-low latency, energy-efficient inference XPUs and GPUs working together, at rack scale, to deliver premium AI experiences,”* [*Sid Sheth*](https://www.linkedin.com/in/sheth/?ref=theleftshift.com)*, d-Matrix Founder and CEO, said.*

> *“With NVIDIA AI infrastructure deployed across cloud and on-premises data centers worldwide, NVLink Fusion gives partners like d-Matrix a path to integrate seamlessly with NVIDIA compute platforms — expanding accelerator choice for customers building the next generation of AI factories,”* [*Jensen Huang*](https://www.linkedin.com/in/jenhsunhuang/?ref=theleftshift.com)*, NVIDIA Founder and CEO, added.*

The partnership comes as AI providers increasingly explore heterogeneous infrastructure, using different processors for different stages of inference. d-Matrix said its rack architecture can pair its Raptor XPUs with [NVIDIA](https://www.theleftshift.com/nvidia-unveils-ai-model-to-help-robotaxis-handle-complex-real-world-driving-scenarios/) GPUs, allowing operators to optimise workloads based on performance and latency requirements.

For AI coding applications, for example, GPUs can handle the compute-intensive prefill stage, while Raptor XPUs can accelerate the decode stage, where response latency is critical.

[AM Intelligence Orders 9,000 NVIDIA Rubin GPUs for Hyderabad AI FactoryAMI said the Hyderabad facility is expected to become one of the first frontier AI compute clusters in Asia.![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/icon/assets_task_01jvq6mt6ee65vskk3794wghn2_1747756766_img_3-1-0cce2b03-5ba8-4988-a8ff-3c058a710d4e.webp)The Left ShiftThe Left Shift Bureau![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/thumbnail/Image-Low-Res-1-37fa6438-6345-446d-84d0-1235d720be96.jpeg)](https://www.theleftshift.com/am-intelligence-orders-9-000-nvidia-rubin-gpus-for-hyderabad-ai-factory/)

Raptor is the successor to d-Matrix’s Corsair XPU and uses the company’s memory-centric architecture. Its key technology combines a DRAM memory chip with an SRAM compute chip in a single package using a 3D stacking approach.

Raptor is expected to tape out before the end of 2026, with initial availability of Raptor XPUs integrated into NVIDIA MGX racks expected in Q4 2027\. The company said the platform is already being evaluated by AI hyperscalers and frontier labs and is backed by more than 100 patents.

The collaboration gives d-Matrix access to NVIDIA’s established rack architecture and supply chain while giving customers another option for scaling inference workloads alongside NVIDIA GPUs.

Earlier this year, [d-Matrix acquired GigaIO’s data centre business](https://www.theleftshift.com/d-matrix-acquires-gigaios-data-centre-business-to-expand-rack-scale-ai-capabilities/), a systems engineering organisation with deep expertise in rack-scale infrastructure and high-performance interconnects.

[The AI Scaling Problem Enterprises Aren’t Talking AboutProduction AI demands a fundamentally different environment, one where compute, networking, storage, power and cooling are designed to work together rather than treated as separate technology layers.![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/icon/assets_task_01jvq6mt6ee65vskk3794wghn2_1747756766_img_3-1-9024c1b4-a951-46be-ba00-785a8bf77c12.webp)The Left ShiftPritam Bordoloi![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/thumbnail/The-AI-Scaling-Problem-Enterprises-Aren---t-Talking-About-e97abfb3-3dba-42f0-b729-0d94f88e7b54.png)](https://www.theleftshift.com/the-ai-scaling-problem-enterprises-arent-talking-about/)

[Just 1% of Customers Drive 80% of OpenAI, Anthropic’s Enterprise RevenueThe companies making up the top 1% of customers are heavily concentrated in technology, AI products, and AI services.![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/icon/assets_task_01jvq6mt6ee65vskk3794wghn2_1747756766_img_3-1-22274247-4373-4e77-8db8-587087937f3c.webp)The Left ShiftThe Left Shift Bureau![](https://storage.ghost.io/c/fd/78/fd78b498-7373-46b7-884c-46f69e81b24d/content/images/thumbnail/Just-1--of-Customers-Drive-80--of-OpenAI--Anthropic---s-Enterprise-Revenue-acbaf33f-a5b0-4d89-a6d0-1d9bd5186e54.jpg)](https://www.theleftshift.com/just-1-of-customers-drive-80-of-openai-anthropics-enterprise-revenue/)