d-Matrix가 차세대 Raptor XPU 추론 가속기를 랙 레벨 배포로 확장하기 위해 NVIDIA NVLink Fusion 상호 연결 아키텍처의 채택을 발표했습니다.
AI inference chipmaker d-Matrix announced today it will use NVIDIA NVLink Fusion to connect its next-generation Raptor XPUs to NVIDIA's AI infrastructure platform, joining a growing roster of ecosystem partners.
By connecting Raptor to NVIDIA NVLink scale-up and Spectrum-X scale-out networking through the NVIDIA MGX rack architecture, NVLink Fusion gives d-Matrix an accelerated, lower-risk path from custom silicon to large-scale deployment.
"Demand for inference is soaring, but capital, time and energy remain finite," said Sid Sheth, cofounder and CEO of d-Matrix. "With NVLink Fusion and MGX, we can integrate our Raptor XPUs into a broadly deployed, liquid-cooled architecture, giving customers a faster, lower-risk path to deploy and scale ultralow-latency inference."
The NVIDIA AI platform is vertically integrated and horizontally open. NVLink Fusion extends this openness to XPUs and CPUs, allowing silicon companies to focus on processor innovation while using NVIDIA infrastructure for AI factory-scale deployment.
Building an XPU is only the first step. Deploying it at scale requires a complete platform spanning networking, rack architecture, power, cooling, software and a proven supply chain. Each integration point—sourcing chips, adding high-speed interfaces, validating scale-up networking, designing and certifying rack architecture—adds time, cost and risk.
NVLink Fusion lets silicon innovators connect directly into NVIDIA's proven platform via high-bandwidth, low-latency technology that links custom XPUs and CPUs to the NVIDIA stack. By adopting NVLink Fusion, d-Matrix taps into the NVIDIA MGX ecosystem's mature, validated rack designs, supply chain, power and cooling infrastructure. Standardizing on a common rack enables data centers to support GPUs, CPUs and XPUs without requiring separate architectures for each processor type.
d-Matrix plans to connect its XPUs in a single high-bandwidth, low-latency scale-up domain using NVIDIA NVLink. Its racks can also work alongside NVIDIA GPU systems like the NVIDIA Vera Rubin NVL72 for disaggregated inference. The company also plans to integrate NVIDIA Vera CPUs, NVIDIA ConnectX-9 SuperNICs, NVIDIA BlueField-4 DPUs and NVIDIA Spectrum-X Ethernet networking—technologies that together provide a proven foundation to deploy specialized inference alongside NVIDIA systems within flexible, unified AI factories.
The NVIDIA full-stack AI factory platform—including NVIDIA Vera Rubin NVL72, Groq 3 LPX, Vera CPU rack, BlueField-4 STX storage and Spectrum-6 SPX Ethernet—is designed to run every AI workload and model architecture with best-in-class performance per watt and lowest cost per token. NVLink Fusion gives customers the flexibility to match the right compute to each workload within this common platform, opening access to NVIDIA networking, systems, software and global supply chain. Silicon innovators like d-Matrix can use NVLink Fusion to increase performance, accelerate time to market and reduce deployment risk for semi-custom AI factories.