Friday, September 11, 2026
AI 인프라 · 뉴스 & 분석
반도체·하드웨어리포트
반도체·하드웨어 · 리포트

d-Matrix가 차세대 Raptor XPU 추론 가속기를 랙 레벨 배포로 확장하기 위해 NVIDIA NVLink Fusion 상호 연결 아키텍처의 채택을 발표했습니다.

NVIDIA 에코시스템을 채택하는 추론 가속화 스타트업(NVLink Fusion + MGX 랙 표준)은 NVIDIA 에코시스템 락인을 검증하며; XPU 배포 타이밍은 H100과 H200에 대한 경쟁력 있는 추론 대안을 가능하게 한다.
공식 공시Slicast · September 11, 2026 · 미국 · 출처: NVIDIA Blog
중요도 75

AI inference chipmaker d-Matrix announced today it will use NVIDIA NVLink Fusion to connect its next-generation Raptor XPUs to NVIDIA's AI infrastructure platform, joining a growing roster of ecosystem partners.

By connecting Raptor to NVIDIA NVLink scale-up and Spectrum-X scale-out networking through the NVIDIA MGX rack architecture, NVLink Fusion gives d-Matrix an accelerated, lower-risk path from custom silicon to large-scale deployment.

"Demand for inference is soaring, but capital, time and energy remain finite," said Sid Sheth, cofounder and CEO of d-Matrix. "With NVLink Fusion and MGX, we can integrate our Raptor XPUs into a broadly deployed, liquid-cooled architecture, giving customers a faster, lower-risk path to deploy and scale ultralow-latency inference."

The NVIDIA AI platform is vertically integrated and horizontally open. NVLink Fusion extends this openness to XPUs and CPUs, allowing silicon companies to focus on processor innovation while using NVIDIA infrastructure for AI factory-scale deployment.

Building an XPU is only the first step. Deploying it at scale requires a complete platform spanning networking, rack architecture, power, cooling, software and a proven supply chain. Each integration point—sourcing chips, adding high-speed interfaces, validating scale-up networking, designing and certifying rack architecture—adds time, cost and risk.

NVLink Fusion lets silicon innovators connect directly into NVIDIA's proven platform via high-bandwidth, low-latency technology that links custom XPUs and CPUs to the NVIDIA stack. By adopting NVLink Fusion, d-Matrix taps into the NVIDIA MGX ecosystem's mature, validated rack designs, supply chain, power and cooling infrastructure. Standardizing on a common rack enables data centers to support GPUs, CPUs and XPUs without requiring separate architectures for each processor type.

d-Matrix plans to connect its XPUs in a single high-bandwidth, low-latency scale-up domain using NVIDIA NVLink. Its racks can also work alongside NVIDIA GPU systems like the NVIDIA Vera Rubin NVL72 for disaggregated inference. The company also plans to integrate NVIDIA Vera CPUs, NVIDIA ConnectX-9 SuperNICs, NVIDIA BlueField-4 DPUs and NVIDIA Spectrum-X Ethernet networking—technologies that together provide a proven foundation to deploy specialized inference alongside NVIDIA systems within flexible, unified AI factories.

The NVIDIA full-stack AI factory platform—including NVIDIA Vera Rubin NVL72, Groq 3 LPX, Vera CPU rack, BlueField-4 STX storage and Spectrum-6 SPX Ethernet—is designed to run every AI workload and model architecture with best-in-class performance per watt and lowest cost per token. NVLink Fusion gives customers the flexibility to match the right compute to each workload within this common platform, opening access to NVIDIA networking, systems, software and global supply chain. Silicon innovators like d-Matrix can use NVLink Fusion to increase performance, accelerate time to market and reduce deployment risk for semi-custom AI factories.

원문 보기