Sunday, October 4, 2026
AI 인프라 · 뉴스 & 분석
홈 › 반도체·하드웨어 › 리포트
반도체·하드웨어 · 리포트

Nvidia는 추론 가속기 Rubin을 공개했으며, 이전 세대 대비 추론 비용을 최대 10배 절감할 수 있다고 발표했다.

10x 추론 비용 감소는 AI 워크로드 경제를 재편할 수 있으며, 엣지 추론 배포를 경제적으로 실행 가능하게 하고 Nvidia의 진출 가능 시장을 확대할 수 있습니다.
업계 전문지Slicast · 2026년 10월 3일 04:18 UTC · 미국 · 출처: NewsBytes
중요도 80

NVIDIA introduced Rubin, its latest AI supercomputing platform, at CES 2026. The system targets a critical market challenge: making advanced AI development and deployment significantly more accessible and cost-effective.

Rubin delivers substantial performance advantages. The platform achieves up to 10 times lower inference token costs compared to previous generations and requires only one-fourth the graphics cards that NVIDIA's Blackwell system needs to train mixture-of-experts models. The platform integrates six chips into a single system, featuring an 88-core Vera CPU and a GPU capable of up to 50 petaflops of performance.

The architecture prioritizes infrastructure efficiency through improved design and faster interconnections, directly addressing the cost barriers that have historically limited AI adoption among enterprises and cloud providers. AWS, Google Cloud, and Microsoft are positioned to benefit from these cost reductions, broadening access to advanced AI capabilities.

Rubin is scheduled to roll out to partners in the second half of 2026.

원문 보기
Nvidia는 추론 가속기 Rubin을 공개했으며, 이전 세대 대비 추론 비용을… · Slicast