NVIDIA has announced the global expansion and ramp-up of its Vera Rubin platform, designed for large-scale AI factories. This system, built with extreme co-design across seven chips and five rack trays, aims to deliver high performance per watt and the lowest token cost. Early benchmarks from partners like CoreWeave show a significant improvement in throughput compared to previous generations. The Vera Rubin platform is being deployed with major cloud providers and is also powering new AI infrastructure initiatives in Europe, including a partnership with Microsoft and Mistral. AI
IMPACT Accelerates the deployment of large-scale AI infrastructure, potentially lowering token costs and improving performance per watt for AI factories.
RANK_REASON NVIDIA's announcement of its Vera Rubin platform, detailing its architecture, performance claims, and partner deployments, constitutes a significant infrastructure and product release from a frontier AI lab. [lever_c_demoted from frontier_release: ic=4 ai=0.7]
- Azure
- ConnectX-9
- CoreWeave
- DeepSeek-R1
- Google.Cloud
- Grace Blackwell NVL72
- Groq 3 LPX
- Microsoft
- NVIDIA Vera CPU
- NVIDIA Vera Rubin
- NVLink
- Olympus core
- Oracle Cloud
- SpaceXAI
- Spectrum 6 (1025)
- Spectrum-6 SPX
- Spectrum-X Ethernet
- Vera BlueField-4 STX
- Vera Rubin NVL72
- NVIDIA
- OpenAI
- Vera Rubin
AI-generated summary · Google Gemini · from 4 sources. How we write summaries →