Fireworks AI is now offering NVIDIA's Nemotron 3 Ultra model on its inference platform, providing day-zero support for the new model. Nemotron 3 Ultra is designed for complex, long-running tasks such as coding agents and deep research, featuring a hybrid Transformer-Mamba architecture and up to 1 million token context. NVIDIA claims the model offers five times faster inference and up to 30% lower costs for agentic tasks compared to similar open models. AI
IMPACT Accelerates deployment of advanced agentic AI for complex, long-running tasks.
RANK_REASON This is a product launch for an inference platform offering a specific model, not a core model release from a frontier lab.
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →