PulseAugur
EN
LIVE 07:14:59

Together AI offers MiniMax M3 multimodal model with 1M context

Together AI has announced the availability of the MiniMax M3 API, an open-weight, native multimodal model. This model boasts a 1 million token context window, enhanced by MiniMax Sparse Attention, and features distinct "thinking" and "non-thinking" modes for complex reasoning and latency-sensitive tasks, respectively. Together AI is highlighted as MiniMax's preferred cloud partner, offering optimized inference for the M3 model. AI

IMPACT This release provides developers with a powerful multimodal model featuring a large context window and optimized inference, potentially accelerating development in areas requiring complex reasoning and handling of diverse data types.

RANK_REASON The cluster announces the availability of a new multimodal model with advanced features like a 1M context window and sparse attention, which falls under research and product release.

Read on X — Together (inference / OSS) →

AI-generated summary · Google Gemini · from 8 sources. How we write summaries →

Together AI offers MiniMax M3 multimodal model with 1M context

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
The cluster announces the availability of a new multimodal model with advanced features like a 1M context window and sparse attention, which falls under research and product release.
Source corroboration
8 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
89 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

Full methodology in our editorial standards.

COVERAGE [8]

  1. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    Highlights:

    Highlights: 👉 Stronger real-world coding performance over Kimi K2.6 👉 ~30% lower thinking-token usage for better token efficiency 👉 256K context for larger repos and longer coding sessions 👉 Interleaved thinking and multi-step tool calling for coding agents

  2. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    Try it now: https://t.co/5w9OrtVrvx

    Try it now: https://t.co/5w9OrtVrvx

  3. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    Try it now: https://t.co/V6tQnswpBw

    Try it now: https://t.co/V6tQnswpBw

  4. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    Highlights:

    Highlights: 👉 Native multimodality across text, image, and video 👉 1M context with MiniMax Sparse Attention for long-context workloads 👉 9x prefill and 15x decode speedups vs M2 at 1M context 👉 Thinking mode for complex reasoning, non-thinking mode for latency-sensitive chat

  5. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    MiniMax-M3 from @MiniMax_AI is now available on Together AI.

    MiniMax-M3 from @MiniMax_AI is now available on Together AI. It’s an open-weight native multimodal model with 1M context, MiniMax Sparse Attention, and thinking / non-thinking modes. Together AI is MiniMax’s preferred cloud partner, with inference optimizations delivering up to…

  6. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    Highlights:

    Highlights: 👉 Native multimodality across text, image, and video 👉 1M context with MiniMax Sparse Attention for long-context workloads 👉 9x prefill and 15x decode speedups vs M2 at 1M context 👉 Thinking mode for complex reasoning, non-thinking mode for latency-sensitive chat

  7. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    Try it now: https://t.co/V6tQnswpBw

    Try it now: https://t.co/V6tQnswpBw

  8. X — Together (inference / OSS) TIER_1 English(EN) · togethercompute ·

    MiniMax-M3 from @minimax_ai is now available on Together AI.

    MiniMax-M3 from @minimax_ai is now available on Together AI. It’s an open-weight native multimodal model with 1M context, MiniMax Sparse Attention, and thinking / non-thinking modes. Together AI is MiniMax’s preferred cloud partner, with inference optimizations delivering up ht…