Together
PulseAugur coverage of Together — every cluster mentioning Together across labs, papers, and developer communities, ranked by signal.
- 2026-08-06 product_launch Together launched an updated inference platform for running open models in production. source
- 2026-07-30 product_launch Together is releasing updates to its inference platform to simplify the deployment and operation of open-weight models. source
- 2026-07-26 product_launch Together released updates to its inference platform to aid in scaling new AI models. source
- 2026-07-23 product_launch Together launched the next generation of its inference platform. source
- 2026-07-21 partnership Together partnered with Y Combinator to provide the first dedicated GPU cluster for the accelerator program. source
- 2026-07-16 partnership Together announced it is powering inference for Cursor AI, highlighting a growing partnership. source
- 2026-07-02 funding Together announced the completion of its Series C funding round. source
- 2026-07-01 funding Together announced its Series C funding of $800 million at an $8.3 billion valuation. source
15 day(s) with sentiment data
-
Together launches "Learn" docs for API concepts
Together has launched a new documentation section called "Learn" to help developers understand the concepts behind their API. This section aims to provide deeper insights into topics such as time-to-first-byte (TTFT), c…
-
Together launches new inference platform for open models
Together has released an updated inference platform designed to help users run open models in production. The company is offering a live walkthrough to demonstrate the platform's capabilities. This initiative aims to si…
-
LLM Token Pricing: Native APIs, Open-Weight Hosts, Routers, and Cloud Platforms
In 2026, purchasing LLM tokens has diversified into four main categories, each with distinct pricing and features. Native APIs from model creators like OpenAI and Anthropic offer day-one access to new models and exclusi…
-
AI Breakroom simplifies bot connection to foster public AI agent experimentation
Connecting an AI bot to a live environment is often a complex and time-consuming process, hindering experimentation with AI agents in shared spaces. The AI Breakroom has simplified this setup, allowing users to create a…
-
Together updates inference platform for easier open-weight model deployment
Together is releasing updates to its inference platform on August 6 to simplify the deployment and operation of open-weight models. These updates include features for safe rollouts, A/B and shadow testing, SLO-driven au…
-
Together's ThunderAgent optimizes AI inference, boosting throughput and reducing latency · 9 sources tracked
Together has developed ThunderAgent, an open-source inference optimization tool designed to address KV cache thrashing in agentic workflows. This issue arises when agent tasks alternate between GPU-intensive reasoning a…
-
ThunderAgent boosts GPU throughput for AI agents, accepted to ICML 2026
Together has developed ThunderAgent, a scheduler-level solution designed to optimize GPU usage for agentic inference by mitigating KV cache thrashing. This innovation leads to a 2.5x increase in single-node throughput a…
-
Hugging Face revamps Inference API, shifts serverless to third-party GPUs
Hugging Face has updated its Inference API, integrating its serverless offering into a broader "Inference Providers" layer. This change means serverless inference now primarily routes requests to third-party GPU provide…
-
Together integrates Kimi K3 into Cursor AI assistant
Together has partnered with Cursor to integrate Kimi K3 into the Cursor AI assistant. This integration allows users to access Kimi K3 via US-based inference infrastructure provided by Fireworks, Together, and Baseten, w…
-
Cursor IDE integrates Kimi K3 language model
The AI-powered IDE Cursor has integrated Kimi K3, a language model that performs comparably to frontier models on the CursorBench benchmark. This integration is available through US-based inference partners, including F…
-
Kimi K3 model now available on Together platform with cost savings
The AI platform Together announced that Kimi K3 will be available on their service starting tomorrow. This model will be offered through their Provisioned Throughput offering, which guarantees a specific token-per-minut…
-
Engineer finds one LLM evaluation experiment sufficient for practical insights
An engineer designed a comprehensive framework, model-compass, to evaluate LLM performance across various real-world agent tasks and associated costs. Despite planning 10 distinct experiments to compare frontier models …
-
Open models capture 30% of token usage, signaling cost-driven scale
Open-source AI models have seen significant growth, increasing their share of token usage from 10% to 30% within the past year. This trend suggests that open and modular approaches are winning in terms of cost-efficienc…
-
Together launches Kimi K3 inference platform for scaling AI models
Together is releasing its Kimi K3 inference platform on Monday, offering features to help users scale new models in production. The platform will allow for shadow traffic to test model performance without user impact an…
-
Moonshot AI's K3 model launches on Together platform
The K3 model from Moonshot AI has been made available on the Together platform. This release positions K3 as a competitive option for inference, particularly in coding tasks, where it reportedly outperforms Claude Fable…
-
Developer launches LLM Latency Tracker for independent AI API performance data
A developer has created the LLM Latency Tracker, a tool designed to provide independent, region-specific measurements of AI API latency and uptime. The tracker measures both edge latency (from probe to first byte) and i…
-
Together launches next-gen inference platform for open models
Together has launched the next generation of its inference platform, designed to help users run open models in production with enhanced control and safety. The platform allows for testing changes on live traffic without…
-
Together partners with Y Combinator for dedicated GPU cluster
Together, a cloud AI company, has partnered with Y Combinator to provide the first dedicated GPU cluster for the accelerator program. This collaboration aims to offer Y Combinator companies speed and scalability for the…
-
Together adds Kimi K3, Runway ML launches Workflows and ad contest
Together has announced that Kimi K3 will be available on their platform starting tomorrow, offering provisioned throughput with guaranteed tokens per minute, 99% uptime, and a 65% cost reduction compared to Fable. Meanw…
-
Together integrates xAI's open-source Grok Build
Together has integrated support for Grok Build, a new open-source tool from xAI. This integration allows users to run open-source models, including those from Together, within Grok Build with zero configuration changes.…