Claude Sonnet 4.5
PulseAugur coverage of Claude Sonnet 4.5 — every cluster mentioning Claude Sonnet 4.5 across labs, papers, and developer communities, ranked by signal.
- developed by Anthropic 100%
- instance of Claude Fable-5 90%
- instance of Claude Opus 4.1 90%
- competes with DagsHub 70%
- used by ScienceCast 70%
- competes with alphaXiv 70%
- used by Gotit.pub 70%
- used by Claude API 70%
- competes with Gotit.pub 70%
- competes with DeepSeek-R1 70%
- competes with Opus 4.8 70%
- competes with ScienceCast 50%
- 2026-09-05 product_launch Anthropic released Claude Sonnet 4.5 with a 200K token context window and extended thinking mode. source
- 2026-06-08 product_launch A startup encountered critical API system failures after upgrading to Claude Sonnet 4.5. source
- 2026-05-26 product_launch Anthropic removed the Claude Sonnet 4.5 model from its claude.ai interface. source
- 2026-05-25 research_milestone Claude Sonnet 4.5 outperformed GPT-4.1 and Gemini 2.5 Pro in a real-world coding benchmark. source
- 2026-05-15 product_launch Anthropic is decommissioning the Sonnet 4.5 model. source
- 2026-05-12 product_launch Claude Sonnet 4.5 is being retired from the claude.ai model selector.
17 day(s) with sentiment data
-
New LLM predicts pension enrollment in China using policy cues
Researchers have developed FlexPension-LLM, a specialized large language model designed to predict pension enrollment among flexible workers in China. This model integrates policy-grounded cues, such as marginal effects…
-
Anthropic's Claude Sonnet 4.5 debuts with 200K context and extended thinking
Anthropic has released Claude Sonnet 4.5, featuring a 200K token context window and a new "extended thinking mode." This mode allows the AI to interleave reasoning with action, pausing to reflect on intermediate results…
-
Gemini Advanced review highlights 1M context window utility
A user has found Gemini Advanced to be a powerful tool for analyzing large documents, codebases, and research papers due to its 1 million token context window. While it excels in handling extensive data and multimodal u…
-
New framework SimGuide enhances AI agent planning with multi-context user representations
Researchers have developed SimGuide, a framework designed to improve how AI agents understand and plan based on user preferences and contexts. This framework utilizes typed multi-context representations and explicit con…
-
OpenClaw 2.0 launches with simplified setup and multiplayer AI sessions · 8 sources tracked
The OpenClaw Foundation has released OpenClaw 2.0, its most significant update to date, incorporating over 16,000 pull requests. This new version simplifies setup by automatically detecting existing AI subscriptions and…
-
New middleware DRL enhances enterprise NL2SQL reliability
A new research paper introduces DRL, a Deterministic Relational Middleware Layer designed to improve the reliability of Natural Language to SQL (NL2SQL) systems in enterprise environments. DRL addresses the challenge of…
-
New AI agent ReproAgent turns research papers into executable code
Researchers have developed ReproAgent, a novel four-stage pipeline designed to automatically convert scientific research papers into executable code repositories. This system addresses the challenge of lost or implicit …
-
AI models struggle with multilingual and meme-based hate speech detection
Researchers are exploring advanced methods to improve AI's ability to detect hate speech, particularly in multilingual and multimodal contexts. One study focuses on training-time explainability to align AI reasoning wit…
-
Intent Engine translates natural language to SLOs, reducing errors
A new architecture called Intent Engine has been developed to translate natural-language intents into validated Service-Level Objectives (SLOs) for compute continuum service placement. This system aims to overcome the a…
-
Anthropic fixes 10% cost under-count for US-only Claude inference
Anthropic has updated its Claude Code CLI to version 2.1.239, addressing a bug that caused a 10% under-counting of inference costs for US-only workspaces. This fix ensures that the CLI accurately reflects the 1.1x premi…
-
Claude AI users report phantom usage drains and model discrepancies
Users of Anthropic's Claude AI are reporting unexpected usage drains on their 5-hour limits, even when not actively prompting the model. This "phantom draw" appears to occur simply by navigating the interface, reviewing…
-
Small language models challenge cloud AI dominance, Stanford paper finds
A recent Stanford research paper indicates that small language models (SLMs) are becoming competitive with large, cloud-based frontier models across various tasks. The study found that SLMs, runnable on local hardware, …
-
Local AI models now trail frontier tech by only 9 months
Local large language models are rapidly closing the gap with frontier models, now only about nine months behind. This advancement suggests that the cost of many AI-driven tasks will significantly decrease in the near fu…
-
Pairwise ranking beats RL for LLM explanation selection in recommendation systems
Researchers have developed a new method for selecting explanations from large language models (LLMs) in recommendation systems, significantly reducing serving costs and latency. By pre-generating a pool of explanations …
-
GenAI models outperform students on OOP assessments, but still struggle with advanced concepts
A new study published on arXiv evaluates the performance of five leading generative AI systems, including ChatGPT 5.2, DeepSeek-V3, Gemini 2.5 Flash, Claude Sonnet 4.5, and M365 Copilot, on introductory object-oriented …
-
Anthropic updates Claude system prompts, excluding API users
Anthropic is updating the system prompts for its Claude models, which are used in its web interface and mobile applications. These updates aim to provide more current information, such as the date, and to encourage spec…
-
Open-source LLMs show promise for emergency department decision support
A new benchmark study evaluated eight open-source small language models (SLMs) for emergency department (ED) decision support, comparing them against commercial models like Claude Haiku 4.5 and Claude Sonnet 4.5. The re…
-
New framework TTP-R1 enhances cyber threat intelligence analysis
Researchers have developed TTP-R1, a novel two-stage framework designed to improve the extraction of attack techniques from cyber threat intelligence (CTI) text. This framework combines retrieval-augmented supervised fi…
-
Prompt Caching Slashes Claude API Costs by 85%
A developer has detailed a prompt caching strategy that significantly reduced their API costs for Anthropic's Claude 3.5 Sonnet model. By implementing prompt caching, which stores and reuses common prompt prefixes, the …
-
Clinician input steers AI toward accurate and harmful medical recommendations
A new study published on arXiv investigated how clinician input influences the recommendations of large language models (LLMs) in clinical settings. Researchers found that clinician reasoning significantly increased the…