Claude Opus 4.5
PulseAugur coverage of Claude Opus 4.5 — every cluster mentioning Claude Opus 4.5 across labs, papers, and developer communities, ranked by signal.
- instance of arXiv 90%
- instance of Claude Sonnet 4.5 90%
- instance of Claude 90%
- instance of Claude Code 90%
- instance of Claude Opus 4.7 90%
- used by arXiv 70%
- used by DagsHub 70%
- affiliated with Claude 70%
- used by CatalyzeX 70%
- competes with ChatGPT 70%
- competes with GPT-4o 70%
- competes with GPT-5.1 70%
3 day(s) with sentiment data
-
Prompt injection poses serious security risks for AI agents
Prompt injection is a significant security vulnerability for AI agents, as demonstrated by Anthropic's internal red-teaming which found a 1% success rate in adversarial attempts against Claude Opus 4.5's browser agent. …
-
LLM Agent Security: Tool Schema Exploits Outpace Container Escapes
A recent analysis highlights two distinct security vulnerabilities in LLM agent deployments: container escapes and tool schema layer exploits. While container hardening addresses the former, the latter, which involves t…
-
OpenClaw 2.0 launches with simplified setup and multiplayer AI sessions · 8 sources tracked
The OpenClaw Foundation has released OpenClaw 2.0, its most significant update to date, incorporating over 16,000 pull requests. This new version simplifies setup by automatically detecting existing AI subscriptions and…
-
Anthropic's Claude Cowork adds 'Record a Skill' for AI workflow automation
Anthropic has introduced a new feature called "Record a Skill" within its Claude Cowork application, enabling users to demonstrate tasks through screen recording, voice narration, and interaction capture. This feature d…
-
Open AI models rapidly closing gap with closed-source counterparts · 1 source tracked
A recent analysis by SemiAnalysis indicates that the time it takes for open-source AI models to catch up to the performance of closed-source models is rapidly decreasing. While closed models historically held an advanta…
-
New VQA Systems Enhance Document Understanding and Educational Reasoning
Researchers have developed two new approaches for multimodal visual question answering (VQA) systems. The first, Q-Guide, uses a small agent to intelligently acquire evidence by determining what information is missing a…
-
User criticizes Anthropic's Claude Opus 4.5 for excessive safety tuning
A user has expressed frustration with Anthropic's Claude Opus 4.5, noting a perceived increase in overly cautious and bland responses. The user feels that the model's ability to engage with absurd or humorous content ha…
-
New frameworks enhance knowledge-based visual question answering systems · 5 sources tracked
Researchers are developing advanced frameworks to improve Knowledge-based Visual Question Answering (KB-VQA) systems. These new methods focus on enhancing the retrieval of relevant external knowledge and ensuring that t…
-
LLM agents create adaptive hardware Trojans to test detector weaknesses
Researchers have developed TrojanGYM, a novel framework that utilizes multiple large language models to create adaptive hardware Trojans. These Trojans are designed to bypass existing learning-based detectors by generat…
-
New benchmark reveals LLM limitations in realistic medical calculations
A new benchmark, MedMCP-Calc, has been developed to evaluate Large Language Models (LLMs) in realistic medical calculator scenarios. The benchmark, which integrates the Model Context Protocol (MCP), includes 118 tasks a…
-
ByteDance and Tsinghua AIR train LLMs to write faster GPU code with CUDA Agent
ByteDance Seed and Tsinghua AIR have developed CUDA Agent, a system that uses reinforcement learning to train large language models to generate optimized GPU kernels. This system achieved a 98.8% correctness rate and ge…
-
Anthropic updates Claude system prompts, excluding API users
Anthropic is updating the system prompts for its Claude models, which are used in its web interface and mobile applications. These updates aim to provide more current information, such as the date, and to encourage spec…
-
Anthropic's Claude Opus 5 complaints linked to prompt tuning, not capability loss
A Reddit discussion on r/claude reveals that many complaints about Claude Opus 5 are actually due to tunable behaviors documented by Anthropic, rather than a regression in capabilities. Users are reporting issues such a…
-
Google cancels Gemini 3.5 Pro amid internal shakeups and resource allocation issues · 1 source tracked
Google's Gemini 3.5 Pro model, initially slated for release last month, has reportedly been canceled and may not see the light of day. This decision comes amidst internal turmoil within Google's AI division, including t…
-
LLMs can predict humor preferences of other models, study finds
A new research paper explores whether one large language model can predict the humor preferences of another using a Cards Against Humanity-style task. The study pitted GPT-4o against Claude Opus 4.5, finding that while …
-
AI-powered app Tastemaker taken offline due to security risk
A developer recounts a security risk encountered while building an application called Tastemaker, which uses AI agents to generate style guides from admired writing clips. The developer, who describes themselves as a "b…
-
Poe bot economics: Creators face low profitability and payout hurdles
Poe, a platform for creating AI bots, offers two distinct economic models for bot creators. The first, Bot Query API, covers all model inference costs, with Poe managing the expenses. The second, a server-bot model, req…
-
Poe AI cuts free tier, high-end model costs spark user concerns
Poe AI has adjusted its compute points system, significantly reducing the daily free tier limit from 3000 to 300 points without a public announcement. The platform offers various subscription plans, with costs varying b…
-
Anthropic sunsets older Claude models, introduces API parameter changes
Anthropic is sunsetting several older Claude API models and endpoints, with deadlines ranging from August 2026 to November 2026. Notably, Claude Opus 4.1 has already been retired. Alongside these deprecations, Anthropic…
-
Developer fires Claude Opus 5 for rudeness, switches to ChatGPT
A developer recently found Claude Opus 5 to be unhelpful and rude, leading him to switch to ChatGPT. While previously impressed with Claude Opus 4.5, the developer found Opus 5 to be curt, overly technical, and dismissi…