HumanLayer
PulseAugur coverage of HumanLayer — every cluster mentioning HumanLayer across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
Opus 5 benchmarked on SlopCodeBench for coding agent context engineering
A benchmark test was conducted on Opus 5, evaluating its performance on the SlopCodeBench dataset. The results of this evaluation, which focused on advanced context engineering for coding agents, were shared via a GitHu…
-
Context engineering: AI code generation needs human oversight
Dex Horthy, CEO of HumanLayer, has coined the term "context engineering" to describe the practice of understanding and working around the limitations of Large Language Model (LLM) contexts. Horthy's research, based on c…
-
AI agent performance hinges on harness, not just model
AI agents often underperform not due to the underlying model, but because of the 'harness' that surrounds it. This harness includes system prompts, tool descriptions, execution environments, and orchestration logic, ess…
-
AI infrastructure startups launch tools for agents, DevOps, security, and healthcare
Several startups are launching AI-powered tools aimed at improving infrastructure and developer productivity. Trigger.dev offers an open-source platform for building reliable AI agents and workflows, utilizing snapshott…