langsmith
PulseAugur coverage of langsmith — every cluster mentioning langsmith across labs, papers, and developer communities, ranked by signal.
- acquired by Promptfoo 95%
- partners with Fireworks 90%
- developed by langchain-openai 90%
- affiliated with langchain-core 90%
- competes with Langfuse 70%
- competes with Arize Phoenix 70%
- used by Promptfoo 70%
- competes with Braintrust Ai 70%
- competes with Phoenix 70%
- used by vcrpy 70%
- affiliated with langchain-openai 70%
- used by langchain-core 70%
- 2026-08-25 product_launch Fireworks AI and LangChain are co-hosting a workshop on August 25th focused on building custom evaluation models for LangSmith. source
- 2026-08-25 product_launch Fireworks and LangChain are co-hosting a workshop to demonstrate building custom evaluation models for LangSmith. source
- 2026-05-28 product_launch AWS and LangChain released a guide detailing how to use LangSmith on AWS for evaluating AI agents. source
6 day(s) with sentiment data
-
Maxim AI leads LLM observability tools for silent AI failures · 2 sources tracked
Maxim AI has emerged as the leading platform for LLM observability in 2026, according to recent analyses. This is due to its ability to integrate production monitoring with pre-deployment simulation and evaluation, brid…
-
LangSmith platform offers AI application observability and monitoring
LangSmith is a new platform designed for monitoring and observing AI applications, developed by the creators of LangChain and LangGraph. It distinguishes between observability, which tracks inputs and outputs of individ…
-
Developer shares essential AI product building toolkit
A developer shared a list of tools they use to rapidly build AI products. The toolkit includes Vercel AI SDK for streaming and tool calls, LangSmith for LLM tracing and evaluations, and Helicone for usage analytics. Oth…
-
LangChain simplifies agent creation with Deep Agents; Cognous adds guardrails
LangChain has introduced Managed Deep Agents, a framework designed to simplify the creation of AI agents by providing a "harness" for tool usage, context management, and task delegation. This system aims to save develop…
-
AI agent frameworks lack critical audit fields for evidence, study finds
A recent audit of six popular AI agent frameworks revealed significant gaps in their logging capabilities for evidentiary purposes. On September 4, 2026, researchers found that the median framework records only five out…
-
LangGraph Guide Explains Building Multi-Agent AI Systems
This repository provides a comprehensive guide to building multi-agent systems using LangGraph, starting from fundamental concepts to a functional two-agent pipeline. It details the setup process, including wiring crede…
-
LLM observability must track RAG evidence pipelines, not just model calls
Observability for retrieval-augmented generation (RAG) systems needs to go beyond standard LLM traces to include the full evidence path. Current LLM observability often focuses on model calls, masking failures in the re…
-
LLM observability tools capture data but fail to judge agent output quality
The article discusses the evolution of LLM infrastructure, moving from direct vendor SDKs to LLM gateways and dedicated observability stacks. While gateways like LiteLLM and Portkey simplify model switching, observabili…
-
Top 5 LLM Evaluation Frameworks for Release Engineering Ranked
A recent analysis highlights Promptfoo as the leading LLM evaluation framework for release engineering, particularly for its CI/CD integration that can block builds on failed tests. DeepEval is recommended for Python-ba…
-
Fireworks and LangChain Host Workshop on LangSmith Evaluation Models
Fireworks and LangChain are co-hosting a workshop on August 25th focused on building custom evaluation models for LangSmith. The event aims to help attendees scale their observability stacks to enhance performance. A ro…
-
ZizkaDB launches to debug LLM agent decision chains
ZizkaDB has launched as an open-source operational database designed to address the debugging challenges of LLM agents. Unlike traditional tracing tools that provide a span tree of events, ZizkaDB stores agent decisions…
-
LangChain updates Anthropic integration with new features and fixes
LangChain has released two new versions of its Anthropic integration: 1.6.1 and 1.6.0. Version 1.6.1 includes a fix for filtering invalid tool calls from Anthropic's v1 content. Version 1.6.0, released prior to 1.6.1, i…
-
LangSmith adds Tuned Evaluators to improve AI agent quality feedback
LangSmith has introduced Tuned Evaluators, a feature designed to enhance the quality feedback loop for AI agents by attaching specific quality metrics to production traces. Initially, this feature focuses on 'Perceived …
-
LLM observability tools diverge, focusing on distinct core problems
The LLM observability landscape is diversifying, with tools like LangSmith, Langfuse, Braintrust Ai, and Helicone each focusing on different core problems rather than competing directly. LangSmith emphasizes tracing Lan…
-
LangChain updates OpenAI integration to v1.5.2 with new features and fixes
LangChain has released version 1.5.2 of its langchain-openai package, introducing several fixes and features. Key updates include preserving reasoning item boundaries, extracting gateway metadata from response headers, …
-
LangChain updates OpenAI integration to v1.5.1, fixing reasoning bug
LangChain has released version 1.5.1 of its langchain-openai integration, following closely on the heels of version 1.5.0. The latest update, 1.5.1, focuses on fixing an issue where streamed encrypted reasoning was not …
-
Fireworks AI and LangChain Host Workshop on Custom Eval Models
Fireworks AI is hosting a workshop on August 25th in collaboration with LangChain. The event will focus on building custom evaluation models for LangSmith and enhancing observability stacks to improve performance. Atten…
-
LangChain releases updates to core components and main library
LangChain has released updates to its core components and the main library. Langchain-core version 1.5.5 includes fixes for pydantic validation, merging chunks, and handling Anthropic content blocks, alongside explicit …
-
LLM observability platforms diverge on advanced features as market booms
The LLM observability and evaluation platform market is rapidly expanding, with projections reaching $9.26 billion by 2030. Platforms are diversifying into AI-native tools, open-source evaluation libraries, AI gateways,…
-
LLM observability tools capture traces but limit assertion granularity
Observability tools for LLM agents, such as Langfuse, LangSmith, and Phoenix, offer ways to capture production traces, but their default configurations for defining inputs and assertions can be limiting. The author argu…