LangSmith

Observability and evaluation platform for LLM agents, with tracing, monitoring, and dataset-based testing; framework-agnostic but built by the LangChain team.

Works with AI agents:llms.txt
Ask your AI about this, with this page as the source:ChatGPT ↗Claude ↗Perplexity ↗

LangSmith is LangChain's hosted platform for watching and testing LLM agents. It records each run of your app as a trace of nested steps (model calls, tools, retrieval) so you can see why an agent did what it did, then turns those runs into datasets and evaluations so you can check that a prompt or model change made things better rather than worse.

You instrument code with its SDKs for Python, TypeScript, Go or Java, typically by wrapping your LLM client or marking functions as traceable; LangChain and LangGraph apps switch it on with an environment variable. It also accepts OpenTelemetry traces on an OTLP endpoint, and the Vercel AI SDK has a dedicated integration. In JavaScript, runs are sent in the background by default so tracing doesn't add latency.

For a small team the useful parts are trace search with dashboards and alerts, online evaluators (LLM-as-a-judge) that score production traffic, offline experiments over datasets, and a prompt playground. The same account can also host agent deployments. The managed cloud runs in US, EU and APAC regions on GCP and a US region on AWS, and you pick the region at sign-up.

Traces are kept 14 days at base retention, with longer retention billed separately. Self-hosting is only available as an Enterprise add-on, and an organization can't move between regions after sign-up.

Where it fits

How LangSmith itself is built

1 tools, from its own code, website and Product Hunt page.

Who uses it

11 makers' products, each linked to the source that shows it, and 49 open-source projects that declare it in their code.

The maker says so 10Subprocessor list 1Declared in code 49How evidence is collected →
AstridThe personal stylist programmed just for you

“We absolutely could not build this product with LangSmith for traces, datasets, annotation queues, and evals. We've become big power users! I use to build AI apps before nice tracing tools like this existed and it's like night and day having a tool like this. We considered Langfuse, but I already had experience with LangSmith so for speed purposes we went with LangSmith.”

LLM Observability & EvalsMaker says so · source ↗
GitLawAgent + templates + workflow. Making legal documents free.

“Evals / tracing / monitoring is awesome and out of the box”

LLM Observability & EvalsMaker says so · source ↗
IntrycAI scores tickets to your SOPs with 90% precision

“Shoutout for all the observability and prompt management capabilities!”

LLM Observability & EvalsMaker says so · source ↗
PigAutomate Your Windows Computer with AI

“I deployed the manage Pig agent on LangGraph and it's been smooth sailing!”

LLM Observability & EvalsMaker says so · source ↗
Watchman AICapturing invisible B2B buyers with AI agents

“LangSmith’s real-time analytics and versioning keep our AI agents rock-solid -- so everything just works better.”

LLM Observability & EvalsMaker says so · source ↗
Meet-TingAI That Gives Your Schedule a Brain | Availability Agent

“We started building Ting without LangSmith and we were flying blind. Diagnosing issues downstream felt like guesswork. Since bringing LangSmith in, we’ve been able to trace problems through the entire prompt path and actually understand what’s going on. We’re still evolving how we use it, but even early on it’s helped us move faster and get closer to that “this feels good” moment.”

LLM Observability & EvalsMaker says so · source ↗
Promptius AIBuild production-grade AI agents using natural language

“Promptius <3 Langchain + Langgraph + Langsmith, this combination has been instrumental in building Promptius in the past 5 months! The abstractions provided by Langchain and Langgraph make building agents as easy as writing a prose. Langsmith provides unmatched observability and also help track our costs.”

LLM Observability & EvalsMaker says so · source ↗
Aix-DBAix-DB 基于 LangChain/LangGraph 框架,结合 MCP Skills 多智能体协作架构,实现自然语言到数据洞察的端到端转换。LLM Observability & EvalsIn its code · source ↗

Open source: a project that declares LangSmith as a dependency in its public code — verifiable, but not necessarily a live product.

What makers say

8 makers on why they use LangSmith, in their own words on Product Hunt.

We absolutely could not build this product with LangSmith for traces, datasets, annotation queues, and evals. We've become big power users! I use to build AI apps before nice tracing tools like this existed and it's like night and day having a tool like this. We considered Langfuse, but I already had experience with LangSmith so for speed purposes we went with LangSmith.
AstridSep 2026 ↗
LangSmith’s real-time analytics and versioning keep our AI agents rock-solid -- so everything just works better.
Watchman AISep 2026 ↗
You can't build AI agents without monitoring. Metadata filtering is strong.
DryMergeSep 2026 ↗
I deployed the manage Pig agent on LangGraph and it's been smooth sailing!
PigSep 2026 ↗
Our preferred way to track and analyze tool usage. Really easy to handle.
lmChatGPTtfySep 2026 ↗

Loved and watch-outs

Themes that recur in makers' words and Hacker News comments, each linked to what it summarises, with how Product Hunt tags its reviews.

Most loved
  • Tracing through the whole prompt path replaces guesswork when diagnosing agent issues. PH
  • Traces, datasets, annotation queues and evals live in one platform widely used by enterprises. PHHN
  • Pairs tightly with LangChain and LangGraph for debugging stateful agent workflows. PHHNHN 2
Watch-outs
  • It keeps pulling users toward the LangChain platform, and LangChain docs push LangSmith, which feels like lock-in. HNHN 2HN 3HN 4
  • Traces show which agent failed but not why, so root-cause analysis stays manual. HNHN 2HN 3HN 4
  • Viewing your own traces requires a cloud account, with no local-first option. HNHN 2
  • The prompt engineering workflow is laborious, gets expensive fast and handles multi-turn prompts poorly. HNHN 2
On Product Hunt 4.8★ · 19 reviews
monitoring AI model performance 3chain sequence debugging 2evals 2
Read the reviews on Product Hunt ↗

Reliability and open issues

Incidents on its status page0 in the last 90 days · 0 in the last yearFrom its own status page, 2026-10-02. Vendors decide what they post; one page can cover several products.

Who switches

Public pull requests on GitHub since Oct 2024 whose title says "X to Y" — real code changes moving a project from one tool to another, by developers in general. Open a row to see the pull requests.

LangSmith → Langfuse9 PRs

Alternatives to LangSmith

All alternatives by situation →

Questions makers ask about LangSmith

Do I have to use LangChain to use it?

No. The SDKs trace any code and are documented as usable without LangChain; LangChain and LangGraph apps only get a shortcut, enabling tracing with a single environment variable. source ↗

Can I keep my data in the EU?

Yes. LangSmith has an EU instance in the Netherlands (GCP europe-west4), plus US and APAC instances, available on every plan including the free one. Billing data stays in the US, and you can't switch an organization between regions later. source ↗

Is it SOC 2 and HIPAA compliant?

LangChain states LangSmith is SOC 2 Type 2 certified, HIPAA compliant and GDPR compliant, and offers a DPA through support. It has no EU legal entity for contracting. source ↗

Can I send OpenTelemetry traces instead of using its SDK?

Yes. LangSmith accepts OTLP traces at its /otel endpoint, and you can use an OpenTelemetry Collector to send the same spans to LangSmith and another backend such as Datadog or Grafana. source ↗

Will tracing slow down my serverless functions?

The JS SDK sends traces in the background by default. In serverless runtimes you either turn background tracing off or await pending trace batches before the function returns, so data isn't lost when the runtime freezes. source ↗

How long are traces kept?

Base traces are kept 14 days and extended traces 180 days. Some actions, like online evaluators and automation rules, can upgrade a trace to extended retention; after retention ends, trace inputs and outputs are deleted. source ↗

Can I self-host it?

Only as an add-on to the Enterprise plan, installed on Kubernetes with a license key from LangChain's sales team. source ↗

Is LangSmith free?

Yes — there is a free tier a small product can run on; paid use starts at $39/seat/mo. source ↗

Is LangSmith open source or self-hostable?

Not open source, and hosted only.

Can AI coding agents work with LangSmith?

It serves an llms.txt docs index.

Who uses LangSmith?

11 makers' products we track, each with a source, and 49 open-source projects declare it in their code. source ↗