LLM observability & evaluation · LangChain
LangSmith
Tracing, evaluation, and deployment platform for LLM applications, built by the team behind LangChain and LangGraph.
LangSmith is LangChain's platform for tracing, debugging, evaluating, and monitoring LLM applications in development and production. It captures full traces of chains, agents, and tool calls (including from LangChain/LangGraph or any instrumented app), letting a team inspect prompts, intermediate steps, and token usage for a given run. Its evaluation side supports offline test suites, dataset-based regression testing, and human or LLM-as-judge scoring, distinct from the tracing/monitoring side that watches live production traffic. Usage is billed on LangChain Compute Units (LCU) and Storage Units (LSU) consumed by tracing, evaluators, and deployment, on top of a per-seat platform fee. Cloud, hybrid, and self-hosted/on-prem deployment are available on Enterprise plans; smaller teams use the hosted cloud service.
At a glance
| Vendor | LangChain |
|---|---|
| Pricing model | Usage-based |
| Free tier | Yes |
| Deployment | Cloud, Self-hosted |
| Open source | No |
| Best for | Teams already building on LangChain/LangGraph that want tracing and evaluation from the same vendor. |
Pricing
Free Developer tier includes a monthly trace allowance per seat, then usage-based charges for compute (LCU) and storage (LSU); Plus adds paid seats and deployment features; Enterprise is custom with self-hosted options.
| Plan | Price | Notes |
|---|---|---|
| Developer | $0/seat/month | Up to 5,000 base traces/month, then usage charges; single seat; community support |
| Plus | $39/seat/month | Up to 10,000 base traces/month, then usage charges; unlimited seats; deployment and advanced features |
| Enterprise | Custom | Self-hosted/hybrid deployment, custom SSO/RBAC, support SLA |
Prices read from the vendor's own page on September 21, 2026. Vendors change prices; check the source before you budget.
Features
- Full request tracing for chains, agents, and tool calls
- Offline evaluation with datasets and regression test suites
- Human annotation queues and LLM-as-judge scoring
- Production monitoring dashboards and alerting
- Prompt playground and prompt version management
- Self-hosted and hybrid deployment on Enterprise
Integrations
Profile last reviewed September 21, 2026