langsmith

Trace, evaluate, and deploy AI agents and LLM applications with LangSmith. Use when adding observability, running evaluations, engineering prompts, or…

npx skills add https://github.com/langchain-ai/docs --skill langsmith

LangSmith

LangSmith is a framework-agnostic platform for building, debugging, and deploying AI agents and LLM applications. Trace requests, evaluate outputs, test prompts, and manage deployments all in one place at smith.langchain.com.

When to use

Use LangSmith when you need to:

  • Trace and debug LLM calls, agent steps, retrieval, and tool use
  • Evaluate LLM outputs with automated or human-in-the-loop scoring
  • Engineer prompts with a visual playground and version control
  • Deploy agents to production with the LangGraph-based agent server
  • Monitor production systems with dashboards, alerts, and cost tracking

When NOT to use

  • To build agent logic or LLM pipelines, use LangChain, LangGraph, or Deep Agents instead
  • LangSmith is the platform layer that complements these frameworks

Quick setup

Set two environment variables to enable tracing from any supported framework:

export LANGSMITH_TRACING=true
export LANGSMITH_API_KEY="your-api-key"  # from smith.langchain.com/settings

Install the SDK

# Python
pip install langsmith

# JavaScript/TypeScript
npm install langsmith

Verify tracing

from langsmith import traceable

@traceable
def my_function(query: str) -> str:
    # Your LLM logic here—all calls inside are traced automatically
    return "result"

Core capabilities

CapabilityDescription
ObservabilityTrace every step of your LLM app with automatic or manual instrumentation
EvaluationRun evaluations with code, LLM-as-judge, or composite evaluators
Prompt engineeringCreate, version, and test prompts in a visual playground
Agent deploymentDeploy LangGraph agents with streaming, human-in-the-loop, and durable execution
MonitoringDashboards, alerts, and cost tracking for production workloads

Key documentation

API reference

For SDK class and method details, use the LangChain API Reference site:

  • Browse: https://reference.langchain.com/python/langsmith
  • MCP server: https://reference.langchain.com/mcp

Related skills

  • langchain—Build agents with prebuilt architecture and model integrations
  • langgraph—Orchestrate stateful, durable agent workflows
  • deep-agents—Batteries-included agent harness with planning and subagents

More skills from langchain-ai

langgraph-docs
langchain-ai
Access LangGraph documentation to build stateful agents and multi-agent workflows. Fetches official LangGraph Python docs covering state machines, graph-based agent design, and human-in-the-loop patterns Prioritizes relevant documentation by query type: implementation guides for how-to questions, concept pages for theory, tutorials for end-to-end examples, and API references for technical details Automatically selects 2–4 most relevant documentation URLs and retrieves their content to answer...
official
langgraph-human-in-the-loop
langchain-ai
Pause graph execution for human review, approval, or validation, then resume with their input. Requires three components: a checkpointer (InMemorySaver or PostgresSaver), a thread ID in config, and JSON-serializable interrupt payloads interrupt(value) pauses and surfaces data; Command(resume=value) resumes and returns that value to the paused node All code before interrupt() re-executes on resume, so side effects must be idempotent (use upsert, not insert) Supports approval workflows,...
official
web-research
langchain-ai
Use this skill for requests related to web research; it provides a structured approach to conducting comprehensive web research
official
langchain-oss-primer
langchain-ai
ALWAYS START HERE for any LangChain, Deep Agents, or LangGraph agent building project. Required starting point before choosing other skills or writing any…
official
skill-creator
langchain-ai
Guide for creating effective skills that extend agent capabilities with specialized knowledge, workflows, or tool integrations. Use this skill when the user…
official
social-media
langchain-ai
Drafts platform-specific social media posts with research-backed content and generated companion images. Supports LinkedIn posts (1,300 characters with professional tone) and Twitter/X threads (280 characters per tweet with 1/🧵 format) Requires delegating research to a subagent before writing, then reading findings to ensure accuracy and relevance Generates eye-catching social images automatically using generate_social_image tool with bold, high-contrast compositions optimized for small...
official
deep-agents-memory
langchain-ai
Pluggable memory and file backends for Deep Agents with ephemeral, persistent, and hybrid routing options. Four backend types: StateBackend (thread-scoped, ephemeral), StoreBackend (cross-session persistent), FilesystemBackend (real disk access for local dev), and CompositeBackend (route different paths to different backends) FilesystemMiddleware provides six file operation tools: ls , read_file , write_file , edit_file , glob , grep CompositeBackend uses longest-prefix matching to route...
official
deep-agents-orchestration
langchain-ai
Orchestrate subagents, plan multi-step tasks, and require human approval for sensitive operations. Delegate work to specialized subagents via the task tool; custom subagents support isolated tool sets and system prompts, while the default "general-purpose" subagent inherits main agent configuration Plan and track complex workflows with write_todos , organizing tasks across pending, in-progress, and completed states; requires a thread_id for persistence across invocations Implement...
official