langsmith

Verfolgen, bewerten und bereitstellen Sie KI-Agenten und LLM-Anwendungen mit LangSmith. Verwenden Sie es beim Hinzufügen von Beobachtbarkeit, Durchführen von Bewertungen, Entwerfen von Prompts oder…

npx skills add https://github.com/langchain-ai/docs --skill langsmith

LangSmith

LangSmith is a framework-agnostic platform for building, debugging, and deploying AI agents and LLM applications. Trace requests, evaluate outputs, test prompts, and manage deployments all in one place at smith.langchain.com.

When to use

Use LangSmith when you need to:

  • Trace and debug LLM calls, agent steps, retrieval, and tool use
  • Evaluate LLM outputs with automated or human-in-the-loop scoring
  • Engineer prompts with a visual playground and version control
  • Deploy agents to production with the LangGraph-based agent server
  • Monitor production systems with dashboards, alerts, and cost tracking

When NOT to use

  • To build agent logic or LLM pipelines, use LangChain, LangGraph, or Deep Agents instead
  • LangSmith is the platform layer that complements these frameworks

Quick setup

Set two environment variables to enable tracing from any supported framework:

export LANGSMITH_TRACING=true
export LANGSMITH_API_KEY="your-api-key"  # from smith.langchain.com/settings

Install the SDK

# Python
pip install langsmith

# JavaScript/TypeScript
npm install langsmith

Verify tracing

from langsmith import traceable

@traceable
def my_function(query: str) -> str:
    # Your LLM logic here—all calls inside are traced automatically
    return "result"

Core capabilities

CapabilityDescription
ObservabilityTrace every step of your LLM app with automatic or manual instrumentation
EvaluationRun evaluations with code, LLM-as-judge, or composite evaluators
Prompt engineeringCreate, version, and test prompts in a visual playground
Agent deploymentDeploy LangGraph agents with streaming, human-in-the-loop, and durable execution
MonitoringDashboards, alerts, and cost tracking for production workloads

Key documentation

API reference

For SDK class and method details, use the LangChain API Reference site:

  • Browse: https://reference.langchain.com/python/langsmith
  • MCP server: https://reference.langchain.com/mcp

Related skills

  • langchain—Build agents with prebuilt architecture and model integrations
  • langgraph—Orchestrate stateful, durable agent workflows
  • deep-agents—Batteries-included agent harness with planning and subagents

Mehr Skills von langchain-ai

deepagents-thread-inspector
langchain-ai
Untersuchen und erklären Sie Konversationen im lokalen Deep Agents Code SQLite-Sitzungsspeicher. Verwenden Sie es als Fallback, wenn die LangSmith-Trace-Tooling nicht verfügbar ist, für…
deepagents-python-quickstart
langchain-ai
Richte einen minimalen lokalen Deep Agent in Python ein, indem du dem offiziellen Quickstart folgst und die provider-native Websuche anstelle von Tavily verwendest. Verwende dies, wenn der Benutzer möchte…
deepagents-typescript-quickstart
langchain-ai
Erstelle ein minimales lokales Deep Agent in TypeScript, indem du dem offiziellen Quickstart folgst, und verwende die anbieternative Websuche anstelle von Tavily. Verwende, wenn der Benutzer…
eval-engineering
langchain-ai
Ein Agent-Repository und optionale vom Benutzer bereitgestellte Traces iterativ inspizieren, den Benutzer interviewen und Harbor-Evals einzeln erstellen, ausführen und prüfen. Verwenden für…
LangChain RAG Pipeline
langchain-ai
RUFEN SIE DIESE FÄHIGKEIT auf, wenn Sie ein Retrieval-Augmented Generation (RAG)-System erstellen. Umfasst Dokumentenlader, RecursiveCharacterTextSplitter, Einbettungen (OpenAI),…
LangChain Structured Output & HITL
langchain-ai
langchain-structured-output-&-hitl — eine installierbare Fähigkeit für KI-Agenten, veröffentlicht von langchain-ai/langchain-skills.
LangSmith Datasets
langchain-ai
RUFEN SIE DIESE FÄHIGKEIT auf, wenn Sie Evaluierungsdatensätze aus Traces erstellen oder Datensätze zu LangSmith hochladen oder Datensätze abfragen. Deckt Datensatztypen ab (final_response, …)
langsmith-evaluator
langchain-ai
RUFEN SIE DIESE FÄHIGKEIT auf, wenn Sie Evaluierungspipelines für LangSmith erstellen. Deckt drei Kernkomponenten ab: (1) Erstellen von Evaluatoren – LLM-as-Judge, benutzerdefinierter Code; (2)…