langsmith

Rastrea, evalúa y despliega agentes de IA y aplicaciones LLM con LangSmith. Úsalo al agregar observabilidad, ejecutar evaluaciones, diseñar prompts o…

npx skills add https://github.com/langchain-ai/docs --skill langsmith

LangSmith

LangSmith is a framework-agnostic platform for building, debugging, and deploying AI agents and LLM applications. Trace requests, evaluate outputs, test prompts, and manage deployments all in one place at smith.langchain.com.

When to use

Use LangSmith when you need to:

  • Trace and debug LLM calls, agent steps, retrieval, and tool use
  • Evaluate LLM outputs with automated or human-in-the-loop scoring
  • Engineer prompts with a visual playground and version control
  • Deploy agents to production with the LangGraph-based agent server
  • Monitor production systems with dashboards, alerts, and cost tracking

When NOT to use

  • To build agent logic or LLM pipelines, use LangChain, LangGraph, or Deep Agents instead
  • LangSmith is the platform layer that complements these frameworks

Quick setup

Set two environment variables to enable tracing from any supported framework:

export LANGSMITH_TRACING=true
export LANGSMITH_API_KEY="your-api-key"  # from smith.langchain.com/settings

Install the SDK

# Python
pip install langsmith

# JavaScript/TypeScript
npm install langsmith

Verify tracing

from langsmith import traceable

@traceable
def my_function(query: str) -> str:
    # Your LLM logic here—all calls inside are traced automatically
    return "result"

Core capabilities

CapabilityDescription
ObservabilityTrace every step of your LLM app with automatic or manual instrumentation
EvaluationRun evaluations with code, LLM-as-judge, or composite evaluators
Prompt engineeringCreate, version, and test prompts in a visual playground
Agent deploymentDeploy LangGraph agents with streaming, human-in-the-loop, and durable execution
MonitoringDashboards, alerts, and cost tracking for production workloads

Key documentation

API reference

For SDK class and method details, use the LangChain API Reference site:

  • Browse: https://reference.langchain.com/python/langsmith
  • MCP server: https://reference.langchain.com/mcp

Related skills

  • langchain—Build agents with prebuilt architecture and model integrations
  • langgraph—Orchestrate stateful, durable agent workflows
  • deep-agents—Batteries-included agent harness with planning and subagents

Más skills de langchain-ai

deepagents-thread-inspector
langchain-ai
Inspecciona y explica conversaciones en el almacén de sesiones SQLite local de Deep Agents Code. Úsalo como respaldo cuando la herramienta de trazado de LangSmith no esté disponible, para…
deepagents-python-quickstart
langchain-ai
Crear un agente local mínimo de Deep Agent en Python siguiendo la guía de inicio rápido oficial, utilizando la búsqueda web nativa del proveedor en lugar de Tavily. Úsalo cuando el usuario quiera…
deepagents-typescript-quickstart
langchain-ai
Crear un agente Deep local mínimo en TypeScript siguiendo la guía de inicio rápido oficial, usando la búsqueda web nativa del proveedor en lugar de Tavily. Usar cuando el usuario…
eval-engineering
langchain-ai
Inspecciona de forma iterativa un repositorio de agentes y los traces opcionales proporcionados por el usuario, entrevista al usuario, y crea, ejecuta y audita los evals de Harbor uno a la vez. Úsalo para…
LangChain RAG Pipeline
langchain-ai
INVOCA ESTA HABILIDAD al construir CUALQUIER sistema de generación aumentada por recuperación (RAG). Cubre cargadores de documentos, RecursiveCharacterTextSplitter, embeddings (OpenAI),…
LangChain Structured Output & HITL
langchain-ai
langchain-structured-output-&-hitl — una habilidad instalable para agentes de IA, publicada por langchain-ai/langchain-skills.
LangSmith Datasets
langchain-ai
INVOCA ESTA HABILIDAD al crear conjuntos de datos de evaluación a partir de trazas O al subir conjuntos de datos a LangSmith O al consultar conjuntos de datos. Cubre tipos de conjuntos de datos (final_response,…
langsmith-evaluator
langchain-ai
INVOCA ESTA HABILIDAD al construir pipelines de evaluación para LangSmith. Cubre tres componentes principales: (1) Creación de Evaluadores - LLM como juez, código personalizado; (2)…