langsmith

Rastreie, avalie e implante agentes de IA e aplicações LLM com LangSmith. Use ao adicionar observabilidade, executar avaliações, projetar prompts ou…

npx skills add https://github.com/langchain-ai/docs --skill langsmith

LangSmith

LangSmith is a framework-agnostic platform for building, debugging, and deploying AI agents and LLM applications. Trace requests, evaluate outputs, test prompts, and manage deployments all in one place at smith.langchain.com.

When to use

Use LangSmith when you need to:

  • Trace and debug LLM calls, agent steps, retrieval, and tool use
  • Evaluate LLM outputs with automated or human-in-the-loop scoring
  • Engineer prompts with a visual playground and version control
  • Deploy agents to production with the LangGraph-based agent server
  • Monitor production systems with dashboards, alerts, and cost tracking

When NOT to use

  • To build agent logic or LLM pipelines, use LangChain, LangGraph, or Deep Agents instead
  • LangSmith is the platform layer that complements these frameworks

Quick setup

Set two environment variables to enable tracing from any supported framework:

export LANGSMITH_TRACING=true
export LANGSMITH_API_KEY="your-api-key"  # from smith.langchain.com/settings

Install the SDK

# Python
pip install langsmith

# JavaScript/TypeScript
npm install langsmith

Verify tracing

from langsmith import traceable

@traceable
def my_function(query: str) -> str:
    # Your LLM logic here—all calls inside are traced automatically
    return "result"

Core capabilities

CapabilityDescription
ObservabilityTrace every step of your LLM app with automatic or manual instrumentation
EvaluationRun evaluations with code, LLM-as-judge, or composite evaluators
Prompt engineeringCreate, version, and test prompts in a visual playground
Agent deploymentDeploy LangGraph agents with streaming, human-in-the-loop, and durable execution
MonitoringDashboards, alerts, and cost tracking for production workloads

Key documentation

API reference

For SDK class and method details, use the LangChain API Reference site:

  • Browse: https://reference.langchain.com/python/langsmith
  • MCP server: https://reference.langchain.com/mcp

Related skills

  • langchain—Build agents with prebuilt architecture and model integrations
  • langgraph—Orchestrate stateful, durable agent workflows
  • deep-agents—Batteries-included agent harness with planning and subagents

Mais skills de langchain-ai

deepagents-thread-inspector
langchain-ai
Inspecione e explique conversas no armazenamento de sessões SQLite local do Deep Agents Code. Use como fallback quando a ferramenta de rastreamento LangSmith não estiver disponível, para…
deepagents-python-quickstart
langchain-ai
Estruture um Deep Agent local mínimo em Python seguindo o quickstart oficial, usando busca web nativa do provedor em vez de Tavily. Use quando o usuário quiser…
deepagents-typescript-quickstart
langchain-ai
Estruture um Deep Agent local mínimo em TypeScript seguindo o quickstart oficial, usando busca web nativa do provedor em vez de Tavily. Use quando o usuário…
eval-engineering
langchain-ai
Inspecione iterativamente um repositório de agente e traces opcionais fornecidos pelo usuário, entreviste o usuário e crie, execute e audite evals do Harbor um por vez. Use para…
LangChain RAG Pipeline
langchain-ai
INVOQUE ESTA HABILIDADE ao construir QUALQUER sistema de geração aumentada por recuperação (RAG). Abrange carregadores de documentos, RecursiveCharacterTextSplitter, embeddings (OpenAI),…
LangChain Structured Output & HITL
langchain-ai
langchain-structured-output-&-hitl — uma skill instalável para agentes de IA, publicada por langchain-ai/langchain-skills.
LangSmith Datasets
langchain-ai
INVOQUE ESTA HABILIDADE ao criar conjuntos de dados de avaliação a partir de rastreamento OU ao fazer upload de conjuntos de dados para o LangSmith OU ao consultar conjuntos de dados. Abrange tipos de conjuntos de dados (final_response,…
langsmith-evaluator
langchain-ai
INVOQUE ESTA HABILIDADE ao construir pipelines de avaliação para LangSmith. Abrange três componentes principais: (1) Criação de Avaliadores - LLM como Juiz, código personalizado; (2)…