langsmith

Theo dõi, đánh giá và triển khai các tác nhân AI và ứng dụng LLM với LangSmith. Sử dụng khi thêm khả năng quan sát, chạy đánh giá, thiết kế prompt, hoặc…

npx skills add https://github.com/langchain-ai/docs --skill langsmith

LangSmith

LangSmith is a framework-agnostic platform for building, debugging, and deploying AI agents and LLM applications. Trace requests, evaluate outputs, test prompts, and manage deployments all in one place at smith.langchain.com.

When to use

Use LangSmith when you need to:

  • Trace and debug LLM calls, agent steps, retrieval, and tool use
  • Evaluate LLM outputs with automated or human-in-the-loop scoring
  • Engineer prompts with a visual playground and version control
  • Deploy agents to production with the LangGraph-based agent server
  • Monitor production systems with dashboards, alerts, and cost tracking

When NOT to use

  • To build agent logic or LLM pipelines, use LangChain, LangGraph, or Deep Agents instead
  • LangSmith is the platform layer that complements these frameworks

Quick setup

Set two environment variables to enable tracing from any supported framework:

export LANGSMITH_TRACING=true
export LANGSMITH_API_KEY="your-api-key"  # from smith.langchain.com/settings

Install the SDK

# Python
pip install langsmith

# JavaScript/TypeScript
npm install langsmith

Verify tracing

from langsmith import traceable

@traceable
def my_function(query: str) -> str:
    # Your LLM logic here—all calls inside are traced automatically
    return "result"

Core capabilities

CapabilityDescription
ObservabilityTrace every step of your LLM app with automatic or manual instrumentation
EvaluationRun evaluations with code, LLM-as-judge, or composite evaluators
Prompt engineeringCreate, version, and test prompts in a visual playground
Agent deploymentDeploy LangGraph agents with streaming, human-in-the-loop, and durable execution
MonitoringDashboards, alerts, and cost tracking for production workloads

Key documentation

API reference

For SDK class and method details, use the LangChain API Reference site:

  • Browse: https://reference.langchain.com/python/langsmith
  • MCP server: https://reference.langchain.com/mcp

Related skills

  • langchain—Build agents with prebuilt architecture and model integrations
  • langgraph—Orchestrate stateful, durable agent workflows
  • deep-agents—Batteries-included agent harness with planning and subagents

Thêm skills từ langchain-ai

deepagents-thread-inspector
langchain-ai
Kiểm tra và giải thích các cuộc hội thoại trong kho lưu trữ phiên SQLite cục bộ của Deep Agents Code. Sử dụng như phương án dự phòng khi công cụ theo dõi LangSmith không khả dụng, cho…
deepagents-python-quickstart
langchain-ai
Tạo khung một Deep Agent cục bộ tối thiểu bằng Python bằng cách làm theo hướng dẫn khởi động nhanh chính thức, sử dụng tìm kiếm web gốc của nhà cung cấp thay vì Tavily. Sử dụng khi người dùng muốn…
deepagents-typescript-quickstart
langchain-ai
Tạo khung một Deep Agent tối thiểu cục bộ bằng TypeScript theo hướng dẫn khởi động nhanh chính thức, sử dụng tìm kiếm web gốc của nhà cung cấp thay vì Tavily. Sử dụng khi người dùng…
eval-engineering
langchain-ai
Lặp lại kiểm tra kho lưu trữ agent và các dấu vết do người dùng cung cấp tùy chọn, phỏng vấn người dùng, và tạo, chạy, cũng như kiểm toán các đánh giá Harbor từng cái một. Sử dụng cho…
LangChain RAG Pipeline
langchain-ai
GỌI KỸ NĂNG NÀY khi xây dựng BẤT KỲ hệ thống tạo sinh tăng cường truy xuất (RAG) nào. Bao gồm bộ tải tài liệu, RecursiveCharacterTextSplitter, embeddings (OpenAI),…
LangChain Structured Output & HITL
langchain-ai
langchain-structured-output-&-hitl — một kỹ năng có thể cài đặt cho các tác nhân AI, được xuất bản bởi langchain-ai/langchain-skills.
LangSmith Datasets
langchain-ai
GỌI KỸ NĂNG NÀY khi tạo bộ dữ liệu đánh giá từ trace HOẶC tải bộ dữ liệu lên LangSmith HOẶC truy vấn bộ dữ liệu. Bao gồm các loại bộ dữ liệu (final_response,…
langsmith-evaluator
langchain-ai
GỌI KỸ NĂNG NÀY khi xây dựng pipeline đánh giá cho LangSmith. Bao gồm ba thành phần cốt lõi: (1) Tạo Evaluator - LLM-as-Judge, mã tùy chỉnh; (2)…