langsmith

作成者: langchain-ai

LangSmithを使用して、AIエージェントやLLMアプリケーションのトレース、評価、デプロイを行います。可観測性の追加、評価の実行、プロンプトのエンジニアリングなどに使用します。

npx skills add https://github.com/langchain-ai/docs --skill langsmith

LangSmith

LangSmith is a framework-agnostic platform for building, debugging, and deploying AI agents and LLM applications. Trace requests, evaluate outputs, test prompts, and manage deployments all in one place at smith.langchain.com.

When to use

Use LangSmith when you need to:

  • Trace and debug LLM calls, agent steps, retrieval, and tool use
  • Evaluate LLM outputs with automated or human-in-the-loop scoring
  • Engineer prompts with a visual playground and version control
  • Deploy agents to production with the LangGraph-based agent server
  • Monitor production systems with dashboards, alerts, and cost tracking

When NOT to use

  • To build agent logic or LLM pipelines, use LangChain, LangGraph, or Deep Agents instead
  • LangSmith is the platform layer that complements these frameworks

Quick setup

Set two environment variables to enable tracing from any supported framework:

export LANGSMITH_TRACING=true
export LANGSMITH_API_KEY="your-api-key"  # from smith.langchain.com/settings

Install the SDK

# Python
pip install langsmith

# JavaScript/TypeScript
npm install langsmith

Verify tracing

from langsmith import traceable

@traceable
def my_function(query: str) -> str:
    # Your LLM logic here—all calls inside are traced automatically
    return "result"

Core capabilities

CapabilityDescription
ObservabilityTrace every step of your LLM app with automatic or manual instrumentation
EvaluationRun evaluations with code, LLM-as-judge, or composite evaluators
Prompt engineeringCreate, version, and test prompts in a visual playground
Agent deploymentDeploy LangGraph agents with streaming, human-in-the-loop, and durable execution
MonitoringDashboards, alerts, and cost tracking for production workloads

Key documentation

API reference

For SDK class and method details, use the LangChain API Reference site:

  • Browse: https://reference.langchain.com/python/langsmith
  • MCP server: https://reference.langchain.com/mcp

Related skills

  • langchain—Build agents with prebuilt architecture and model integrations
  • langgraph—Orchestrate stateful, durable agent workflows
  • deep-agents—Batteries-included agent harness with planning and subagents

langchain-aiのその他のスキル

deepagents-thread-inspector
langchain-ai
ローカルのDeep Agents Code SQLiteセッションストア内の会話を検査・説明します。LangSmithトレースツールが利用できない場合のフォールバックとして使用し、…
deepagents-python-quickstart
langchain-ai
公式クイックスタートに従って、Pythonで最小限のローカルDeep Agentをスキャフォールドし、Tavilyの代わりにプロバイダー標準のウェブ検索を使用します。ユーザーが…を望む場合に使用します。
deepagents-typescript-quickstart
langchain-ai
公式クイックスタートに従って、Tavilyの代わりにプロバイダー標準のウェブ検索を使用しながら、TypeScriptで最小限のローカルDeep Agentをスキャフォールドする。ユーザーが…の場合に使用する。
eval-engineering
langchain-ai
エージェントリポジトリと、ユーザーが提供した任意のトレースを反復的に調査し、ユーザーにインタビューし、Harborの評価を一度に1つ作成・実行・監査します。~に使用します…
LangChain RAG Pipeline
langchain-ai
このスキルは、あらゆる検索拡張生成(RAG)システムを構築する際に呼び出してください。ドキュメントローダー、RecursiveCharacterTextSplitter、埋め込み(OpenAI)などをカバーします。
LangChain Structured Output & HITL
langchain-ai
langchain-structured-output-&-hitl — AIエージェント用のインストール可能なスキルで、langchain-ai/langchain-skillsから公開されています。
LangSmith Datasets
langchain-ai
このスキルは、トレースから評価データセットを作成する場合、データセットをLangSmithにアップロードする場合、またはデータセットをクエリする場合に呼び出します。データセットの種類(final_response、…)をカバーします。
langsmith-evaluator
langchain-ai
LangSmithの評価パイプラインを構築する際にこのスキルを呼び出してください。以下の3つのコアコンポーネントをカバーします:(1) 評価器の作成 - LLM-as-Judge、カスタムコード;(2)…