cosmos3-setup

作成者: nvidia

Cosmos3のインストール、環境設定、チェックポイントのダウンロード、および検証を通じてユーザーをガイドします。ユーザーが「cosmos3のインストール方法は?」や「どうやって…」と尋ねたときに使用します。

npx skills add https://github.com/nvidia/cosmos-framework --skill cosmos3-setup

Cosmos3 Setup

When to use this skill

  • Use when a user wants to install Cosmos3 or set up a development environment
  • Use when a user asks about system requirements, CUDA versions, or GPU compatibility
  • Use when a user needs to download model checkpoints or configure HuggingFace auth
  • Use when a user wants to run Cosmos3 inside a Docker container or NGC container
  • For errors during setup, hand off to the cosmos3-env-troubleshoot skill

Path convention

All paths below are relative to the cosmos3 package root (../../../ from this skill file). All uv run / python commands should also be run from there.

Where to find answers

The canonical setup reference is docs/setup.md. The README (README.md § Setup) has the shortest quickstart.

User questionGo to
What are the system requirements?docs/setup.md § System Requirements
How do I install with uv? (sync, pip venv, pip system)docs/setup.md § Virtual Environment
How do I install with Docker?docs/setup.md § Docker Container
Custom torch/CUDA versions or attention backends?docs/setup.md § Advanced
Which CUDA version? (cu130 vs cu128)docs/setup.md § CUDA Variants, docs/faq.md § Which CUDA version?
How do I download checkpoints?docs/setup.md § Downloading Base Checkpoints
NGC container issues?docs/setup.md § PyTorch Import Issue
Any installation error../cosmos3-env-troubleshoot/SKILL.md

Setup steps at a glance

  1. Clone the repository and cd into the project root (the directory containing pyproject.toml)
  2. System deps: sudo apt-get install -y --no-install-recommends curl ffmpeg git-lfs libx11-dev tree wget
  3. Install uv: curl -LsSf https://astral.sh/uv/install.sh | sh && source $HOME/.local/bin/env
  4. Install package: uv sync --all-extras --group=cu130-train && source .venv/bin/activate && export LD_LIBRARY_PATH= (use cu128-train on older drivers; the inference-only cu130 / cu128 groups omit the training extras)
  5. Checkpoints: auto-downloaded during inference; requires HuggingFace auth (see docs)
  6. Verify: uv run --all-extras --group=cu130-train python -c "import cosmos_framework; print('ok')"

Things not obvious from the docs

  • NGC container caveat: you must run export LD_LIBRARY_PATH='' before any Python imports when inside an NGC PyTorch container. Easy to miss.
  • CUDA version alignment: the major CUDA version from nvidia-smi must match torch.version.cuda. Mismatches cause cryptic shared-library errors.
  • HF_HOME: controls where checkpoints are cached (default: ~/.cache/huggingface). Set this if disk space is tight or you want a shared cache.
  • Conflicting env vars: stale HF_TOKEN or HUGGING_FACE_HUB_TOKEN env vars can silently override CLI auth. Check with printenv | grep HF_.

Related skills

SkillWhen to use
../cosmos3-inference/SKILL.mdRunning inference after setup is complete
../cosmos3-codebase-nav/SKILL.mdFinding files, parameters, and configs in code
../cosmos3-env-troubleshoot/SKILL.mdDebugging environment and runtime errors

nvidiaのその他のスキル

compileiq-debug
nvidia
何かがおかしいときに使用:Search()がハングする、すべての評価がINVALID_SCOREを返す、スコアが改善しない、すべての設定が同じ数値を返す、ptxasエラー…
create-github-pr
nvidia
gh CLIを使用してGitHubのプルリクエストを作成します。ユーザーが新しいPRを作成したい、コードをレビューに提出したい、またはプルリクエストを開きたい場合に使用します。トリガーキーワード -…
nemoclaw-maintainer-cross-issue-sweep
nvidia
他のオープンなIssueをスキャンし、特定のPRが修正する可能性があるものや、誤って壊す可能性があるものを見つけます。隣接修正の機会や矛盾リスクをfile:line…と共に出力します。
fhir-basics
nvidia
エージェントにFHIR R4 APIの動作方法、利用可能なリソース、検索パラメータを使ったクエリ方法、およびすべてのレスポンス形式を正しく解析する方法を教えます…
compileiq-validate-result
nvidia
検索が完了した後、かつスピードアップの申請やACFの発送の前に使用します。dump_results CSVを読み込み、トップK候補(単一目的)を抽出します…
changelog-audit
nvidia
リリース前にWarp CHANGELOG.mdを監査:失われたエントリを復元、ユーザー影響で並べ替え、エントリの文言を洗練、行折り返し、および(リリースブランチモードで)比較をバンプ…
maintain-dynamic-plugins
nvidia
NeMo Relayの動的プラグインローダー、マニフェスト、RustネイティブSDK、gRPCワーカープロトコル、PythonワーカーSDK、ドキュメント、テスト、およびリリースワークフローのカバレッジを維持する
dgx-diagnose
nvidia
一般的なDGX Station GB300の問題(CUDAクラッシュ、誤ったGPUターゲット、vLLM/SGLangコンテナのバグ、MIG状態の問題、NVLink/Fabric Managerエラーなど)を診断します。