nemo-relay-get-started

par nvidia

Utilisez cette compétence lorsque les utilisateurs de NeMo Relay pour la première fois souhaitent essayer Relay, choisir le démarrage rapide pris en charge le moins complexe, ou vérifier la valeur initiale via la CLI, a…

npx skills add https://github.com/nvidia/nemo-relay --skill nemo-relay-get-started

Get Started With NeMo Relay

Guide a new user to visible Relay value with the least complicated applicable trial. Do not begin with production deployment or Relay's full architecture. Keep the first run focused on one observable success path.

Choose A Try-Now Path

Evaluate these paths in order. Use the first one that fits the user's stated goal and existing environment.

  1. CLI try-now (default): choose this for a generic "try Relay" request or when the user wants value without modifying application code. Run Codex or Claude Code through the local CLI wrapper. Read CLI Try-Now.
  2. Hermes Agent native path: choose this when Hermes Agent owns the execution boundary. Explain that NeMo Relay is built in and requires no separate Relay installation, observability plugin, or Relay CLI setup. State explicitly: "Hermes Agent understands NeMo Relay plugin configurations." Stop after confirming the native integration. Do not provide installation or observability-configuration instructions, and do not continue into the generic plugin progression.
  3. Built-in integrations try-now: choose this when an existing LangChain, LangGraph, Deep Agents, or OpenClaw application owns the execution boundary. Prefer the maintained supported integration over manual wrapping. Read Built-In Integrations Try-Now.
  4. Language-specific manual try-now: choose this when the user's Python, Node.js, or Rust application directly owns its tool or LLM call sites and no maintained integration is the better boundary. Read Manual Language Try-Now.

Do not ask the user to choose among all four when their request, manifest, or framework already identifies the boundary. For an unspecified request, use the CLI path. When more than one CLI agent is available, ask one concise question to select the agent.

Resolve Installation Without Looping

Select the try-now path before choosing an install package.

  • CLI path -> verify nemo-relay --version; if missing, use nemo-relay-install for the CLI outcome.
  • Hermes Agent native path -> do not install or configure Relay separately.
  • Built-in integration path -> use nemo-relay-install for the named framework or harness package.
  • Manual language path -> use nemo-relay-install for the detected language package.

If installation already succeeded, preserve the chosen path and continue from its next step. Do not ask the install-path question again or bounce between the install and get-started skills.

Apply The Common First-Value Contract

Do not apply this section to the Hermes Agent native path.

Follow the selected reference, then:

  1. Inspect the target environment and existing Relay configuration before proposing changes. Treat repository-local .nemo-relay/config.toml and .nemo-relay/plugins.toml as unsupported project configuration. Do not create, edit, merge, or trust those files as active Relay configuration.
  2. Explain the attachment boundary and show the exact minimal change.
  3. Obtain confirmation before writing configuration, modifying application code, or launching a model-consuming run.
  4. Use a non-sensitive, read-only trial and exercise one representative tool and LLM path when the selected surface exposes both.
  5. Verify observable evidence from the selected path rather than treating a successful application result as proof that Relay is active.
  6. Summarize the captured root, tool, and model relationships without dumping prompts, credentials, or complete event payloads.

Explain only the concepts visible in the result: the chosen attachment boundary, scopes and parentage, captured lifecycle events, and how subscribers or the Observability plugin make those events inspectable. Keep instrumentation and export distinct.

Continue With One Plugin

Do not apply this section to the Hermes Agent native path.

Stop the initial try-now workflow when the selected path's success checks pass. Then make one additional built-in plugin the primary suggested next step.

Explain Relay's core progression: instrument an execution boundary once, then change or extend behavior through plugin configuration without repeatedly rewriting those call sites. Easy plugin configuration and reconfiguration is the main value to demonstrate after the first observable proof.

If the selected path did not use plugin-managed Observability, add it first to establish the reusable plugin path. If Observability already produced the proof, ask what outcome matters next and recommend exactly one plugin:

  • Adaptive -> adaptive runtime behavior and optimization
  • NeMo Guardrails -> policy checks around managed execution
  • PII Redaction -> sanitization of sensitive observability payloads
  • Model Pricing -> cost estimates for managed LLM responses

Use the plugin overview to select the next component. Preview its smallest configuration, obtain confirmation, and verify its behavior before layering in another plugin.

Use another handoff only after the user accepts or declines plugin progression, or when the demonstrated boundary does not yet cover the intended workflow:

  • Direct application expansion -> nemo-relay-instrument-calls
  • Unsupported framework or harness integration -> nemo-relay-integrate-upstream
  • Additional exporters or durable observability configuration -> nemo-relay-plugin-observability
  • A different package or supported integration -> nemo-relay-install
  • Persistent Claude Code or Codex loading -> nemo-relay-install; for Codex Desktop, complete its recovery-note safety gate before changing global config
  • Missing hooks, gateway traffic, or events -> nemo-relay doctor, nemo-relay doctor --json, or nemo-relay-debug-runtime-integration

Do not configure production OTLP backends, model pricing, guardrails, adaptive tuning, custom plugins, Go or FFI examples during the quick start. Mention optional plugins only after the initial proof and add only the one the user selects.

Public Entry Points

Use these public entry points for current product documentation:

Plus de skills de nvidia

compileiq-debug
nvidia
Utilisez quand quelque chose ne va pas : Search() bloque, toutes les évaluations retournent INVALID_SCORE, les scores ne s'améliorent pas, chaque configuration retourne le même nombre, erreurs ptxas…
create-github-pr
nvidia
Créer des pull requests GitHub en utilisant l'interface en ligne de commande gh. Utiliser lorsque l'utilisateur souhaite créer une nouvelle PR, soumettre du code pour révision, ou ouvrir une pull request. Mots-clés de déclenchement -…
nemoclaw-maintainer-cross-issue-sweep
nvidia
Analyse les autres problèmes ouverts pour trouver ceux qu’une PR donnée pourrait également corriger ou casser accidentellement. Génère des opportunités de correctifs adjacents et des risques de contradiction avec fichier:ligne…
fhir-basics
nvidia
Apprend aux agents comment fonctionnent les API FHIR R4, quelles ressources sont disponibles, comment les interroger avec des paramètres de recherche, et comment analyser correctement tous les formats de réponse…
compileiq-validate-result
nvidia
Utiliser APRÈS qu'une recherche soit terminée et AVANT de réclamer un accélérateur ou d'expédier un ACF. Charge le CSV dump_results, extrait les K meilleurs candidats (mono-objectif)…
changelog-audit
nvidia
Auditer le CHANGELOG.md de Warp avant une publication : récupérer les entrées perdues, trier par impact utilisateur, affiner le langage des entrées, ajuster les retours à la ligne et (en mode branche de publication) mettre à jour la comparaison…
maintain-dynamic-plugins
nvidia
Maintenir les chargeurs de plugins dynamiques NeMo Relay, les manifestes, les SDK natifs Rust, le protocole worker gRPC, le SDK worker Python, la documentation, les tests et la couverture du workflow de publication
dgx-diagnose
nvidia
Diagnostiquer les problèmes courants du DGX Station GB300 — plantages CUDA, ciblage incorrect du GPU, bugs de conteneur vLLM/SGLang, problèmes d'état MIG, erreurs NVLink/Fabric Manager,…