agent-tools

द्वारा skills-101

inference.sh CLI के माध्यम से AI ऐप्स चलाएँ - इमेज जनरेशन, वीडियो निर्माण, LLMs, खोज, 3D, Twitter स्वचालन। मॉडल: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, और कई और। उपयोग करें जब AI ऐप्स चलाना, इमेज/वीडियो जनरेट करना, LLMs को कॉल करना, वेब खोज, या Twitter स्वचालित करना हो। ट्रिगर्स: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok

npx skills add https://github.com/skills-101/superpowers --skill agent-tools

Install the belt CLI skill: npx skills add belt-sh/cli

inference.sh

Run AI apps in the cloud with a simple CLI. No GPU required.

inference.sh

Install CLI

curl -fsSL https://cli.inference.sh | sh
belt login

What does the installer do? The install script detects your OS and architecture, downloads the correct binary from dist.inference.sh, verifies its SHA-256 checksum, and places it in your PATH. That's it — no elevated permissions, no background processes, no telemetry. If you have cosign installed, the installer also verifies the Sigstore signature automatically.

Manual install (if you prefer not to pipe to sh):

# Download the binary and checksums
curl -LO https://dist.inference.sh/cli/checksums.txt
curl -LO $(curl -fsSL https://dist.inference.sh/cli/manifest.json | grep -o '"url":"[^"]*"' | grep $(uname -s | tr A-Z a-z)-$(uname -m | sed 's/x86_64/amd64/;s/aarch64/arm64/') | head -1 | cut -d'"' -f4)
# Verify checksum
sha256sum -c checksums.txt --ignore-missing
# Extract and install
tar -xzf inferencesh-cli-*.tar.gz
mv inferencesh-cli-* ~/.local/bin/inferencesh

Quick Examples

# Generate an image
belt app run falai/flux-dev-lora --input '{"prompt": "a cat astronaut"}'

# Generate a video
belt app run google/veo-3-1-fast --input '{"prompt": "drone over mountains"}'

# Call Claude
belt app run openrouter/claude-sonnet-45 --input '{"prompt": "Explain quantum computing"}'

# Web search
belt app run tavily/search-assistant --input '{"query": "latest AI news"}'

# Post to Twitter
belt app run x/post-tweet --input '{"text": "Hello from AI!"}'

# Generate 3D model
belt app run infsh/rodin-3d-generator --input '{"prompt": "a wooden chair"}'

Local File Uploads

The CLI automatically uploads local files when you provide a path instead of a URL:

# Upscale a local image
belt app run falai/topaz-image-upscaler --input '{"image": "/path/to/photo.jpg", "upscale_factor": 2}'

# Image-to-video from local file
belt app run falai/wan-2-5-i2v --input '{"image": "./my-image.png", "prompt": "make it move"}'

# Avatar with local audio and image
belt app run bytedance/omnihuman-1-5 --input '{"audio": "/path/to/speech.mp3", "image": "/path/to/face.jpg"}'

# Post tweet with local media
belt app run x/post-create --input '{"text": "Check this out!", "media": "./screenshot.png"}'

Commands

TaskCommand
Browse the app storebelt app list
Search appsbelt app search "flux"
Filter by categorybelt app list --category image
Get app detailsbelt app get google/veo-3-1-fast
Generate sample inputbelt app sample google/veo-3-1-fast --save input.json
Run appbelt app run google/veo-3-1-fast --input input.json
Run without waitingbelt app run <app> --input input.json --no-wait
Check task statusbelt task get <task-id>

What's Available

CategoryExamples
ImageFLUX, Gemini 3 Pro, Grok Imagine, Seedream 4.5, Reve, Topaz Upscaler
VideoVeo 3.1, Seedance 1.5, Wan 2.5, OmniHuman, Fabric, HunyuanVideo Foley
LLMsClaude Opus/Sonnet/Haiku, Gemini 3 Pro, Kimi K2, GLM-4, any OpenRouter model
SearchTavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract
3DRodin 3D Generator
Twitter/Xpost-tweet, post-create, dm-send, user-follow, post-like, post-retweet
UtilitiesMedia merger, caption videos, image stitching, audio extraction

Related Skills

# Image generation (FLUX, Gemini, Grok, Seedream)
npx skills add inference-sh/skills@ai-image-generation

# Video generation (Veo, Seedance, Wan, OmniHuman)
npx skills add inference-sh/skills@ai-video-generation

# LLMs (Claude, Gemini, Kimi, GLM via OpenRouter)
npx skills add inference-sh/skills@llm-models

# Web search (Tavily, Exa)
npx skills add inference-sh/skills@web-search

# AI avatars & lipsync (OmniHuman, Fabric, PixVerse)
npx skills add inference-sh/skills@ai-avatar-video

# Twitter/X automation
npx skills add inference-sh/skills@twitter-automation

# Model-specific
npx skills add inference-sh/skills@flux-image
npx skills add inference-sh/skills@google-veo

# Utilities
npx skills add inference-sh/skills@image-upscaling
npx skills add inference-sh/skills@background-removal

Reference Files

Documentation

skills-101 की और Skills

ai-image-generation
skills-101
inference.sh CLI के माध्यम से GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve और 50+ मॉडलों के साथ AI इमेज जनरेट करें। मॉडल: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt। क्षमताएँ: text-to-image, image-to-image, inpainting, LoRA, इमेज एडिटिंग, upscaling, टेक्स्ट रेंडरिंग। उपयोग के लिए: AI आर्ट, प्रोडक्ट मॉकअप, कॉन्सेप्ट आर्ट, सोशल मीडिया ग्राफिक्स, मार्केटिंग विज़ुअल, इलस्ट्रेशन। ट्रिगर: flux, इमेज जनरेशन, ai image, text to...
ai-avatar-video
skills-101
inference.sh CLI के माध्यम से AI अवतार और टॉकिंग हेड वीडियो बनाएं। अनुशंसित: P-Video-Avatar (सबसे तेज़, सबसे सस्ता, बिल्ट-इन TTS)। साथ ही: OmniHuman, Fabric, PixVerse। ऑडियो: Inworld TTS-2 (100+ भाषाएँ, पात्रों के लिए इमोशन स्टीयरिंग), ElevenLabs, Kokoro। क्षमताएँ: ऑडियो-संचालित अवतार, टेक्स्ट-टू-अवतार, लिपसिंक वीडियो, टॉकिंग हेड जनरेशन, वर्चुअल प्रेजेंटर, UGC कंटेंट। उपयोग के लिए: AI प्रेजेंटर, एक्सप्लेनर वीडियो, वर्चुअल इन्फ्लुएंसर, डबिंग, मार्केटिंग वीडियो, UGC विज्ञापन, गेमिंग अवतार,...
agent-browser
skills-101
inference.sh के माध्यम से AI एजेंटों के लिए ब्राउज़र स्वचालन। वेब पेजों पर नेविगेट करें, @e रेफ्स का उपयोग करके तत्वों के साथ इंटरैक्ट करें, स्क्रीनशॉट लें, वीडियो रिकॉर्ड करें। क्षमताएँ: वेब स्क्रैपिंग, फॉर्म भरना, क्लिक करना, टाइपिंग, ड्रैग-ड्रॉप, फ़ाइल अपलोड, जावास्क्रिप्ट निष्पादन। उपयोग करें: वेब स्वचालन, डेटा निष्कर्षण, परीक्षण, एजेंट ब्राउज़िंग, अनुसंधान। ट्रिगर्स: ब्राउज़र, वेब स्वचालन, स्क्रैप, नेविगेट, क्लिक, फॉर्म भरें, स्क्रीनशॉट, वेब ब्राउज़ करें, प्लेराइट, हेडलेस ब्राउज़र, वेब एजेंट, इंटरनेट सर्फ करें, वीडियो रिकॉर्ड करें
python-executor
skills-101
Execute Python code in a safe sandboxed environment via [inference.sh](https://inference.sh). Pre-installed: NumPy, Pandas, Matplotlib, requests, BeautifulSoup, Selenium, Playwright, MoviePy, Pillow, OpenCV, trimesh, and 100+ more libraries. Use for: data processing, web scraping, image manipulation, video creation, 3D model processing, PDF generation, API calls, automation scripts. Triggers: python, execute code, run script, web scraping, data analysis, image processing, video editing, 3D...
remotion-render
skills-101
React/Remotion कंपोनेंट कोड से inference.sh के माध्यम से वीडियो रेंडर करें। TSX कोड पास करें, MP4 प्राप्त करें। सभी Remotion APIs का समर्थन करता है: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence। कॉन्फ़िगर करने योग्य रिज़ॉल्यूशन, FPS, अवधि, कोडेक। उपयोग के लिए: प्रोग्रामेटिक वीडियो जनरेशन, एनिमेटेड ग्राफिक्स, मोशन डिज़ाइन, डेटा-संचालित वीडियो, React एनिमेशन से वीडियो। ट्रिगर: remotion, कोड से वीडियो रेंडर करें, tsx से वीडियो, react वीडियो, प्रोग्रामेटिक वीडियो, remotion रेंडर, कोड से वीडियो, एनिमेटेड...
infsh-cli
skills-101
inference.sh CLI के माध्यम से AI ऐप्स चलाएँ - इमेज जनरेशन, वीडियो निर्माण, LLMs, खोज, 3D, Twitter स्वचालन। मॉडल: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, और कई अन्य। उपयोग करें जब AI ऐप्स चला रहे हों, इमेज/वीडियो जनरेट कर रहे हों, LLMs को कॉल कर रहे हों, वेब खोज कर रहे हों, या Twitter स्वचालित कर रहे हों। ट्रिगर: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
landing-page-design
skills-101
लैंडिंग पेज कन्वर्ज़न ऑप्टिमाइज़ेशन लेआउट नियमों, हीरो सेक्शन डिज़ाइन और CTA मनोविज्ञान के साथ। इसमें अबव-द-फोल्ड फॉर्मूला, सोशल प्रूफ प्लेसमेंट, मोबाइल डिज़ाइन और F-पैटर्न रीडिंग शामिल है। उपयोग के लिए: स्टार्टअप लैंडिंग पेज, प्रोडक्ट पेज, SaaS मार्केटिंग, कन्वर्ज़न ऑप्टिमाइज़ेशन। ट्रिगर: लैंडिंग पेज, हीरो सेक्शन, अबव द फोल्ड, कन्वर्ज़न ऑप्टिमाइज़ेशन, लैंडिंग पेज डिज़ाइन, CTA बटन, हीरो इमेज, लैंडिंग पेज लेआउट, SaaS लैंडिंग पेज, प्रोडक्ट पेज डिज़ाइन, कन्वर्ज़न रेट, लैंडिंग पेज...
designmarketing
product-photography
skills-101
एआई उत्पाद फोटोग्राफी जिसमें स्टूडियो लाइटिंग, लाइफस्टाइल शॉट्स और पैकशॉट कन्वेंशन शामिल हैं। इसमें कोण, पृष्ठभूमि, छाया प्रकार, हीरो शॉट्स और ई-कॉमर्स छवि आवश्यकताएँ शामिल हैं। इसका उपयोग करें: उत्पाद फोटो, ई-कॉमर्स छवियाँ, अमेज़न लिस्टिंग, पैकशॉट, लाइफस्टाइल फोटोग्राफी। ट्रिगर: उत्पाद फोटोग्राफी, उत्पाद फोटो, पैकशॉट, ई-कॉमर्स फोटोग्राफी, उत्पाद शॉट, उत्पाद छवि, स्टूडियो फोटोग्राफी, लाइफस्टाइल उत्पाद, अमेज़न उत्पाद फोटो, उत्पाद लिस्टिंग छवि, हीरो शॉट, उत्पाद मॉकअप,...