infsh-cli

Chạy ứng dụng AI qua CLI inference.sh - tạo hình ảnh, tạo video, LLM, tìm kiếm, 3D, tự động hóa Twitter. Mô hình: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter và nhiều mô hình khác. Sử dụng khi chạy ứng dụng AI, tạo hình ảnh/video, gọi LLM, tìm kiếm web hoặc tự động hóa Twitter. Từ khóa kích hoạt: inference.sh, infsh, mô hình ai, chạy ai, serverless ai, ai api, flux, veo, claude api, tạo hình ảnh, tạo video, openrouter, tavily, exa search, twitter api, grok

npx skills add https://github.com/skills-101/superpowers --skill infsh-cli

Install the belt CLI skill: npx skills add belt-sh/cli

inference.sh

Run AI apps in the cloud with a simple CLI. No GPU required.

inference.sh

Install CLI

curl -fsSL https://cli.inference.sh | sh
belt login

What does the installer do? The install script detects your OS and architecture, downloads the correct binary from dist.inference.sh, verifies its SHA-256 checksum, and places it in your PATH. That's it — no elevated permissions, no background processes, no telemetry. If you have cosign installed, the installer also verifies the Sigstore signature automatically.

Manual install (if you prefer not to pipe to sh):

# Download the binary and checksums
curl -LO https://dist.inference.sh/cli/checksums.txt
curl -LO $(curl -fsSL https://dist.inference.sh/cli/manifest.json | grep -o '"url":"[^"]*"' | grep $(uname -s | tr A-Z a-z)-$(uname -m | sed 's/x86_64/amd64/;s/aarch64/arm64/') | head -1 | cut -d'"' -f4)
# Verify checksum
sha256sum -c checksums.txt --ignore-missing
# Extract and install
tar -xzf inferencesh-cli-*.tar.gz
mv inferencesh-cli-* ~/.local/bin/inferencesh

Quick Examples

# Generate an image
belt app run falai/flux-dev-lora --input '{"prompt": "a cat astronaut"}'

# Generate a video
belt app run google/veo-3-1-fast --input '{"prompt": "drone over mountains"}'

# Call Claude
belt app run openrouter/claude-sonnet-45 --input '{"prompt": "Explain quantum computing"}'

# Web search
belt app run tavily/search-assistant --input '{"query": "latest AI news"}'

# Post to Twitter
belt app run x/post-tweet --input '{"text": "Hello from AI!"}'

# Generate 3D model
belt app run infsh/rodin-3d-generator --input '{"prompt": "a wooden chair"}'

Local File Uploads

The CLI automatically uploads local files when you provide a path instead of a URL:

# Upscale a local image
belt app run falai/topaz-image-upscaler --input '{"image": "/path/to/photo.jpg", "upscale_factor": 2}'

# Image-to-video from local file
belt app run falai/wan-2-5-i2v --input '{"image": "./my-image.png", "prompt": "make it move"}'

# Avatar with local audio and image
belt app run bytedance/omnihuman-1-5 --input '{"audio": "/path/to/speech.mp3", "image": "/path/to/face.jpg"}'

# Post tweet with local media
belt app run x/post-create --input '{"text": "Check this out!", "media": "./screenshot.png"}'

Commands

TaskCommand
Browse the app storebelt app list
Search the storebelt app search "flux"
Filter by categorybelt app list --category image
List your appsbelt app list
Get app detailsbelt app get google/veo-3-1-fast
Generate sample inputbelt app sample google/veo-3-1-fast --save input.json
Run appbelt app run google/veo-3-1-fast --input input.json
Run without waitingbelt app run <app> --input input.json --no-wait
Check task statusbelt task get <task-id>

What's Available

CategoryExamples
ImageFLUX, Gemini 3 Pro, Grok Imagine, Seedream 4.5, Reve, Topaz Upscaler
VideoVeo 3.1, Seedance 2.0, Wan 2.5, OmniHuman, Fabric, HunyuanVideo Foley
LLMsClaude Opus/Sonnet/Haiku, Gemini 3 Pro, Kimi K2, GLM-4, any OpenRouter model
SearchTavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract
3DRodin 3D Generator
Twitter/Xpost-tweet, post-create, dm-send, user-follow, post-like, post-retweet
UtilitiesMedia merger, caption videos, image stitching, audio extraction

Related Skills

# Image generation (FLUX, Gemini, Grok, Seedream)
npx skills add inference-sh/skills@ai-image-generation

# Video generation (Veo, Seedance, Wan, OmniHuman)
npx skills add inference-sh/skills@ai-video-generation

# LLMs (Claude, Gemini, Kimi, GLM via OpenRouter)
npx skills add inference-sh/skills@llm-models

# Web search (Tavily, Exa)
npx skills add inference-sh/skills@web-search

# AI avatars & lipsync (OmniHuman, Fabric, PixVerse)
npx skills add inference-sh/skills@ai-avatar-video

# Twitter/X automation
npx skills add inference-sh/skills@twitter-automation

# Model-specific
npx skills add inference-sh/skills@flux-image
npx skills add inference-sh/skills@google-veo

# Utilities
npx skills add inference-sh/skills@image-upscaling
npx skills add inference-sh/skills@background-removal

Reference Files

Documentation

Thêm skills từ skills-101

ai-image-generation
skills-101
Tạo hình ảnh AI với GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve và 50+ mô hình qua CLI inference.sh. Mô hình: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Khả năng: text-to-image, image-to-image, inpainting, LoRA, chỉnh sửa hình ảnh, upscaling, hiển thị văn bản. Sử dụng cho: nghệ thuật AI, mô phỏng sản phẩm, concept art, đồ họa mạng xã hội, hình ảnh tiếp thị, minh họa. Kích hoạt: flux, tạo hình ảnh, hình ảnh ai, văn bản đến...
ai-avatar-video
skills-101
Tạo video AI avatar và video người nói chuyện qua CLI inference.sh. Đề xuất: P-Video-Avatar (nhanh nhất, rẻ nhất, có TTS tích hợp). Ngoài ra: OmniHuman, Fabric, PixVerse. Âm thanh: Inworld TTS-2 (hơn 100 ngôn ngữ, điều chỉnh cảm xúc cho nhân vật), ElevenLabs, Kokoro. Tính năng: avatar điều khiển bằng âm thanh, văn bản thành avatar, video khớp môi, tạo người nói chuyện, người dẫn ảo, nội dung UGC. Dùng cho: người dẫn AI, video giải thích, người ảnh hưởng ảo, lồng tiếng, video tiếp thị, quảng cáo UGC, avatar game,...
agent-browser
skills-101
Tự động hóa trình duyệt cho các tác nhân AI qua inference.sh. Điều hướng trang web, tương tác với các phần tử bằng tham chiếu @e, chụp ảnh màn hình, ghi video. Khả năng: thu thập dữ liệu web, điền biểu mẫu, nhấp chuột, gõ chữ, kéo-thả, tải tệp lên, thực thi JavaScript. Sử dụng cho: tự động hóa web, trích xuất dữ liệu, kiểm thử, duyệt web bằng tác nhân, nghiên cứu. Kích hoạt: trình duyệt, tự động hóa web, thu thập dữ liệu, điều hướng, nhấp chuột, điền biểu mẫu, chụp ảnh màn hình, duyệt web, playwright, trình duyệt không đầu, tác nhân web, lướt internet, ghi video
agent-tools
skills-101
Chạy các ứng dụng AI qua CLI inference.sh - tạo hình ảnh, tạo video, LLM, tìm kiếm, 3D, tự động hóa Twitter. Các mô hình: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter, và nhiều hơn nữa. Sử dụng khi chạy ứng dụng AI, tạo hình ảnh/video, gọi LLM, tìm kiếm web, hoặc tự động hóa Twitter. Kích hoạt: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
python-executor
skills-101
Execute Python code in a safe sandboxed environment via [inference.sh](https://inference.sh). Pre-installed: NumPy, Pandas, Matplotlib, requests, BeautifulSoup, Selenium, Playwright, MoviePy, Pillow, OpenCV, trimesh, and 100+ more libraries. Use for: data processing, web scraping, image manipulation, video creation, 3D model processing, PDF generation, API calls, automation scripts. Triggers: python, execute code, run script, web scraping, data analysis, image processing, video editing, 3D...
remotion-render
skills-101
Kết xuất video từ mã thành phần React/Remotion thông qua inference.sh. Truyền mã TSX, nhận MP4. Hỗ trợ tất cả API của Remotion: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence. Có thể cấu hình độ phân giải, FPS, thời lượng, codec. Sử dụng cho: tạo video theo chương trình, đồ họa động, thiết kế chuyển động, video dựa trên dữ liệu, hoạt ảnh React thành video. Kích hoạt: remotion, kết xuất video từ mã, tsx sang video, react video, video theo chương trình, remotion render, mã sang video, hoạt ảnh...
landing-page-design
skills-101
Tối ưu hóa chuyển đổi trang đích với quy tắc bố cục, thiết kế phần hero và tâm lý học CTA. Bao gồm công thức above-the-fold, vị trí đặt bằng chứng xã hội, thiết kế di động và cách đọc theo mẫu chữ F. Sử dụng cho: trang đích khởi nghiệp, trang sản phẩm, tiếp thị SaaS, tối ưu hóa chuyển đổi. Kích hoạt: trang đích, phần hero, above the fold, tối ưu hóa chuyển đổi, thiết kế trang đích, nút cta, hình ảnh hero, bố cục trang đích, trang đích saas, thiết kế trang sản phẩm, tỷ lệ chuyển đổi, trang đích...
designmarketing
product-photography
skills-101
Chụp ảnh sản phẩm bằng AI với ánh sáng studio, ảnh phong cách sống và quy ước chụp packshot. Bao gồm góc chụp, phông nền, loại bóng, ảnh hero và yêu cầu hình ảnh thương mại điện tử. Sử dụng cho: ảnh sản phẩm, hình ảnh thương mại điện tử, danh sách Amazon, packshot, chụp ảnh phong cách sống. Kích hoạt: chụp ảnh sản phẩm, ảnh sản phẩm, packshot, chụp ảnh thương mại điện tử, chụp sản phẩm, hình ảnh sản phẩm, chụp ảnh studio, sản phẩm phong cách sống, ảnh sản phẩm Amazon, hình ảnh danh sách sản phẩm, ảnh hero, mockup sản phẩm,...