D

Doany Ai技能

ace-step
doany-ai
Generate, inpaint, and outpaint music with ACE Step on RunComfy via the `runcomfy` CLI. ACE Step is StepFun-AI's open-weights music foundation model — tag-driven composition (genre, mood, instruments), multilingual lyrics with section markers, 5 s to 4 min stereo output, $0.0002–0.0003 per second (≈ 27× cheaper than ElevenLabs Music). Four endpoints: ACE Step text-to-audio (the default), ACE Step 1.5 text-to-audio (50+ language lyrics, refined structured-lyric handling), ACE Step...
creativemediaaudio
ai-avatar-video
doany-ai
Create AI avatar, talking-head, and lip-sync videos on RunComfy via the `runcomfy` CLI. Routes across ByteDance OmniHuman (audio-driven full-body avatar), Wan-AI Wan 2-7 (audio-driven mouth sync via `audio_url` on a portrait), HappyHorse 1.0 (Arena #1 t2v / i2v with in-pass audio), and Seedance v2 Pro (multi-modal cinematic with reference audio + reference subject). Picks the right model for the user's actual intent — UGC voiceover, virtual presenter, dubbed product demo, lip-synced...
videocreativemedia
ai-image-generation
doany-ai
Generate and edit images on RunComfy via the `runcomfy` CLI — a smart router across the full image-model catalog: FLUX 2 (Klein 9B/4B, Pro, Dev, Flash, Turbo, Max), Google Nano Banana 2 / Pro, OpenAI GPT Image 2, ByteDance Seedream 5 / 4-5 / 4-0 and Dreamina 4-0, Alibaba Qwen Image and Z-Image Turbo, Wan 2-7. Covers both text-to-image (t2i) and image-to-image / edit (i2i) endpoints — the skill picks the right model for the user's actual intent (typography precision, photoreal portraits,...
creativemediaimage
ai-music
doany-ai
我們要將提供的英文描述翻譯成繁體中文。注意事項:保留產品名稱、協議名稱、URL、數字和技術術語。不要添加聲明、解釋、Markdown、項目符號、鏈接、標籤、前綴或額外評論。目標語言是繁體中文。目錄項類型是agent skill,名稱是ai-music,但不要包含名稱除非它在源文本中出現。源文本中沒有出現ai-music,所以不用管。 翻譯內容:描述RunComfy上的AI音樂生成,通過runcomfy CLI,智能路由到不同音樂模型。包括ElevenLabs AI Music Generation(高級44.1 kHz立體聲人聲軌道,5秒到5分鐘,$0.0083/秒)和ACE Step / ACE Step 1.5(StepFun-AI開放權重,標籤驅動作曲,多語言歌詞,$0.0002–0.0003/秒,約便宜27倍),還有ACE Step音頻修復
creativemediaaudio
ai-video-generation
doany-ai
Generate AI videos on RunComfy via the `runcomfy` CLI — a smart router across the full video-model catalog: HappyHorse 1.0 (Arena #1, native in-pass audio), Wan-AI Wan 2-7 (open weights, audio-driven lip-sync), ByteDance Seedance v2 / 1-5 / 1-0 (multi-modal cinematic), Kling 3.0 / 2-6, Google Veo 3-1, MiniMax Hailuo 2-3, ByteDance Dreamina 3-0. Covers text-to-video (t2v), image-to-video (i2v), and Veo's video-extend endpoint. The skill picks the right model for the user's intent (Arena-#1...
creativevideomedia
controlnet-pose
doany-ai
Pose-conditioned generation on RunComfy via the `runcomfy` CLI. Routes across Kling 2-6 Motion Control Pro / Standard (transfer the motion / blocking of a reference video onto a target character), community Wan 2-2 Animate (audio-driven character animation with pose conditioning), and Z-Image Turbo ControlNet LoRA (pose-conditioned image generation from an OpenPose / DWPose / canny / depth control image). Picks the right route based on video vs still and stylized vs photoreal. Triggers on...
creativemediavideo
elevenlabs-music-generation
doany-ai
Generate full songs and instrumental tracks with ElevenLabs Music on RunComfy via the `runcomfy` CLI. ElevenLabs Music turns a style description plus structured lyrics into studio-quality 44.1 kHz stereo audio — 5 seconds to 5 minutes — with section-level control (Intro / Verse / Chorus / Bridge), multilingual vocals, and commercial-friendly output. Generate a backing track, a full vocal song, a jingle, a podcast intro, a game loop, or an instrumental bed. Calls `runcomfy run...
creativeaudiomedia
face-swap
doany-ai
我們需要將提供的英文文本翻譯成繁體中文。文本描述了一個名為 face-swap 的 agent skill,內容是關於透過 runcomfy CLI 在 RunComfy 上進行臉部/角色替換到影片或圖片中。文本列舉了多個模型/工具:Wan 2-2 Animate、GPT Image 2 Edit、Nano Banana Edit、Flux Kontext、Kling 2-6 Motion Control Pro,並簡述了各自的功能。注意:不要包含名稱 face-swap 除非它出現在原文中,但原文第一句就有 "face-swap",所以應該保留。但指令說「不要包含名稱除非它出現在源文本中」,所以如果源文本有,就可以保留。另外,不要添加標籤如「描述」等。直接翻譯<text>內的內容。 翻譯時注意專有名詞(如 RunComfy, runcomfy CLI, Wan 2-2 Animate 等)保持原樣。技術術語如 "audio-driven character animation", "identity
creativevideoimage
flux-2-klein
doany-ai
在 RunComfy 上使用 Flux 2 Klein(Black Forest Labs 的 Flux 2 蒸餾快速變體)生成圖像 — 內建該模型的提示模式文檔,使技能比直接對同一模型進行簡單提示能獲得更高品質的輸出。記錄了 Flux 2 Klein 的優勢(亞秒級延遲、多參考品牌風格、宣告式主體優先提示)、步數策略(快速迭代用 4–8 步,精修用約 25 步)、9B 與 4B 變體的取捨,以及何時應轉向 Flux 2 Pro /...
creativeimagemedia
flux-kontext
doany-ai
在 RunComfy 上使用 Flux 1 Kontext Pro(Black Forest Labs 的精確局部影像編輯模型)編輯圖片 — 內建該模型文件中的提示模式,使技能產出比直接對同一模型進行簡單提示更精準的結果。說明 Flux Kontext 的優勢(單一參考精確局部編輯、強大的提示控制、一致的高保真輸出)、架構(單一圖片 + 提示),以及何時應改用 Nano Banana Edit / GPT Image 2 edit / Flux 2 Klein。呼叫...
creativeimagedocument
gpt-image-2
doany-ai
使用 OpenAI GPT Image 2(ChatGPT Images 2.0)在 RunComfy 上生成和編輯圖像。說明 GPT Image 2 的優勢(嵌入文字、標誌、多語言排版、指令精確度)、其三種固定尺寸、保留編輯的語言,以及何時轉向同類工具(Flux 2 / Nano Banana Pro / Seedream)。透過本機 RunComfy CLI 呼叫
creativeimageapi
gpt-image-edit
doany-ai
Edit images with OpenAI GPT Image 2 (the `/edit` endpoint of ChatGPT Images 2.0) on RunComfy — bundled with the model's documented prompting patterns so the skill gets sharper output than naive prompting against the same model. Documents GPT Image Edit's strengths (preservation language, multilingual in-image text editing, multi-reference up to 10 images, layout / typography precision), the schema, and when to route to Nano Banana Edit / Flux Kontext / GPT Image 2 t2i instead. Calls...
imagecreativeapi
happyhorse-1-0
doany-ai
在 RunComfy 上使用 HappyHorse 1.0 生成文字轉影片。說明 HappyHorse 1.0 的優勢(在 Artificial Analysis Video Arena 排名第一、原生 1080p 搭配同步音軌、多鏡次
creativevideomedia
image-edit
doany-ai
在 RunComfy 上編輯圖片——此技能為智能路由器,能將用戶意圖匹配至 RunComfy 目錄中的正確編輯模型。可選用 Nano Banana Edit(批次最多 20 張,預設保留身份)、OpenAI GPT Image 2 Edit(多語言圖內文字重寫、多參考合成、佈局精準)、Flux Kontext Pro(單參考高保真局部編輯)或 Z-Image Turbo Inpaint(遮罩驅動精確區域編輯)。整合各模型的提示模式文檔,使技能能...
creativeimagemedia
image-inpainting
doany-ai
We need to translate the given text from English to Traditional Chinese. The text describes an agent skill for image inpainting. We must preserve product names, protocol names, URLs, numbers, technical terms. The name "image-inpainting" is to be preserved but not included unless it appears in source text. The source text does not include the name "image-inpainting" as a standalone; it appears as part of "image inpainting" in the first sentence. But the instruction says "Do not include the name unless it appears in the source text." The name "image-inpainting" is the directory item type, but we are translating the text inside <text>. The text includes "image inpainting" (without hyphen) as a phrase. We should translate that phrase naturally. Also preserve "RunComfy", "runcomfy CLI", "Tongyi MAI Z-Image Turbo Inpainting", "Nano Banana 2 Edit", "GPT Image 2 Edit", "FLUX Kontext Pro". Translate the rest. Let's break down
creativeimagemedia
image-outpainting
doany-ai
We need to translate the given text from English to Traditional Chinese. The text describes an image outpainting skill on RunComfy via the runcomfy CLI. It mentions extending a still beyond its original canvas, filling in missing parts, changing aspect ratio while preserving content. It lists several routes: Nano Banana 2 Edit, GPT Image 2 Edit, FLUX Kontext Pro, and brand edit endpoints (Seedream / Dreamina / Qwen / FLUX 2). The text ends with "Picks the..." which is incomplete. We should translate the entire provided text, preserving names like "RunComfy", "runcomfy", "Nano Banana 2 Edit", "GPT Image 2 Edit", "FLUX Kontext Pro", "Seedream", "Dreamina", "Qwen", "FLUX 2". Also preserve "CLI". Translate the rest naturally. Note: The instruction says "Do not include the name unless it appears in the source text." The name "image-outpainting" is not in the source text, so we don't
creativeimagemedia
image-to-video
doany-ai
We need to translate the given text from English to Traditional Chinese. The text describes a skill called "image-to-video" but the instruction says to preserve the name only if it appears in the source text. The name "image-to-video" does appear in the source? Actually the source text starts with "Animate any still image on RunComfy — this skill is a smart router..." The name "image-to-video" is not explicitly in the source text; it's given as the directory item name. The instruction says "Do not include the name unless it appears in the source text." So we should not add "image-to-video" in the translation. We only translate the text inside <text>. Also preserve product names, protocol names, URLs, numbers, technical terms. So "RunComfy", "HappyHorse 1.0 I2V", "Arena #1", "Wan 2.7", "audio_url", "Seedance 2.0 Pro" should remain as is. Translate the rest into Traditional Chinese. Be careful
creativevideomedia
kling-3-0
doany-ai
在 RunComfy 上進行 Kling 3.0 影片生成。Kling 3.0(也稱為 Kling V3.0)是快手科技的第三代多鏡頭影片模型,具備原生同步音訊及跨鏡頭一致的角色身份。此技能涵蓋所有六個 Kling 3.0 端點,橫跨三種渲染層級(標準、專業、4K)與兩種模式(文字轉影片、圖片轉影片)。透過本機 RunComfy CLI 執行 runcomfy run kling/kling-3.0/ /。觸發詞為「kling」、「kling 3.0」、「kling v3」、「kling pro」等。
videocreativemedia
lipsync
doany-ai
We need to translate the given text from English to Traditional Chinese. The text describes a skill for lip-syncing a face to an audio track using various tools. We must preserve names like "lipsync", "RunComfy", "runcomfy", "ByteDance OmniHuman", "Sync Labs sync v2 / Pro", "Kling lipsync", "Creatify lipsync". Also preserve technical terms like "CLI", "audio-driven full-body avatar", "portrait + audio", "mouth sync", "audio-to-video", "text-to-video", "mouth-swap". The translation should be in Traditional Chinese (zh-TW). Do not add any extra commentary or labels. Just output the translated text inside the <text> tag? Actually the instruction says "Translate only the text inside <text>. Do not include the name unless it appears in the source text." But the source text is given inside <text> tags. So we need to output the translated version of that text, presumably without the <text> tags? The instruction says
creativevideomedia
nano-banana-2
doany-ai
使用 RunComfy 上的 Google Nano Banana 2(Gemini 系列 flash 層級文字轉圖像)生成圖片——內建該模型的提示模式,使技能輸出比直接對同一模型進行簡單提示更為精準。說明 Nano Banana 2 的優勢(快速迭代、圖像內文字渲染、可預測構圖、可選網路接地上下文)、解析度層級定價、安全容忍度調節,以及何時轉向 Nano Banana Pro / GPT Image 2 / Flux 2 / Seedream……
creativeimageresearch
nano-banana-edit
doany-ai
Edit images with Google Nano Banana 2 (image-to-image edit endpoint) on RunComfy. Documents Nano Banana Edit's strengths (preserve subject identity, swap background, localize edits with spatial language, multi-image batch edits up to 20 inputs), the schema, and when to route to GPT Image 2 edit / Flux Kontext / Nano Banana 2 t2i instead. Calls `runcomfy run google/nano-banana-2/edit` through the local RunComfy CLI. Triggers on "nano banana edit", "edit with nano banana", "image edit nano...
creativeimageapi
relight
doany-ai
Relight a still image — change the lighting setup, color temperature, direction, or mood — on RunComfy via the `runcomfy` CLI. Routes to Qwen Edit 2509's dedicated `relight` LoRA endpoint for purpose-built relighting, with fallback to identity-preserving edit endpoints (Nano Banana 2 Edit, GPT Image 2 Edit, FLUX Kontext Pro) when prose lighting language is enough. Use for product relighting (studio softbox → window light), portrait mood shift (overcast → golden hour), or color-grade change....
creativeimagemedia
runcomfy-cli
doany-ai
從命令列在 Run
creativemediaapi
seedance-v2
doany-ai
We are asked to translate the text inside <text> into Traditional Chinese. The target language is 繁體中文. We must preserve the name "seedance-v2" if it appears, but it does not appear in the source text. The source text mentions "Seedance 2.0 Pro", "RunComfy", "HappyHorse 1.0", "Wan 2.7", "Kling", and the command "runcomfy run bytedance/seedance-v2/pro". We need to preserve these product names, protocol names, URLs, numbers, and technical terms. Do not add any extra commentary or labels. Just output the translated text. Let's translate the text: "Generate cinematic short-form video with ByteDance Seedance 2.0 Pro on RunComfy. Documents Seedance 2.0 Pro's strengths (multi-modal references — up to 9 images, 3 videos, 3 audio — synchronized in-pass audio with natural lip-sync, cinematic motion refinement), the 4–15s duration schema, and
videocreativemedia
video-edit
doany-ai
在 RunComfy 上編輯現有影片 — 此技能為智慧路由器,可將使用者意圖匹配至 RunComfy 目錄中的正確編輯模型。選用 Wan 2.7 Edit-Video(通用重製風格 / 背景替換 / 包裝替換,保留身份與動作)、Kling 2.6 Pro Motion Control(將參考影片的精確動作轉移至目標角色),或 Lucy Edit Restyle(輕量級身份穩定重製風格 / 服裝替換)。整合各模型的提示模式,使技能...
videocreativemedia
video-extend
doany-ai
我们要求翻译一段文本,目标语言是繁体中文。需要保留产品名称、协议名称、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称本身,除非名称出现在源文本中。不要添加标签如"description"等。 源文本是英文,描述了一个agent skill,名为video-extend。翻译时注意保持技术术语如"RunComfy", "runcomfy CLI", "Google Veo 3-1", "extend-video", "fast/extend-video"等不变。注意"Veo clip"中的Veo是产品名,保留。另外"prompt"可以翻译为"提示"或保留?技术术语通常保留英文,但这里上下文是描述,可以翻译为"提示"以符合中文习惯。但要求保留技术术语,所以"prompt"可以保留或翻译?谨慎起见,保留英文"prompt"可能更安全,但中文中常用"提示
videocreativemedia
video-inpainting
doany-ai
Region edits across video frames on RunComfy via the `runcomfy` CLI — remove an object that appears across many frames, clean up wires or watermarks, replace a region with matching motion. Routes across Wan 2-7 edit-video (default, prompt-driven region edits with spatial language), Lucy Edit Restyle (identity-stable region-aware restyle), and Seedream 4-0 edit-sequential (when treating the clip as a frame stack). Picks the right route based on whether the change is prose-driven,...
videocreativemedia
video-outpainting
doany-ai
We need to translate the given text from English to Traditional Chinese. The text describes a video outpainting skill on RunComfy via the runcomfy CLI. It explains extending spatial canvas, changing aspect ratio, adding environment, preserving central action, using Wan 2-7 edit-video, and dedicated ComfyUI workflows. Also mentions triggers. We must preserve product names, protocol names, URLs, numbers, technical terms. So "RunComfy", "runcomfy CLI", "Wan 2-7 edit-video", "ComfyUI" should remain as is. Also "video-outpainting" is the name to preserve but it's not in the text? Actually the instruction says "Name to preserve: video-outpainting" but the text does not contain that exact string. The text has "video outpainting" (lowercase, no hyphen). We should preserve the term as it appears? The instruction says "Do not include the name unless it appears in the source text." So we only translate the text as given. The name "video-out
videocreativemedia
wan-2-7
doany-ai
Generate text-to-video with Wan 2.7 (Wan-AI's flagship motion model) on RunComfy. Documents Wan 2.7's strengths (multi-reference conditioning, audio-driven lip-sync via `audio_url`, smoother transitions, prompt expansion), the duration / resolution / aspect-ratio schema, and when to route to HappyHorse 1.0 / Seedance 2.0 / Kling / LTX 2 instead. Calls `runcomfy run wan-ai/wan-2-7/text-to-video` through the local RunComfy CLI. Triggers on "wan", "wan 2.7", "wan-2-7", "wan video", or any...
creativevideomedia