媒体技能

ace-step
agentspace-so
Generate, inpaint, and outpaint music with ACE Step on RunComfy via the `runcomfy` CLI. ACE Step is StepFun-AI's open-weights music foundation model — tag-driven composition (genre, mood, instruments), multilingual lyrics with section markers, 5 s to 4 min stereo output, $0.0002–0.0003 per second (≈ 27× cheaper than ElevenLabs Music). Four endpoints: ACE Step text-to-audio (the default), ACE Step 1.5 text-to-audio (50+ language lyrics, refined structured-lyric handling), ACE Step...
creativeaudiomedia
ace-step
doany-ai
Generate, inpaint, and outpaint music with ACE Step on RunComfy via the `runcomfy` CLI. ACE Step is StepFun-AI's open-weights music foundation model — tag-driven composition (genre, mood, instruments), multilingual lyrics with section markers, 5 s to 4 min stereo output, $0.0002–0.0003 per second (≈ 27× cheaper than ElevenLabs Music). Four endpoints: ACE Step text-to-audio (the default), ACE Step 1.5 text-to-audio (50+ language lyrics, refined structured-lyric handling), ACE Step...
creativemediaaudio
ai-avatar-video
agentspace-so
We need to translate the given English text into Simplified Chinese. The instruction says to preserve product names, protocol names, URLs, numbers, and technical terms. The name "ai-avatar-video" is not in the text, so we don't include it. We must not add any labels or extra commentary. Just translate the text inside <text>. The text describes creating AI avatar videos using runcomfy CLI, mentioning various models: ByteDance OmniHuman, Wan-AI Wan 2-7, HappyHorse 1.0, Seedance v2 Pro. Also mentions user intents like UGC voiceover, virtual presenter, etc. We need to translate accurately, keeping technical terms and names as is. For example, "talking-head" might be translated as "说话头像" or keep as "talking-head"? Probably keep as "talking-head" or translate? The instruction says preserve technical terms, but "talking-head" is a common term. I'll translate it as "说话头像" but maybe keep English? Better to translate common terms
videocreativemedia
ai-avatar-video
qu-skills
通过inference.sh CLI创建AI虚拟形象和说话头像视频。推荐:P-Video-Avatar(最快、最便宜、内置TTS)。其他选项:OmniHuman、Fabric、PixVerse。音频:Inworld TTS-2(支持100多种语言、角色情感控制)、ElevenLabs、Kokoro。功能:音频驱动虚拟形象、文本转虚拟形象、唇形同步视频、说话头像生成、虚拟主持人、UGC内容。用途:AI主持人、解说视频、虚拟网红、配音、营销视频、UGC广告、游戏虚拟形象……
videocreativemedia
ai-avatar-video
runcomfy-com
Create AI avatar, talking-head, and lip-sync videos on RunComfy via the `runcomfy` CLI. Routes across ByteDance OmniHuman (audio-driven full-body avatar), Wan-AI Wan 2-7 (audio-driven mouth sync via `audio_url` on a portrait), HappyHorse 1.0 (Arena #1 t2v / i2v with in-pass audio), and Seedance v2 Pro (multi-modal cinematic with reference audio + reference subject). Picks the right model for the user's actual intent — UGC voiceover, virtual presenter, dubbed product demo, lip-synced...
videocreativemedia
ai-avatar-video
doany-ai
We need to translate the given English text into Simplified Chinese. The instruction says to preserve product names, protocol names, URLs, numbers, and technical terms. The name "ai-avatar-video" is not in the text, so we don't include it. We must not add any labels or extra commentary. Just translate the text inside <text> tags. The text describes creating AI avatar videos using runcomfy CLI, mentioning various models: ByteDance OmniHuman, Wan-AI Wan 2-7, HappyHorse 1.0, Seedance v2 Pro. Also mentions user intents like UGC voiceover, virtual presenter, etc. We need to ensure technical terms like "t2v", "i2v", "in-pass audio", "multi-modal cinematic" are preserved or appropriately translated? The instruction says preserve technical terms, so we can keep them as is or translate if common? But "t2v" and "i2v" are likely abbreviations for text-to-video and image-to-video, which are common. However,
videocreativemedia
ai-avatar-video
halt-catch-fire
通过inference.sh CLI创建AI虚拟形象和说话头像视频。推荐:P-Video-Avatar(最快、最便宜、内置TTS)。其他选项:OmniHuman、Fabric、PixVerse。音频:Inworld TTS-2(支持100+语言、角色情感控制)、ElevenLabs、Kokoro。功能:音频驱动虚拟形象、文本转虚拟形象、唇形同步视频、说话头像生成、虚拟主持人、UGC内容。用途:AI主持人、解说视频、虚拟网红、配音、营销视频、UGC广告、游戏虚拟形象……
videocreativemedia
ai-avatar-video
101-skills
通过inference.sh CLI创建AI虚拟形象和说话头像视频。推荐:P-Video-Avatar(最快、最便宜、内置TTS)。其他选项:OmniHuman、Fabric、PixVerse。音频:Inworld TTS-2(支持100+语言、角色情感控制)、ElevenLabs、Kokoro。功能:音频驱动虚拟形象、文本转虚拟形象、唇形同步视频、说话头像生成、虚拟主持人、UGC内容。用途:AI主持人、解说视频、虚拟网红、配音、营销视频、UGC广告、游戏虚拟形象等。
creativevideomedia
ai-image-generation
qu-skills
通过 inference.sh CLI 使用 GPT-Image-2、FLUX、Gemini、Grok、Seedream、Reve 及 50 多个模型生成 AI 图像。模型包括:GPT-Image-2、FLUX Dev LoRA、FLUX.2 Klein LoRA、Gemini 3 Pro Image、Grok Imagine、Seedream 4.5、Reve、ImagineArt。功能:文本转图像、图像转图像、修复、LoRA、图像编辑、放大、文本渲染。用途:AI 艺术、产品模型、概念艺术、社交媒体图形、营销视觉、插图。触发词:flux、图像生成、AI 图像、文本转...
creativemediaimage
ai-image-generation
agentspace-so
Generate and edit images on RunComfy via the `runcomfy` CLI — a smart router across the full image-model catalog: FLUX 2 (Klein 9B/4B, Pro, Dev, Flash, Turbo, Max), Google Nano Banana 2 / Pro, OpenAI GPT Image 2, ByteDance Seedream 5 / 4-5 / 4-0 and Dreamina 4-0, Alibaba Qwen Image and Z-Image Turbo, Wan 2-7. Covers both text-to-image (t2i) and image-to-image / edit (i2i) endpoints — the skill picks the right model for the user's actual intent (typography precision, photoreal portraits,...
creativemediaimage
ai-image-generation
doany-ai
我们要求翻译一段文本,目标语言是简体中文。文本内容是关于一个名为"ai-image-generation"的agent skill的描述。需要保留产品名称、协议名称、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称本身,除非名称出现在源文本中。注意:源文本中提到了"ai-image-generation"吗?没有直接出现,但目录项类型是agent skill,名称是ai-image-generation,但翻译指令说"不要包含名称,除非它出现在源文本中"。源文本中没有出现"ai-image-generation",所以不翻译它。只翻译<text>内的英文。 文本内容:描述通过runcomfy CLI在RunComfy上生成和编辑图像,是一个智能路由器,涵盖完整的图像模型目录:FLUX 2 (Klein 9B/4B, Pro, Dev, Flash, Turbo, Max), Google Nano Banana 2 / Pro, OpenAI GPT Image 2, ByteDance
creativemediaimage
ai-image-generation
runcomfy-com
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名称、协议名称、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容。注意不要包含名称"ai-image-generation"除非它在源文本中出现。源文本中并没有出现该名称,所以不翻译。 文本内容:描述一个agent skill,通过runcomfy CLI在RunComfy上生成和编辑图像,是一个智能路由器,涵盖多个图像模型。需要翻译成中文,保留模型名称如FLUX 2 (Klein 9B/4B, Pro, Dev, Flash, Turbo, Max), Google Nano Banana 2 / Pro, OpenAI GPT Image 2, ByteDance Seedream 5 / 4-5 / 4-0 and Dreamina 4-0, Alibaba Qwen Image and Z-Image Turbo, Wan 2-7。以及术语如text-to-image (t2i), image-to-image / edit (
creativemediaimage
ai-image-generation
halt-catch-fire
通过 inference.sh CLI 使用 GPT-Image-2、FLUX、Gemini、Grok、Seedream、Reve 及 50 多个模型生成 AI 图像。模型包括:GPT-Image-2、FLUX Dev LoRA、FLUX.2 Klein LoRA、Gemini 3 Pro Image、Grok Imagine、Seedream 4.5、Reve、ImagineArt。功能:文本转图像、图像转图像、图像修复、LoRA、图像编辑、图像放大、文本渲染。用途:AI 艺术、产品模型、概念艺术、社交媒体图形、营销视觉、插图。触发词:flux、图像生成、AI 图像、文本转...
creativemediaimage
ai-image-generation
101-skills
通过 inference.sh CLI 使用 GPT-Image-2、FLUX、Gemini、Grok、Seedream、Reve 及 50 多个模型生成 AI 图像。模型包括:GPT-Image-2、FLUX Dev LoRA、FLUX.2 Klein LoRA、Gemini 3 Pro Image、Grok Imagine、Seedream 4.5、Reve、ImagineArt。功能:文生图、图生图、图像修复、LoRA、图像编辑、图像放大、文本渲染。用途:AI 艺术、产品模型、概念艺术、社交媒体图形、营销视觉、插图。触发词:flux、图像生成、AI 图像、文本转...
creativemediaimage
ai-music
runcomfy-com
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名、协议名、URL、数字、技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容。注意不要包含名称"ai-music"除非它在源文本中出现。源文本中出现了"ai-music"吗?没有,所以不包含。直接翻译。 文本内容:描述RunComfy上的AI音乐生成,通过runcomfy CLI,智能路由到多个模型。包括ElevenLabs AI Music Generation(高级44.1 kHz立体声人声轨道,5秒到5分钟,$0.0083/秒)和ACE Step / ACE Step 1.5(StepFun-AI开放权重,标签驱动作曲,多语言歌词,$0.0002–0.0003/秒,约便宜27倍),还有ACE Step音频内补(在现有轨道内重新生成一段时间范围)和音频外补(在轨道前后扩展)。选择
creativeaudiomedia
ai-music
doany-ai
Generate AI music on RunComfy via the `runcomfy` CLI — a smart router across the music-model catalog. Routes to ElevenLabs AI Music Generation (premium 44.1 kHz stereo vocal tracks, 5 s–5 min, $0.0083/s) and ACE Step / ACE Step 1.5 (StepFun-AI open-weights, tag-driven composition, multilingual lyrics, $0.0002–0.0003/s, ~27× cheaper), plus ACE Step audio-inpaint (regenerate a time range inside an existing track) and ACE Step audio-outpaint (extend a track before or after). Picks the right...
creativemediaaudio
ai-video-generation
qu-skills
通过 inference.sh CLI,使用 Google Veo、Seedance 2.0、HappyHorse、Wan、Grok 及 40 多个模型生成 AI 视频。模型包括:Veo 3.1、Veo 3、Seedance 2.0、HappyHorse 1.0、Wan 2.5、Grok Imagine Video、OmniHuman、Fabric、HunyuanVideo。功能涵盖:文本转视频、图像转视频、参考转视频、视频编辑、唇形同步、虚拟形象动画、视频增强、拟音音效。适用于:社交媒体视频、营销内容、解说视频、产品演示、AI 虚拟形象。触发词:视频生成、AI 视频……
videocreativemedia
ai-video-generation
agentspace-so
Generate AI videos on RunComfy via the `runcomfy` CLI — a smart router across the full video-model catalog: HappyHorse 1.0 (Arena #1, native in-pass audio), Wan-AI Wan 2-7 (open weights, audio-driven lip-sync), ByteDance Seedance v2 / 1-5 / 1-0 (multi-modal cinematic), Kling 3.0 / 2-6, Google Veo 3-1, MiniMax Hailuo 2-3, ByteDance Dreamina 3-0. Covers text-to-video (t2v), image-to-video (i2v), and Veo's video-extend endpoint. The skill picks the right model for the user's intent (Arena-#1...
creativevideomedia
ai-video-generation
runcomfy-com
Generate AI videos on RunComfy via the `runcomfy` CLI — a smart router across the full video-model catalog: HappyHorse 1.0 (Arena #1, native in-pass audio), Wan-AI Wan 2-7 (open weights, audio-driven lip-sync), ByteDance Seedance v2 / 1-5 / 1-0 (multi-modal cinematic), Kling 3.0 / 2-6, Google Veo 3-1, MiniMax Hailuo 2-3, ByteDance Dreamina 3-0. Covers text-to-video (t2v), image-to-video (i2v), and Veo's video-extend endpoint. The skill picks the right model for the user's intent (Arena-#1...
creativevideomedia
ai-video-generation
halt-catch-fire
通过 inference.sh CLI,使用 Google Veo、Seedance 2.0、HappyHorse、Wan、Grok 及 40 多个模型生成 AI 视频。模型包括:Veo 3.1、Veo 3、Seedance 2.0、HappyHorse 1.0、Wan 2.5、Grok Imagine Video、OmniHuman、Fabric、HunyuanVideo。功能涵盖:文生视频、图生视频、参考视频生成、视频编辑、唇形同步、虚拟人动画、视频增强、拟音音效。适用于:社交媒体视频、营销内容、解说视频、产品演示、AI 虚拟人。触发词:视频生成、AI 视频……
creativevideomedia
ai-video-generation
101-skills
通过 inference.sh CLI,使用 Google Veo、Seedance 2.0、HappyHorse、Wan、Grok 及 40 多个模型生成 AI 视频。模型包括:Veo 3.1、Veo 3、Seedance 2.0、HappyHorse 1.0、Wan 2.5、Grok Imagine Video、OmniHuman、Fabric、HunyuanVideo。功能涵盖:文生视频、图生视频、参考视频生成、视频编辑、唇形同步、虚拟人动画、视频增强、拟音音效。适用于:社交媒体视频、营销内容、解说视频、产品演示、AI 虚拟人。触发词:视频生成、AI 视频……
creativevideomedia
ai-video-generation
doany-ai
We need to translate the given text from English to Simplified Chinese. The text describes an agent skill for AI video generation using RunComfy CLI. It lists various models and capabilities. We must preserve product names, protocol names, URLs, numbers, technical terms. No extra commentary. The name "ai-video-generation" is not in the text, so we don't include it. We translate only the text inside <text>. The text includes a list of models with versions and descriptions. We need to translate the descriptions but keep model names and version numbers as is. Also keep "RunComfy", "CLI", "Arena", "open weights", "audio-driven lip-sync", "multi-modal cinematic", "text-to-video (t2v)", "image-to-video (i2v)", "video-extend endpoint". For Chinese, we can use common translations for terms like "native in-pass audio" -> "原生内嵌音频", "smart router" -> "智能路由器", "covers" -> "涵盖", etc. Ensure the translation is natural and
creativevideomedia
character-design-sheet
halt-catch-fire
利用参考图和LoRA技术实现AI生成图像中角色的一致性。涵盖转面视图、表情表、配色方案及风格一致性技巧。适用于:角色设计、游戏美术、插画、动画、漫画、视觉小说。触发词:角色设计、角色表、角色一致性、角色参考、转面图、表情表、角色美术、一致角色、角色概念、参考图、角色创建、OC设计……
creativedesignmedia
character-design-sheet
101-skills
利用参考图和LoRA技术实现AI生成图像中角色的一致性。涵盖转面视图、表情表、配色方案及风格一致性技巧。适用于:角色设计、游戏美术、插画、动画、漫画、视觉小说。触发词:角色设计、角色表、角色一致性、角色参考、转面图、表情表、角色美术、一致角色、角色概念、参考图、角色创建、OC设计……
creativedesignmedia
character-design-sheet
qu-skills
利用参考图和LoRA技术实现AI生成图像中角色的一致性。涵盖转面视图、表情表、配色方案及风格一致性技巧。适用于:角色设计、游戏美术、插画、动画、漫画、视觉小说。触发词:角色设计、角色表、角色一致性、角色参考、转面图、表情表、角色美术、一致角色、角色概念、参考图、角色创建、OC设计……
creativedesignmedia
ckm:design
nextlevelbuilder
全面设计技能:品牌识别、设计令牌、UI样式、标志生成(55种风格,Gemini AI)、企业识别方案(50项交付物,CIP样机)、HTML演示文稿(Chart.js)、横幅设计(22种风格,社交/广告/网页/印刷)、图标设计(15种风格,SVG,Gemini 3.1 Pro)、社交照片(HTML→截图,多平台)。操作:设计标志、创建CIP、生成样机、制作幻灯片、设计横幅、生成图标、创建社交照片、社交媒体图像、品牌……
designcreativemedia
controlnet-pose
doany-ai
We need to translate the given text from English to Simplified Chinese. The instruction says to preserve product names, protocol names, URLs, numbers, and technical terms. The name "controlnet-pose" is to be preserved but not included unless it appears in the source text. The source text does not include "controlnet-pose" explicitly, so we don't add it. We translate only the text inside <text>. No extra commentary, labels, etc. The text describes a pose-conditioned generation system on RunComfy via CLI, with routes across Kling 2-6 Motion Control Pro/Standard, community Wan 2-2 Animate, and Z-Image Turbo ControlNet LoRA. It picks the right route based on video vs still and stylized vs photoreal. Ends with "Triggers on..." which is incomplete but we translate as is. We need to preserve: "RunComfy", "runcomfy", "Kling 2-6 Motion Control Pro / Standard", "Wan 2-2 Animate", "Z-Image Turbo Control
creativemediavideo
controlnet-pose
runcomfy-com
We need to translate the given text from English to Simplified Chinese. The instruction says to preserve product names, protocol names, URLs, numbers, and technical terms. The name "controlnet-pose" is to be preserved but not included unless it appears in the source text. The source text does not include "controlnet-pose" explicitly, so we don't add it. We translate only the text inside <text>. No extra commentary, labels, etc. The text describes a pose-conditioned generation system on RunComfy via CLI, with routes across Kling 2-6 Motion Control Pro/Standard, community Wan 2-2 Animate, and Z-Image Turbo ControlNet LoRA. It picks the right route based on video vs still and stylized vs photoreal. Ends with "Triggers on..." which is incomplete. We need to translate accurately, preserving terms like "RunComfy", "CLI", "Kling 2-6 Motion Control Pro / Standard", "Wan 2-2 Animate", "Z-Image Turbo ControlNet Lo
creativemediavideo
design
nextlevelbuilder
We need to translate the given text from English to Simplified Chinese. The instruction says to preserve product names, protocol names, URLs, numbers, and technical terms. The name "design" is to be preserved if it appears in the source text, but it does appear as part of "design skill" and elsewhere. However, the instruction says "Do not include the name unless it appears in the source text." So we should translate the text as is, keeping "design" as is when it appears. Also preserve terms like "Gemini AI", "Chart.js", "SVG", "Gemini 3.1 Pro", "HTML", "CIP", etc. Numbers and styles should be kept. The text is a description of a comprehensive design skill. Translate naturally. Let me break down the text: "Comprehensive design skill: brand identity, design tokens, UI styling, logo generation (55 styles, Gemini AI), corporate identity program (50 deliverables, CIP mockups), HTML presentations (Chart.js), banner design (22 styles, social/ads/web/print), icon design
designcreativemedia
elevenlabs-music-generation
runcomfy-com
Generate full songs and instrumental tracks with ElevenLabs Music on RunComfy via the `runcomfy` CLI. ElevenLabs Music turns a style description plus structured lyrics into studio-quality 44.1 kHz stereo audio — 5 seconds to 5 minutes — with section-level control (Intro / Verse / Chorus / Bridge), multilingual vocals, and commercial-friendly output. Generate a backing track, a full vocal song, a jingle, a podcast intro, a game loop, or an instrumental bed. Calls `runcomfy run...
creativeaudiomedia
elevenlabs-music-generation
doany-ai
Generate full songs and instrumental tracks with ElevenLabs Music on RunComfy via the `runcomfy` CLI. ElevenLabs Music turns a style description plus structured lyrics into studio-quality 44.1 kHz stereo audio — 5 seconds to 5 minutes — with section-level control (Intro / Verse / Chorus / Bridge), multilingual vocals, and commercial-friendly output. Generate a backing track, a full vocal song, a jingle, a podcast intro, a game loop, or an instrumental bed. Calls `runcomfy run...
creativeaudiomedia
elevenlabs-music-generation
agentspace-so
Generate full songs and instrumental tracks with ElevenLabs Music on RunComfy via the `runcomfy` CLI. ElevenLabs Music turns a style description plus structured lyrics into studio-quality 44.1 kHz stereo audio — 5 seconds to 5 minutes — with section-level control (Intro / Verse / Chorus / Bridge), multilingual vocals, and commercial-friendly output. Generate a backing track, a full vocal song, a jingle, a podcast intro, a game loop, or an instrumental bed. Calls `runcomfy run...
creativeaudiomedia
embedded-captions
heygen-com
为说话人视频添加字幕。一个包含32种视觉风格的目录(CATALOG.md),基于两种引擎:列流(字幕合成到场景中——遮罩遮挡+混合模式;奶油/墨水/编辑/主题演讲/纪录片/响亮/霓虹/故障/铬/速度)和主题构成(锚点/军械/终端/霓虹灯/星尘/跺脚/记分牌/交通/VHS/街机/档案/激光/雷声/全息/生物光/极光/光谱/剪纸/弹出/黑板/涂鸦/画笔/水墨/勒索/最后一页/夜城...
videocreativemedia
extension-camera
caffeinelabs
支持网络摄像头。
mediavideoimage
extension-object-storage
caffeinelabs
通用文件/对象存储,适用于图片、视频、文件、文档等批量数据。非常适合图片库、视频库及其他文件或对象管理。支持超过IC限制的大文件,可通过浏览器缓存的HTTP URL访问。
developmentmedia
extension-qr-code
caffeinelabs
使用摄像头的二维码扫描器。
productivitymediaimage
faceless-explainer
heygen-com
faceless-explainer 视频工作流程 - 任意文本(文章/笔记/主题/简报)→ narrator_scripts.json + 音频(语音+背景音乐)+ section_plan.md → 排版/抽象图形/图表/数据可视化视频。典型时长约3分钟以内(最佳时长约30-90秒);真正较长的内容属于通用视频,不适用此工作流程。生成自身的旁白(TTS)——不与用户提供/预先录制的画外音同步(那是通用视频)。不涉及网站抓取,无真实产品截图……
videocreativemedia
flux-2-klein
doany-ai
在RunComfy上使用Flux 2 Klein(Black Forest Labs的Flux 2蒸馏快速变体)生成图像——该技能内置了模型文档中的提示模式,因此相比对同一模型进行简单提示,能获得更清晰的输出。文档说明了Flux 2 Klein的优势(亚秒级延迟、多参考品牌风格、声明式主语优先提示)、步数策略(快速迭代用4–8步,精修用约25步)、9B与4B变体的权衡,以及何时转向Flux 2 Pro/...
creativeimagemedia
general-video
heygen-com
用作自定义HyperFrames HTML视频合成创作的后备方案,适用于无专门工作流程的场景。涵盖较长或多场景作品、品牌/宣传片、蒙太奇、标题卡、动态海报、静态循环以及任意长度或格式的自由创作。不适用于营销产品推广(product-launch-video)、通用网站转视频(website-to-video)、主题解说(faceless-explainer)、GitHub PR视频(pr-to-video)、为现有素材添加字幕等场景。
videocreativemedia
gpt-image
qu-skills
通过inference.sh CLI使用OpenAI GPT-Image-2生成和编辑图像。模型:GPT-Image-2。功能:文本转图像、图像编辑、修复、基于遮罩的编辑、多图像参考、批量生成。用途:产品模型、营销视觉、图像编辑、概念艺术、修复、照片处理。触发词:gpt image、gpt-image-2、openai image、chatgpt image、dall-e、dalle、openai image generation、gpt image edit、gpt inpainting、openai dall-e、gpt 4o image
creativeimagemedia
gpt-image
101-skills
通过inference.sh CLI使用OpenAI GPT-Image-2生成和编辑图像。模型:GPT-Image-2。功能:文本转图像、图像编辑、修复、基于遮罩的编辑、多图像参考、批量生成。用途:产品模型图、营销视觉素材、图像编辑、概念艺术、修复、照片处理。触发词:gpt image、gpt-image-2、openai image、chatgpt image、dall-e、dalle、openai image generation、gpt image edit、gpt inpainting、openai dall-e、gpt 4o image
creativemediaimage
gpt-image-2
runcomfy-com
Generate and edit images with OpenAI GPT Image 2 (ChatGPT Images 2.0) on RunComfy. Documents GPT Image 2's strengths (embedded text, logos, multilingual typography, instruction precision), its 3 fixed sizes, edit-with-preservation language, and when to route to a sibling (Flux 2 / Nano Banana Pro / Seedream) instead. Calls `runcomfy run openai/gpt-image-2/text-to-image` or `/edit` through the local RunComfy CLI. Triggers on "gpt image 2", "gpt-image-2", "ChatGPT Images 2", "image 2", or any...
creativeimagemedia
happyhorse
qu-skills
通过inference.sh命令行界面,使用阿里巴巴HappyHorse 1.0模型生成和编辑视频。模型包括:HappyHorse T2V、I2V、R2V、Video Edit。功能:文本转视频、图像转视频、参考转视频、自然语言视频编辑、角色保留、720P/1080P分辨率、最长15秒。用途:物理逼真视频、视频编辑、角色一致性内容、产品演示、社交媒体。触发词:happyhorse、happy horse、alibaba video、happyhorse 1.0、dashscope video、alibaba...
creativevideomedia
happyhorse
101-skills
通过inference.sh命令行界面,使用阿里巴巴HappyHorse 1.0模型生成和编辑视频。模型包括:HappyHorse T2V、I2V、R2V、Video Edit。功能:文本转视频、图像转视频、参考转视频、自然语言视频编辑、角色保留、720P/1080P分辨率、最长15秒。用途:物理逼真视频、视频编辑、角色一致性内容、产品演示、社交媒体。触发词:happyhorse、happy horse、alibaba video、happyhorse 1.0、dashscope video、alibaba...
videocreativemedia
happyhorse-1-0
agentspace-so
在RunComfy上使用HappyHorse 1.0生成文本到视频。介绍HappyHorse 1.0的优势(在Artificial Analysis Video Arena排名第一,原生1080p并同步音频,多镜头角色一致性,支持6种语言提示),时长/宽高比/分辨率方案,以及何时转向Wan
creativevideomedia
happyhorse-1-0
doany-ai
在 RunComfy 上使用 HappyHorse 1.0 生成文本到视频。介绍 HappyHorse 1.0 的优势(在 Artificial Analysis Video Arena 排名第一,原生 1080p 并带有同步音频,多镜头角色一致性,支持 6 种语言提示)、时长
creativevideomedia
happyhorse-1-0
runcomfy-com
在RunComfy上使用HappyHorse 1.0生成文本到视频。介绍HappyHorse 1.0的优势(在Artificial Analysis Video Arena排名第一,原生1080p并同步音频,多
creativevideomedia
higgsfield-generate
higgsfield-ai
通过Higgsfield AI生成图像/视频。默认:GPT Image 2用于图像/设计/文字,Seedance 2.0用于视频,Nano Banana 2/Pro用于角色/参考图像工作,Marketing Studio用于包含头像/产品/钩子的广告、设置,以及Soul V2/Cinema/Cast/Location和Kling 3.0。使用场景:“生成一张图像”、“制作一段视频”、“让这张照片动起来”、“图像转视频”、“编辑/风格化/混音这张图像”、“制作一个短片”、“创建一则广告”、“制作UGC视频”、“产品演示”、“开箱”、“品牌视频”……
creativemediavideo
higgsfield-soul-id
higgsfield-ai
训练一个灵魂角色——基于个人面部特征的个性化模型,用于Higgsfield生成身份保真的图像和视频。使用场景:“创建我的灵魂”、“训练我的面部”、“制作我的数字分身”、“为我构建头像”、“学习我的外貌”、“创建我的角色”、“为视频设置身份”、“我希望我的脸出现在生成的
creativemediavideo
hyperframes
heygen-com
对于任何制作、创建、编辑、动画或渲染视频、动画或动态图形的请求——包括宣传片、解说视频、带字幕的片段、标题卡、叠加层或任何合成内容——请先阅读此说明。HyperFrames 通过 HTML 渲染视频;这是入口技能,也是智能体创作或编辑视频的默认方式。它会将请求路由到正确的专业工作流程,并指向 HyperFrames 领域技能,因此在尝试任何其他视频或动画技能之前,请先阅读此说明,而不是猜测工作流程。
creativevideomedia
hyperframes-core
heygen-com
HyperFrames HTML 组合合约。用于组合结构、数据属性、剪辑、轨道、子组合、变量、媒体播放、确定性渲染规则以及最小可渲染项目的验证。
developmentmediacreative
hyperframes-media
heygen-com
为HyperFrames合成提供资产预处理——多供应商TTS(HeyGen / ElevenLabs / Kokoro本地)、多供应商BGM(Google Lyria / 本地MusicGen)、Whisper转录、背景移除及字幕编写。用于npx hyperframes tts、bgm、transcribe、remove-background、语音/供应商选择、音乐情绪提示、字幕/副标题/歌词/卡拉OK/逐词样式。
mediaaudiovideo
hyperframes-read-first
heygen-com
对于任何制作、创建、生成、编辑、动画化或渲染视频、动画、动态图形、解说视频、标题卡、叠加层、带字幕视频、产品宣传片、网站视频、公关或更新日志视频、数据蒙太奇、动态海报或HyperFrames HTML合成的请求,请从此处开始。当用户希望HyperFrames创作或渲染完成的MP4/网络视频、选择工作流程,或在产品发布视频、无脸解说视频、网站转视频等之间路由时,请在其他视频或动画技能之前使用此功能。
creativevideomedia
Image Enhancer
composiohq
该技能可处理您的图片和截图,使其更清晰、更锐利、更专业。
media
image-edit
agentspace-so
在RunComfy上编辑图像——该技能是一个智能路由,能将用户意图匹配到RunComfy目录中的正确编辑模型。可选择Nano Banana Edit(最多批量处理20张,默认保留身份特征)、OpenAI GPT Image 2 Edit(多语言图像内文本重写、多参考合成、布局精准)、Flux Kontext Pro(单参考高保真局部编辑)或Z-Image Turbo Inpaint(遮罩驱动的精准区域编辑)。该技能整合了每个模型的文档化提示模式,从而...
creativeimagemedia
image-edit
doany-ai
在RunComfy上编辑图像——该技能是一个智能路由,能将用户意图匹配到RunComfy目录中的正确编辑模型。可选择Nano Banana Edit(最多批量处理20张,默认保留身份特征)、OpenAI GPT Image 2 Edit(多语言图像内文本重写、多参考合成、布局精准)、Flux Kontext Pro(单参考高保真局部编辑)或Z-Image Turbo Inpaint(遮罩驱动的精准区域编辑)。该技能整合了每个模型的文档化提示模式,从而...
creativeimagemedia
image-edit
runcomfy-com
在RunComfy上编辑图像——该技能是一个智能路由,能将用户意图匹配到RunComfy目录中的正确编辑模型。可选择Nano Banana Edit(最多批量处理20张,默认保留身份特征)、OpenAI GPT Image 2 Edit(多语言图像内文本重写、多参考合成、布局精准)、Flux Kontext Pro(单参考高保真局部编辑)或Z-Image Turbo Inpaint(遮罩驱动的精准区域编辑)。该技能整合了每个模型文档中的提示模式,从而...
creativeimagemedia
image-inpainting
agentspace-so
Mask-driven image inpainting on RunComfy via the `runcomfy` CLI. Routes to Tongyi MAI Z-Image Turbo Inpainting (the dedicated inpainting endpoint with mask, strength, and control-scale) and to identity-preserving edit models (Nano Banana 2 Edit, GPT Image 2 Edit, FLUX Kontext Pro) when a mask isn't available and the region must be described instead. Use for object removal, watermark removal, region replacement, blemish cleanup, and any controlled local edit where a binary mask defines the...
creativeimagemedia
image-inpainting
doany-ai
We need to translate the given English text into Simplified Chinese. The text describes an agent skill for image inpainting. We must preserve product names, protocol names, URLs, numbers, technical terms. The name "image-inpainting" is to be preserved if it appears in the source text? The instruction says: "Do not include the name unless it appears in the source text." The source text does not contain "image-inpainting" as a standalone term; it says "Mask-driven image inpainting" so "image inpainting" appears. But the name to preserve is "image-inpainting" with a hyphen. The source has "image inpainting" without hyphen. I think we should preserve the term as it appears in the source. The instruction says "Name to preserve: image-inpainting" but then says "Do not include the name unless it appears in the source text." So we only include it if it appears. It appears as "image inpainting" (two words). I'll translate that as "图像修复" but keep the
creativeimagemedia
image-inpainting
runcomfy-com
Mask-driven image inpainting on RunComfy via the `runcomfy` CLI. Routes to Tongyi MAI Z-Image Turbo Inpainting (the dedicated inpainting endpoint with mask, strength, and control-scale) and to identity-preserving edit models (Nano Banana 2 Edit, GPT Image 2 Edit, FLUX Kontext Pro) when a mask isn't available and the region must be described instead. Use for object removal, watermark removal, region replacement, blemish cleanup, and any controlled local edit where a binary mask defines the...
creativeimagemedia
image-outpainting
agentspace-so
Image outpainting on RunComfy via the `runcomfy` CLI — extend a still beyond its original canvas, fill in what the camera didn't capture, change aspect ratio (square → 16:9, portrait → landscape) while preserving the original content. Routes across Nano Banana 2 Edit (default, spatial-language driven), GPT Image 2 Edit (multi-ref with reference-style matching), FLUX Kontext Pro (single-shot maximum-preservation), and the brand edit endpoints (Seedream / Dreamina / Qwen / FLUX 2). Picks the...
creativeimagemedia
image-outpainting
runcomfy-com
Image outpainting on RunComfy via the `runcomfy` CLI — extend a still beyond its original canvas, fill in what the camera didn't capture, change aspect ratio (square → 16:9, portrait → landscape) while preserving the original content. Routes across Nano Banana 2 Edit (default, spatial-language driven), GPT Image 2 Edit (multi-ref with reference-style matching), FLUX Kontext Pro (single-shot maximum-preservation), and the brand edit endpoints (Seedream / Dreamina / Qwen / FLUX 2). Picks the...
creativeimagemedia
image-outpainting
doany-ai
Image outpainting on RunComfy via the `runcomfy` CLI — extend a still beyond its original canvas, fill in what the camera didn't capture, change aspect ratio (square → 16:9, portrait → landscape) while preserving the original content. Routes across Nano Banana 2 Edit (default, spatial-language driven), GPT Image 2 Edit (multi-ref with reference-style matching), FLUX Kontext Pro (single-shot maximum-preservation), and the brand edit endpoints (Seedream / Dreamina / Qwen / FLUX 2). Picks the...
creativeimagemedia
image-to-video
runcomfy-com
Animate any still image on RunComfy — this skill is a smart router that matches the user's intent to the right i2v model in the RunComfy catalog. Picks HappyHorse 1.0 I2V (Arena #1, native audio, identity preservation) for general animations, Wan 2.7 with `audio_url` for custom-voiceover lip-sync, or Seedance 2.0 Pro for multi-modal animation from image + reference video + reference audio. Bundles each model's documented prompting patterns so the caller gets sharper output without burning...
creativevideomedia
image-to-video
doany-ai
We need to translate the given text from English to Simplified Chinese. The text describes a skill called "image-to-video" but the instruction says to preserve the name only if it appears in the source text. The name "image-to-video" does appear in the source? Actually the source text starts with "Animate any still image on RunComfy — this skill is a smart router..." The name "image-to-video" is not explicitly in the source text; it's given as the directory item name. The instruction says "Do not include the name unless it appears in the source text." So we should not add "image-to-video" to the translation. We only translate the text inside <text>. Also preserve product names, protocol names, URLs, numbers, technical terms. So "RunComfy", "HappyHorse 1.0 I2V", "Arena #1", "Wan 2.7", "audio_url", "Seedance 2.0 Pro" should remain as is. Translate the rest naturally. The text ends with
creativevideomedia
image-to-video
qu-skills
静态转视频指南:模型选择、运动提示与镜头移动。涵盖Wan 2.5 i2v、Seedance、Fabric、Grok Video及各自适用场景。用途:图像动画化、从静态图像创建视频、添加运动效果、产品动画。触发词:image to video、i2v、animate image、still to video、add motion to image、image animation、photo to video、animate still、wan i2v、image2video、bring image to life、animate photo、motion from image
creativevideomedia
image-to-video
101-skills
静态转视频指南:模型选择、运动提示与镜头移动。涵盖Wan 2.5 i2v、Seedance、Fabric、Grok Video及各自适用场景。用途:图像动画化、从静态图像创建视频、添加运动效果、产品动画。触发词:image to video、i2v、animate image、still to video、add motion to image、image animation、photo to video、animate still、wan i2v、image2video、bring image to life、animate photo、motion from image
creativevideomedia
image-to-video
agentspace-so
在 RunComfy 上为任意静态图像制作动画——该技能是一个智能路由器,能将用户意图匹配到 RunComfy
creativevideomedia
kling-3-0
agentspace-so
RunComfy上的Kling 3.0视频生成。Kling 3.0(也称Kling V3.0)是快手科技第三代多镜头视频模型,具备原生同步音频和跨镜头一致的角色身份。该技能涵盖全部六个Kling 3.0端点,覆盖三种渲染等级(标准、专业、4K)和两种模式(文生视频、图生视频)。通过本地RunComfy CLI调用runcomfy run kling/kling-3.0/ /。触发词为"kling"、"kling 3.0"、"kling v3"、"kling pro"等。
creativevideomedia
kling-3-0
doany-ai
RunComfy上的Kling 3.0视频生成。Kling 3.0(也称Kling V3.0)是快手科技第三代多镜头视频模型,具备原生同步音频及跨镜头一致的角色身份。该技能覆盖全部六个Kling 3.0端点,涵盖三种渲染级别(标准、专业、4K)和两种模式(文生视频、图生视频)。通过本地RunComfy CLI调用runcomfy run kling/kling-3.0/ /。触发词为"kling"、"kling 3.0"、"kling v3"、"kling pro"等。
videocreativemedia
kling-3-0
runcomfy-com
RunComfy上的Kling 3.0视频生成。Kling 3.0(也称Kling V3.0)是快手科技第三代多镜头视频模型,具备原生同步音频和跨镜头一致的角色身份。该技能涵盖全部六个Kling 3.0端点,覆盖三种渲染等级(标准、专业、4K)和两种模式(文生视频、图生视频)。通过本地RunComfy CLI调用runcomfy run kling/kling-3.0/ /。触发词为"kling"、"kling 3.0"、"kling v3"、"kling pro"等。
videocreativemedia
lipsync
runcomfy-com
Lip-sync a face to a specific audio track on RunComfy via the `runcomfy` CLI. Routes across ByteDance OmniHuman (audio-driven full-body avatar from a portrait + audio), Sync Labs sync v2 / Pro (state-of-the-art mouth sync onto a video), Kling lipsync (audio-to- video and text-to-video with synced speech), and Creatify lipsync. The skill picks the right endpoint for the user's actual intent — portrait still + audio (avatar-style), source video + audio (mouth- swap on existing footage), or...
creativevideomedia
lipsync
doany-ai
We need to translate the given text from English to Simplified Chinese. The text describes a skill for lip-syncing a face to an audio track using various services. We must preserve product names, protocol names, URLs, numbers, and technical terms. The name "lipsync" is to be preserved if it appears in the source text, but it does appear in the first line: "Lip-sync a face..." so we should keep "lipsync" as is? Actually the instruction says "Name to preserve: lipsync" but also "Do not include the name unless it appears in the source text." The name "lipsync" appears in the source text as part of "Lip-sync" (hyphenated) and later "lipsync" (no hyphen) in "Kling lipsync" and "Creatify lipsync". So we should preserve those occurrences. Also preserve "RunComfy", "runcomfy", "ByteDance OmniHuman", "Sync Labs sync v2 / Pro", "Kling lipsync", "Creatify
creativevideomedia
media-use
heygen-com
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名称、协议名称、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容。注意不要包含名称"media-use"除非它在源文本中出现。源文本中出现了"Agent Media OS",需要保留。另外"HeyGen"是产品名,保留。"BGM, SFX, image, icon"这些术语保留英文或翻译?技术术语通常保留英文,但中文中常用"背景音乐"、"音效"等。但要求保留技术术语,所以可以保留BGM、SFX,或者翻译?指令说"preserve technical terms",但中文环境下BGM和SFX也是常见缩写。为了准确,保留BGM和SFX。另外"frozen local file"中的"frozen"可能指"冻结"或"固定",但结合上下文是生成一个本地文件并记录。翻译为"冻结的本地文件"或"固定的
mediacreativeproductivity
mediabunny
remotion-dev
使用Mediabunny库进行多媒体处理
audiomediaofficial
nano-banana-2
agentspace-so
使用RunComfy上的Google Nano Banana 2(Gemini系列闪级文生图模型)生成图像——该技能内置了模型文档中的提示模式,因此相比直接使用同一模型进行简单提示,能获得更精准的输出。文档说明了Nano Banana 2的优势(快速迭代、图像内文字渲染、可预测构图、可选网络上下文支持)、分辨率层级定价、安全容忍度调节,以及何时转向Nano Banana Pro / GPT Image 2 / Flux 2 / Seedream等模型。
creativeimagemedia
p-image
qu-skills
通过 inference.sh CLI 使用 Pruna P-Image 模型生成图像。模型:P-Image、P-Image-LoRA、P-Image-Edit、P-Image-Edit-LoRA。功能:文本到图像、图像编辑、LoRA 风格、多图像合成、快速推理。Pruna 在不损失质量的前提下优化模型速度。触发词:pruna、p-image、pruna image、fast image generation、optimized flux、pruna ai、p image、fast ai image、economic image generation、cheap image generation
creativeimagemedia
p-video
qu-skills
通过 inference.sh CLI 使用 Pruna P-Video 和 WAN 模型生成视频。模型:P-Video、WAN-T2V、WAN-I2V。功能:文本转视频、图像转视频、音频支持、720p/1080p、快速推理。Pruna 在不损失质量的前提下优化模型速度。触发词:pruna video、p-video、pruna ai video、fast video generation、optimized video、wan t2v、wan i2v、economic video generation、cheap video generation、pruna text to video、pruna image to video
videocreativemedia
p-video
101-skills
通过 inference.sh CLI 使用 Pruna P-Video 和 WAN 模型生成视频。模型:P-Video、WAN-T2V、WAN-I2V。功能:文本转视频、图像转视频、音频支持、720p/1080p、快速推理。Pruna 在不损失质量的前提下优化模型速度。触发词:pruna video、p-video、pruna ai video、fast video generation、optimized video、wan t2v、wan i2v、economic video generation、cheap video generation、pruna text to video、pruna image to video
videocreativemedia
p-video-avatar
qu-skills
通过inference.sh CLI使用Pruna P-Video-Avatar生成说话头像视频。将肖像图像转化为逼真的说话视频,内置TTS功能。速度比竞品快18倍,成本低6倍。模型:P-Video-Avatar、P-Image(用于肖像生成)。功能:文本转头像、音频驱动头像、30种语音、10种语言、720p/1080p、内置TTS、动态背景、全身控制。用途:AI主播、产品演示、解说视频、虚拟网红、营销……
videocreativemedia
p-video-avatar
101-skills
使用Pruna P-Video-Avatar通过inference.sh CLI生成说话头像视频。将肖像图像转换为逼真的说话视频,内置TTS功能。速度比竞品快18倍,成本低6倍。模型:P-Video-Avatar、P-Image(用于肖像生成)。功能:文本转头像、音频驱动头像、30种语音、10种语言、720p/1080p、内置TTS、动态背景、全身控制。用途:AI主播、产品演示、解说视频、虚拟网红、营销等。
videocreativemedia
pexo-agent
pexoai
AI视频生成技能,自动在Seedance 2、Kling 3.0、HappyHorse及10余个模型间选择。可从文本、图片、URL、脚本或音频生成完整多镜头视频(5–120秒),包含AI音乐、唇形同步和多镜头序列。无需编写提示词,无需选择模型。用途:视频制作、AI视频、制作视频、产品视频、品牌视频、宣传短片、解说视频、短视频、TikTok视频、Instagram Reel、YouTube Short、产品广告……
creativevideomedia
relight
agentspace-so
Relight a still image — change the lighting setup, color temperature, direction, or mood — on RunComfy via the `runcomfy` CLI. Routes to Qwen Edit 2509's dedicated `relight` LoRA endpoint for purpose-built relighting, with fallback to identity-preserving edit endpoints (Nano Banana 2 Edit, GPT Image 2 Edit, FLUX Kontext Pro) when prose lighting language is enough. Use for product relighting (studio softbox → window light), portrait mood shift (overcast → golden hour), or color-grade change....
creativeimagemedia
relight
doany-ai
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名、协议名、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容。注意不要包含名称"relight"除非它在源文本中出现。源文本中第一句就有"Relight",所以需要翻译。但注意指令说"不要包含名称除非它出现在源文本中",所以"Relight"作为动词应该翻译,但作为产品名或技能名?实际上"relight"在文本中作为动词出现,但也是技能名称。指令说"Name to preserve: relight",但翻译时只翻译文本,不额外添加名称。所以"Relight"作为动词应该翻译为"重新打光"或类似。但注意保留技术术语如"LoRA"、"CLI"等。另外"RunComfy"、"Qwen Edit 2509"、"Nano Banana 2 Edit"等是产品名,保留。翻译要
creativeimagemedia
relight
runcomfy-com
Relight a still image — change the lighting setup, color temperature, direction, or mood — on RunComfy via the `runcomfy` CLI. Routes to Qwen Edit 2509's dedicated `relight` LoRA endpoint for purpose-built relighting, with fallback to identity-preserving edit endpoints (Nano Banana 2 Edit, GPT Image 2 Edit, FLUX Kontext Pro) when prose lighting language is enough. Use for product relighting (studio softbox → window light), portrait mood shift (overcast → golden hour), or color-grade change....
creativeimagemedia
remotion-captions
remotion-dev
处理 Remotion 中的字幕
developmentmediaofficial
runcomfy-cli
agentspace-so
Run any model on RunComfy from the command line. The `runcomfy` CLI is one binary, one auth, hundreds of model endpoints — image generation, image edit, video generation, image-to-video, lip-sync, face swap, video edit, inpainting, outpainting, extend, ControlNet, relight, upscale, LoRA training and more. Submit a request, poll for status, download the output. This skill teaches the agent how to install, authenticate, discover model schemas, invoke models, stream / poll / no-wait, script in...
creativemediaapi
runcomfy-cli
runcomfy-com
我们需将提供的英文文本翻译成简体中文。注意:保留产品名runcomfy-cli(但文本中未出现,所以不处理)。保留runcomfy、CLI、ControlNet、LoRA等专有名词。翻译要准确,不添加额外内容。 文本内容:描述runcomfy CLI的功能,可以从命令行运行任何模型,支持多种模型端点,以及提交请求、轮询状态、下载输出等。最后一句说这个技能教agent如何安装、认证、发现模型模式、调用模型、流式/轮询/无等待、脚本等。 翻译时注意技术术语:image generation(图像生成)、image edit(图像编辑)、video generation(视频生成)、image-to-video(图像转视频)、lip-sync(唇形同步)、face swap(换脸)、video edit(视频编辑)、inpainting(修复)、outpainting(扩展)、extend(扩展?但outpainting也是扩展,注意区分)、ControlNet(保留)、relight(重打光)、upscale(放大)、
creativemediaapi
runcomfy-cli
doany-ai
Run any model on RunComfy from the command line. The `runcomfy` CLI is one binary, one auth, hundreds of model endpoints — image generation, image edit, video generation, image-to-video, lip-sync, face swap, video edit, inpainting, outpainting, extend, ControlNet, relight, upscale, LoRA training and more. Submit a request, poll for status, download the output. This skill teaches the agent how to install, authenticate, discover model schemas, invoke models, stream / poll / no-wait, script in...
creativemediaapi
seedance
qu-skills
通过 inference.sh CLI 使用字节跳动 Seedance 2.0 生成视频。统一模型支持文生视频、图生视频和参考生视频,同步音频,最高1080p,时长4-15秒。提供专业版和快速版。工作室版配备私有资产库,实现人像一致性。适用于:社交媒体视频、音乐视频、产品演示、动画内容、带声音的AI视频。触发词:seedance, seedance 2, bytedance video, seedance t2v, seedance i2v, seedance r2v, video with audio,...
creativevideomedia
seedance
101-skills
通过 inference.sh CLI 使用字节跳动 Seedance 2.0 生成视频。统一模型支持文生视频、图生视频和参考生视频,同步音频,最高1080p,时长4-15秒。提供专业版和快速版。工作室版配备私有资产库,实现人像一致性。适用于:社交媒体视频、音乐视频、产品演示、动画内容、带声音的AI视频。触发词:seedance, seedance 2, bytedance video, seedance t2v, seedance i2v, seedance r2v, video with audio,...
creativevideomedia
seedance-v2
agentspace-so
在 RunComfy 上使用字节跳动 Seedance 2.0 Pro 生成电影级短视频。介绍 Seedance 2.0 Pro 的优势(多模态参考——最多 9 张图片、3 个视频、3 个音频——同步音频内自然唇形同步、电影级运动优化)、4–15 秒时长方案,以及何时改用 HappyHorse 1.0 / Wan 2.7 / Kling
creativevideomedia
seedance-v2
doany-ai
我们被要求将一段英文文本翻译成简体中文。文本描述了一个名为"seedance-v2"的agent skill。需要保留名称"seedance-v2"(但注意指令说"不要包含名称,除非它在源文本中出现"——源文本中出现了"seedance-v2"在最后一行?实际上最后一行是"seedance...",但前面有"seedance-v2"在路径中。指令说"Preserve product names, protocol names, URLs, numbers, and technical terms." 所以"seedance-v2"应该保留。但注意指令说"Translate only the text inside <text>. Do not include the name unless it appears in the source text." 所以如果名称出现在源文本中,就保留。这里"seedance-v2"出现在"runcomfy run bytedance/seedance-v2/pro"中,所以保留。另外"Seedance 2.0 Pro"也是产品名,保留。"RunComfy"保留。"HappyHorse 1.0 / Wan 2.7 /
videocreativemedia
seedance-v2
runcomfy-com
We are asked to translate the text inside <text> into Simplified Chinese. The target language is 简体中文. We must preserve the name "seedance-v2" if it appears, but it does not appear in the source text. The source text mentions "Seedance 2.0 Pro", "RunComfy", "HappyHorse 1.0", "Wan 2.7", "Kling", and the command "runcomfy run bytedance/seedance-v2/pro". We need to keep these as is. Also preserve numbers, URLs, technical terms. Do not add any extra commentary or labels. The text is a description of an agent skill. We need to translate the entire content naturally. Let's break it down: "Generate cinematic short-form video with ByteDance Seedance 2.0 Pro on RunComfy." -> "在RunComfy上使用字节跳动Seedance 2.0 Pro生成电影级短视频。" "Documents Seedance 2.0 Pro's strengths (multi-modal references — up to 9
creativevideomedia
storyboard-creation
qu-skills
影视与视频分镜设计,涵盖镜头术语、连续性规则及分格布局。包括镜头类型、摄影角度、运镜方式、180度法则及标注格式。适用于:视频策划、电影前期制作、广告分镜、音乐视频策划、动画制作。触发词:分镜、分镜设计、镜头列表、电影策划、视频策划、前期制作、镜头构图、摄影角度、场景规划、视觉脚本、动态分镜、分镜画格、视频分镜
creativemediavideo
storyboard-creation
halt-catch-fire
影视与视频分镜设计,涵盖镜头术语、连续性规则及分镜板布局。包括镜头类型、摄影角度、运镜、180度法则及标注格式。适用于:视频策划、电影前期制作、广告分镜、音乐视频策划、动画制作。触发词:分镜、分镜设计、镜头列表、电影策划、视频策划、前期制作、镜头构图、摄影角度、场景策划、视觉脚本、动态分镜、分镜板、视频分镜
creativevideomedia
storyboard-creation
101-skills
影视与视频分镜设计,涵盖镜头术语、连续性规则及分格布局。包括镜头类型、摄影角度、运镜方式、180度法则及标注格式。适用于:视频策划、电影前期制作、广告分镜、音乐视频策划、动画制作。触发词:分镜、分镜设计、镜头列表、电影策划、视频策划、前期制作、镜头构图、摄影角度、场景策划、视觉脚本、动态分镜、分镜面板、视频分镜
creativemediavideo
text-to-lottie
diffusionstudio
编写一个可在本地Skia播放器中渲染的Lottie(Bodymovin)JSON动画。当用户要求创建、生成、编辑或修复Lottie动画,或要求加载“动画”时使用。
creativedesignmedia
video
coreyhaines31
当用户希望使用AI工具或程序化框架创建、生成或制作视频内容时。也包括用户提及“视频制作”、“AI视频”、“Remotion”、“Hyperframes”、“HeyGen”、“Synthesia”、“Veo”、“Sora”、“Runway”、“Kling”、“Seedance”、“Hailuo”、“MiniMax”、“Pika”、“Hunyuan”、“Wan”、“视频生成”、“AI数字人”、“说话头像视频”、“程序化视频”、“视频模板”、“解说视频”、“产品演示视频”、“视频流水线”或“给我做个视频”等场景。使用...
videocreativemedia
video-edit
agentspace-so
编辑RunComfy上的现有视频——此技能是一个智能路由器,将用户意图匹配到RunComfy目录中的正确编辑模型。选择Wan 2.7 Edit-Video(通用重风格化/背景替换/包装替换,保留身份+动作)、Kling 2.6 Pro Motion Control(将参考视频的精确动作迁移到目标角色)或Lucy Edit Restyle(轻量级身份稳定重风格化/服装替换)。整合每个模型记录的提示模式,使该技能...
videocreativemedia
video-edit
runcomfy-com
编辑RunComfy上的现有视频——此技能是一个智能路由器,将用户意图匹配到RunComfy目录中的正确编辑模型。选择Wan 2.7 Edit-Video(通用重风格化/背景替换/包装替换,保留身份与动作)、Kling 2.6 Pro Motion Control(将参考视频的精确动作迁移至目标角色)或Lucy Edit Restyle(轻量级身份稳定重风格化/服装替换)。整合各模型文档化的提示模式,使该技能...
videocreativemedia
video-edit
doany-ai
编辑RunComfy上的现有视频——此技能是一个智能路由,能将用户意图匹配到RunComfy目录中的正确编辑模型。选择Wan 2.7 Edit-Video(通用重风格化/背景替换/包装替换,保留身份与动作)、Kling 2.6 Pro Motion Control(将参考视频的精确动作迁移至目标角色)或Lucy Edit Restyle(轻量级身份稳定重风格化/服装替换)。整合各模型文档化的提示模式,使该技能...
videocreativemedia
video-extend
agentspace-so
我们要求翻译一段文本,目标语言是简体中文。文本内容是关于一个名为"video-extend"的agent skill的描述。需要保留产品名称、协议名称、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称本身,除非名称出现在源文本中。注意:源文本中出现了"video-extend"吗?没有直接出现,但目录项名称是"video-extend",但要求不包含名称除非在源文本中。源文本中没有"video-extend"这个词,所以不翻译它。另外注意"RunComfy"、"runcomfy"、"Google Veo 3-1"、"extend-video"、"fast/extend-video"等专有名词和技术术语要保留。翻译要准确流畅。 文本内容:描述如何通过runcomfy CLI在RunComfy上扩展或延续现有视频片段。路由到Google Veo 3-1的extend-video和fast
videocreativemedia
video-extend
doany-ai
We need to translate the given text from English to Simplified Chinese. The instruction says: preserve product names, protocol names, URLs, numbers, technical terms. Do not add claims, explanations, Markdown, bullets, links, labels, prefixes, or extra commentary. Also, do not include the name "video-extend" unless it appears in the source text. The source text does not include "video-extend" as a standalone name; it appears only in the instruction as "Name to preserve: video-extend" but not in the <text>. So we should not add it. The text inside <text> is the description. We need to translate it naturally while keeping terms like "RunComfy", "runcomfy CLI", "Google Veo 3-1", "extend-video", "fast/extend-video", "Veo clip" etc. Also numbers like "3-1" should be preserved. The translation should be accurate and concise. Let's break down the text: "Extend or continue an existing video clip on RunComfy via
videocreativemedia
video-extend
runcomfy-com
We need to translate the given text from English to Simplified Chinese. The instruction says: preserve product names, protocol names, URLs, numbers, technical terms. Do not add claims, explanations, Markdown, bullets, links, labels, prefixes, or extra commentary. Also, do not include the name "video-extend" unless it appears in the source text. The source text does not include "video-extend" as a standalone name; it appears only in the instruction as "Name to preserve: video-extend" but not in the <text>. So we should not add it. The text describes an agent skill: extending a video clip on RunComfy via CLI, using Google Veo 3-1 endpoints. We need to translate the description naturally while keeping terms like "RunComfy", "runcomfy CLI", "Google Veo 3-1", "extend-video", "fast/extend-video" as is. Also numbers like "3-1" should be preserved. Let's translate sentence by sentence: "Extend or continue an existing
videocreativemedia
video-inpainting
agentspace-so
Region edits across video frames on RunComfy via the `runcomfy` CLI — remove an object that appears across many frames, clean up wires or watermarks, replace a region with matching motion. Routes across Wan 2-7 edit-video (default, prompt-driven region edits with spatial language), Lucy Edit Restyle (identity-stable region-aware restyle), and Seedream 4-0 edit-sequential (when treating the clip as a frame stack). Picks the right route based on whether the change is prose-driven,...
videocreativemedia
video-inpainting
runcomfy-com
Region edits across video frames on RunComfy via the `runcomfy` CLI — remove an object that appears across many frames, clean up wires or watermarks, replace a region with matching motion. Routes across Wan 2-7 edit-video (default, prompt-driven region edits with spatial language), Lucy Edit Restyle (identity-stable region-aware restyle), and Seedream 4-0 edit-sequential (when treating the clip as a frame stack). Picks the right route based on whether the change is prose-driven,...
videocreativemedia
video-inpainting
doany-ai
Region edits across video frames on RunComfy via the `runcomfy` CLI — remove an object that appears across many frames, clean up wires or watermarks, replace a region with matching motion. Routes across Wan 2-7 edit-video (default, prompt-driven region edits with spatial language), Lucy Edit Restyle (identity-stable region-aware restyle), and Seedream 4-0 edit-sequential (when treating the clip as a frame stack). Picks the right route based on whether the change is prose-driven,...
videocreativemedia
video-outpainting
agentspace-so
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名称、协议名称、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称除非它在源文本中出现。不要添加"description"、"server name"或"skill name"等标签。 源文本是英文,描述了一个agent skill:video-outpainting。翻译时注意保持技术术语如"RunComfy"、"runcomfy CLI"、"Wan 2-7 edit-video"、"ComfyUI"等不变。注意"video-outpainting"是名称,但源文本中出现了"Video outpainting"作为开头,所以需要翻译"Video outpainting"为"视频外扩"或类似?但名称要保留,但这里"Video outpainting"是描述的一部分,不是作为名称标签。根据指令"Name to preserve: video-outpainting",但源文本中"Video outpainting"是首
videocreativemedia
video-outpainting
doany-ai
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名称、协议名称、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称除非它在源文本中出现。不要添加"description"、"server name"、"skill name"等标签。 源文本是英文,描述了一个agent skill:video-outpainting。翻译时注意保持技术术语如"RunComfy"、"CLI"、"Wan 2-7 edit-video"、"ComfyUI"等不翻译。注意"video-outpainting"是名称,但源文本中出现了"Video outpainting"作为开头,所以需要翻译"Video outpainting"为"视频外扩"或类似?但要求保留名称,但名称是"video-outpainting",而源文本中写的是"Video outpainting",大小写不同。根据指令"Name to preserve: video-outpainting",但源文本中出现的
videocreativemedia
video-outpainting
runcomfy-com
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名称、协议名称、URL、数字和技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称除非它在源文本中出现。不要添加"description"、"server name"、"skill name"等标签。 源文本是英文,描述了一个agent skill:video-outpainting。翻译时注意保持技术术语如"RunComfy"、"runcomfy CLI"、"Wan 2-7 edit-video"、"ComfyUI"等不变。注意"video-outpainting"是名称,但源文本中出现了"Video outpainting"作为开头,所以需要翻译"Video outpainting"为"视频外扩"或类似?但要求保留名称,但名称是"video-outpainting",而源文本中写的是"Video outpainting",大小写不同。根据指令"Name to preserve: video-outpainting",但源文本中出现的
videocreativemedia
wan-2-7
agentspace-so
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名、协议名、URL、数字、技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称除非在源文本中出现。不要添加"description"等标签。 源文本是英文,描述了一个agent skill,关于使用Wan 2.7生成文本到视频。需要翻译成中文,保留"Wan 2.7"、"RunComfy"、"HappyHorse 1.0"、"Seedance 2.0"、"Kling"、"LTX 2"、"audio_url"、"runcomfy run wan-ai/wan-2-7/text-to-video"等专有名词和技术术语。注意"wan-2-7"是名称,但源文本中出现了,所以保留。 翻译要准确,保持原意。注意"multi-reference conditioning"翻译为"多参考条件","audio-driven lip-sync"翻译
creativevideomedia
wan-2-7
runcomfy-com
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名、协议名、URL、数字、技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称除非在源文本中出现。不要添加"description"等标签。 源文本是英文,描述了一个agent skill,关于使用Wan 2.7生成文本到视频。需要翻译成中文,保留"Wan 2.7"、"RunComfy"、"HappyHorse 1.0"、"Seedance 2.0"、"Kling"、"LTX 2"、"audio_url"、"runcomfy run wan-ai/wan-2-7/text-to-video"等专有名词和技术术语。注意"Wan 2.7"可能指代模型名称,保留原样。另外"Wan-AI"也保留。注意"multi-reference conditioning"翻译为"多参考条件","audio-driven lip-sync
creativevideomedia
wan-2-7
doany-ai
我们要求翻译一段文本,目标语言是简体中文。需要保留产品名、协议名、URL、数字、技术术语。不要添加声明、解释、Markdown、项目符号、链接、标签、前缀或额外评论。只翻译<text>内的内容,不包括名称除非在源文本中出现。不要添加"description"等标签。 源文本是英文,描述了一个agent skill,关于使用Wan 2.7生成文本到视频。需要翻译成中文,保留"Wan 2.7"、"RunComfy"、"HappyHorse 1.0"、"Seedance 2.0"、"Kling"、"LTX 2"、"audio_url"、"runcomfy run wan-ai/wan-2-7/text-to-video"等专有名词和技术术语。注意"wan-2-7"是名称,但源文本中出现了,所以保留。 翻译要准确,保持原意。注意"Documents Wan 2.7's strengths..."这部分是描述文档内容,需要翻译。最后
creativevideomedia
watch
bradautomates
观看视频(URL或本地路径)。使用yt-dlp下载,用ffmpeg提取自动缩放帧,从字幕中获取转录(或Whisper API备用),并将结果交给Claude,以便它能够回答关于视频内容的问题。
videomediaresearch
wonda-cli
degausai
通过终端使用Wonda CLI生成图像、视频、音乐和音频,并进行LinkedIn、Reddit和X/Twitter的研究与自动化操作
creativemediaresearch
YouTube Transcript Downloader
michalparkola
当用户提供YouTube网址或要求从YouTube下载/获取/抓取转录文本时,下载YouTube视频转录文本。也用于用户想要转录或获取YouTube视频的字幕/隐藏式字幕时。
mediayoutubefeatured