character-design-sheet

作者: qu-skills

透過參考圖與LoRA技術,確保AI生成角色圖像的一致性。涵蓋轉面視圖、表情圖集、色票與風格一致性技巧。適用於:角色設計、遊戲美術、插畫、動畫、漫畫、視覺小說。觸發詞:角色設計、角色圖集、角色一致性、角色參考、轉面圖、表情圖集、角色美術、一致角色、角色概念、參考圖、角色創作、原創角色設計...

npx skills add https://github.com/qu-skills/skills --skill character-design-sheet

Install the belt CLI skill: npx skills add belt-sh/cli

Character Design Sheet

Create consistent characters across multiple AI-generated images via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate a character concept
belt app run falai/flux-dev-lora --input '{
  "prompt": "character design reference sheet, front view of a young woman with short red hair, green eyes, wearing a blue jacket and white t-shirt, full body, white background, clean lines, concept art style, character turnaround",
  "width": 1024,
  "height": 1024
}'

The Consistency Problem

AI image generation produces different-looking characters every time, even with the same prompt. This is the #1 challenge in AI art for any project requiring the same character across multiple images.

Solutions (Ranked by Effectiveness)

TechniqueConsistencyEffortBest For
FLUX LoRA (trained on character)Very highHigh (requires training data)Ongoing projects, many images
Detailed description anchorMedium-highLowQuick projects, few images
Same seed + similar promptMediumLowVariations of single pose
Image-to-image refinementMediumMediumRefining existing images
Reference image in promptVariesLowWhen model supports it

Reference Sheet Types

1. Turnaround Sheet

Shows the character from multiple angles:

┌────────┬────────┬────────┬────────┐
│        │        │        │        │
│ FRONT  │  3/4   │  SIDE  │  BACK  │
│  VIEW  │  VIEW  │  VIEW  │  VIEW  │
│        │        │        │        │
└────────┴────────┴────────┴────────┘
# Generate front view
belt app run falai/flux-dev-lora --input '{
  "prompt": "character design, front view, young woman with short asymmetric red hair, bright green eyes, wearing navy blue bomber jacket over white graphic tee, dark jeans, red sneakers, standing in neutral pose, full body, clean white background, concept art, sharp details",
  "width": 768,
  "height": 1024
}' --no-wait

# Generate 3/4 view (same description)
belt app run falai/flux-dev-lora --input '{
  "prompt": "character design, three-quarter view, young woman with short asymmetric red hair, bright green eyes, wearing navy blue bomber jacket over white graphic tee, dark jeans, red sneakers, standing, full body, clean white background, concept art, sharp details",
  "width": 768,
  "height": 1024
}' --no-wait

# Generate side view
belt app run falai/flux-dev-lora --input '{
  "prompt": "character design, side profile view, young woman with short asymmetric red hair, bright green eyes, wearing navy blue bomber jacket over white graphic tee, dark jeans, red sneakers, standing, full body, clean white background, concept art, sharp details",
  "width": 768,
  "height": 1024
}' --no-wait

# Generate back view
belt app run falai/flux-dev-lora --input '{
  "prompt": "character design, back view, young woman with short asymmetric red hair, wearing navy blue bomber jacket over white graphic tee, dark jeans, red sneakers, standing, full body, clean white background, concept art, sharp details",
  "width": 768,
  "height": 1024
}' --no-wait

# Stitch into reference sheet
belt app run infsh/stitch-images --input '{
  "images": ["front.png", "three-quarter.png", "side.png", "back.png"],
  "direction": "horizontal"
}'

2. Expression Sheet

Shows the character's face with different emotions:

┌────────┬────────┬────────┐
│NEUTRAL │ HAPPY  │ ANGRY  │
│        │        │        │
├────────┼────────┼────────┤
│  SAD   │SURPRISE│THINKING│
│        │        │        │
└────────┴────────┴────────┘

Minimum 6 expressions: neutral, happy, angry, sad, surprised, thinking.

# Neutral
belt app run falai/flux-dev-lora --input '{
  "prompt": "character portrait, close-up face, young woman with short red hair and green eyes, neutral calm expression, clean white background, concept art, consistent character design",
  "width": 512,
  "height": 512
}' --no-wait

# Happy
belt app run falai/flux-dev-lora --input '{
  "prompt": "character portrait, close-up face, young woman with short red hair and green eyes, warm genuine smile, happy expression, clean white background, concept art, consistent character design",
  "width": 512,
  "height": 512
}' --no-wait

# Angry
belt app run falai/flux-dev-lora --input '{
  "prompt": "character portrait, close-up face, young woman with short red hair and green eyes, furrowed brows, angry determined expression, clean white background, concept art, consistent character design",
  "width": 512,
  "height": 512
}' --no-wait

# (Continue for sad, surprised, thinking...)

3. Outfit/Costume Sheet

Multiple outfits for the same character:

OutfitDescription
CasualBomber jacket, t-shirt, jeans
WorkBlazer, button-down, slacks
AthleticSports bra, leggings, running shoes
FormalEvening dress, heels

4. Color Palette Sheet

Document exact colors for consistency:

CHARACTER: Maya Chen

Skin:    ████ #F5D0A9 (warm beige)
Hair:    ████ #C0392B (auburn red)
Eyes:    ████ #27AE60 (emerald green)
Jacket:  ████ #2C3E50 (navy blue)
T-shirt: ████ #ECF0F1 (off-white)
Jeans:   ████ #34495E (dark slate)
Shoes:   ████ #E74C3C (bright red)

The Description Anchor Technique

The most practical consistency technique: write a 50+ word detailed description and reuse it exactly in every prompt.

Template

[age] [gender] with [hair: color, length, style], [eye color] eyes,
[skin tone], [facial features: any distinctive marks],
wearing [top: specific color and style], [bottom: specific color and style],
[shoes: specific color and style], [accessories: specific items]

Example

young woman in her mid-twenties with short asymmetric auburn red hair
swept to the right side, bright emerald green eyes, light warm skin
with a small beauty mark below her left eye, wearing a fitted navy
blue bomber jacket with silver zipper over a white crew-neck t-shirt,
dark slate slim jeans, and bright red canvas sneakers, small silver
stud earrings

Use this exact block in EVERY prompt for this character, only changing the action/pose/scene.

Proportion Guide

StyleHead-to-Body RatioBest For
Realistic7.5 : 1Film, photorealistic
Heroic8 : 1Superheroes, action
Anime/Manga5-6 : 1Japanese animation style
Stylized4-5 : 1Western animation
Chibi/Super-deformed2-3 : 1Cute, comedic, mascots

Include proportion style in your prompts: "realistic proportions" vs "anime style proportions" vs "chibi proportions"

Using LoRA for Consistency

For projects requiring many images of the same character, train a LoRA:

# Use FLUX with a character LoRA
belt app run falai/flux-dev-lora --input '{
  "prompt": "maya_chen character, sitting at a cafe reading a book, warm afternoon light, candid photography style",
  "loras": [{"path": "path/to/maya-chen-lora.safetensors", "scale": 0.8}]
}'

LoRA Training Tips:

  • Need 10-20 reference images of the character (consistent style)
  • Train on specific trigger word (e.g., "maya_chen")
  • Scale 0.7-0.9 balances consistency with prompt flexibility
  • Lower scale = more creative freedom, higher = more strict matching

Common Consistency Failures

IssueWhy It HappensMitigation
Hair color driftModel interprets "red hair" differently each timeUse specific shade: "auburn red #C0392B"
Eye color changeLow priority in generationMention eye color early in prompt
Outfit inconsistencyModel fills in details creativelyDescribe every clothing item explicitly
Age shiftVague age descriptionUse "mid-twenties" not "young"
Face structure changeDifferent generations = different facesUse LoRA or same seed base
Proportion shiftStyle interpretation variesSpecify "7.5 head proportions"

Character Bible Template

For ongoing projects, maintain a character bible document:

# Character: Maya Chen

## Visual Description (use in all prompts)
young woman in her mid-twenties with short asymmetric auburn red hair...
[full 50+ word anchor description]

## Color Palette
- Skin: #F5D0A9
- Hair: #C0392B
- Eyes: #27AE60
- Primary outfit: Navy #2C3E50
- Accent: Red #E74C3C

## Personality Notes (for expression/pose choices)
- Confident but approachable
- Default expression: slight curious smile
- Gestures: talks with hands, leans forward when interested

## Style Keywords
concept art, clean lines, sharp details, [art style reference]

## LoRA (if trained)
Path: ./loras/maya-chen-v2.safetensors
Trigger: maya_chen
Recommended scale: 0.8

Common Mistakes

MistakeProblemFix
Vague descriptionsDifferent character every time50+ word detailed anchor
Inconsistent prompt structureVarying emphasis = varying resultsSame structure, only change action/scene
Generating one view onlyCan't use character in different contextsCreate full turnaround reference
No color documentationColors drift across generationsRecord exact hex codes
Skipping expression sheetCharacter feels one-dimensionalGenerate 6+ expressions
Not using LoRA for big projectsInconsistency compoundsTrain LoRA for 10+ image projects

Related Skills

npx skills add inference-sh/skills@ai-image-generation
npx skills add inference-sh/skills@flux-image
npx skills add inference-sh/skills@prompt-engineering

Browse all apps: belt app store

來自 qu-skills 的更多技能

ai-video-generation
qu-skills
透過 inference.sh CLI 使用 Google Veo、Seedance 2.0、HappyHorse、Wan、Grok 及 40 多種模型生成 AI 影片。模型:Veo 3.1、Veo 3、Seedance 2.0、HappyHorse 1.0、Wan 2.5、Grok Imagine Video、OmniHuman、Fabric、HunyuanVideo。功能:文字轉影片、圖片轉影片、參考轉影片、影片編輯、唇形同步、虛擬人物動畫、影片放大、擬音音效。用途:社群媒體影片、行銷內容、解說影片、產品展示、AI 虛擬人物。觸發條件:影片生成、AI 影片、...
videocreativemedia
remotion-render
qu-skills
透過 inference.sh 從 React/Remotion 元件程式碼渲染影片。傳入 TSX 程式碼,取得 MP4。支援所有 Remotion API:useCurrentFrame、useVideoConfig、spring、interpolate、AbsoluteFill、Sequence。可設定解析度、FPS、時長、編碼器。用途:程式化影片生成、動畫圖形、動態設計、資料驅動影片、React 動畫轉影片。觸發詞:remotion、從程式碼渲染影片、tsx 轉影片、react 影片、程式化影片、remotion 渲染、程式碼轉影片、動畫...
developmentvideocreative
ai-image-generation
qu-skills
透過 inference.sh CLI 使用 GPT-Image-2、FLUX、Gemini、Grok、Seedream、Reve 及 50 多種模型生成 AI 圖像。模型包括:GPT-Image-2、FLUX Dev LoRA、FLUX.2 Klein LoRA、Gemini 3 Pro Image、Grok Imagine、Seedream 4.5、Reve、ImagineArt。功能涵蓋:文字轉圖像、圖像轉圖像、修補、LoRA、圖像編輯、放大、文字渲染。適用於:AI 藝術、產品模型、概念藝術、社交媒體圖形、行銷視覺、插圖。觸發詞:flux、圖像生成、AI 圖像、文字轉...
creativemediaimage
ai-avatar-video
qu-skills
透過 inference.sh CLI 建立 AI 虛擬人偶與說話頭影片。推薦:P-Video-Avatar(最快、最便宜、內建 TTS)。另可選:OmniHuman、Fabric、PixVerse。音訊:Inworld TTS-2(100 多種語言、角色情感引導)、ElevenLabs、Kokoro。功能:音訊驅動虛擬人偶、文字轉虛擬人偶、唇形同步影片、說話頭生成、虛擬主持人、UGC 內容。用途:AI 主持人、解說影片、虛擬網紅、配音、行銷影片、UGC 廣告、遊戲虛擬人偶……
videocreativemedia
twitter-automation
qu-skills
透過 inference.sh CLI 自動化 Twitter/X 的發文、互動與用戶管理。應用程式:x/post-tweet、x/post-create(含媒體)、x/post-like、x/post-retweet、x/dm-send、x/user-follow。功能:發推文、排程內容、按讚、轉推、發送私訊、追蹤用戶、取得個人資料。用途:社群媒體自動化、內容排程、互動機器人、受眾成長、X API。觸發條件:twitter api、x api、推文自動化、發文至 twitter、twitter 機器人、社群媒體自動化、x...
api
agent-browser
qu-skills
透過 inference.sh 為 AI 代理提供瀏覽器自動化功能。可導覽網頁、使用 @e 參考與元素互動、擷取螢幕截圖、錄製影片。功能包括:網頁抓取、表單填寫、點擊、打字、拖放、檔案上傳、執行 JavaScript。適用於:網頁自動化、資料擷取、測試、代理瀏覽、研究。觸發詞:瀏覽器、網頁自動化、抓取、導覽、點擊、填寫表單、螢幕截圖、瀏覽網頁、playwright、無頭瀏覽器、網頁代理、上網、錄製影片
browser-automationweb-scrapingtesting
web-search
qu-skills
透過 inference.sh CLI 使用 Tavily 和 Exa 進行網路搜尋與內容擷取。應用:Tavily 搜尋、Tavily 擷取、Exa 搜尋、Exa 回答、Exa 擷取。功能:AI 驅動搜尋、內容擷取、直接回答、研究。用途:研究、RAG 管線、事實查核、內容彙整、代理程式。觸發詞:網路搜尋、tavily、exa、搜尋 API、內容擷取、研究、網際網路搜尋、AI 搜尋、搜尋助理、網頁抓取、rag、perplexity 替代方案
researchweb-scrapingapi
agent-tools
qu-skills
透過 inference.sh CLI 執行 250 多個 AI 應用程式 - 圖片生成、影片創作、大型語言模型、搜尋、3D、Twitter 自動化。模型:FLUX、Veo、Gemini、Grok、Claude、Seedance、OmniHuman、Tavily、Exa、OpenRouter 等。用於執行 AI 應用程式、生成圖片/影片、呼叫大型語言模型、網路搜尋或自動化 Twitter 時觸發。觸發詞:inference.sh、infsh、ai model、run ai、serverless ai、ai api、flux、veo、claude api、image generation、video generation、openrouter、tavily、exa search、twitter api、grok
developmentapicreative