youtube-thumbnail-design

작성자: qu-skills

YouTube 썸네일 디자인: 특정 치수, 대비 규칙, 모바일 미리보기 최적화 포함. 안전 영역, 텍스트 배치, 표정 심리학, A/B 테스트를 다룹니다. 용도: YouTube 썸네일, 동영상 커버 이미지, 클릭률 최적화. 트리거: youtube thumbnail, thumbnail design, video thumbnail, click through rate, ctr optimization, youtube cover, video cover image, thumbnail maker, thumbnail tips, youtube design, video preview image

npx skills add https://github.com/qu-skills/skills --skill youtube-thumbnail-design

Install the belt CLI skill: npx skills add belt-sh/cli

YouTube Thumbnail Design

Create high-CTR YouTube thumbnails with AI image generation via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate a thumbnail
belt app run falai/flux-dev-lora --input '{
  "prompt": "YouTube thumbnail style, close-up of a person with surprised excited expression looking at a glowing laptop screen, vibrant blue and orange color scheme, dramatic studio lighting, shallow depth of field, high contrast, cinematic",
  "width": 1280,
  "height": 720
}'

Specifications

SpecValue
Dimensions1280 x 720 px (minimum)
Recommended1920 x 1080 px
Aspect ratio16:9
Max file size2 MB
FormatsJPG, GIF, PNG

The 120px Test

Your thumbnail appears at roughly 120px wide on mobile — that's how most viewers first see it.

At 120px, viewers must be able to identify:

  1. The mood/emotion (from colors and expression)
  2. The general subject (from composition)
  3. The text (if any — only if large enough)

Test: view your thumbnail at 120px width. If it's a muddy blur, redesign.

Safe Zones

┌─────────────────────────────────────────────┐
│                                             │
│   ✅ SAFE FOR TEXT AND KEY ELEMENTS         │
│                                             │
│                                             │
│                                             │
│                                             │
│                                       ┌───┐ │
│                                       │ ⏱ │ │ ← Timestamp overlay
│                              ┌────────┴───┘ │    (bottom-right)
│   ┌────┐                     │  DURATION    │
│   │ CH │ Chapter marker      └──────────────│
└───┴────┴────────────────────────────────────┘
     ↑ Bottom-left: chapter/progress markers

Avoid placing critical elements in:

  • Bottom-right corner (video duration timestamp)
  • Bottom-left corner (chapter markers, progress bar)
  • Extreme edges (cropping varies by device)

Color Strategy

High-Contrast Pairs That Work

CombinationMoodBest For
Yellow + BlackUrgency, attentionTech, business, lists
Red + WhiteEnergy, excitementEntertainment, reactions
Blue + OrangeProfessional contrastEducation, tutorials
Green + WhiteGrowth, moneyFinance, success stories
Purple + YellowPremium, creativeDesign, art, creativity
White + DarkClean, minimalLuxury, minimalist channels

Color Rules

  • Background and text/subject should be complementary or high-contrast
  • Avoid same-temperature colors touching (red on orange = mud)
  • Use 3 colors maximum per thumbnail
  • Saturate more than real life — thumbnails compete with bright UI

Text on Thumbnails

When to Use Text

  • Lists/numbers: "7 Tips", "Top 10"
  • Strong opinions: "STOP Doing This"
  • Results: "$10K in 30 Days"
  • Comparisons: "vs" between two things

When NOT to Use Text

  • The video title already says it (redundant)
  • The emotion/visual tells the story
  • You can't make it large enough to read at 120px

Text Rules

RuleReason
Max 6 wordsReadability at thumbnail size
Min 60pt equivalentMust be legible at 120px width
Bold sans-serif fontThin fonts disappear at small sizes
Contrast stroke/shadowEnsures readability on any background
No small textIf it's not readable small, cut it

Face Expression Psychology

Thumbnails with faces get higher CTR than faceless thumbnails. Expression matters:

ExpressionCTR ImpactBest For
Surprise/shockHighestReaction, reveal, discovery content
CuriosityHighTutorial, how-to, tips
ExcitementHighUnboxing, reviews, announcements
Concern/worryMedium-highWarning, mistake, problem content
ConfidenceMediumExpert advice, authority content
NeutralLowestAvoid unless your brand is minimalist

Face Composition Rules

  • Face should fill 30-50% of the thumbnail
  • Eyes looking toward the text or subject (directs viewer attention)
  • Eyes looking at camera = connection. Eyes looking at object = curiosity.
  • Place face on one side (usually left), text or subject on the other
# Generate a face-forward thumbnail
belt app run falai/flux-dev-lora --input '{
  "prompt": "close-up portrait of a man with genuinely surprised expression, mouth slightly open, raised eyebrows, looking at camera, left side of frame, vibrant teal background, dramatic rim lighting, YouTube thumbnail style, high contrast, cinematic",
  "width": 1280,
  "height": 720
}'

# Generate a face-looking-at-subject thumbnail
belt app run bytedance/seedream-4-5 --input '{
  "prompt": "person looking amazed at a glowing holographic chart showing upward growth, dramatic blue and green lighting, right side profile view, dark background, tech aesthetic, high energy",
  "size": "2K"
}'

Thumbnail Patterns by Content Type

Tutorial / How-To

belt app run falai/flux-dev-lora --input '{
  "prompt": "overhead flat lay of organized workspace with laptop showing code editor, colorful sticky notes, coffee cup, clean bright background, professional setup, tutorial style composition, warm lighting",
  "width": 1280,
  "height": 720
}'

Before/After

belt app run falai/flux-dev-lora --input '{
  "prompt": "split composition, left side dark and messy disorganized desk, right side bright clean organized minimalist workspace, dramatic contrast between chaos and order, clear dividing line in center, high contrast",
  "width": 1280,
  "height": 720
}'

Product Review / Comparison

belt app run falai/flux-dev-lora --input '{
  "prompt": "two products facing each other with dramatic lighting and sparks between them, competition battle concept, dark background with colorful rim lighting, versus comparison style, high energy, product photography",
  "width": 1280,
  "height": 720
}'

Listicle / Number

belt app run falai/flux-dev-lora --input '{
  "prompt": "dynamic arrangement of 7 different colorful objects floating in space against dark gradient background, each item distinct and clearly separated, energetic composition, vibrant saturated colors, studio lighting",
  "width": 1280,
  "height": 720
}'

A/B Testing

Test one variable at a time:

VariableTest A vs B
Face vs No faceSame composition, with/without person
ExpressionSurprise vs curiosity
Color schemeWarm vs cool palette
Text vs No textWith/without text overlay
BackgroundBright vs dark
CompositionLeft-facing vs right-facing subject
# Generate variant A
belt app run falai/flux-dev-lora --input '{
  "prompt": "..., bright yellow background, ...",
  "width": 1280, "height": 720
}' --no-wait

# Generate variant B (same prompt, different background)
belt app run falai/flux-dev-lora --input '{
  "prompt": "..., dark navy background, ...",
  "width": 1280, "height": 720
}' --no-wait

Thumbnail Checklist

  • 1280x720 minimum (1920x1080 preferred)
  • Under 2MB file size
  • Passes the 120px squint test
  • No critical elements in bottom-right (timestamp) or bottom-left (chapter)
  • Max 3 colors, high contrast
  • Text (if any) is max 6 words, bold, with contrast stroke
  • Face expression matches content energy (if applicable)
  • Doesn't duplicate the video title
  • Stands out from surrounding thumbnails (check your niche)
  • Works on both light and dark YouTube backgrounds

Common Mistakes

MistakeProblemFix
Too much textUnreadable at thumbnail sizeMax 6 words or no text
Low contrastDisappears in the feedUse complementary colors
Cluttered compositionEye doesn't know where to lookOne focal point
Generic stock photo feelNo personality, gets skippedAuthentic expressions, unique angles
Tiny detailsLost at 120pxBold, simple shapes
Same style every videoViewer fatigueVary within brand guidelines
Misleading thumbnailKills trust, hurts retentionMatch the actual content

Related Skills

npx skills add inference-sh/skills@ai-image-generation
npx skills add inference-sh/skills@image-upscaling
npx skills add inference-sh/skills@prompt-engineering

Browse all apps: belt app store

qu-skills의 다른 스킬

ai-video-generation
qu-skills
inference.sh CLI를 통해 Google Veo, Seedance 2.0, HappyHorse, Wan, Grok 및 40개 이상의 모델로 AI 비디오를 생성합니다. 모델: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. 기능: 텍스트-비디오, 이미지-비디오, 참조-비디오, 비디오 편집, 립싱크, 아바타 애니메이션, 비디오 업스케일링, 폴리 사운드. 사용처: 소셜 미디어 비디오, 마케팅 콘텐츠, 설명 비디오, 제품 데모, AI 아바타. 트리거: 비디오 생성, AI 비디오,...
videocreativemedia
remotion-render
qu-skills
React/Remotion 컴포넌트 코드를 inference.sh를 통해 비디오로 렌더링합니다. TSX 코드를 전달하면 MP4를 받을 수 있습니다. useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence 등 모든 Remotion API를 지원합니다. 해상도, FPS, 지속 시간, 코덱을 설정할 수 있습니다. 사용 사례: 프로그래매틱 비디오 생성, 애니메이션 그래픽, 모션 디자인, 데이터 기반 비디오, React 애니메이션을 비디오로 변환. 트리거: remotion, 코드에서 비디오 렌더링, tsx를 비디오로, react 비디오, 프로그래매틱 비디오, remotion 렌더, 코드를 비디오로, 애니메이션...
developmentvideocreative
ai-image-generation
qu-skills
GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve 및 inference.sh CLI를 통해 50개 이상의 모델로 AI 이미지를 생성합니다. 모델: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. 기능: 텍스트-이미지, 이미지-이미지, 인페인팅, LoRA, 이미지 편집, 업스케일링, 텍스트 렌더링. 용도: AI 아트, 제품 목업, 컨셉 아트, 소셜 미디어 그래픽, 마케팅 비주얼, 일러스트레이션. 트리거: flux, 이미지 생성, ai 이미지, 텍스트 투...
creativemediaimage
ai-avatar-video
qu-skills
inference.sh CLI를 통해 AI 아바타 및 토킹 헤드 영상을 생성합니다. 권장: P-Video-Avatar (가장 빠르고 저렴하며 TTS 내장). 추가: OmniHuman, Fabric, PixVerse. 오디오: Inworld TTS-2 (100개 이상 언어, 캐릭터 감정 조절), ElevenLabs, Kokoro. 기능: 오디오 기반 아바타, 텍스트-투-아바타, 립싱크 영상, 토킹 헤드 생성, 가상 프레젠터, UGC 콘텐츠. 용도: AI 프레젠터, 설명 영상, 가상 인플루언서, 더빙, 마케팅 영상, UGC 광고, 게이밍 아바타,...
videocreativemedia
twitter-automation
qu-skills
inference.sh CLI를 통해 게시, 참여, 사용자 관리로 Twitter/X를 자동화합니다. 앱: x/post-tweet, x/post-create (미디어 포함), x/post-like, x/post-retweet, x/dm-send, x/user-follow. 기능: 트윗 게시, 콘텐츠 예약, 게시물 좋아요, 리트윗, DM 전송, 사용자 팔로우, 프로필 조회. 사용처: 소셜 미디어 자동화, 콘텐츠 예약, 참여 봇, 오디언스 성장, X API. 트리거: twitter api, x api, tweet automation, post to twitter, twitter bot, social media automation, x...
api
agent-browser
qu-skills
inference.sh를 통한 AI 에이전트용 브라우저 자동화. @e 참조를 사용하여 웹 페이지 탐색, 요소와 상호작용, 스크린샷 촬영, 비디오 녹화. 기능: 웹 스크래핑, 양식 작성, 클릭, 타이핑, 드래그 앤 드롭, 파일 업로드, JavaScript 실행. 용도: 웹 자동화, 데이터 추출, 테스트, 에이전트 브라우징, 연구. 트리거: 브라우저, 웹 자동화, 스크래핑, 탐색, 클릭, 양식 작성, 스크린샷, 웹 브라우징, Playwright, 헤드리스 브라우저, 웹 에이전트, 인터넷 서핑, 비디오 녹화
browser-automationweb-scrapingtesting
web-search
qu-skills
Tavily와 Exa를 통해 inference.sh CLI로 웹 검색 및 콘텐츠 추출. 앱: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract. 기능: AI 기반 검색, 콘텐츠 추출, 직접 답변, 리서치. 용도: 리서치, RAG 파이프라인, 사실 확인, 콘텐츠 수집, 에이전트. 트리거: 웹 검색, tavily, exa, search api, 콘텐츠 추출, 리서치, 인터넷 검색, ai 검색, 검색 어시스턴트, 웹 스크래핑, rag, perplexity 대안
researchweb-scrapingapi
agent-tools
qu-skills
inference.sh CLI를 통해 250개 이상의 AI 앱 실행 - 이미지 생성, 비디오 제작, LLM, 검색, 3D, Twitter 자동화. 모델: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter 등. AI 앱 실행, 이미지/비디오 생성, LLM 호출, 웹 검색 또는 Twitter 자동화 시 사용. 트리거: inference.sh, infsh, ai model, run ai, serverless ai, ai api, flux, veo, claude api, image generation, video generation, openrouter, tavily, exa search, twitter api, grok
developmentapicreative