youtube-thumbnail-design

Diseño de miniaturas de YouTube con dimensiones específicas, reglas de contraste y optimización para vista previa móvil. Cubre zonas seguras, colocación de texto, psicología de expresiones faciales y pruebas A/B. Útil para: miniaturas de YouTube, imágenes de portada de video, optimización de clics. Desencadenantes: miniatura de YouTube, diseño de miniatura, miniatura de video, tasa de clics, optimización de CTR, portada de YouTube, imagen de portada de video, creador de miniaturas, consejos para miniaturas, diseño de YouTube, imagen de vista previa de video

npx skills add https://github.com/qu-skills/skills --skill youtube-thumbnail-design

Install the belt CLI skill: npx skills add belt-sh/cli

YouTube Thumbnail Design

Create high-CTR YouTube thumbnails with AI image generation via inference.sh CLI.

Quick Start

Requires inference.sh CLI (belt). Install instructions

belt login

# Generate a thumbnail
belt app run falai/flux-dev-lora --input '{
  "prompt": "YouTube thumbnail style, close-up of a person with surprised excited expression looking at a glowing laptop screen, vibrant blue and orange color scheme, dramatic studio lighting, shallow depth of field, high contrast, cinematic",
  "width": 1280,
  "height": 720
}'

Specifications

SpecValue
Dimensions1280 x 720 px (minimum)
Recommended1920 x 1080 px
Aspect ratio16:9
Max file size2 MB
FormatsJPG, GIF, PNG

The 120px Test

Your thumbnail appears at roughly 120px wide on mobile — that's how most viewers first see it.

At 120px, viewers must be able to identify:

  1. The mood/emotion (from colors and expression)
  2. The general subject (from composition)
  3. The text (if any — only if large enough)

Test: view your thumbnail at 120px width. If it's a muddy blur, redesign.

Safe Zones

┌─────────────────────────────────────────────┐
│                                             │
│   ✅ SAFE FOR TEXT AND KEY ELEMENTS         │
│                                             │
│                                             │
│                                             │
│                                             │
│                                       ┌───┐ │
│                                       │ ⏱ │ │ ← Timestamp overlay
│                              ┌────────┴───┘ │    (bottom-right)
│   ┌────┐                     │  DURATION    │
│   │ CH │ Chapter marker      └──────────────│
└───┴────┴────────────────────────────────────┘
     ↑ Bottom-left: chapter/progress markers

Avoid placing critical elements in:

  • Bottom-right corner (video duration timestamp)
  • Bottom-left corner (chapter markers, progress bar)
  • Extreme edges (cropping varies by device)

Color Strategy

High-Contrast Pairs That Work

CombinationMoodBest For
Yellow + BlackUrgency, attentionTech, business, lists
Red + WhiteEnergy, excitementEntertainment, reactions
Blue + OrangeProfessional contrastEducation, tutorials
Green + WhiteGrowth, moneyFinance, success stories
Purple + YellowPremium, creativeDesign, art, creativity
White + DarkClean, minimalLuxury, minimalist channels

Color Rules

  • Background and text/subject should be complementary or high-contrast
  • Avoid same-temperature colors touching (red on orange = mud)
  • Use 3 colors maximum per thumbnail
  • Saturate more than real life — thumbnails compete with bright UI

Text on Thumbnails

When to Use Text

  • Lists/numbers: "7 Tips", "Top 10"
  • Strong opinions: "STOP Doing This"
  • Results: "$10K in 30 Days"
  • Comparisons: "vs" between two things

When NOT to Use Text

  • The video title already says it (redundant)
  • The emotion/visual tells the story
  • You can't make it large enough to read at 120px

Text Rules

RuleReason
Max 6 wordsReadability at thumbnail size
Min 60pt equivalentMust be legible at 120px width
Bold sans-serif fontThin fonts disappear at small sizes
Contrast stroke/shadowEnsures readability on any background
No small textIf it's not readable small, cut it

Face Expression Psychology

Thumbnails with faces get higher CTR than faceless thumbnails. Expression matters:

ExpressionCTR ImpactBest For
Surprise/shockHighestReaction, reveal, discovery content
CuriosityHighTutorial, how-to, tips
ExcitementHighUnboxing, reviews, announcements
Concern/worryMedium-highWarning, mistake, problem content
ConfidenceMediumExpert advice, authority content
NeutralLowestAvoid unless your brand is minimalist

Face Composition Rules

  • Face should fill 30-50% of the thumbnail
  • Eyes looking toward the text or subject (directs viewer attention)
  • Eyes looking at camera = connection. Eyes looking at object = curiosity.
  • Place face on one side (usually left), text or subject on the other
# Generate a face-forward thumbnail
belt app run falai/flux-dev-lora --input '{
  "prompt": "close-up portrait of a man with genuinely surprised expression, mouth slightly open, raised eyebrows, looking at camera, left side of frame, vibrant teal background, dramatic rim lighting, YouTube thumbnail style, high contrast, cinematic",
  "width": 1280,
  "height": 720
}'

# Generate a face-looking-at-subject thumbnail
belt app run bytedance/seedream-4-5 --input '{
  "prompt": "person looking amazed at a glowing holographic chart showing upward growth, dramatic blue and green lighting, right side profile view, dark background, tech aesthetic, high energy",
  "size": "2K"
}'

Thumbnail Patterns by Content Type

Tutorial / How-To

belt app run falai/flux-dev-lora --input '{
  "prompt": "overhead flat lay of organized workspace with laptop showing code editor, colorful sticky notes, coffee cup, clean bright background, professional setup, tutorial style composition, warm lighting",
  "width": 1280,
  "height": 720
}'

Before/After

belt app run falai/flux-dev-lora --input '{
  "prompt": "split composition, left side dark and messy disorganized desk, right side bright clean organized minimalist workspace, dramatic contrast between chaos and order, clear dividing line in center, high contrast",
  "width": 1280,
  "height": 720
}'

Product Review / Comparison

belt app run falai/flux-dev-lora --input '{
  "prompt": "two products facing each other with dramatic lighting and sparks between them, competition battle concept, dark background with colorful rim lighting, versus comparison style, high energy, product photography",
  "width": 1280,
  "height": 720
}'

Listicle / Number

belt app run falai/flux-dev-lora --input '{
  "prompt": "dynamic arrangement of 7 different colorful objects floating in space against dark gradient background, each item distinct and clearly separated, energetic composition, vibrant saturated colors, studio lighting",
  "width": 1280,
  "height": 720
}'

A/B Testing

Test one variable at a time:

VariableTest A vs B
Face vs No faceSame composition, with/without person
ExpressionSurprise vs curiosity
Color schemeWarm vs cool palette
Text vs No textWith/without text overlay
BackgroundBright vs dark
CompositionLeft-facing vs right-facing subject
# Generate variant A
belt app run falai/flux-dev-lora --input '{
  "prompt": "..., bright yellow background, ...",
  "width": 1280, "height": 720
}' --no-wait

# Generate variant B (same prompt, different background)
belt app run falai/flux-dev-lora --input '{
  "prompt": "..., dark navy background, ...",
  "width": 1280, "height": 720
}' --no-wait

Thumbnail Checklist

  • 1280x720 minimum (1920x1080 preferred)
  • Under 2MB file size
  • Passes the 120px squint test
  • No critical elements in bottom-right (timestamp) or bottom-left (chapter)
  • Max 3 colors, high contrast
  • Text (if any) is max 6 words, bold, with contrast stroke
  • Face expression matches content energy (if applicable)
  • Doesn't duplicate the video title
  • Stands out from surrounding thumbnails (check your niche)
  • Works on both light and dark YouTube backgrounds

Common Mistakes

MistakeProblemFix
Too much textUnreadable at thumbnail sizeMax 6 words or no text
Low contrastDisappears in the feedUse complementary colors
Cluttered compositionEye doesn't know where to lookOne focal point
Generic stock photo feelNo personality, gets skippedAuthentic expressions, unique angles
Tiny detailsLost at 120pxBold, simple shapes
Same style every videoViewer fatigueVary within brand guidelines
Misleading thumbnailKills trust, hurts retentionMatch the actual content

Related Skills

npx skills add inference-sh/skills@ai-image-generation
npx skills add inference-sh/skills@image-upscaling
npx skills add inference-sh/skills@prompt-engineering

Browse all apps: belt app store

Más skills de qu-skills

ai-video-generation
qu-skills
Genera videos de IA con Google Veo, Seedance 2.0, HappyHorse, Wan, Grok y más de 40 modelos a través de la CLI de inference.sh. Modelos: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Capacidades: texto a video, imagen a video, referencia a video, edición de video, sincronización de labios, animación de avatares, mejora de resolución de video, sonido foley. Uso para: videos para redes sociales, contenido de marketing, videos explicativos, demostraciones de productos, avatares de IA. Disparadores: generación de video, video de IA,...
videocreativemedia
remotion-render
qu-skills
Renderiza videos desde código de componentes React/Remotion a través de inference.sh. Envía código TSX, obtén MP4. Soporta todas las APIs de Remotion: useCurrentFrame, useVideoConfig, spring, interpolate, AbsoluteFill, Sequence. Resolución, FPS, duración y códec configurables. Útil para: generación programática de videos, gráficos animados, diseño en movimiento, videos basados en datos, animaciones de React a video. Disparadores: remotion, renderizar video desde código, tsx a video, react video, video programático, remotion render, código a video, animado...
developmentvideocreative
ai-image-generation
qu-skills
Genera imágenes de IA con GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve y más de 50 modelos a través de la CLI de inference.sh. Modelos: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Capacidades: texto a imagen, imagen a imagen, inpainting, LoRA, edición de imágenes, escalado, renderizado de texto. Usos: arte de IA, maquetas de productos, arte conceptual, gráficos para redes sociales, imágenes de marketing, ilustraciones. Disparadores: flux, generación de imágenes, imagen de IA, texto a...
creativemediaimage
ai-avatar-video
qu-skills
Crea videos de avatar AI y videos de cabeza parlante a través de la CLI de inference.sh. Recomendado: P-Video-Avatar (el más rápido, más barato, TTS integrado). También: OmniHuman, Fabric, PixVerse. Audio: Inworld TTS-2 (más de 100 idiomas, control de emociones para personajes), ElevenLabs, Kokoro. Capacidades: avatares impulsados por audio, texto a avatar, videos de sincronización de labios, generación de cabeza parlante, presentadores virtuales, contenido UGC. Usar para: presentadores AI, videos explicativos, influencers virtuales, doblaje, videos de marketing, anuncios UGC, avatares de juegos,...
videocreativemedia
twitter-automation
qu-skills
Automatiza Twitter/X con publicación, interacción y gestión de usuarios mediante la CLI de inference.sh. Aplicaciones: x/post-tweet, x/post-create (con medios), x/post-like, x/post-retweet, x/dm-send, x/user-follow. Capacidades: publicar tweets, programar contenido, dar like a publicaciones, retweetear, enviar DMs, seguir usuarios, obtener perfiles. Uso para: automatización de redes sociales, programación de contenido, bots de interacción, crecimiento de audiencia, API de X. Disparadores: twitter api, x api, automatización de tweets, publicar en twitter, bot de twitter, automatización de redes sociales, x...
api
agent-browser
qu-skills
Automatización del navegador para agentes de IA a través de inference.sh. Navega por páginas web, interactúa con elementos usando referencias @e, toma capturas de pantalla, graba video. Capacidades: raspado web, llenado de formularios, clics, escritura, arrastrar y soltar, carga de archivos, ejecución de JavaScript. Usos: automatización web, extracción de datos, pruebas, navegación de agentes, investigación. Disparadores: navegador, automatización web, raspar, navegar, hacer clic, llenar formulario, captura de pantalla, navegar por la web, playwright, navegador sin cabeza, agente web, navegar por internet, grabar video
browser-automationweb-scrapingtesting
web-search
qu-skills
Búsqueda web y extracción de contenido con Tavily y Exa a través de la CLI de inference.sh. Aplicaciones: Tavily Search, Tavily Extract, Exa Search, Exa Answer, Exa Extract. Capacidades: búsqueda impulsada por IA, extracción de contenido, respuestas directas, investigación. Uso para: investigación, pipelines RAG, verificación de datos, agregación de contenido, agentes. Disparadores: búsqueda web, tavily, exa, api de búsqueda, extracción de contenido, investigación, búsqueda en internet, búsqueda IA, asistente de búsqueda, web scraping, rag, alternativa a Perplexity
researchweb-scrapingapi
agent-tools
qu-skills
Ejecuta más de 250 aplicaciones de IA mediante la CLI de inference.sh: generación de imágenes, creación de videos, LLMs, búsqueda, 3D, automatización de Twitter. Modelos: FLUX, Veo, Gemini, Grok, Claude, Seedance, OmniHuman, Tavily, Exa, OpenRouter y muchos más. Úsalo al ejecutar aplicaciones de IA, generar imágenes/videos, llamar a LLMs, buscar en la web o automatizar Twitter. Disparadores: inference.sh, infsh, modelo de IA, ejecutar IA, IA sin servidor, API de IA, flux, veo, API de Claude, generación de imágenes, generación de videos, openrouter, tavily, búsqueda en Exa, API de Twitter, grok
developmentapicreative