cover-image-generation

Gere imagens de capa para o programa de rádio usando Gemini 3 Flash Image (com fallback para Pro).

npx skills add https://github.com/google-gemini/gemini-managed-agents-templates --skill cover-image-generation

Image Generation Skill

This skill generates an appropriate cover image for the radio show based on a prompt, using the Gemini 3 Flash Image model, with a fallback to the Gemini 3 Pro Image model ("Nano Banana Pro") if needed.

Requirements

  • Python 3.10+
  • google-genai Python package (>= 2.0.1)

Instructions

  1. Generate an image (Recommended: use the metadata file directly):

    python3 skills/cover-image-generation/scripts/generate_image.py \
      --workspace ./workspace \
      --metadata ./workspace/data/show_notes.json
    

    Note: Using --metadata will automatically extract the show title and apply a random, high-quality prompt template. This is the preferred method.

    Alternative (Manual prompt):

    python3 skills/cover-image-generation/scripts/generate_image.py \
      --workspace ./workspace \
      --prompt "A prompt describing the image"
    
  2. Output:

    • The image will be saved to {workspace}/images/cover.png.

Model

  • Primary Model: gemini-3-flash-image-preview
  • Fallback Model: gemini-3-pro-image-preview
  • Resolution: 1:1 (default)

Prompting rules

When using the --metadata option, this skill uses a set of predefined prompt templates and selects one at random to generate cover images. It dynamically inserts the show title into the selected template.

  • Example Prompt: "A professional podcast cover image for a show titled 'AI Talk Radio' on the 'AI Talk Radio' station. The design features the text 'AI Talk Radio' in a bold, stylish white font centered on the cover. The background is a vibrant purple with a textured water ripple effect that covers the entire frame, creating a dynamic and clean aesthetic."

If you choose to use the --prompt option instead, you must construct the prompt yourself. In that case, follow these rules:

Forbidden themes

  • Do not ask for futuristic, cyberpunk or neon themes
  • Do not include any text other than the show title

Mais skills de google-gemini

agent-tui
google-gemini
Main Agents: Do NOT use this skill directly. If you need to test the TUI, invoke the `tui_tester` subagent. Drive terminal UI (TUI) applications…
gemini-api-cli
google-gemini
Guia para usar a ferramenta de linha de comando da API Gemini. Use quando precisar interagir com a API Gemini via linha de comando, gerenciar agentes ou gerar mídia (imagens,…
behavioral-evals
google-gemini
Orientação para criar, executar, corrigir e promover avaliações comportamentais. Use ao verificar a lógica de decisão do agente, depurar falhas, depurar prompt…
gemini-live-api-dev
google-gemini
We need to translate the given text from English to Brazilian Portuguese. The text describes a real-time bidirectional streaming capability with Gemini over WebSockets. It mentions audio, video, text, audio input/output specs, video frames, automatic transcriptions, voice activity detection, interruption handling, native audio features (affective dialog, proactive audio, thinking mode), function calling, Google Search grounding, session management with context compression, resumption, etc. We must preserve product names, protocol names, URLs, numbers, technical terms. The name "gemini-live-api-dev" is not in the text, so we don't include it. We only translate the text inside <text>. No extra commentary, no labels. The text ends with "and..." so we keep that as is. Translation: "Streaming bidirecional em tempo real com Gemini via WebSockets para conversas de áudio, vídeo e texto. Suporta entrada/saída de áudio (PCM 16 kHz), quadros de vídeo, texto e transcrições automáticas com detecção
gemini-omni-flash-api
google-gemini
Use esta habilidade para edição generativa de vídeo, texto para vídeo, geração de vídeo com referência de imagem e animações de transição de primeiro quadro para vídeo usando o…
gemini-api-dev
google-gemini
We need to translate the given text from English to Brazilian Portuguese. The text describes building applications with Google's Gemini models. We must preserve product names, protocol names, URLs, numbers, technical terms. The name "gemini-api-dev" is not in the text, so we don't include it. We translate only the text inside <text>. No extra commentary, no labels. The text: "Build applications with Google's Gemini models, supporting multimodal content, function calling, and structured outputs across Python, JavaScript, Go, and Java. Access current Gemini 3 models (Pro, Flash, Pro Image) with 1M token context; legacy Gemini 2.x and 1.5 models are deprecated Supports text generation, image/audio/video understanding, function calling, structured JSON output, code execution, context caching, and embeddings Official SDKs available: google-genai (Python),..." We need to translate fluently but preserve terms like "Gemini models", "multimodal content", "function calling", "structured outputs", "Python", "JavaScript", "Go",
gemini-interactions-api
google-gemini
Interface unificado para modelos e agentes Gemini com estado no servidor, streaming e orquestração de ferramentas. Suporta vários modelos atuais (gemini-3-flash-preview, gemini-3-pro-preview, gemini-2.5-flash/pro) e o agente Deep Research; substitui automaticamente IDs de modelos obsoletos por alternativas atuais. Descarrega o histórico de conversas para o servidor via previous_interaction_id para interações multi-turno com estado, sem gerenciamento manual de histórico. Orquestração de ferramentas integrada, incluindo...
deliver
google-gemini
Publica uma versão condensada do briefing em um webhook de entrada do Google Chat ou Slack, para que a execução diária se entregue sozinha — pula silenciosamente quando não há webhook…