script-writing

โดย google-gemini

สร้างสคริปต์รายการวิทยุ AI Talk Radio จากงานวิจัยโดยใช้ Interactions API

npx skills add https://github.com/google-gemini/gemini-managed-agents-templates --skill script-writing

Script Writing

Generate a naturalistic, multi-character radio show script directly via the LLM. The script features host Paul taking calls from people around the world, with advanced audio tags.

Embedded Script

python3 skills/script-writing/scripts/generate_script.py --workspace ./workspace --style debate

Arguments

ArgumentDefaultDescription
--workspaceworkspaceRoot workspace directory
--styledefaultShow format (see styles below)
--context""Additional tone/style notes inferred from user prompts (e.g., "emphasize the technical details", "make it sound like a late night show"). Applies to ALL styles. Keep brief. Do NOT use to specify stories/topics.

Styles

StyleCallersFormatUse when user says...
debate2 per topic, opposing viewsStructured back-and-forth, Paul moderates"opposing views", "debate", "both sides"
roundtable3-4, different anglesCollaborative discussion, callers build on each other"roundtable", "panel", "discussion"
interview1-2 with direct experienceQ&A format, Paul asks probing questions"interview", "deep dive", "expert"
explainer2-3, each explains an aspectTeach the audience, Paul asks clarifying questions"explain", "break it down", "what is"
defaultSame as debateDefault when no style specified(anything else)

Examples

# HN discussion with opposing views, applying context
python3 skills/script-writing/scripts/generate_script.py \
  --workspace ./workspace --style debate --context "make it sound like a late night show"

# Panel discussion about a GitHub repo with context notes
python3 skills/script-writing/scripts/generate_script.py \
  --workspace ./workspace --style roundtable --context "emphasize the performance improvements"

# Interview with an expert about a paper
python3 skills/script-writing/scripts/generate_script.py \
  --workspace ./workspace --style interview

# Explain a new technology
python3 skills/script-writing/scripts/generate_script.py \
  --workspace ./workspace --style explainer

What it does

  1. Reads all research from {workspace}/data/research/*.md.
  2. Builds a system prompt with the base format rules + style-specific caller/structure instructions.
  3. Calls Gemini via the Interactions API (client.interactions.create()).
  4. Saves the script to {workspace}/data/script.md.

Dependencies

  • google-genai (>= 2.0.0)

Output

  • Script File: {workspace}/data/script.md
  • Format: Plain text, each line starting with Speaker: dialogue

Cast

SpeakerRolePersonality
PaulHostProfessional but relatable British moderator
CallersCall-insAmateur, rough, natural — from various cities around the world

Script Format Rules

  • Every line: SpeakerName: [audio tag] Dialogue text
  • Must include audio tags in square brackets (e.g., [sighs], [excitedly]).
  • Callers should sound amateurish (include "uh", "like", natural pauses).
  • Short punchy sentences — spoken word, not prose.
  • ~450-500 words total (~3 minutes when spoken).

Skills เพิ่มเติมจาก google-gemini

agent-tui
google-gemini
Main Agents: Do NOT use this skill directly. If you need to test the TUI, invoke the `tui_tester` subagent. Drive terminal UI (TUI) applications…
gemini-api-cli
google-gemini
คู่มือการใช้เครื่องมือ CLI ของ Gemini API ใช้เมื่อคุณต้องการโต้ตอบกับ Gemini API ผ่านทางบรรทัดคำสั่ง จัดการเอเจนต์ หรือสร้างสื่อ (รูปภาพ, …)
behavioral-evals
google-gemini
คำแนะนำสำหรับการสร้าง การรัน การแก้ไข และการส่งเสริมการประเมินพฤติกรรม ใช้เมื่อตรวจสอบตรรกะการตัดสินใจของตัวแทน การแก้ไขข้อบกพร่อง การดีบักพรอมต์…
gemini-live-api-dev
google-gemini
การสตรีมแบบสองทิศทางแบบเรียลไทม์กับ Gemini ผ่าน WebSockets สำหรับการสนทนาด้วยเสียง วีดีโอ และข้อความ รองรับการป้อน/ส่งออกเสียง (16 kHz PCM), เฟรมวีดีโอ, ข้อความ และการถอดความอัตโนมัติพร้อมการตรวจจับกิจกรรมเสียงเพื่อจัดการการขัดจังหวะ รวมถึงคุณสมบัติเสียงแบบเนทีฟ: การสนทนาที่มีอารมณ์, เสียงเชิงรุก และโหมดการคิด; การเรียกใช้ฟังก์ชันสำหรับการใช้เครื่องมือแบบซิงโครนัสและอะซิงโครนัส; และการอ้างอิง Google Search มีการจัดการเซสชันด้วยการบีบอัดบริบท, การกลับมาดำเนินการต่อ และ...
gemini-omni-flash-api
google-gemini
ใช้ทักษะนี้สำหรับการตัดต่อวิดีโอเชิงสร้างสรรค์ การสร้างวิดีโอจากข้อความ การสร้างวิดีโอโดยอ้างอิงจากภาพ และแอนิเมชันเปลี่ยนผ่านจากเฟรมแรกสู่วิดีโอ โดยใช้…
gemini-api-dev
google-gemini
สร้างแอปพลิเคชันด้วยโมเดล Gemini ของ Google รองรับเนื้อหาหลายรูปแบบ การเรียกใช้ฟังก์ชัน และผลลัพธ์ที่มีโครงสร้างใน Python, JavaScript, Go และ Java เข้าถึงโมเดล Gemini 3 ปัจจุบัน (Pro, Flash, Pro Image) พร้อมบริบท 1M โทเค็น; โมเดล Gemini 2.x และ 1.5 รุ่นเก่าถูกเลิกใช้งานแล้ว รองรับการสร้างข้อความ การทำความเข้าใจรูปภาพ/เสียง/วิดีโอ การเรียกใช้ฟังก์ชัน ผลลัพธ์ JSON ที่มีโครงสร้าง การเรียกใช้โค้ด การแคชบริบท และการฝังเวกเตอร์ SDK อย่างเป็นทางการ: google-genai (Python),...
gemini-interactions-api
google-gemini
อินเทอร์เฟซแบบรวมสำหรับโมเดล Gemini และเอเจนต์ พร้อมสถานะฝั่งเซิร์ฟเวอร์ การสตรีม และการจัดระเบียบเครื่องมือ รองรับโมเดลปัจจุบันหลายรุ่น (gemini-3-flash-preview, gemini-3-pro-preview, gemini-2.5-flash/pro) และเอเจนต์ Deep Research; แทนที่รหัสโมเดลที่เลิกใช้งานโดยอัตโนมัติด้วยทางเลือกปัจจุบัน ถ่ายโอนประวัติการสนทนาไปยังเซิร์ฟเวอร์ผ่าน previous_interaction_id สำหรับการโต้ตอบแบบหลายเทิร์นที่มีสถานะโดยไม่ต้องจัดการประวัติด้วยตนเอง มีการจัดระเบียบเครื่องมือในตัวรวมถึง...
deliver
google-gemini
โพสต์เวอร์ชันย่อของสรุปข้อมูลไปยัง Google Chat หรือ Slack incoming webhook เพื่อให้การส่งรายงานประจำวันเกิดขึ้นได้เอง — จะข้ามอย่างเงียบ ๆ เมื่อไม่มี webhook…