CSV Data Summarizer

작성자: coffeefuelbump

CSV 파일을 자동으로 분석하고 시각화를 포함한 포괄적인 인사이트를 생성하는 강력한 클로드 스킬입니다. CSV를 업로드하면 원하는 내용을 묻지 않고 즉시 지능적인 분석을 제공합니다!

npx skills add https://github.com/coffeefuelbump/csv-data-summarizer-claude-skill --skill csv-data-summarizer

CSV Data Summarizer

This Skill analyzes CSV files and provides comprehensive summaries with statistical insights and visualizations.

When to Use This Skill

Claude should use this Skill whenever the user:

  • Uploads or references a CSV file
  • Asks to summarize, analyze, or visualize tabular data
  • Requests insights from CSV data
  • Wants to understand data structure and quality

How It Works

⚠️ CRITICAL BEHAVIOR REQUIREMENT ⚠️

DO NOT ASK THE USER WHAT THEY WANT TO DO WITH THE DATA. DO NOT OFFER OPTIONS OR CHOICES. DO NOT SAY "What would you like me to help you with?" DO NOT LIST POSSIBLE ANALYSES.

IMMEDIATELY AND AUTOMATICALLY:

  1. Run the comprehensive analysis
  2. Generate ALL relevant visualizations
  3. Present complete results
  4. NO questions, NO options, NO waiting for user input

THE USER WANTS A FULL ANALYSIS RIGHT AWAY - JUST DO IT.

Automatic Analysis Steps:

The skill intelligently adapts to different data types and industries by inspecting the data first, then determining what analyses are most relevant.

  1. Load and inspect the CSV file into pandas DataFrame

  2. Identify data structure - column types, date columns, numeric columns, categories

  3. Determine relevant analyses based on what's actually in the data:

    • Sales/E-commerce data (order dates, revenue, products): Time-series trends, revenue analysis, product performance
    • Customer data (demographics, segments, regions): Distribution analysis, segmentation, geographic patterns
    • Financial data (transactions, amounts, dates): Trend analysis, statistical summaries, correlations
    • Operational data (timestamps, metrics, status): Time-series, performance metrics, distributions
    • Survey data (categorical responses, ratings): Frequency analysis, cross-tabulations, distributions
    • Generic tabular data: Adapts based on column types found
  4. Only create visualizations that make sense for the specific dataset:

    • Time-series plots ONLY if date/timestamp columns exist
    • Correlation heatmaps ONLY if multiple numeric columns exist
    • Category distributions ONLY if categorical columns exist
    • Histograms for numeric distributions when relevant
  5. Generate comprehensive output automatically including:

    • Data overview (rows, columns, types)
    • Key statistics and metrics relevant to the data type
    • Missing data analysis
    • Multiple relevant visualizations (only those that apply)
    • Actionable insights based on patterns found in THIS specific dataset
  6. Present everything in one complete analysis - no follow-up questions

Example adaptations:

  • Healthcare data with patient IDs → Focus on demographics, treatment patterns, temporal trends
  • Inventory data with stock levels → Focus on quantity distributions, reorder patterns, SKU analysis
  • Web analytics with timestamps → Focus on traffic patterns, conversion metrics, time-of-day analysis
  • Survey responses → Focus on response distributions, demographic breakdowns, sentiment patterns

Behavior Guidelines

✅ CORRECT APPROACH - SAY THIS:

  • "I'll analyze this data comprehensively right now."
  • "Here's the complete analysis with visualizations:"
  • "I've identified this as [type] data and generated relevant insights:"
  • Then IMMEDIATELY show the full analysis

✅ DO:

  • Immediately run the analysis script
  • Generate ALL relevant charts automatically
  • Provide complete insights without being asked
  • Be thorough and complete in first response
  • Act decisively without asking permission

❌ NEVER SAY THESE PHRASES:

  • "What would you like to do with this data?"
  • "What would you like me to help you with?"
  • "Here are some common options:"
  • "Let me know what you'd like help with"
  • "I can create a comprehensive analysis if you'd like!"
  • Any sentence ending with "?" asking for user direction
  • Any list of options or choices
  • Any conditional "I can do X if you want"

❌ FORBIDDEN BEHAVIORS:

  • Asking what the user wants
  • Listing options for the user to choose from
  • Waiting for user direction before analyzing
  • Providing partial analysis that requires follow-up
  • Describing what you COULD do instead of DOING it

Usage

The Skill provides a Python function summarize_csv(file_path) that:

  • Accepts a path to a CSV file
  • Returns a comprehensive text summary with statistics
  • Generates multiple visualizations automatically based on data structure

Example Prompts

"Here's sales_data.csv. Can you summarize this file?"

"Analyze this customer data CSV and show me trends."

"What insights can you find in orders.csv?"

Example Output

Dataset Overview

  • 5,000 rows × 8 columns
  • 3 numeric columns, 1 date column

Summary Statistics

  • Average order value: $58.2
  • Standard deviation: $12.4
  • Missing values: 2% (100 cells)

Insights

  • Sales show upward trend over time
  • Peak activity in Q4 (Attached: trend plot)

Files

  • analyze.py - Core analysis logic
  • requirements.txt - Python dependencies
  • resources/sample.csv - Example dataset for testing
  • resources/README.md - Additional documentation

Notes

  • Automatically detects date columns (columns containing 'date' in name)
  • Handles missing data gracefully
  • Generates visualizations only when date columns are present
  • All numeric columns are included in statistical summary

관련 스킬

azure-storage-file-datalake-py
microsoft
Azure Data Lake Storage Gen2 SDK for Python. 계층적 파일 시스템, 빅데이터 분석, 파일/디렉터리 작업에 사용합니다. 트리거: "data lake", "DataLakeServiceClient", "FileSystemClient", "ADLS Gen2", "hierarchical namespace".
database
investigate-first
juliusbrussee
모호한 실패를 편집 전에 진단하세요. 원인 불명, 간헐적 동작, 성능 회귀, 또는 증거 기반 순위 가설이 필요한 조사에 사용합니다.
runtime-cache
vercel
Vercel Runtime Cache API 안내 — 태그 기반 무효화가 가능한 임시적인 지역별 키-값 캐시입니다. Functions, Routing Middleware 및 Builds에서 공유됩니다.…
pr-name
remotion-dev
PR에 대한 올바른 명명
client-review
anthropic
클라이언트 리뷰 미팅을 준비하며 포트폴리오 성과 요약, 자산 배분 분석, 논의 포인트, 실행 항목을 정리합니다. 계정 데이터를 종합하여...
uml-and-software-architecture-visualization
openai
UML 및 UML 유사 소프트웨어 다이어그램을 설계, 비평, 읽기, 작성, 렌더링 및 구현합니다. 사용자가 UML, 시퀀스 다이어그램, 클래스 다이어그램 등을 언급할 때 사용하세요.
holohub-module-lifecycle
nvidia
재사용 가능한 Holoscan Module 작업에 ./holohub와 함께 사용: 스캐폴드, 테스트, 편집 가능한 설치, DEB/WHEEL 패키징, 클린 컨슈머 검증.
dbs-decision
dontbesilent2025
dontbesilent 개인 의사결정 시스템. 장기적으로 추적이 필요한 모든 영역(업무, 관계, 건강, 직업, 학습, 투자 등)을 로컬 지식 프로젝트로 만듭니다: 4계층 구조, 출처 태그, 수정 불가 스냅샷, 패턴을 학습하는 개념 라이브러리. 트리거: /dbs-decision, /决策系统, /决策立案, /结果回填, /状态画像
productivitydata-analysisresearch