pandas-dataframe-analyzer
SkillDocs & knowledgeAutomated DataFrame analysis skill for statistical summaries, missing value detection, data type inference, and memory optimization recommendations.
Instructions available. Your AI can read the instructions. Execution depends on the setup they require.
Account requirements not reviewed. Check the skill instructions before use; ahel provides instructions and does not run this skill.
Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.
Then ask your AI: use the pandas-dataframe-analyzer skill
What this skill tells your AI
The instructions your AI receives, as published by a5c-ai/babysitter in library/specializations/data-science-ml/skills/pandas-dataframe-analyzer/SKILL.md and read by ahel’s review.
Overview
Automated DataFrame analysis skill for statistical summaries, missing value detection, data type inference, and memory optimization recommendations using pandas and profiling libraries.
Capabilities
- Statistical profiling of DataFrames
- Missing value pattern detection
- Data type optimization suggestions
- Memory footprint analysis
- Duplicate detection and handling
- Distribution analysis and visualization
- Correlation matrix computation
- Cardinality analysis for categorical features
Target Processes
- Exploratory Data Analysis (EDA) Pipeline
- Data Collection and Validation Pipeline
- Feature Engineering Design and Implementation
Tools and Libraries
- pandas
- pandas-profiling / ydata-profiling
- numpy
- scipy (for statistical tests)
Input Schema
{
"type": "object",
"required": ["dataPath"],
"properties": {
"dataPath": {
"type": "string",
"description": "Path to the data file (CSV, Parquet, JSON)"
},
"sampleSize": {
"type": "integer",
"description": "Number of rows to sample for analysis",
"default": 10000
},
"profileType": {
"type": "string",
"enum": ["minimal", "standard", "full"],
"default": "standard"
},
"outputFormat": {
"type": "string",
"enum": ["json", "html", "markdown"],
"default": "json"
}
}
}
Output Schema
{
"type": "object",
"required": ["summary", "columns", "recommendations"],
"properties": {
"summary": {
"type": "object",
"properties": {
"rowCount": { "type": "integer" },
"columnCount": { "type": "integer" },
"memoryUsageMB": { "type": "number" },
"duplicateRows": { "type": "integer" },
"missingCells": { "type": "integer" },
"missingCellsPercent": { "type": "number" }
}
},
"columns": {
"type": "array",
"items": {
"type": "object",
"properties": {
"name": { "type": "string" },
"dtype": { "type": "string" },
"nullCount": { "type": "integer" },
"uniqueCount": { "type": "integer" },
"stats": { "type": "object" }
}
}
},
"recommendations": {
"type": "array",
"items": {
"type": "object",
"properties": {
"type": { "type": "string" },
"column": { "type": "string" },
"suggestion": { "type": "string" },
"impact": { "type": "string" }
}
}
}
}
}
Usage Example
{
kind: 'skill',
title: 'Analyze training dataset',
skill: {
name: 'pandas-dataframe-analyzer',
context: {
dataPath: 'data/train.csv',
profileType: 'full',
outputFormat: 'json'
}
}
}
Signals
- GitHub stars
- 2k
- Forks
- 112
- Last commit
- Sep 2026
Advanced
- Item type
- skill
- Key
pandas-dataframe-analyzer- Source
- github.com/a5c-ai/babysitter
Related picks
Skill · thedaviddias
The pick for JavaScriptmodern-javascript-patterns
Skill · wshobson
The pick for JavaScriptpython-performance-optimization
Skill · wshobson
The pick for Pythonpython-pro
Skill · jeffallan
The pick for Pythonrseng-notebooks
Skill · fdiblen
The pick for Notebooksexecute
Skill · brycewang-stanford
The pick for Notebooks