Onboarding Clean
Description
--- description: Write and run data cleaning code to produce partitioned parquet output argument-hint: <dataset_slug> <raw_data_path> [output_path] --- Spawn the `cleaner` agent with: $ARGUMENTS The agent will read architecture tables from Drive, inspect the raw files, write Python cleaning code, validate a subset, and scale to full data after user confirmation.
Installation
Installs to ~/.claude/skills/basedosdados-pipelines-onboarding-clean/SKILL.md
mkdir -p ~/.claude/skills/basedosdados-pipelines-onboarding-clean && curl -fsSL https://raw.githubusercontent.com/basedosdados/pipelines/HEAD/.claude/skills/onboarding-clean.md -o ~/.claude/skills/basedosdados-pipelines-onboarding-clean/SKILL.md Restart Claude Code, or start a new session, for it to be picked up.
Full documentation available on GitHub
View Source RepositoryRelated Skills
mcp-server-postgres
Read-only PostgreSQL database access.
Data mcp-server-sqlite
SQLite database interaction and querying.
Data mcp-server-google-maps
Google Maps integration for location data.
Data Bitbucket Data Center
---
Data Csv Data Summarizer
Automatically analyze CSV files and generate comprehensive insights with visualizations
Data Financial Services
Reference agents, skills, and data connectors for the financial-services workflows we see most — investment ba
Data