yonatangross

Data Pipeline Engineer — Data & AI agent for Claude Code

Data & AI community

Data pipeline specialist who generates embeddings, implements chunking strategies, manages vector indexes, and transforms raw data for AI consumption.

How to install Data Pipeline Engineer

Installs to ~/.claude/agents/yonatangross-create-yg-app-data-pipeline-engineer.md

Terminal
mkdir -p ~/.claude/agents && curl -fsSL https://raw.githubusercontent.com/yonatangross/create-yg-app/HEAD/.claude/agents/data-pipeline-engineer.md -o ~/.claude/agents/yonatangross-create-yg-app-data-pipeline-engineer.md

Restart Claude Code, or start a new session, for it to be picked up.

What Data Pipeline Engineer does


name: data-pipeline-engineer color: emerald description: Data pipeline specialist who generates embeddings, implements chunking strategies, manages vector indexes, and transforms raw data for AI consumption. Ensures data quality and optimizes batch processing for production scale model: sonnet max_tokens: 16000 tools: Bash, Read, Write, Edit, Grep, Glob

Directive

Generate embeddings, implement chunking strategies, and manage vector indexes for AI-ready data pipelines at production sc

Alternatives in Data & AI

  • Migration Agent — Creates safe, reversible database migrations with proper indexes, constraints, and zero-downtime strategies 652 ★
  • Data Collector — Pure MCP and web data fetching for portfolio skills 139 ★
  • Supabase Inference Explorer — Use proactively for exploring creative, non-obvious applications of Supabase Edge Functions, AI (embeddings an 64 ★

Full documentation available on GitHub

View Source Repository