Python Document Processor — Development agent for Claude Code
Production-grade Python document processor for PDF/DOCX/Markdown/TXT extraction with robust error handling, text cleaning, metadata extraction, and format-specific optimizations.
How to install Python Document Processor
Installs to ~/.claude/agents/dimitritholen-raggy-python-document-processor.md
mkdir -p ~/.claude/agents && curl -fsSL https://raw.githubusercontent.com/dimitritholen/raggy/HEAD/.claude/agents/python-document-processor.md -o ~/.claude/agents/dimitritholen-raggy-python-document-processor.md Restart Claude Code, or start a new session, for it to be picked up.
What Python Document Processor does
name: python-document-processor description: Production-grade Python document processor for PDF/DOCX/Markdown/TXT extraction with robust error handling, text cleaning, metadata extraction, and format-specific optimizations. Eliminates 400 lines of code duplication and implements Strategy pattern for extensible document parsing. tools: [Read, Write, Edit, Bash, Glob, Grep, WebSearch] model: claude-sonnet-4-5 color: green
IDENTITY
You are a **Production-Grade Python Document Processor*
Alternatives in Development
- JavaScript Testing Patterns — Comprehensive guide for implementing robust testing strategies in JavaScript/TypeScript applications 31.9k ★
- Liaison — You are a specialized reviewer for integrations and API implementations 3.6k ★
- Identify Programming Language — find -name ".go" -o -name ".c" -o -name ".py" -o -name ".rs" -o -name ".cs" ls /Makefile /CMakeLists.txt /go.m 216 ★
Full documentation available on GitHub
View Source RepositoryRelated Agents
Lite Drafter
Produces a short, triage-grade paper summary — Key Takeaways, Background, Main Idea & Summary, Critique. Invok
Patent Creator
Drafts complete patent applications autonomously through 6-phase workflow (estimated 55-80 min). Produces mark
PDF Docx Generator
WRITE-access implementer for the resume/export domain — PDF/DOCX rendering, layout, fonts, pagination, golden
Docx Reader
Покровский — факты из текстовых документов (DOCX/XLSX/PPTX/RTF/HTML и PDF с текстовым слоем, route=text-pdf) ч
PDF Reader
Гольмстен — читатель скан-PDF (route=scan) по OCR-тексту Apple Vision из сайдкаров ocr_dir/page_NNN.txt, сегме
Dev Document
Generation of documents (PDF, DOCX, XLSX, PPTX). Use to create a document, generate a report, export to PDF/Wo
Related Skills
Umwelten
CLI tool for evaluating and comparing AI models across Google, Ollama, OpenRouter, LM Studio, LlamaBarn, and G
Create Worktrees
by evmts - Creates git worktrees for all open PRs or specific branches, handling branches with slashes, cleani
Tukuy
Tukuy is a robust, extensible data transformation library that leverages a flexible plugin system. It simplifi