Extraction Pipeline
Description
# Extraction Pipeline — Core Domain Logic — GenAI IDP Accelerator ## Processing Flow ``` Documents → S3 Input Bucket → EventBridge → Queue Sender Lambda → SQS → Queue Processor Lambda → Step Functions → [OCR → Classification → Extraction → Assessment → Rule Validation → Summarization] → S3 Output Bucket ``` ## Two Processing Modes Controlled by `use_bda` flag in configuration: ### Pipeline Mode (`use_bda: false`) — Default 1. **OCR**: Amazon Textract (or Bedrock OCR backend) 2. **Classificati
Installation
Installs to ~/.claude/skills/aws-solutions-library-samples-accelerated-intelligent-document-processing-on-aws-extraction-pipeline/SKILL.md
mkdir -p ~/.claude/skills/aws-solutions-library-samples-accelerated-intelligent-document-processing-on-aws-extraction-pipeline && curl -fsSL https://raw.githubusercontent.com/aws-solutions-library-samples/accelerated-intelligent-document-processing-on-aws/HEAD/.claude/skills/extraction-pipeline.md -o ~/.claude/skills/aws-solutions-library-samples-accelerated-intelligent-document-processing-on-aws-extraction-pipeline/SKILL.md Restart Claude Code, or start a new session, for it to be picked up.
Full documentation available on GitHub
View Source RepositoryRelated Skills
mcp-server-postgres
Read-only PostgreSQL database access.
Data mcp-server-sqlite
SQLite database interaction and querying.
Data mcp-server-google-maps
Google Maps integration for location data.
Data Bitbucket Data Center
---
Data Csv Data Summarizer
Automatically analyze CSV files and generate comprehensive insights with visualizations
Data Financial Services
Reference agents, skills, and data connectors for the financial-services workflows we see most — investment ba
Data