MCP Web Scrape β AI skill for Claude Code
π mcp-web-scrape β Clean, cache-aware web content fetcher for AI agents.
How to install MCP Web Scrape
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open mukul975/mcp-web-scrape and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What MCP Web Scrape does
π mcp-web-scrape β Clean, cache-aware web content fetcher for AI agents. Fetch any URL β extract readable content β return Markdown/JSON with citations. β‘ Fast caching, π€ robots.txt compliant, π Markdown-ready output, οΏ½οΏ½ works with ChatGPT/Claude Desktop.
Alternatives in AI
- Riteway β Simple, readable, helpful unit tests 1.2k β
- Qiaomu Markdown Proxy β Fetch any URL as clean Markdown via proxy services (r.jina.ai / defuddle.md) or built-in scripts 426 β
- PromptEnhance β Rewrite the following prompt to be clearer, more specific, and more effective for an AI assistant 250 β
README
π·οΈ MCP Web Scrape
Clean, cached web content for agentsβMarkdown + citations, robots-aware, ETag/304 caching.
[](https://www.npmjs.com/package/mcp-web-scrape) [](https://opensource.org/licenses/MIT) [](https://github.com/mukul975/mcp-web-scrape)
π¦ Version
**Current Version:** `1.0.7`
π Quick Start Demo
# Extract content from any webpage
npx mcp-web-scrape@1.0.7
# Example: Extract from a news article
> extract_content https://news.ycombinator.com
β
Extracted 1,247 words with 5 citations
π Clean Markdown ready for your AI agent
π― Tool Examples
# Extract all forms from a webpage
> extract_forms https://example.com/contact
β
Found 3 forms with 12 input fields
# Parse tables into structured data
> extract_tables https://example.com/data --format json
β
Extracted 5 tables with 247 rows
# Find social media profiles
> extract_social_media https://company.com
β
Found Twitter, LinkedIn, Facebook profiles
# Analyze sentiment of content
> sentiment_analysis https://blog.example.com/article
β
Sentiment: Positive (0.85), Emotional tone: Optimistic
# Extract named entities
> extract_entities https://news.example.com/article
β
Found 12 people, 8 organizations, 5 locations
# Check for security vulnerabilities
> scan_vulnerabilities https://mysite.com
β
No XSS vulnerabilities found, 2 header improvements suggested
# Analyze competitor SEO
> analyze_competitors ["https://competitor1.com", "https://competitor2.com"]
β
Competitor analysis complete: keyword gaps identified
# Monitor uptime and performance
> monitor_uptime https://mysite.com --interval 300
β
Uptime: 99.9%, Average response: 245ms
# Generate comprehensive report
> generate_reports https://website.com --metrics ["seo", "performance", "sec
Related Skills
Tech SEO Toolkit
A Claude Code (AI agent) skill that audits & fixes the technical SEO of any existing website β JSON-LD schema,
Firecrawl Power App
Powerful web UI for Firecrawl API + MCP Server docs. Scrape, Crawl, Map, Search, Extract & AI Agent. Built wit
Web Scraping Automation With AI Agent
A Firecrawl-based AI automation to source, scrape, and verify research based on a topic before publication. Th
Web Fetch
LLM-neutral skill that fetches a URL to a temp file so agents can pipe the path through rg / jq / awk instead
Agent Ready SEO
A Claude Code skill for making a site readable and citable by AI answer engines as well as search crawlers: th
Aeo Check
Quick AEO readiness check for a URL β can AI engines crawl it, extract answers from it, and trust it enough to
Related Agents
Aeo Foundations Architect
Expert in AI Engine Optimization infrastructure β implements llms.txt, AI-aware robots.txt, token-budgeted con
Wp Audit Rankmath
Rank Math SEO installer, configurator, and SEO data seeder β modules, schema, breadcrumbs, meta, llms.txt, rob
Scrapling Data Engineer
Web scraping & data collection specialist. Writes production-grade Scrapling code (stealth fetchers, adaptive