Video To Text — AI skill for Claude Code
Turns YouTube and Twitter/X videos into readable articles (PT-BR + English) — local transcription via Whisper, translation via Claude/Gemma, static HTML with SEO, LLMO and on-demand Markdown for AI ag.
How to install Video To Text
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open adhenawer/video-to-text and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Video To Text does
Turns YouTube and Twitter/X videos into readable articles (PT-BR + English) — local transcription via Whisper, translation via Claude/Gemma, static HTML with SEO, LLMO and on-demand Markdown for AI agents via Cloudflare Worker
Alternatives in AI
- Repomix — 📦 Repomix is a powerful tool that packs your entire repository into a single, AI-friendly file 22.7k ★
- Azure OpenAI .net — GPT-4, embeddings, DALL-E, and Whisper client 1.8k ★
- Qiaomu Markdown Proxy — Fetch any URL as clean Markdown via proxy services (r.jina.ai / defuddle.md) or built-in scripts 426 ★
README
video-to-text
[](https://opensource.org/licenses/MIT) [](https://github.com/adhenawer/video-to-text/releases) [](https://adhenawer.net/) [](https://www.python.org/) [](https://github.com/ml-explore/mlx) [](https://workers.cloudflare.com/)
🇧🇷 **[Leia em português brasileiro](README-pt_br.md)** · 🇺🇸 You are reading in English
Turns YouTube and Twitter/X videos and podcasts into readable posts — organized by sections and published as static HTML.
Why
Long-form video is hard to skim, quote, search, or reread. This project turns videos into structured articles you can actually read.
- Local transcription via Whisper — no API cost for audio
- LLM-organized into thematic sections — not a chronological wall of text
- Bilingual out of the box — PT-BR and English with
hreflangalternates - Agent-friendly — Cloudflare Worker on free tier serves Markdown to AI crawlers via content negotiation (75% fewer tokens than HTML)
- Zero frontend build — static HTML deployable to GitHub Pages
Live at [adhenawer.net](https://adhenawer.net/) · [Blog](https://adhenawer.net/blog/)
How it works
The pipeline auto-detects the provider from the URL and uses the right strategy to fetch the transcript:
Video URL (YouTube, Twitter/X)
↓
src/providers/ — detect provider, capture transcript
├── youtube.py — captions via youtube-transcript-api
└──
Related Skills
Roughcut AI Local First Editor
Open-source, local-first AI video editor: on-device transcription (whisper.cpp), local LLM rough cuts (Gemma v
Yt Transcriber
CLI app- Give it a YouTube URL and you get a transcription with possible speaker identification and optional s
Openclaw Knowledge Distiller
Open CLAW Knowledge Distiller · 龍蝦知識蒸餾器 — Turn YouTube/Bilibili videos into structured knowledge articles. Loc
AI Social Media Content
Create AI-powered social media content for TikTok, Instagram, YouTube, Twitter/X. Generate: images, videos, re
Dettivo Linux
Private, offline speech-to-text for Linux: push-to-talk dictation into any app, meeting transcription with spe
ADHX
/plugin marketplace add itsmemeworks/adhx or curl -sL https://raw.githubusercontent.com/itsmemeworks/adhx/main
Related Agents
Editor Videos
Genera videos verticales para TikTok (edits de juegos/pelis y brain rots) con el pipeline local ffmpeg + Polli
YouTube Cue Sheet Builder
Builds cue-sheet.json + cue-sheet.md — the single source of truth mapping every script.md line/tag to a real,
Content To Video
Bridge agent that combines Gemini research analysis with video synthesis. Analyzes content with video-research