Douyin Video Distiller — Documentation skill for Claude Code
抖音视频蒸馏 Skill:链接 → 无水印下载 → NVIDIA Omni 原生视频理解 → 带时间戳证据的知识卡片,直接落盘 Markdown/Wiki.
How to install Douyin Video Distiller
This entry records only its repository, not the path inside it, so there is no
exact command to give. Open 3378925604-a11y/douyin-video-distiller and copy the folder into
~/.claude/skills/, or the file into ~/.claude/agents/.
What Douyin Video Distiller does
抖音视频蒸馏 Skill:链接 → 无水印下载 → NVIDIA Omni 原生视频理解 → 带时间戳证据的知识卡片,直接落盘 Markdown/Wiki.
Alternatives in Documentation
- Notion Knowledge Capture — Convert conversations into structured, searchable Notion wiki entries 14.6k ★
- Save — Save the current conversation or a specific insight into the wiki vault as a structured note 3.3k ★
- Hyperresearch — Agent-driven research knowledge base 2.4k ★
README
douyin-video-distiller — 抖音视频蒸馏 Skill
把抖音短视频**一步到位**蒸馏成可检索的知识卡片:带时间戳的原视频证据、模型归纳、关键概念、行动项,直接落盘为 Markdown/Wiki 格式。
支持 Claude Code / OpenCode 等任何读取 `SKILL.md` 规范的 agent。
致谢与来源
本 skill 的下载环节基于 **[aehyok/douyin-video-download](https://github.com/aehyok/douyin-video-download)**(无水印视频/图集解析下载)。本仓库**不包含**该下载器的任何源码,仅作为上游依赖调用它——请先自行安装。
在此基础上的新增/改动:
- 一步到位的正确流程:链接 → 下载 → 视频理解 → 证据分级 → 知识卡片 → 临时文件清理,失败即如实报告、保留现场,不产生"假摘要"。
- 本地多模态模型直接吃完整视频:调用本地 Qwen2.5-Omni-7B,视频以带时间戳的帧序列原生送入模型(不是抽帧当图片看),还可加
--audio连音轨一起理解;不联网、无额度限制、无请求体大小上限。自带scripts/analyze_omni.py。 - 长视频的正确处理:靠采样帧率(
--fps)和--max-pixels控制耗时与显存,脚本给出 FFmpeg 切段压缩命令(150s/段、720p、CRF32),长视频也能稳定跑通。 - 证据分级输出规范:
原视频证据 / 蒸馏结论 / 待确认三层分离 + 时间戳,防止模型把单条视频观点升级成客观事实。
前置条件
| 依赖 | 必需 | 说明 |
|---|---|---|
| Qwen2.5-Omni-7B 权重 | ✅ | 本地推理;4bit 量化后权重常驻约 6GB 显存,8G 显卡实测可跑 |
| Python 3.10+ | ✅ | 需 torch / transformers>=4.57 / bitsandbytes / qwen-omni-utils / torchvision>=0.19 |
| douyin-video-download | ✅ | 链接下载环节的上游 skill |
| FFmpeg | 建议 | 长视频切段压缩用 |
设置模型路径:
export OMNI_MODEL_PATH="/path/to/Qwen2.5-Omni-7B" # Linux / macOS
$env:OMNI_MODEL_PATH="D:\Models\Qwen2.5-Omni-7B" # Windows PowerShell
安装
把本目录整个放进你的 skills 目录:
# Claude Code
~/.claude/skills/douyin-video-distiller/
# OpenCode
~/.config/opencode/skills/douyin-video-distiller/
使用
对 agent 说:
把这个抖音视频蒸馏成知识卡片:https://v.douyin.com/xxxxx/
或直接处理本地视频:
python scripts/analyze_omni.py video.mp4 "请完整分析视频并输出带时间戳的转录、画面文字、时间线、明确证据、模型归纳和待确认事项。"
# 口播类视频:连音轨一起听
python scripts/analyze_omni.py video.mp4 "把口播逐字转写" --audio
# 纯音频文件会自动走音频通道
python scripts/analyze_omni.py audio.wav "这段音频里有人说话吗?"
输出格式
---
title: ""
type: learning
source: ["原始链接"]
confidence: high | medium | low
status: draft
---
## 一句话结论 / 原视频证据 / 蒸馏结论 / 关键概念 / 可执行行动 / 待确认事项
已知限制
- 抖音反爬持续升级:上游下载器失效时本 skill 无法下载,此
Related Skills
Wiki Douyin
Generate 抖音 oral script from a knowledge source
Video Default
Provide a comprehensive scene-by-scene breakdown of this video. Organize the analysis into a Markdown table wi
Remotion Video
English 使用 Remotion 框架编程式创建视频的 Claude Code Skill。
ThinkWiki
Install once, talk to your agent. ThinkWiki turns documents and notes into a local Markdown wiki with inbox re
Blink Query
A typed wiki for LLMs — markdown on disk, resolution in the library.
Compile Main System
You are the knowledge-base compiler for a personal markdown wiki. Use Read, Glob, and Grep to inspect existing
Related Agents
Ren Distiller
Worker-class batch miner for the #60 wiki-distiller. Spawned by /ren:distill with a batch of pre-escaped L1 na
Wiki Page Writer
Claude Code-only worker. One page-plan entry in, one grounded Markdown draft out.
Nvidia Cuda Engineer
You are the system GPU Engineer — a senior NVIDIA platform specialist embedded in the system Distributed AI Ho