Imported from YT-smart/video-parse-transcribe (
SKILL.md). Install upstream withnpx skills add YT-smart/video-parse-transcribe. Copyright stays with the author.
Video Parse Transcribe
Core Workflow
Use the bundled script with paths relative to this skill folder. Do not hard-code the creator's local absolute path.
From the skill directory:
python scripts/video_parse_transcribe.py --url "<share text or URL>" --mode parse
python scripts/video_parse_transcribe.py --url "<share text or URL>" --mode both --model tiny
If running from another directory, resolve the script path from the current skill folder instead of using a machine-specific path.
Modes:
parse: return JSON metadata only, includingvideo_url,cover_url,title,author, and image/live-photo data when available.both: parse the share link, download the video to a temporary directory, extract audio withffmpeg, transcribe with local Whisper, and print raw transcript text.transcribe: use when the input URL is already a direct media/play URL.
Default behavior:
- Whisper model is
tinyfor speed. - Temporary video/audio files are automatically cleaned up.
- No workspace output folder is created unless
--keep-tempis passed. - Use
--jsononly when both parse metadata and transcript are needed in one payload.
Response Rules
For transcript/copy requests:
- Run
--mode both --model tiny. - Treat script output as raw ASR, not final copy.
- Correct obvious ASR errors using context before replying.
- Common examples:
等于->抖音,克服->客服,利益变->裂变,富盘->复盘. - Keep the speaker's wording and structure. Do not rewrite into a polished article unless the user asks.
- Common examples:
- Output only the corrected文案 when the user asks for文案/逐字稿. Do not include file paths, tool logs, or process notes unless requested.
For parse-only requests:
- Run
--mode parse. - Return the useful fields: title, author, video URL, cover URL, image URLs if present.
- Mention that platform media URLs are usually time-limited.
Platform Policy
Advertise and rely on stable/common platforms only:
- Douyin / 抖音
- Xiaohongshu / 小红书
- Bilibili / B站
- Kuaishou / 快手
- Direct media or play URLs, such as
.mp4,.m3u8, or already-resolved platform play URLs
Do not advertise broad support for every parser inherited from parse-video-py. Other platforms may be tried as best effort only when the user explicitly asks, but report instability clearly.
Known boundaries:
- Douyin app short links like
https://v.douyin.com/...work. - Douyin PC links like
https://www.douyin.com/jingxuan?modal_id=...work in current testing. - Already-resolved play URLs such as
https://www.douyin.com/aweme/v1/play/...should be treated as direct media URLs instead of parsed as share pages. - Image-only posts may return no
video_url; report images instead of attempting ASR.
Dependencies
The script vendors parse-video-py under scripts/vendor/parse_video_py, so it does not need the external parse_video MCP tool.
Runtime requirements:
- Python 3.10+
ffmpegon PATH for transcription- Python packages used by the vendored parser:
httpx,fake-useragent,pyyaml,parsel,lxml,jmespath,aiohttp - Python package for transcription:
whisper
If imports fail, install the missing package in the active Python environment. If network access fails in the sandbox, rerun the same command with escalation.
