mcpbeat Sign in

Youtube Transcribe Skill for Claude

Extract subtitles or transcripts from YouTube URLs and save normalized timestamped text locally. Use when the user asks for YouTube subtitles, captions, transcripts, video-to-text, 视频字幕, 字幕提取, YouTube 转文字, or 提取字幕.

4k tokens
context cost
the whole folder, loaded on every use
4
files
ships runnable scripts
0
copies elsewhere
how many repositories repackaged it
230
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/feiskyer/claude-code-settings --skill youtube-transcribe-skill

The instruction itself

6 sections, as written by the author

YouTube Transcript Extraction

Use the bundled scripts/transcribe_youtube.py wrapper for the primary workflow. It downloads subtitles with yt-dlp, selects the preferred language, converts WebVTT cues into timestamped plain text, and reports the output path and line count.

Requirements

  • Install yt-dlp and confirm yt-dlp --version succeeds.
  • Resolve the absolute directory containing this SKILL.md; refer to it as <skill-dir>.
  • Keep the output in the user's current working directory unless they request another location.

The Python wrapper itself uses only the standard library. It can display --help without yt-dlp being installed.

Primary workflow

Run without browser cookies first:

python3 "<skill-dir>/scripts/transcribe_youtube.py" \
  "<youtube-url>" \
  --output-dir "$PWD"

The default language preference is Simplified Chinese, Traditional Chinese, other Chinese variants, then English. Override it when the user requests another language:

python3 "<skill-dir>/scripts/transcribe_youtube.py" \
  "<youtube-url>" \
  --languages "ja.*,en.*" \
  --output-dir "$PWD"

The wrapper supports normal watch URLs plus youtu.be, Shorts, live, and embed URLs. A successful run writes one .txt file whose non-empty lines use this shape:

00:03 Subtitle text
00:08 Next subtitle text

Do not read browser cookies by default. If the initial command fails specifically because the video requires sign-in, age verification, membership, or another authenticated session:

  • Explain that the retry will let yt-dlp read YouTube cookies from a local browser profile.
  • Ask the user which browser profile they approve using.
  • Retry only after approval:
python3 "<skill-dir>/scripts/transcribe_youtube.py" \
  "<youtube-url>" \
  --cookies-from-browser chrome \
  --output-dir "$PWD"

Never print, copy, or persist browser cookies in the transcript or logs.

Browser fallback

Use browser automation only when yt-dlp is unavailable or cannot obtain subtitles and an appropriate browser-control capability is available.

  • Inspect the tools available in the current session; do not assume fixed MCP server or tool names.
  • Open the video, reveal the transcript panel, and extract visible timestamp/text pairs.
  • Save the result as <sanitized-video-title>.txt in the requested directory.
  • Close pages or sessions opened solely for this task.

If neither CLI extraction nor browser automation is available, report the missing capability instead of claiming success.

Completion report

Return:

  • Absolute transcript path
  • Selected subtitle language
  • Number of transcript lines
  • Whether browser cookies or browser automation were used

Other skills for the same job

different authors, same section of the catalogue
Canvas Design
by anthropics
vendor ×13

Create beautiful visual art in .png and .pdf documents using design philosophy. You should use this skill when the user asks to create a poster, piece of art, design, or other static piece. Create original visual designs, never copying existing artists' work to avoid copyright violations.

1388k tokens
Algorithmic Art
by anthropics
vendor ×10

Creating algorithmic art using p5.js with seeded randomness and interactive parameter exploration. Use this when users request creating art using code, generative art, algorithmic art, flow fields, or particle systems. Create original algorithmic art rather than copying existing artists' work to avoid copyright violations.

15k tokens scripts
Image Enhancer
by frostant
×6

Improves the quality of images, especially screenshots, by enhancing resolution, sharpness, and clarity. Perfect for preparing images for presentations, documentation, or social media posts.

635 tokens
Video Downloader
by CommandCodeAI
×4

Downloads videos from YouTube and other platforms for offline viewing, editing, or archival. Handles various formats and quality options.

671 tokens
Histolab
by christophacham
×3

Lightweight WSI tile extraction and preprocessing. Use for basic slide processing tissue detection, tile extraction, stain normalization for H&E images. Best for simple pipelines, dataset preparation, quick tile-based analysis. For advanced spatial proteomics, multiplexed imaging, or deep learning pipelines use pathml.

18k tokens
Omero Integration
by christophacham
×3

Microscopy data management platform. Access images via Python, retrieve datasets, analyze pixels, manage ROIs/annotations, batch processing, for high-content screening and microscopy workflows.

32k tokens
Pydicom
by christophacham
×3

Python library for working with DICOM (Digital Imaging and Communications in Medicine) files. Use this skill when reading, writing, or modifying medical imaging data in DICOM format, extracting pixel data from medical images (CT, MRI, X-ray, ultrasound), anonymizing DICOM files, working with DICOM metadata and tags, converting DICOM images to other formats, handling compressed DICOM data, or processing medical imaging datasets. Applies to tasks involving medical image analysis, PACS systems, radiology workflows, and healthcare imaging applications.

13k tokens scripts
Transformers
by christophacham
×3

This skill should be used when working with pre-trained transformer models for natural language processing, computer vision, audio, or multimodal tasks. Use for text generation, classification, question answering, translation, summarization, image classification, object detection, speech recognition, and fine-tuning models on custom datasets.

13k tokens

How to use it

Copy the folder

Take feiskyer/youtube-transcribe-skill from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.