mcpbeat

Media Skills

3 019 media skills from 496 authors. They create and process images, video and sound. Half of them fit into 1 966 tokens or less — that is what one costs your context window when the agent loads it. 768 ship runnable scripts rather than instructions alone. 30 of them cannot work without an MCP server, most often rube. We also found 321 copies of these same skills sitting in other people's repositories — counted once here, not 321 times.

3 019 unique 496 authors 1 925 updated this month 160 from vendors

1 966
tokens, median
what a typical one costs in context
768
ship scripts
code that runs, not instructions alone
30
need a server
most often rube
321
copies elsewhere
counted once here, not once per repository

2 785–2 832 of 3 019

page 59 of 63
Downloader
next-open-ai

Downloads files, images, or media from URLs using curl, wget, or python scripts. Can handle bulk downloads and save them to a specified directory.

255 tokens
Create Assistant
VapiAI

Create Vapi voice AI assistant payloads or assistants through the Vapi API. Use when building phone or web call agents, generating assistant JSON, choosing safe default model/voice/transcriber settings, attaching existing Vapi tool IDs, adding assistant hooks, configuring HIPAA/compliance only when explicitly requested, or fixing Vapi assistant API validation errors.

3k tokens
Setup API Key
VapiAI

Guide users through obtaining and configuring a Vapi API key. Use when the user needs to set up Vapi, when API calls fail due to missing keys, or when the user mentions needing access to Vapi's voice AI platform.

812 tokens
Create Squad
VapiAI

Create multi-assistant squads in Vapi with handoffs between specialized voice agents. Use when building complex voice workflows that need multiple assistants with different roles, like triage-to-booking or sales-to-support handoffs.

3k tokens
Create Call
VapiAI

Create outbound phone calls, web calls, and batch calls using the Vapi API. Use when making automated calls, testing voice assistants, scheduling call campaigns, or initiating conversations programmatically.

2k tokens
Create Phone Number
VapiAI

Set up and manage phone numbers in Vapi for inbound and outbound voice AI calls. Use when importing Twilio, Vonage, or Telnyx numbers, buying Vapi numbers, or configuring phone numbers for assistants.

1k tokens
Dataverse Web API
DanielKerridge

> Use when programmatically creating, modifying, or querying Dataverse schema and metadata via the Web API (OData v4.0). Covers table/column/relationship definitions, solution ALM, form and view XML construction, app module composition, global option sets, business rules, "entitydefinitions", "web api schema", "create dataverse table", "create dataverse column", "fetchxml", "formxml", "layoutxml", "dataverse solution", "dataverse relationship", "odata dataverse", "metadata api", "publish customizations", "dataverse alm", "grid control", "editable grid", "business rule", "rich text", "auto-number", "file column", "image column", "pcf control", "security role", "column security", "environment variable", "custom api", "data migration", "solution import".

48k tokens
Record Screen
DanielKerridge

Record a Chrome browser tab to video via CLI. Use when the user wants to capture a screen recording of a browser tab. Supports list, start, stop, caption, and status subcommands. No debug mode required — uses a Chrome extension.

21043k tokens scripts
Download Anything
hAcKlyc

> Find and download virtually any digital resource from the internet — ebooks, academic papers, movies, TV shows, music, software, images, fonts, courses, and more. Covers both English and Chinese internet ecosystems. Includes CLI tool workflows (yt-dlp, aria2, gallery-dl, spotdl), resource site directories, cloud drive search engines (百度/阿里/夸克网盘搜索), and search from a URL, (2) find and download an ebook or academic paper, (3) find and download software, (4) search for any digital resource, (5) batch download images or media from a gallery/site, (6) download torrents or magnet links, (7) find free stock assets (images, video, audio, fonts), (8) search Chinese cloud drives for resources, or (9) any task involving finding or downloading digital content from the internet.

19k tokens scripts
Video Generator
panaversity

AI video production workflow using Remotion. Use when creating videos, short films, commercials, or motion graphics. Triggers on requests to make promotional videos, product demos, social media videos, animated explainers, or any programmatic video content. Produces polished motion graphics, not slideshows.

7k tokens scripts
Nano Banana2
BENZEMA216

> Nano Banana2 多模态 AI - 支持文生图、图生图、Vision 分析。发送"生图 柴犬"生成图片,发送图片+描述做风格化重绘。底层调用 Gemini,支持 10 种宽高比。关键词:生图、画图、AI画图、文生图、图生图、generate image、Gemini、重绘

3k tokens scripts zh
Upload Media
BENZEMA216

上传视频/音频/图片到即梦(Jimeng) VOD 平台并获取 Vid。支持 mp4/mp3/jpg/png 等格式,自动处理 STS 鉴权和分块上传。当需要上传媒体文件、获取视频ID、发布素材到字节跳动平台时使用。关键词:上传、即梦、VOD、视频ID、媒体文件

4k tokens scripts zh
Xhs Note Creator
BENZEMA216

小红书笔记素材创作技能。当用户需要创建小红书笔记素材时使用这个技能。技能包含:根据用户的需求和提供的资料,撰写小红书笔记内容(标题+正文),生成图片卡片(封面+正文卡片),以及发布小红书笔记。

45k tokens scripts zh
Technical Analyst
ajeeshworkspace

This skill should be used when analyzing weekly price charts for Indian stocks (NSE/BSE), indices (Nifty 50, Bank Nifty, Sensex), or any other instrument. Use this skill when the user provides chart images and requests technical analysis, trend identification, support/resistance levels, scenario planning, or probability assessments based purely on chart data without consideration of news or fundamental factors.

7k tokens
Hinge Profile Optimizer
b1rdmania

Comprehensive, research-backed Hinge dating profile optimization. Use when someone wants to improve their Hinge profile, audit an existing profile, write better prompts/captions, select and order photos strategically, or understand why they're not getting quality matches. This is the thorough process (~45 mins) - discovery interview, honest market math, photo strategy, copy creation, settings cleanup, and implementation support. Grounded in peer-reviewed behavioral research, platform data, and signaling theory.

32k tokens
Scoop
EricTechPro

Competitor content gap analysis — extracts ALL comments from competitor and your own YouTube videos, identifies what audiences want but aren't getting, ranks gaps by engagement, and proposes video ideas. Use when the user says "scoop," "content gap," "analyze competitor videos," "what should my next video be about," "find content gaps," or "competitor analysis.

7k tokens
Concert Playlist Builder
aparente

| Extract artists from concert/festival lineups and create playlists to preview their music.

2k tokens
Tagore
apurvrdx1

| Write or rewrite prose so it sounds like a human wrote it — not a frontier model. Named in homage to Rabindranath Tagore, whose prose carried what a 29-pattern catalog of AI tells (from humanizer) plus an 8-rule operating system with an 8-dimension scoring gate (extending stop-slop). Use when emails. Detects and removes inflated symbolism, promotional language, superficial -ing analyses, vague attributions, em dash overuse, rule of three, AI vocabulary, passive voice, negative parallelisms, filler phrases, inanimate-verb constructions, narrator-from-a-distance voice, and metronomic specificity, restraint, varied rhythm, and trust in the reader.

375k tokens scripts
P12a Contemplation Right Speech
gmaxxxie

把话说真,而不是把问题说顺——语言清洗、纪要体检、诚实表达练习

4k tokens zh
Animation Shader
Yuki001

READ this skill when implementing or configuring animation-style shaders (Toon/Cel Shaders) — including outlines, rim lighting, toon shading, MatCap, emission, dissolve, hatching, or any stylized rendering effect. Contains preset styles and feature-to-reference mappings for lilToon, Poiyomi, UTS2, RToon, SToon, and ToonShadingCollection. Works as a domain knowledge plugin alongside workflow skills (OpenSpec, SpecKit) or plan mode of an agent.

46k tokens
Imagemagick CLI
Yuki001

Comprehensive ImageMagick 7 command-line image processing for inspecting, converting, resizing, cropping, compositing, masking, drawing, annotating, color correction, quantization, filtering, morphology, distortion, animation, montage, comparison, metadata, batch automation, and advanced pixel analysis. Use whenever a task mentions ImageMagick, `magick`, `identify`, `mogrify`, raster image conversion or batch image manipulation, or when a reproducible CLI pipeline is preferable to a GUI or language binding. Also use for debugging ImageMagick commands, porting ImageMagick 6 `convert` syntax to version 7, choosing coder/delegate options, and safely processing untrusted or large images. Do not use merely to edit an existing SVG as vector source or when the user explicitly requests a different image tool.

59k tokens scripts
Game Image Generation Rules
Yuki001

Plan, generate, inspect, vision-evaluate, refine, and package images and visual assets for game development through a backend-agnostic production loop. Use for PNG or SVG props, icons, characters, environments, tiles, UI, VFX, transparent cutouts, concept sheets, sprite sheets, or animation frames; especially when the task needs batch generation, native prompts or ComfyUI/Stable Diffusion Danbooru tags, reference consistency, background removal, format validation, or iterative quality scoring. This skill orchestrates available image, image-edit, vision, SVG, MCP, local, and host-native tools without requiring a specific generator.

34k tokens scripts
Openrouter Image Generate
Yuki001

Generate images through OpenRouter's dedicated Image API. Use this skill whenever the user wants to create, render, generate, or save images with OpenRouter, including text-to-image, image-to-image/reference images, choosing OpenRouter image models, listing image models, setting resolution/aspect ratio/quality/output format, or passing provider-specific image options. Prefer the bundled Python script so API options are passed explicitly as command-line arguments instead of hand-written ad hoc curl requests.

6k tokens scripts
Image Generation
guinacio

Generates professional AI images using Google Gemini. ALWAYS invoke this skill when building websites, landing pages, slide decks, presentations, or any task needing visual content. Invoke IMMEDIATELY when you detect image needs - don't wait for the user to ask. This skill handles prompt optimization and aspect ratio selection.

3k tokens
Image Management
chaterm

Docker 镜像管理

1k tokens
Refine Canvas Strokes
machaomc

Design, implement, review, or troubleshoot safe refinement of editable 2D canvas strokes. Use for handwriting smoothing, automatic text layout cleanup, optional font-guided handwriting normalization, freehand cleanup, geometric shape snapping, diagram cleanup, stroke-style normalization, semantic stroke replacement, and previewable or undoable refinement pipelines in whiteboards, note apps, drawing tools, annotation systems, and pen-input SDKs. Do not use for general raster image editing, prose rewriting, or renderer bugs unrelated to stroke transformation.

502k tokens scripts
Qiaomu Icon Generator
joeseesun

| Generate multiple app, favicon, or website icon candidates with either the local QM Icon Studio CLI or Codex built-in image generation with reference icons, then present a contact sheet so the user can choose. Use when building websites or apps and needing reusable icon options.

2500k tokens scripts zh
Video Overview Generator
simonmesmith

Create narrated video overviews from source documents using editable storyboard rows, OpenAI text-to-speech audio, OpenAI image generation, captions, and deterministic ffmpeg assembly. Use when Codex needs to turn PDFs, DOCX files, Markdown, text notes, research, meeting materials, briefs, or folders of documents into a NotebookLM-style static-image video overview with voiceover, or when the user wants to edit/regenerate individual narration or image rows without rebuilding the whole project.

914k tokens scripts
Codex Mood Board
simonmesmith

Build lightweight mood boards, visual territories, creative directions, style boards, reference boards, art-direction boards, campaign mood explorations, brand look-and-feel explorations, image batches for visual ideation, more-like-this rounds from attached images, and follow-up mood-board iterations from annotated or attached images. Use this standalone skill instead of Creative Production when the user asks Codex to create, generate, explore, iterate, or review mood-board-style image directions.

214k tokens scripts
Podcast Generator
simonmesmith

Generate NotebookLM-style podcasts from documents or supplied scripts. Use when Codex needs to turn source material into a reviewable single-host, two-host, or multi-host podcast script; create host-marked dialogue with stable IDs; manage client revisions and pronunciation guidance; generate ElevenLabs v3 Text to Dialogue audio in natural chunks; or merge chunked audio with FFmpeg into a final episode.

517k tokens scripts
Polished Presentations
simonmesmith

Create polished executive or client presentation decks in Codex through a staged workflow: optional research, user direction, narrative outline, subagent critique, wireframe, mood boards, and final full-slide image design. Use when the user asks to make a polished presentation, turn research, notes, or dictated direction into a deck, create an executive/client deck, build a deck with mood boards, or design slides using image generation. Requires the imagegen skill for mood boards and final slide images.

59k tokens
Ml Product Analysis
joomcode

> Analyzes one Mercado Livre product and its competitors. Use this skill when a user wants to size up a single product or the products that compete with it — by a Mercado Livre link, a JoomPulse link, a photo, or a row of data. It returns a product card for the subject plus a ranked table of comparable products, each with price, estimated monthly sales and revenue, logistics, and does this sell", "find similar products", "find competing products", and the pt-BR equivalents "analisar este produto", "quanto vende esse produto", "produtos parecidos", "produtos concorrentes", "análise por foto", "análise por link". For "what new products should I add to my store" across a whole catalog, use the assortment gap-analysis skill instead — this skill analyzes a single product the seller already has in hand.

3k tokens
Pulse Find Exact Same Product
joomcode

> Finds product listings that appear to represent the same real-world product as a reference item. Use this skill when a user wants to find duplicate listings, match a product across listings, identify identical products by title or URL, compare product photos against marketplace listings, or find the same product on Mercado Livre.

894 tokens
Watch Video
Newuxtreme

SLASH-COMMAND-ONLY. Invoke ONLY when the user explicitly types the literal `/watch-video` slash command. Never auto-trigger on phrases like "watch this video," "summarize this video," "take notes on this video," or on a bare YouTube URL. For ad-hoc YouTube questions, fetch the transcript directly (YouTube page or yt-dlp) and skim — do not run this heavyweight pipeline.

100k tokens scripts
Assemblyai Streaming
ratacat

This skill should be used when working with AssemblyAI’s Speech-to-Text and LLM Gateway APIs, especially for streaming/live transcription, meeting notetakers, and voice agents that need low-latency transcripts and audio analysis.

2k tokens
Brave Search
ratacat

Use when user asks to search the web, look something up online, find current/recent/latest information, or needs cited answers. Triggers on "search", "look up", "find out about", "what is the current/latest", image searches, news lookups. NOT for searching code/files—only for web/internet searches.

10k tokens scripts
Every Style Editor 2
ratacat

Use this agent when you need to review and edit text content to conform to Every's specific style guide. This includes reviewing articles, blog posts, newsletters, documentation, or any written content that needs to follow Every's editorial standards. The agent will systematically check for title case in headlines, sentence case elsewhere, company singular/plural usage, overused words, passive voice, number formatting, punctuation rules, and other style guide requirements.

915 tokens
Feature Video
ratacat

Record a video walkthrough of a feature and add it to the PR description

2k tokens
Gemini Imagegen
ratacat

This skill should be used when generating and editing images using the Gemini API (Nano Banana Pro). It applies when creating images from text prompts, editing existing images, applying style transfers, generating logos with text, creating stickers, product mockups, or any image generation/manipulation task. Supports text-to-image, image editing, multi-turn refinement, and composition from multiple reference images.

8k tokens scripts
Rclone
ratacat

Upload, sync, and manage files across cloud storage providers using rclone. Use when uploading files (images, videos, documents) to S3, Cloudflare R2, Backblaze B2, Google Drive, Dropbox, or any S3-compatible storage. Triggers on "upload to S3", "sync to cloud", "rclone", "backup files", "upload video/image to bucket", or requests to transfer files to remote storage.

1k tokens scripts
Markdown Video
jykim

Convert Deckset-format markdown slides with speaker notes to presentation video with TTS narration. Use when user requests to create video from slides, generate presentation video, or convert slides to MP4 format.

29k tokens scripts
Markdown Slides
jykim

Create presentation slides in Markdown format (Deckset/Marp compatible). Use when user requests to create slides, presentations, or convert documents to slide format. Handles image positioning, speaker notes, and proper formatting.

3k tokens
Obsidian Canvas
jykim

Create and manage Obsidian Canvas files with automatic layout generation. Use when creating visual knowledge maps, weekly reading summaries, or project timelines.

3k tokens
Video Add Chapters
jykim

Add chapters to videos by transcribing, analyzing, and generating structured markdown documents with YouTube chapter markers. Optionally generate highlight videos.

17k tokens scripts
Video Full Process
jykim

Unified workflow combining video-clean and video-add-chapters with transcript reuse and chapter remapping

8k tokens scripts
Youtube Transcript Summarizer
jykim

Extract YouTube video transcripts and generate AI-powered summaries in any language. Converts videos to structured markdown documents with summaries, key points, and timelines.

6k tokens scripts
Generate Image
maxedapps

>- Generates AI images through fal.ai HTTP queue workflows from a Bun CLI. Use this skill when a task needs AI image generation, schema inspection, image upload, or queue polling through fal.ai. Do not use for video, audio, 3D, or non-image workflows.

6k tokens scripts
Create Slides
maxedapps

>- Creates and materially redesigns polished, dependency-free HTML slide decks from five styled templates, with stepped and automatic reveal animations driven by named motion presets, a cross-slide title morph, and PDF plus MP4 export at 1080p, 2K and 4K with reveals optionally timed to a narration/caption track. Use this skill when asked to create, build, redesign, restyle, or export a slide deck, presentation, pitch deck, talk, tutorial/video slides, course slides, slide b-roll for a video, or keynote-style slides as a web page. Do not use for ordinary websites or apps, editing proprietary slide files (PowerPoint, Keynote, Google Slides), research-only requests without a deck deliverable, or generating slide images only.

51k tokens scripts