Agent skills

Media & Video skills

Read straight from the source repositories, not from submitted listings. Every skill shows what it does, what is inside, where it came from — and whether attention around its source is actually growing.

Toolclaude-code 28,839codex 6,576cursor 4,239copilot 1,338windsurf 73cline 44
CategoryAI & Agents 4,178Data & Analytics 3,150Code Review & Quality 1,855Backend & API 1,737Security 1,624Workflow & Productivity 1,610Documentation 1,359Design & Presentation 1,342Content & Marketing 1,217Testing & QA 1,077DevOps & Cloud 782Databases 734Frontend 626Business & Finance 444Media & Video 346Other 7,919
509 found
193240 · page 5 / 11
audio-dspAudio DSP skill for filters and real-time processing.a5c-aiwritesaudio-jingleAudio generation skill — jingles, beds, voiceover, and sound effects. Routes music requests to Suno V5 / Udio / Lyria, speech to…nexu-ioAudio ML ValidatorYou are the on-device audio ML specialist for Modcaster's AI-driven audio processing.majiayu000writesaudio-reactive-video-compositionProvider-independent production guidance for translating measured audio features into deterministic video timing and motion. Use…calesthioaudiobook-ingestProcess audiobooks from the inbox: transcribe to text and organize audio for voice training.majiayu000writesaudiobook-productionProduce full-length audiobooks and long-form narration with generative voice tools. Use when the task is to turn a manuscript or…calesthiobeat-sync-reelGenerates Instagram Reels where product image cuts are synced to audio beats. Accepts audio as a local file, URL, or search…gooseworks-aiwritescatsharp-galoisCatSharp Scale Galois Connections between agent-o-rama and Plurigrid ACT via Mazzola's categorical music theorymajiayu000cellcog#1 on DeepResearch Bench (Feb 2026). Any-to-Any AI for agents. Combines deep reasoning with all modalities through sophisticated…majiayu000bio-clip-seq-clip-alignmentAlign CLIP-seq reads to the genome with crosslink site awareness. Use when mapping preprocessed CLIP reads for peak calling.majiayu000clip-aware-embeddingsSemantic image-text matching with CLIP and alternatives. Use for image search, zero-shot classification, similarity matching. NOT…majiayu000writesbio-workflows-clip-pipelineEnd-to-end CLIP-seq analysis from FASTQ to binding sites and motif enrichment. Use when analyzing protein-RNA interactions from…majiayu000create-video-storyboardUse this plugin when the user wants a video concept, storyboard, shot list, prompt pack, or render-ready motion brief for a…nexu-iodeepgram-core-workflow-aImplement production pre-recorded speech-to-text with Deepgram. Use when building audio transcription, batch processing, or…jeremylongshorewritesdeveloper-advocacyWhen the user wants to do developer advocacy activities including conference talks, live coding, podcasts, and building in…sickn33drone-cv-expertExpert in drone systems, computer vision, and autonomous navigation. Specializes in flight control, SLAM, object detection…majiayu000writesElevenLabs AutomationAutomate ElevenLabs text-to-speech workflows -- generate speech from text, browse and inspect voices, check subscription limits…majiayu000elevenlabs-performance-tuningOptimize ElevenLabs TTS latency with model selection, streaming, caching, and audio format tuning. Use when experiencing slow TTS…jeremylongshorewriteselevenlabsConvert documents and text to audio using ElevenLabs text-to-speech. Use this skill when the user wants to create a podcast…majiayu000faion-multimodal-aiMultimodal AI: vision, image/video generation, speech-to-text, text-to-speech, voice synthesis.majiayu000writesfal-aiGenerate images, videos, and audio with fal.ai serverless AI. Use when building AI image generation, video generation, image…majiayu000fal-ai-mediaUnified media generation via fal.ai MCP — image, video, and audio. Covers text-to-image (Nano Banana), text/image-to-video…majiayu000fal-audioText-to-speech and speech-to-text using fal.ai audio modelsmajiayu000fal-audioText-to-speech and speech-to-text using fal.ai audio modelsmajiayu000fal-upscaleUpscale and enhance image and video resolution using AImajiayu000fal-upscaleUpscale and enhance image and video resolution using AImajiayu000game-feelAdd "juice" and game feel that makes actions satisfying — screen shake, hit-stop/freeze frames, tweened/eased motion, squash &…gamedev-skillsgodot-audioPlay and mix audio in Godot 4.x: AudioStreamPlayer (2D/3D variants), audio buses with volume/mute and effects, music vs SFX…gamedev-skillsgranola-common-errorsTroubleshoot common Granola errors \u2014 audio capture failures, transcription\ \ issues,\ncalendar sync problems, and…jeremylongshorewritesheadline-generatorGenerate headline candidates from a story's raw facts: news-style headlines, press-release headlines, and pitch subject lines. A…elvisunHeyGen AutomationAutomate AI video generation, avatar browsing, template-based video creation, and video status tracking through HeyGen's platform…majiayu000higgsfield-videoProduce video (and its supporting stills) through Higgsfield (higgsfield.ai), a multi-model creative platform that fronts…calesthiohumanizeUse whenever the user asks to "humanize", "make this sound more human", "rewrite to avoid AI detection", "make this less…harshaneelklingai-model-catalogBuild explore Kling AI models and their capabilities for video generation. Use when selecting models or understanding features.…majiayu000writeslivestream-event-productionUse this skill to plan, direct, troubleshoot, and hand off provider-independent live or hybrid livestream event productions…calesthiollavaLarge Language and Vision Assistant. Enables visual instruction tuning and image-based conversations. Combines CLIP vision…Orchestra-Researchmakepad-2.0-eventsmakepad event, makepad action, MatchEvent, handle_event, handle_actions, on_click, on_render, on_return, on_startup…ZhangHanDongmatlab-display-imageDisplay images and annotations for image processing, computer vision, and visual inspection. Use when displaying images with…matlabmedia-crawler-douyinCollect Douyin competitive evidence with MediaCrawler using keyword search, exact video detail, comments, media, and creator…tsingyuaimedia-crawler-kuaishouCollect Kuaishou competitive evidence with MediaCrawler using keyword search, exact video detail, comments, media, and creator…tsingyuaiminutes-recordStart or stop recording a meeting, call, or voice memo. Use this whenever the user says "record", "start recording", "capture…silversteinmulerouterGenerates images and videos using MuleRouter or MuleRun multimodal APIs. Text-to-Image, Image-to-Image, Text-to-Video…majiayu000multimodal-mlVision-language model patterns including CLIP, LLaVA, cross-modal alignment, and embedding fusion. Use when building or…majiayu000multimodal-ragCLIP, SigLIP 2, Voyage multimodal-3 patterns for image+text retrieval, cross-modal search, and multimodal document chunking. Use…majiayu000news-collector-agentCollects daily hot stock market issues and top movers for MeowStreet Wars video production. Identifies tickers with significant…majiayu000nvenc-nvdecNVIDIA hardware video encoding/decoding integration. Configure NVENC encoding parameters, set up NVDEC decoding pipelines, handle…a5c-aiwritesnvidia-cosmos-videoSelect, run, and govern NVIDIA Cosmos world-video generation across Cosmos 3 Generator, Predict2.5, Transfer2.5, downloadable…calesthiood-media-generationDefault reference pipeline for image, video, and audio projects — routes through media-image / media-video / media-audio atoms…nexu-io
← Prev5 / 11Next →
How the catalog works
What is an agent skill?

A folder with a SKILL.md inside — instructions, and often scripts and assets, that an AI agent loads when the task matches. Claude Code, Codex, Cursor and Copilot all read the same format, so one skill usually works across them.

Where does this catalog come from?

We read 660 source repositories straight from their file trees rather than from submitted listings — what you see is what is actually published. 98 repositories were rejected because they advertise skills but contain none: link lists, not folders.

Why is there no install counter?

Because install counts live in the registry that serves `npx skills add`, and that is not ours — publishing a number we cannot verify would be worse than showing none. Instead we show where a skill comes from and whether attention around its source is actually growing, measured from our own weekly snapshots.

Do you deduplicate?

Yes, and it matters more than expected. Aggregator repositories republish the same skill in several places — one source carried 6,317 SKILL.md files for 2,001 actual skills. We collapse by folder name and keep the canonical copy, so the catalog counts things, not copies.

Keep going