Tags
Best Video Skills for AI Agents
79 Video skills for AI coding assistants — ranked by popularity.
This collection provides a curated directory of AI-agent skills designed to automate video workflows. If you are a developer using Claude Code, Codex, or Cursor, these skills enable your agents to manipulate, generate, and process video assets directly from your codebase. Whether you need to extract frames with ffmpeg, automate edits via the Jianying API, or produce complex mathematical animations using Manim and Motion Canvas, these tools turn your AI agent into a video editor. The list covers a wide range of needs, from simple format transcoding and YouTube downloads to advanced AI-driven video generation through the Jimeng MCP server. Use these skills to bridge the gap between text-based coding and visual media production. Each entry provides the necessary integration paths to help your agent interact with local and cloud-based video engines, reducing manual setup time while increasing your technical output.
Top Video skills
video-downloader
ComposioHQ
Downloads videos from YouTube and other platforms for offline viewing, editing, or archival. Handles various formats and quality options.
motion-canvas
davila7
Complete production-ready guide for Motion Canvas with ESM/CommonJS workarounds, full setup templates, and troubleshooting for programmatic video creation using TypeScript
jimeng-mcp-skill
wwwzhouhui
使用jimeng-mcp-server进行AI图像和视频生成。当用户请求从文本生成图像、合成多张图片、从文本描述创建视频或为静态图像添加动画时使用此技能。支持四大核心能力:文生图、图像合成、文生视频、图生视频。需要jimeng-mcp-server在本地运行或通过SSE/HTTP访问。
jianying-editor
luoluoluo22
剪映 (JianYing) AI自动化剪辑的高级封装 API (JyWrapper)。提供开箱即用的 Python 接口,支持录屏、素材导入、字幕生成、Web 动效合成及项目导出。
manim
davila7
Comprehensive guide for Manim Community - Python framework for creating mathematical animations and educational videos with programmatic control
video-processor
basher83
Process video files with audio extraction, format conversion (mp4, webm), and Whisper
video-frames
openclaw
Extract frames or short clips from videos using ffmpeg.
yt-dlp
SecKatie
Download audio and video from thousands of websites using yt-dlp. Feature-rich command-line tool supporting format selection, subtitle extraction, playlist handling, metadata embedding, and post-processing. This skill is triggered when the user says things like "download this video", "download from YouTube", "extract audio from video", "download this playlist", "get the mp3 from this video", "download subtitles", or "save this video locally".
vectcut-api
sun-guannan
VectCutAPI is a powerful cloud-based video editing API tool that provides programmatic control over CapCut/JianYing (剪映) for professional video editing. Use this skill when users need to: (1) Create video draft projects programmatically, (2) Add video/audio/image materials with precise control, (3) Add text, subtitles, and captions, (4) Apply effects, transitions, and animations, (5) Add keyframe animations, (6) Process videos in batch, (7) Generate AI-powered videos, (8) Integrate with n8n workflows, (9) Build MCP video editing agents. The API supports HTTP REST and MCP protocols, works with both CapCut (international) and JianYing (China), and provides web preview without downloading.
remotion
davila7
Best practices and comprehensive guide for Remotion - programmatic video creation in React with animations, compositions, and media handling
youtube-clipper
op7418
YouTube 视频智能剪辑工具。下载视频和字幕,AI 分析生成精细章节(几分钟级别), 用户选择片段后自动剪辑、翻译字幕为中英双语、烧录字幕到视频,并生成总结文案。 使用场景:当用户需要剪辑 YouTube 视频、生成短视频片段、制作双语字幕版本时。 关键词:视频剪辑、YouTube、字幕翻译、双语字幕、视频下载、clip video
heygen-best-practices
davila7
Best practices for HeyGen - AI avatar video creation API
youtube-summarizer
sickn33
Extract transcripts from YouTube videos and generate comprehensive, detailed summaries using intelligent analysis frameworks
video-transcript-downloader
steipete
Download videos, audio, subtitles, and clean paragraph-style transcripts from YouTube and any other yt-dlp supported site. Use when asked to “download this video”, “save this clip”, “rip audio”, “get subtitles”, “get transcript”, or to troubleshoot yt-dlp/ffmpeg and formats/playlists.
video-report
remotion-dev
Generate a report about a video
gemini-count-in-video
benchflow-ai
Analyze and count objects in videos using Google Gemini API (object counting, pedestrian detection, vehicle tracking, and surveillance video analysis).
ffmpeg-keyframe-extraction
benchflow-ai
Extract key frames (I-frames) from video files using FFmpeg command line tool. Use this skill when the user needs to pull out keyframes, thumbnails, or important frames from MP4, MKV, AVI, or other video formats for analysis, previews, or processing.
high-dynamic-video-choreographer
amao2001
Analyzes a theme or image to design a sequence of 5 high-energy, logically connected action beats for AI video generation. Features a recursive optimization process to ensure physical momentum and visual diversity.
video-comparer
daymade
This skill should be used when comparing two videos to analyze compression results or quality differences. Generates interactive HTML reports with quality metrics (PSNR, SSIM) and frame-by-frame visual comparisons. Triggers when users mention "compare videos", "video quality", "compression analysis", "before/after compression", or request quality assessment of compressed videos.
klingai-video-extension
jeremylongshore
Execute extend video duration using Kling AI continuation features. Use when creating longer videos from shorter clips or building seamless sequences. Trigger with phrases like 'klingai extend video', 'kling ai video continuation', 'klingai longer video', 'extend klingai clip'.
pdf-to-video
DangJin
Use when user wants to convert a PDF document into a showcase video, extract key points from PDF, or create video presentation from PDF file
sora
davila7
Use when the user asks to generate, remix, poll, list, download, or delete Sora videos via OpenAI’s video API using the bundled CLI (`scripts/sora.py`), including requests like “generate AI video,” “Sora,” “video remix,” “download video/thumbnail/spritesheet,” and batch video generation; requires `OPENAI_API_KEY` and Sora API access.
ffmpeg-media-info
benchflow-ai
Analyze media file properties - duration, resolution, bitrate, codecs, and stream information
acestep
ace-step
Use ACE-Step API to generate music, edit songs, and remix music. Supports text-to-music, lyrics generation, audio continuation, and audio repainting. Use this skill when users mention generating music, creating songs, music production, remix, or audio continuation.
How to choose a Video skill
Evaluate these skills based on your specific production pipeline. Consider the required output format: if you need technical animations, prioritize Manim or Motion Canvas. For post-production and editing tasks, look at the Jianying or FFmpeg implementations. Check the maintenance status and required local dependencies—like whether the tool necessitates a local server or functions via direct API calls. Finally, match the tool’s scope to your task; some skills handle bulk file processing, while others are built for high-fidelity content synthesis.