Canonical media and document processing covering Fal AI media, video editing, videodb, Manim video, Remotion video creation, nutrient document processing, and visa document…
Ingest video, audio, PDF, book, screenshot, and GitHub repo content into the brain. Multi-format handling with entity extraction and backlink propagation.
Localizes audio/video pipeline failures—VFR drift, silent FFmpeg merges, lip-sync no-ops, subtitle lag, Veo speech-vs-song—by checking frames and streams per stage.
ミーティングのトランスクリプト(プレーンテキスト)から Remotion ストーリー型ビデオプロジェクトを生成する。npx remotion preview でローカルプレビューできる状態まで自動構築する。Use when: meeting transcript, meeting summary video, meeting recap video,…
Generate and maintain a multi-channel product feed across Google Merchant Center, Meta Catalog (Instagram + Facebook Shops), TikTok Shop, Pinterest Catalogs, Bing Shopping, and…
Extract and analyze content from video ads using Gemini Vision AI. Supports frame extraction, OCR text detection, audio transcription, and AI-powered scene analysis.
Deconstruct video ad creatives into marketing dimensions using Gemini AI. Extracts hooks, social proof, CTAs, target audience, emotional triggers, urgency tactics, and more.
Use when the user asks for Japanese media-forge video-prompt examples, Japanese prompt patterns, example rewrites, or safe versions of working Japanese video-generation prompts,…
Use when the user asks for Korean media-forge video-prompt examples, Korean prompt patterns, example rewrites, or safe versions of working Korean video-generation prompts, for any…
Use when the user asks for Chinese media-forge video-prompt examples, Chinese prompt patterns, example rewrites, or safe versions of working Chinese video-generation prompts, for…
Use when the user asks to write, improve, translate, compress, or debug a media-forge video prompt; mentions T2V, I2V, V2V, R2V, camera direction, prompt quality, or provides…
Use when the user asks for a compact media-forge video prompt, short Chinese prompt, prompt compression, 30-100 word output, or removal of unnecessary prompt language, for any…
Use when the user asks for Spanish media-forge video-prompt wording, Spanish cinematic vocabulary, or translation of camera, lighting, action, VFX, audio, and production terms…
Use when the user asks for Japanese media-forge video-prompt wording, Japanese cinematic vocabulary, or translation of camera, lighting, action, VFX, audio, and production terms…
Use when the user asks for Korean media-forge video-prompt wording, Korean cinematic vocabulary, or translation of camera, lighting, action, VFX, audio, and production terms into…
Use when the user asks for Russian media-forge video-prompt wording, Russian cinematic vocabulary, or translation of camera, lighting, action, VFX, audio, and production terms…
Use when the user asks for Chinese media-forge video-prompt wording, Mandarin cinematic vocabulary, Chinese prompt compression, or translation of camera, lighting, action, VFX,…
Turn one image of a small creature, mascot, fantasy pet, figurine-like character, or tiny animal into a stable 7–8 second MiniMax H3 image-to-video clip.
MiniMax M-series production wiring patterns for the OpenAI-compatible API at api.minimax.io. Use when wiring MiniMax-M2.7 / MiniMax-M2.7-highspeed / Hailuo into a service…
Plan and generate MiniMax H3 video effectively through current fal.ai Hailuo 03 endpoints. Use when a production selects or considers MiniMax H3 for text-to-video, image-to-video,…
Analyze a product walkthrough, bug report video, Loom, or ScreenPal using Minutes transcription plus visual review.
Canonical short-form-ad audio mix in one FFmpeg pass. VO loudnorm + 3.0× per-clip + 2.0× mix, music 0.13 base + apad+afade, sidechain compress 20:1 @ 0.01, climax line +20%,…
MLT is an open-source LGPL multimedia framework designed for video editing. It provides a toolkit and the melt command-line tool for non-linear video editing, transitions,…
MoviePy is a Python library for video editing — cuts, concatenations, title insertions, compositing, and custom effects.
Create a high-end cinematic product video advertisement starting from a simple product photo. — from majiayu000/claude-skill-registry
Compose video-synced background scores from a storyline JSON using local waveform synthesis (numpy + ffmpeg). Styles: chiptune, ambient, electronic.
Solo-Selbstständige: prüft GEMA, Leistungsschutz, Samples, Podcast-Rechte und Plattformen; mit Abfrage von Tätigkeit, Status, Belegen, Fristen, Geldfolge und konkretem nächstem…
An ASE skill built around the official Mux Node SDK for working with Mux Video and Mux Data from JavaScript or TypeScript.
End-to-end automated Music Video pipeline. Covers songwriting (lyrics/composition), Suno music generation (browser automation), lyrics alignment (stable-ts), video generation (Veo…
Generates 10 video ideas for the channel's niche — each with a working title, one-sentence hook, and content angle suited to the channel's voice.
Turn an mp4 video of a single object (orbited 360°) into both a clean, browser-navigable 3D Gaussian splat and a watertight, 3D-printable STL/GLB mesh - 100% locally on Apple…
OmniHuman1によるAIアバター・リップシンク動画生成ガイド。 Use when: (1) user says「AIアバター」「リップシンク」「OmniHuman」, (2) user wants talking head videos from images, (3) user mentions「アバター動画」「1枚の画像から動画」.
Use when handling external short-video materials such as Douyin/TikTok China links, downloaded videos, transcripts, screenshots, image posts, creator batches, or visual gallery…
AI-powered video generation skill. Use when the user wants to generate videos from text descriptions, browse video recipes, upload assets, or manage video creation workflows.
Generate images & videos with AIsa. Gemini 3 Pro Image (image) + Qwen Wan 2.6 (video) via one API key. — from satoshistackalotto/skills
Generate images & videos with AIsa. Gemini 3 Pro Image (image) + Qwen Wan 2.6 (video) via one API key. — from satoshistackalotto/skills
YouTube SERP Scout for agents. Search top-ranking videos, channels, and trends for content research and competitor tracking. — from satoshistackalotto/skills
Uses PeerTube's REST API and federation-aware platform features to automate video uploads, channel management, moderation queues, and instance operations.
turn photos and images into polished photo slideshow with this photo-video-maker-deutsch skill. Works with JPG, PNG, HEIC, WebP files up to 500MB.
Use when running video data augmentation and auto-labeling workflows on OSMO: flow selection, preflight, submit-time interpolation, monitoring, and output retrieval.
Use when building an autonomous short-film/animation video pipeline — audio-first (narration decides cuts), image-then-animate for character consistency, per-video cost ledger,…
视频处理任务规划工具。从用户输入中提取视频 URLs,生成唯一 VideoId,创建结构化的 todolist.md 追踪待生成文件。支持小红书、抖音、TikTok、B站、YouTube、快手等 11+ 平台。
Turn one source image into a short video through the PopiArt runtime baseline. Use this when the user wants the most direct image-to-video path for motion previews, short teaser…
Generate one short showcase video for PopiStudio Alice with strict character consistency. Use this when a creator agent needs a single Alice teaser shot, demo clip, or animated…
PPT/슬라이드를 나레이션과 자막이 포함된 영상으로 변환합니다. PPTX 파일 또는 slides.json에서 슬라이드 이미지를 추출/렌더링하고, TTS로 나레이션을 생성하며, 자막을 추가하여 최종 MP4 영상을 만듭니다. "PPT를 영상으로 만들어줘", "발표 영상 생성", "자막 포함 영상 만들기" 요청 시 사용합니다.
Prepare multi-platform publishing materials inside the current Codex task from a local video or subtitle file.
Generate AI UGC video ads from any public product URL. Use when the user wants to produce, iterate on, or A/B test video ad creative for e-commerce; when they are evaluating AI…
Use video-recap-skills when an agent should turn source video into a Chinese narration recap with scene understanding, scripting, voiceover, subtitles, and an optional editable…
Turns Claude Code, Codex, or another skills-aware coding agent into a Remotion-based product video operator with shot cards, motion previews, templates, capture scripts, sound…
Create a high-end cinematic product video advertisement starting from a simple product photo. — from SamurAIGPT/Generative-Media-Skills
Record, camera-process, encode, stage, and validate polished Kandev product films, landing-page loops, screenshots, and alternate framing from isolated demo data.
Create, render, QA, and publish IngredientHQ YouTube Shorts end-to-end. Use when: Derrek says 'make a video', 'next short', 'new ingredient video', or any request to…
Generates a valid YouTube video ID and outputs only the ID string without any conversational filler, apologies, or explanations.
通义千问视频分析专业版,面向企业团队与专业用户的高级AI视频内容理解工具。核心能力: - 批量视频分析,支持目录扫描与队列处理 - 自定义模型选择(Qwen 系列多版本) - 结构化分析报告(JSON/Markdown/Excel) - 视频内容审核(违规/敏感内容检测) - 多维度分析(场景/物体/动作/情绪/字幕) - 优先 API 配额与企业级技术支持
Uses Subaligner to retime an existing subtitle file against the final audio track, then outputs a corrected subtitle asset.
Extract a structured cooking recipe from a shared video URL when the user sends `recipe `. Prioritize caption/description and comments via browser automation, then use web…
Turn a raw talking-head recording or voiceover into a tight video essay with word-level timing, clean cuts, captions, and timed image overlays.
Use when generating short videos with RedBox official video API. Produces a detailed shot script first, asks the user to confirm it, then chooses between text-to-video,…
Use before a rendered reel, screen-recording, demo video, or GIF goes public (social, YouTube, a landing page) to catch identity and infrastructure leaks burned into the frames.
Extract video (Reels/TikTok/Shorts/any MP4), analyze frame-by-frame, generate production playbook, then reconstruct in Remotion via video-director skill.