ai-media

Featured

Use when a creative goal must become a finished media file: pick and order generative-media models per modality — AI voiceover, image-to-video clips, score — then glue them with ffmpeg (mux, duck, loudnorm, concat). NOT still-image generation/editing (that is `replicate-images`); NOT code-rendered React compositing (that is `remotion-video`).

AI & Automation 116 stars 9 forks Updated 2 days ago MIT

Install

View on GitHub

Quality Score: 92/100

Stars 20%
69
Recency 20%
100
Frontmatter 20%
70
Documentation 15%
100
Issue Health 10%
80
License 10%
100
Description 5%
100

Skill Content

# ai-media You are the cross-modal director. You decide **which** generative-media model to call per modality, in **what order**, with **what params**, then **assemble** the pieces with ffmpeg into one finished file. You do not own a single provider's API surface and you do not prompt still images — you orchestrate and glue. ## Pipeline shape — decide what the goal needs Map the goal to modalities and an ordered step list, and **lock that plan before you generate a single asset** — media generation is slow and metered, so a re-roll of a 10 s Veo clip or a 90 s music track costs real money and minutes. Fixing the scene list, aspect ratio, target loudness and model per modality *first* is cheaper than discovering at mux time that your clips are 9:16 and your VO is the wrong sample rate. The "delegate to" column is where the actual call mechanics live — you pick the model and params, those skills run the call. | Goal | Needs | Ordered steps | Delegate calls to | |------|-------|---------------|-------------------| | Narrated explainer | stills + img→video + VO + music | script → per-scene stills → clip per scene → VO → music → conform → concat → mix+duck → loudnorm → MP4 | `replicate-images`, `fal`/`replicate` | | Product teaser (1 hero) | 1 still + img→video + music | still → clip → music → mix → loudnorm → MP4 | `replicate-images`, `fal`/`replicate` | | Faceless short | stills + img→video + VO + music + captions | (explainer pipeline) + burn captions | `replicate-images`; ...

Details

Author
ericrisco
Repository
ericrisco/rsc-harness
Created
3 months ago
Last Updated
2 days ago
Language
JavaScript
License
MIT

Similar Skills

Semantically similar based on skill content — not just same category

AI & Automation Listed

ai-generated-media-pipeline

Turn AI-generated clips and stills into production web hero assets — generate via Higgsfield/Runway/Sora/Veo/Kling (video) and Nano Banana/Flux/Seedream (image), ideally through an MCP so the agent produces assets in-loop, then encode them web-ready (AV1/H.264, faststart, poster extraction, frame sequences) within a weight and licensing budget. Use when sourcing a hero video/background loop/scroll frame-sequence from AI tools, or preparing any generated media for the web. Feeds cinematic-hero-sections (which consumes the assets). Triggers on "Higgsfield", "Nano Banana", "AI hero video", "generate a background video", "Runway/Sora/Veo/Kling", "encode video for web", "ffmpeg hero", "loopable clip", "poster frame".

0 Updated 2 months ago
BenMacDeezy
AI & Automation Listed

super-claudiomedia-content-creation

Media content creation skill. Use when the user wants to create, generate, or produce any kind of video, audio, or image. This is the main skill for all media generation tasks. Trigger on video: "I want to make a video", "create a TikTok video", "generate a realistic video", "make a promo video", "animate my photo", "create a video ad", "Remotion", "Higgsfield", "Kling", "Seedance", "Weavy AI", "Hailuo". Trigger on audio: "read this article aloud", "create a voiceover", "text to speech", "TTS", "generate narration in Portuguese", "background music", "create a jingle", "ElevenLabs", "Francisca Neural", "Suno", "Udio", "audio summary". Trigger on image: "generate an image", "create a graphic", "make a diagram", "draw X", "generate a photo of Y", "make an infographic", "Midjourney", "DALL-E", "Flux", "Napkin.ai", "Nano Banana 2", "animate a static image". Also triggers for: marketing creatives, social media visuals, product photos, content creator tools.

4 Updated today
toolbox-playground
AI & Automation Listed

generic-video

Create a genre-agnostic AI video — either recreating a source (handed over by mimic-video) or from a fresh brief. Owns the creative intake (fidelity, theme/style, pacing, duration, stitch-vs-fully-AI, model, voice) and the build flow. References ugc-craft for prompt craft, ai-video-models for per-engine rules, ai-asset-generation for mechanics.

2 Updated today
Nagellabs