music-video

Solid

Generate a 60-second 9:16 vertical music video (YouTube Shorts / TikTok / Reels format) from an operator-supplied music file plus mood keywords. Use when the user provides an audio file (mp3 / wav / m4a / aac) and wants a short-form vertical video with beat-aligned cuts, mood-matched Pexels B-roll, and optional vintage post-shaders (pond ripple / breathing zoom / halation / combo). Music is the primary audio — no narration, no captions.

AI & Automation 15 stars 4 forks Updated today MIT

Install

View on GitHub

Quality Score: 86/100

Stars 20%
40
Recency 20%
100
Frontmatter 20%
70
Documentation 15%
100
Issue Health 10%
80
License 10%
100
Description 5%
100

Skill Content

# music-video Generate a 60-second 9:16 vertical music video from a music file + mood keywords. Designed for YouTube Shorts / TikTok / Reels upload. ## What this produces Given: - A music file (mp3 / wav / m4a / aac) — typically 60–240 seconds of operator-supplied music (Suno-generated, YouTube Audio Library, Pixabay, etc.). Music itself is the *only* audio track in the output. - A short list of mood keywords (3–6 comma-separated phrases like `"rainy street, jazz cafe, vinyl, wet pavement"`). Produces: - A **1080 × 1920 (9:16 vertical)** mp4, exactly 60 seconds. - 8 B-roll clips fetched from Pexels (per-keyword), trimmed and ordered to match phrase boundaries detected via `aubiotrack`. - Drum-onset-aligned glitch micro-edits (`aubioonset`) on static- camera clips only. - Vintage lo-fi processing (film grain, vignetting, zoom-pulse) per v6 defaults; tunable via env vars. - Optional post-shader pass: `pond` (water-surface ripple), `breathing` (5-s scale wave), `halation` (warm bloom), or `combo` (phrase-aware pond + halation envelope). ## How to invoke User-facing invocation: `/music-video <music_file_path> "<comma_separated_keywords>"` Examples: ```text /music-video "assets/music/Rainy Bossa.mp3" "rainy street, jazz cafe, vinyl, wet pavement" /music-video ~/Downloads/track.wav "tokyo neon, vibraphone, late night, shibuya" ``` If the user invokes without a path or without keywords, ask them for the missing input rather than guessing. ## Step-by-s...

Details

Author
MelonS
Repository
MelonS/MelonS-Agents
Created
2 months ago
Last Updated
today
Language
C#
License
MIT

Integrates with

Similar Skills

Semantically similar based on skill content — not just same category

Code & Development Listed

music-to-video

Turn a music track (an audio file, a video to pull audio from, or a track generated from a mood brief) into a beat-synced video — lyric video, slideshow, or kinetic promo. The music drives all pacing; any user-supplied images/videos are cut onto the same beat grid, and a complete video needs zero assets. Narrated pieces → the input-matched workflow (see /hyperframes). Unclear → /hyperframes.

1 Updated today
jpratt9
AI & Automation Listed

ai-music-and-sound

The AI music + sound-design skill for social -- original/licensed audio beds and sound design for Reels/TikToks/Shorts/videos. Use when someone needs background music, a track, or sound effects for a social video, asks which AI music tool is safe to use, or asks "can I use this trending sound/song on my brand video?". The real brief is "audio that won't get muted, claimed, or sued," so it picks the safest licensed source and never uses copyrighted or trending music without a license. Uses the SCORE framework. Reads brand-profile + the video it scores first. The agent briefs the music + sound design, picks the safest licensed source (ElevenLabs Music/SFX or stock libraries over Suno/Udio; paid tier for commercial rights), and advises licensing/Content-ID/disclosure. The tool generates/licenses the audio; the creator bakes it in; WoopSocial publishes the video and does NOT generate music. Pure AI music may not be copyrightable; never "100% legally safe." Pairs with ai-voiceover.

5 Updated 4 days ago
social-media-skills
Code & Development Featured

render-song-mv

Assemble a song-driven music-video ad from a config — a generated sung track carries the whole narration across N tableaux (one keyframe -> one i2v clip per lyric beat) with NO separate voiceover, captions synced to the song's OWN word timings (script-window, never Whisper) and the hook word landing on the chorus drop, closed on a PIL brand end card. This is the FREE deterministic assembly stage (clip cut-to-timeline + captions + end card + FFmpeg composite); the song, keyframes, and clips come from create-music-elevenlabs / create-image-fal / create-video-fal. Use for the song-driven-music-video format.

1,063 Updated 4 days ago
gooseworks-ai