AI Music Video Maker
Build a short AI music video from a song theme — keyframes, animation, and matching music. Needs a MuAPI key.
Installation
- Make sure Claude is on your device and in your terminal.
Skills load from
~/.claude/skills/when Claude Code starts up — so you need it on your machine first. If you don't have it yet, install it once with the command below, then runclaudein any terminal to verify.One-time setupnpm i -g @anthropic-ai/claude-codeAlready have it? Skip ahead.
- Paste into Claude Code or into your terminal.
This copies the whole skill folder into
~/.claude/skills/muapi-music-video-samuraigpt/— the SKILL.md plus any scripts, reference docs, or templates the skill ships with. Safe default: works for every skill.Faster alternative (instruction-only skills)
Skips the clone and grabs only the SKILL.md file. Don't use this if the skill ships Python scripts, reference markdowns, or asset templates — they won't be downloaded and the skill will fail when it tries to load them.
Quick install (SKILL.md only)Sign up to copy - Restart Claude Code.
Quit and reopen Claude Code (or any other agent that loads from
~/.claude/skills/). New skills are picked up on startup. - Just ask Claude.
Skills auto-activate when your request matches the skill's description — no slash command needed. Trigger phrases live in the skill's own frontmatter; you can read them in the “What this skill does” section above.
Prefer to read the source first? Open on GitHub.
When Claude uses it
Build a short music video from a song theme — N keyframes, animate each, generate matching music.
What this skill does
Music Video
Build a short music video from a song theme — N keyframes, animate each, generate matching music.
Inputs
| Name | Type | Required | Default | Description |
|---|---|---|---|---|
theme | text | yes | — | Song / video theme (e.g. "lonely robot finds a friend, hopeful"). |
scenes | int | no | 3 | Number of scenes (each becomes a 5s clip). |
music_style | text | no | ambient cinematic, instrumental, slow tempo, warm | Suno-style tags for the soundtrack. |
visual_style | text | no | cinematic, photoreal, soft volumetric light, 16:9 |
Steps
Build one the plan covering:
- Layer A (parallel) — N keyframes + 1 music track all at once.
- For each scene 1..N:
muapi image generatewith a beat-specific prompt +{{visual_style}}, model=nano-banana-pro (these feed video gen). - One
muapi audio create(kind=music) using{{music_style}}, duration = N × 5 + a 2s tail.
- For each scene 1..N:
- Layer B (parallel, depends on Layer A) — animate each keyframe.
- For each scene:
muapi video from-imagewithimage=$nX.url, model=veo3.1-image-to-video, duration=5, prompt=scene-specific motion direction.
- For each scene:
- Return:
- The scene keyframes (asset ids in order).
- The animation clips (asset ids in order).
- The music track asset id.
- A short summary describing the cut order.
Notes
- Keep character continuity by repeating the character description in every scene prompt verbatim.
- Don't auto-confirm any single video call > 50 cr — those need the user's nod (the loop will prompt automatically).
- If a scene's
muapi video from-imagefails after failover, fall back tomuapi video generate(text-to-video) for that scene only.
Trigger Keywords
music video, mv, video story, song visualization
Notes for the Executing Agent
- This recipe is LLM-orchestrated: read each phase, gather any missing inputs from the user, then call
muapiCLI commands. Usemuapi auth configurefirst ifMUAPI_API_KEYis unset. - For model IDs without a CLI alias yet, fall back to the raw endpoint via
curl -X POST https://api.muapi.ai/api/v1/<endpoint> -H "x-api-key: $MUAPI_API_KEY" -H 'content-type: application/json' -d '{...}'and poll withmuapi predict wait <request_id>. - Substitute
{{input_name}}placeholders with the user's actual inputs before issuing each call.
Related skills
Skill Builder & Optimizer
anthropics
Create, edit, and optimize Claude skills with performance testing and benchmarking.
Org Change Management
alirezarezvani
Guide teams through organizational changes using the ADKAR model and communication strategies.
Audio/Video Transcription
daymade
Transcribe audio and video files to text with fast local or remote processing.
Claude Export Conversation Fixer
daymade
Repair broken line wrapping in Claude Code exported conversation files.