
Generate spoken audio from text with OpenAI voices

Generate and iterate music tracks with MiniMax models from agent workflows
Clone or download it from github.com, then drop the folder into ~/.claude/skills/ (Claude Code) or your agent's skills directory.
Share this product's name and rating in your README or on your website.
Updates may be delayed by image caching.

MiniMax music generation skill for coding agents—prompting, iterating, and wiring audio outputs into product flows.
Use when an agent needs soundtrack or clip generation without hand-holding every API parameter.
Used it? Write the first review.

Generate spoken audio from text with OpenAI voices

Transcribe audio with optional speaker diarization

Build realtime multimodal Gemini Live sessions
Sequence images, audio, and video clips cleanly inside Remotion timelines