Create a short visual clip guided by a song's mood, rhythm, and visual concept. This AI music video clip maker and song-to-video generator turns a music excerpt and visual direction into a cinematic music promo clip, animates approved cover art or a portrait in time with the music, interpolates motion between opening and ending art, or uses audio as a loose mood reference for a new visual concept. Use it for new-song teasers, album promo clips, cover art animation, mood visuals, virtual performer scenes, and social music teasers, with an audio-visual map built from the song's hook, energy, palette, and landing image.
设计与多媒体
short-video-bgm-studio
试用Describe the footage and get an original instrumental track written for it, yours to keep and use commercially. This AI background music generator turns a scene, mood, and tempo feel into royalty-free BGM for short videos, vlogs, product clips, tutorials, livestream and store loops, podcast intros, and slideshow recaps, with an energy arc you choose — a calm or immediate opening, a lift at the moment that matters, a clean ending — room left for narration, and a result you can listen to before you publish.
它能做什么
Describe the footage and get an original instrumental track written for it, yours to keep and use commercially. This AI background music generator turns a scene, mood, and tempo feel into royalty-free BGM for short videos, vlogs, product clips, tutorials, livestream and store loops, podcast intros, and slideshow recaps, with an energy arc you choose — a calm or immediate opening, a lift at the moment that matters, a clean ending — room left for narration, and a result you can listen to before you publish.
技能文档
Short Video BGM Studio
Turn a description of the footage into an original instrumental track built for it: read the scene, shape an energy curve that fits how the piece moves, and return a result the user can hear before publishing.
Scope and routing
Use this Skill when a video, livestream, podcast, store, or brand moment needs background music and no vocal is wanted. It fits short-video posts, vlogs, product and unboxing clips, tutorial and course footage, livestream and store loops, podcast intros and outros, and slideshow or event recaps.
Route a song with sung lyrics to beatra-ai-music-creator. Route finished
lyrics that need a melody to suno-lyrics-to-song, a gift or occasion song to
personalized-song-maker, and a cover or re-arrangement of an existing song to
ai-song-cover-studio. Route a music video built around a finished track to
ai-music-video-clip-maker, and spoken narration to
short-form-voiceover-audio.
Inputs and defaults
The one hard input is what the music is for: the scene, product, or mood it has to sit under. Reuse the platform, footage description, brand tone, target length, and any reference track already present in the conversation.
Ask only when the answer changes the paid result: the intended use, when the request is just "make music" with no scene attached.
Defaults that avoid extra questions:
instrumental: truewith no lyrics, because this route is background music.- A calm-to-lift energy arc with a clean resolved ending, which suits a cut that has to end cleanly.
model: "suno-5.5"for ordinary generation. Never omit the model and never silently useauto; pass a different model only when the user names one.- Room left for a voice, since most short-video BGM sits under narration.
Golden path
Briefing and planning are free. Only the generation call is paid.
- Write a short music card from the footage: use and destination, mood, genre, tempo feel, instrumentation, the energy arc across the cut, the intended length, whether narration sits on top, and anything to avoid.
- Turn the card into one positive prompt that carries genre, mood, tempo feel, instrumentation, structure, and intended use in a single coherent direction.
- Call
beatra.models.listfor the text-to-music capability, or the reference-audio-to-music capability when the user supplied a reference recording, whenever compatibility, controls, or price matter. Read the live card rather than assuming a model, a control, or an input limit. - Confirm before paid work. Show the frozen prompt,
instrumental: true, the title, the model, any accepted model options, the current maximum charge, and one opaque stableclient_request_id. - Submit
beatra.music.generateexactly once, record the task ID immediately, and poll that same task. - Deliver every returned clip in order with its real duration, MIME type, size,
and URL or artifact ID, plus the returned title when present, the resolved
model, the actual usage, and
billing.net_charged_credits. - Review the result against the music card. Read the actual returned duration rather than the requested one, and say plainly what the host Agent could not hear.
Treat requested length, a loop-friendly arrangement, and space for narration as arrangement direction in the prompt. Read the BGM workflow for the music card, prompt shape, payloads, model options, reference-guided tracks, recovery, and delivery review.
How this Skill executes
Use the bundled scripts/mcp_client.py for every remote Beatra operation: the
MCP tool name is the CLI argument after call, and one JSON object goes on
standard input. Never configure or call a host Beatra Connector, and never use
REST/OpenAPI as a fallback. Register the package with
beatra.installations.register on first use. Every creation is an asynchronous
task: submit once, then follow that task to a terminal state.
Decisions that require confirmation
Confirm before submitting: the frozen prompt, the instrumental setting, the title, the model and any model options, and the current maximum charge. A changed prompt, title, model, option, reference track, or instrumental setting is new paid work with a new request ID.
When the user supplies a reference recording, upload it once through the bundled client, state what should carry over and what should change, and keep that as musical direction rather than a promise about melody or arrangement.
Recovery
Save the task ID the moment it returns and poll with beatra.tasks.get;
queued and running mean wait. Replay a create only when its response is
genuinely unknown and every validated argument is byte-equivalent under the same
request ID. If the task ID is lost, use beatra.tasks.list, confirm candidates
with beatra.tasks.get, and recover the original before considering new work.
If the request ID itself is lost, do not invent a new one and do not replay.
Call beatra.tasks.cancel only at the user's request; on 409, keep polling the
original task and report cancellation only when its terminal status is
canceled.
References by task
- BGM workflow: music card, prompt construction, payloads, model options, reference-guided tracks, recovery, and delivery review.
- Installation and authentication and installation registration: first use and shared credentials.
- Tasks and results and billing, errors, and recovery: task, artifact, and billing facts.
- Bundled MCP Client diagnostics: client operation and connection diagnostics; do not configure a host Connector.
- automatic updates and safety: update behaviour and controls.
- uninstall and disconnect: package removal and shared credential cleanup.
Runtime and safe automatic updates
The bundled client silently checks at most once every 24 hours per installation. When a newer release is available, it installs automatically without separate confirmation. It uses only fixed official Beatra discovery and immutable CDN paths for this package, channel, and locale, verifies discovery, archive, manifest, and every packaged file before replacement, and replaces only package-owned files. Update checks, downloads, verification, replacement, and recovery fail open: the current installation remains usable and the original command continues. An update failure never authorizes retrying a paid generation. The choice persists across later commands.
python3 scripts/mcp_client.py update --auto off
python3 scripts/mcp_client.py update --auto on
python3 scripts/mcp_client.py update --check
--auto off disables silent checks, --auto on restores them, and --check
reports the official available version without replacing files. See
automatic updates and safety.
相关技能
Create original songs, AI-generated music, lyrics-to-song tracks, instrumentals, background music, video soundtracks, jingles, multilingual songs, and reference-led arrangements from a clear creative brief. This AI music generator and AI song maker develops genre, mood, structure, singable lyrics, vocal direction, and production style, then creates reviewable audio with Suno 5.5 or another model you explicitly choose. Use it as an AI music studio for songwriting, an AI lyrics writer, text-to-music creation, BGM generation, brand music, podcast themes, game music, bilingual songs, or focused new versions of reference audio. Review vocals, pronunciation, duration, loop points, and arrangement fit after generation, then refine the strongest result.
Create original songs, AI-generated music, lyrics-to-song tracks, instrumentals, background music, video soundtracks, jingles, multilingual songs, and reference-led arrangements from a clear creative brief. This AI music generator and AI song maker develops genre, mood, structure, singable lyrics, vocal direction, and production style, then creates reviewable audio with Suno 5.5 or another model you explicitly choose. Use it as an AI music studio for songwriting, an AI lyrics writer, text-to-music creation, BGM generation, brand music, podcast themes, game music, bilingual songs, or focused new versions of reference audio. Review vocals, pronunciation, duration, loop points, and arrangement fit after generation, then refine the strongest result.
Based on the Volcano Engine Doubao Music Generation API, supports instrumental BGM generation and vocal song generation.
Generate AI video, images, music, and voice-over in one connected creative flow, edit visual results, and keep everything you make easy to find. This all-in-one AI media generator turns text into images, images and references into video, ideas into songs or instrumentals, and scripts into narration, and can build a custom voice when a series needs its own sound. Working as an AI video generator, AI image generator, AI music generator, and AI voice generator in one place, it covers text-to-image, text-to-video, image-to-video, AI video editing, text-to-speech, multilingual voice-over, and voice cloning for social posts, product visuals, ads, courses, podcasts, short films, and cross-media campaigns. Reuse source media and finished files across formats, follow production progress, and keep every finished piece in one place.
AI-assisted short video creation. User selects topic, aspect ratio, and duration. AI guides through video generation using the user's own API key (Kling/Doubao etc.). One-time payment ¥16.90 per creation.