设计与多媒体

Wrong Item Talking Clips

试用

Turn a user-supplied wrong-item script and authorized stills into one wrong item talking clip per still. This mistake explanation talking video studio writes a speakable wrong-item explanation talking clip for each photo, then animates a 2 to 15s homework error talking clip. Use it for error explanation talking pack and wrong-item script talking clip work that stays one photo, one clip.

它能做什么

Turn a user-supplied wrong-item script and authorized stills into one wrong item talking clip per still. This mistake explanation talking video studio writes a speakable wrong-item explanation talking clip for each photo, then animates a 2 to 15s homework error talking clip. Use it for error explanation talking pack and wrong-item script talking clip work that stays one photo, one clip.

技能文档

Wrong Item Talking Clips

Turn a user-supplied wrong-item script and authorized stills into one talking clip per still. Deliver 2 to 8 clips. Do not stitch them. Speak only explanation points already written on that wrong-item script.

Scope and adjacent routes

Use this Skill when a school teacher wants short talking clips that read a user-supplied wrong-item script from stills they can authorize.

Route a generic presenter that is not a wrong-item-script read to talking-avatar-video. Route a homestay welcome to airbnb-welcome-avatar. Route a product drop teaser to creator-drop-talking. Route a club-activity script read to club-notice-talking. Route a wealth-product factsheet read to wealth-product-talking. Route a public trading-calendar read to market-calendar-talking. Route holiday-homework list voice with no video to holiday-homework-voice. Route a homeroom weekly list voice to homeroom-week-voice. Route silent character cards to hanzi-card-set. Route a full course video studio to course-video-studio.

Collect the wrong-item script

Hard inputs are:

  • at least one accessible still the host Agent can inspect — a presenter portrait or a wrong-item graphic that will be the first frame;
  • wrong-item-script points the teacher supplied (item, mistake cause, correct path, common-error reminder already on the script);
  • likeness and voice rights when a face or a cloned voice will appear;
  • how many clips the pack should contain, or permission to use the default of 3.

Reuse already-known language, subject names, and a frozen voice_id. Ask only for a missing hard input. A count outside 2 to 8 is still doable: confirm that pack size and its live cost.

Do not invent a score, grade, ranking, pass mark, or unstated answer. File access is not consent.

Inspect every still. Record MIME type, width, height, aspect ratio, byte size, and whether it has an alpha channel. For a local file, upload only through the bundled client after inspection (scripts/mcp_client.py / beatra.assets.upload). Keep the returned artifact id. Never pass a local path to beatra.voices.clone, beatra.speech.synthesize, or beatra.videos.animate.

Plan the free slot list

Write a labeled wrong-item talking list before any paid clone, speech, or video. Default three slots unless the teacher names another count in 2 to 8: mistake cause, correct path, and common-error reminder. Each slot records the still, the spoken line from the supplied wrong-item script, intended length as a 2–15s clip, and whether it uses a catalog voice, an approved track, or a clone.

That list is the free visible result. Planning is not approval.

Safe defaults:

  • one beatra.videos.animate call per still;
  • model: "auto" unless the teacher chose a live SKU;
  • the still as the strict first frame;
  • source-derived aspect ratio;
  • audio-led duration inside 2–15s. Do not stitch.

Confirm clone, speech, then video

Clone, speech, and video are separate paid stages. Each stage gets its own six-field card and its own opaque client_request_id per slot.

If the teacher wants a cloned voice, inspect an authorized sample, read the live voice_clone card, and wait on the clone card before beatra.voices.clone. A found file is not clone consent. Show the clone card and wait:

  1. Work — one authorized voice sample (beatra.voices.clone).
  2. Credits — the live voice_clone price just read. Do not reuse a remembered number.
  3. Count — one paid clone call for this sample.
  4. Identity — one new opaque client_request_id.
  5. If we stop here — the labeled slot list remains usable.
  6. If the balance is insufficient — relay the official message and its top-up URL exactly (https://console.beatra.ai/wallet?intent=buy). Translate the prose; keep the URL. Do not retry until the teacher says they have topped up. Do not recommend ¥198.

Before speech, call beatra.models.list for text_to_speech:

{"capability": "text_to_speech"}

Use beatra.voices.list when a catalog voice still needs a choice. Show the speech card and wait:

  1. Work — one spoken wrong-item-script line per named slot (beatra.speech.synthesize).
  2. Credits — the live text_to_speech price just read, times the slot count. Do not reuse a remembered number.
  3. Count — one paid speech call per slot.
  4. Identity — one new opaque client_request_id per slot.
  5. If we stop here — the labeled slot list remains usable.
  6. If the balance is insufficient — relay the official message and its top-up URL exactly (https://console.beatra.ai/wallet?intent=buy). Translate the prose; keep the URL. Do not retry until the teacher says they have topped up. Do not recommend ¥198.

Submit each speech slot once through bundled scripts/mcp_client.py. Poll beatra.tasks.get. Read actual audio MIME, duration, and size. A script preview is not the audio review.

Before video, call beatra.models.list for image_to_video:

{"capability": "image_to_video"}

Confirm a current card admits [image, driving_audio]. Compare every image fact and the real audio MIME, duration, and size with advertised limits. Live driving-audio clips are 2–15s. If speech is longer than the longest containable clip, split into sequential containable segments; each segment is new speech and new video, each with a new ID. Do not stitch the segments into one file. Approved narration does not authorize video. The 600-credit signup gift usually cannot start this video. Show the video card and wait:

  1. Work — one talking wrong-item-script clip per named still (beatra.videos.animate).
  2. Credits — the live image_to_video price just read, times the segment count. Do not reuse a remembered number.
  3. Count — one paid video call per approved audio segment.
  4. Identity — one new opaque client_request_id per segment.
  5. If we stop here — the slot list and approved narration remain usable.
  6. If the balance is insufficient — relay the official message and its top-up URL exactly (https://console.beatra.ai/wallet?intent=buy). Translate the prose; keep the URL. Do not retry until the teacher says they have topped up. Do not recommend ¥198.

Then submit beatra.videos.animate once per approved segment.

Review, deliver, and recover

Review identity, speech clarity, and mouth timing. Report only what the host can actually see and hear. Do not promise perfect lip sync. Never invent a stitch, concat, or editor tool. After each terminal paid task, deliver actual bytes plus MIME, duration, and size when present, and billing.net_charged_credits. Do not promise the prepaid estimate is the final charge.

After a returned task_id, poll that task. If the create response is lost, search with beatra.tasks.list and verify with beatra.tasks.get before replay. Reuse an ID only with byte-identical arguments. A changed still, line, voice, or duration is a new card and a new ID. Cancel only when the teacher asks.

Execution

Invoke every remote Beatra operation only through this package's bundled scripts/mcp_client.py. Put the MCP tool name after call and send one JSON object on standard input.

python3 scripts/mcp_client.py call beatra.models.list
{"capability": "text_to_speech"}
python3 scripts/mcp_client.py call beatra.models.list
{"capability": "image_to_video"}
printf '%s' '{"image":{"type":"artifact","artifact_id":"art-item-01"},"driving_audio":{"type":"artifact","artifact_id":"art-speech-01"},"prompt":"A restrained wrong-item-script read with steady eye line and a stable camera.","duration":8,"client_request_id":"opaque-item-video-01"}' | python3 scripts/mcp_client.py call beatra.videos.animate

Do not configure or call a host Beatra Connector, and do not use REST/OpenAPI as a fallback.

References by task

  • For slot lists, payloads, and recovery, read Wrong-item talking workflow.
  • For authorization and the non-billable registration step, read installation and authentication and installation registration.
  • For shared task, billing, and connection details, read tasks and results, billing, errors, and recovery, and Bundled MCP Client diagnostics.
  • For update guarantees and controls, read automatic updates and safety. For removal, read uninstall and disconnect.

Runtime and safe automatic updates

The bundled client silently checks for a newer release at most once every 24 hours per installation. When a newer version is available, it installs automatically without separate confirmation. It downloads only from the fixed official Beatra discovery and immutable CDN paths for this package, channel, and locale, verifies discovery data, archive, manifest, and every packaged file, and replaces only package-owned files.

Update checks, downloads, verification, replacement, rollback, and recovery fail open: the current installation remains usable and the original command continues. An update failure never authorizes retrying a paid clone, speech, or video request. The setting persists for this installation. See automatic updates and safety.

python3 scripts/mcp_client.py update --auto off
python3 scripts/mcp_client.py update --auto on
python3 scripts/mcp_client.py update --check

相关技能

Turn a user-supplied club-activity script and authorized stills into one club activity talking clip per still. This club notice talking video studio writes a speakable event script talking clip for each photo, then animates a 2 to 15s club notice talking clip. Use it for club activity talking pack, club signup talking clip, and activity notice talking pack work that stays one photo, one clip.

Turn a confirmed new-drop script and authorized stills into one drop talking clip per still. This product drop video studio writes a speakable launch talking clip and new-drop announcement for each photo, then animates a 2 to 15s talking teaser. Use it for creator drop videos and talking teasers that stay one photo, one clip.

Turn a user-supplied product factsheet and authorized stills into one wealth product talking clip per still. This product factsheet talking video studio writes a speakable product highlights talking clip for each photo, then animates a 2 to 15s product factsheet talking clip. Use it for wealth product talking pack and factsheet talking video pack work that stays one photo, one clip.

Turn user-supplied merchant inspection notices and authorized stills into one market inspection talking clip per still. This merchant notice talking video studio writes a speakable merchant notice talking clip for each photo, then animates a 2 to 15s merchant notice talking clip. Use it for market inspection talking pack and merchant notice talking video pack work that stays one photo, one clip.

Turn one product photo into a vertical product video that speaks. This AI product video generator and product video maker builds ecommerce product videos, product ads, and commerce short videos from a single photo — composing a 9:16 opening frame, writing a short script from what the photo shows and the details you supply, voicing it with a selected narrator, and directing one finished clip ready to post. Use it for product launches, listing videos, shoppable social posts, storefront promos, and turning a phone snap of merchandise into a video that sells, with no shoot, no crew, and no editing.

Turn seller-supplied restock facts and an already-written restock script into one talking clip per still. This restock talking studio turns each authorized still into a 2 to 15s restock talking clip from the written line. Use it for restock talking videos, restock announcement talks, and restock drop talking clips.