Convert static character images into vivid action videos with Jimeng Dream Actor. 使用即梦 (Jimeng) Dream Actor,将静态人物图片转化为生动的动作视频。
Design & media
数字人视频 即梦 OmniHuman 1.5
Generate a digital human broadcast video from a portrait image plus audio or text, using the Jimeng OmniHuman 1.5 model.
What it does
Generates a digital human broadcast video by calling the Jimeng OmniHuman 1.5 model through the dLazy CLI command `dlazy jimeng-omnihuman-1.5`. Provide a single portrait image (URL or local path), optionally an audio file and a text prompt, and the API returns a generated video URL hosted on files.dlazy.com. Supports 720p or 1080p output, a fast-mode toggle, a dry-run cost preview, and async polling via `--no-wait`. Requires a dLazy API key obtained from the dlazy.com dashboard and configured via `dlazy login` or `dlazy auth set`.
When to use it
- Creating a talking-head video from one portrait image and an audio clip
- Producing a short avatar broadcast with only a text prompt and a face image
- Rendering a 1080p virtual presenter clip for a product demo
- Using --dry-run to preview payload and cost before submitting a render
The skill document
数字人视频 即梦 OmniHuman 1.5
English · 中文
Generate realistic digital human broadcast videos from portrait images and audio/text using Jimeng OmniHuman 1.5.
Trigger Keywords
- digital human
- jimeng omnihuman
- generate digital human video
- virtual human broadcast
Authentication
All requests require a dLazy API key. The recommended way to authenticate is:
dlazy login
This runs a device-code flow (also works in remote shells) and automatically saves your API key to the local CLI config — no manual copy/paste required.
Alternative: Set the Key Manually
If you already have an API key, you can save it directly:
dlazy auth set YOUR_API_KEY
The CLI saves the key in your user config directory (~/.dlazy/config.json on macOS/Linux, %USERPROFILE%\.dlazy\config.json on Windows), with file permissions restricted to your OS user account. You can also supply the key per-invocation via the DLAZY_API_KEY environment variable.
Getting Your API Key Manually
- Sign in or create an account at dlazy.com
- Go to dlazy.com/dashboard/organization/api-key
- Copy the key shown in the API Key section
Each key is scoped to your dLazy organization and can be rotated or revoked at any time from the same dashboard.
About & Provenance
- CLI source code: github.com/dlazyai/cli
- Maintainer: dlazyai
- npm package:
@dlazy/cli(pinned to1.2.3in this skill's install spec) - Homepage: dlazy.com
You can install on demand without persisting a global binary by running:
npx @dlazy/cli@1.2.3
Or, if you prefer a global install, the skill's metadata.clawdbot.install field declares the exact pinned version (npm install -g @dlazy/cli@1.2.3). Review the GitHub source before installing.
How It Works
This skill is a thin client over the dLazy hosted API. When you invoke it:
- Prompts and parameters you provide are sent to the dLazy API endpoint (
api.dlazy.com) for inference. - Any local file paths you pass to image / video / audio fields are uploaded to dLazy's media storage (
files.dlazy.com) so the model can read them — the same flow as any cloud-based generation API. - Generated output URLs returned by the API are hosted on
files.dlazy.com.
This is the standard SaaS pattern; the skill itself does not access network or filesystem resources beyond what the dLazy CLI already handles. See dlazy.com for the full service terms.
Usage
CRITICAL INSTRUCTION FOR AGENT:
Run the dlazy jimeng-omnihuman-1.5 command to get results.
dlazy jimeng-omnihuman-1.5 -h
Options:
--images [images...] Images [image: url or local path] (max 1)
--audio [audio...] Audio [audio: url or local path] (max 1)
--prompt [prompt] Prompt
--resolution [resolution] Resolution [default: 1080p] (choices: "720p", "1080p")
--fast_mode [fast_mode] Fast Mode [default: false]
--dry-run Print payload + cost estimate without calling API
--no-wait Return generateId immediately for async tasks
--timeout Max seconds to wait for async completion (default: "1800")
-h, --help display help for command
Any flag also accepts pipe references —
-(auto-pick from upstream stdin),@N(n-th output),@N.path(jsonpath into output),@*(all primary values),@stdin/@stdin:path(whole envelope). Seedlazy --helpfor details.
Output Format
{
"ok": true,
"result": {
"tool": "jimeng-omnihuman-1.5",
"modelId": "jimeng-omnihuman-1.5",
"outputs": [
{
"type": "image",
"id": "o_xxxxxxxx",
"url": "https://files.dlazy.com/result.png",
"mimeType": "image/png"
}
]
}
}
Async tasks (when
--no-waitis passed) returnoutputs: []and atask: { generateId, status }field instead. Usedlazy status --waitto poll.
Command Examples
# basic call:
dlazy jimeng-omnihuman-1.5 --prompt 'prompt text' --image '/path/to/image.png'
# complex call:
dlazy jimeng-omnihuman-1.5 --prompt 'prompt text' --image 'https://files.dlazy.com/image.png'
Error Handling
| Code | Error Type | Example Message |
|---|---|---|
| 401 | Unauthorized (No API Key) | ok: false, code: "unauthorized", message: "API key is missing or invalid" |
| 501 | Missing required parameter | error: required option '--prompt ' not specified |
| 502 | Local file read error | Error: Image file/Video file not found: C:\path\to\your\file |
| 503 | API request failed (no balance) | ok: false, code: "insufficient_balance" |
| 503 | API request failed (server error) | HTTP status code error (500 server crash) |
| 504 | Asynchronous task execution failed | === Generation Failed === / {Specific error reason returned by backend, for example "Prompt violates safety policy"} |
AGENT CRITICAL INSTRUCTION:
- If the execution result returns
code: "insufficient_balance", you MUST explicitly inform the user that their credits are insufficient and guide them to recharge: https://dlazy.com/dashboard/organization/settings?tab=credits- If the execution result returns
code: "unauthorized"or indicates missing API key, you MUST inform the user to get their API key from https://dlazy.com/dashboard/organization/api-key and save it usingdlazy auth setand resume the task.
Tips
Visit https://dlazy.com for more information.
Related skills
Generate dynamic videos based on a single first frame image and prompts using Jimeng. 使用即梦 (Jimeng) 首帧生视频模型,基于单张首帧图片和提示词生成动态视频。
Generate coherent transition videos using Jimeng's first and tail frame models. 使用即梦 (Jimeng) 首尾帧生视频模型,通过提供的第一帧和最后一帧图片生成连贯的过渡视频。
Text-to-image generation with Jimeng, quickly converting text to high-quality images. 使用即梦 (Jimeng) 进行文生图创作,快速将文字转化为高质量图像。
Generate videos from text prompts and reference images using the Kling v3 Omni model via the dLazy CLI.
Generate high-quality images with Doubao Seedream 4.5. Supports text-to-image and image-to-image. 使用豆包 Seedream 4.5 模型生成高质量图像,支持文生图与图生图。