pixverse-c1
PixVerse C1 video model (strong on action, VFX, and high-motion scenes) — one model covers text-to-video, image-to-video, first/last-frame-to-video, and reference-to-video: t2v when no images, i2v with first frame only, kf2v with first+last frames, r2v with reference images.
PixVerse C1 video model (strong on action, VFX, and high-motion scenes) — one model covers text-to-video, image-to-video, first/last-frame-to-video, and reference-to-video: t2v when no images, i2v with first frame only, kf2v with first+last frames, r2v with reference images.
Overview
| Field | Value |
|---|---|
| Model ID | pixverse-c1 |
| CLI | dlazy pixverse-c1 |
| MCP | pixverse-c1 (Claude Code surfaces this as mcp__dlazy__pixverse-c1) |
| Type | video |
| Execution | Async task; the CLI polls until completion (--no-wait returns generateId immediately) |
| Batch | Supports --batch <n> parallel fan-out |
Parameters
| Arg | Type | Required | Notes |
|---|---|---|---|
prompt | string | Yes | Prompt |
generation_mode | "components" | "frames" | No | Generation Mode(components=Components; frames=Frames); default "components" |
images | array<url> | No | Images; image (accepts URL, local path, or data: URL); max 7 items; only when !(generation_mode="frames") |
firstFrame | url | No | First Frame; image (accepts URL, local path, or data: URL); only when generation_mode="frames" |
lastFrame | url | No | Last Frame; image (accepts URL, local path, or data: URL); only when generation_mode="frames" |
resolution | "360P" | "540P" | "720P" | "1080P" | No | Resolution; default "720P" |
aspectRatio | "16:9" | "4:3" | "1:1" | "3:4" | "9:16" | "3:2" | "2:3" | "21:9" | No | Aspect Ratio; default "16:9" |
duration | "1" | "2" | "3" | "4" | "5" | "6" | "7" | "8" | "9" | "10" | "11" | "12" | "13" | "14" | "15" | No | Duration (s); default "5" |
generate_audio | "true" | "false" | No | Generate Audio; default "false" |
--input @file.jsonor--input '{...}'can supply all args at once; flags take precedence over--inputkeys.
CLI Examples
dlazy pixverse-c1 --help
dlazy pixverse-c1 --prompt "Write your prompt here" --images "https://example.com/reference1.jpg" "https://example.com/reference2.jpg"
dlazy pixverse-c1 --prompt "Write your prompt here" --images "./local-image.png"
dlazy pixverse-c1 --prompt "Write your prompt here" --images "https://example.com/reference1.jpg" "https://example.com/reference2.jpg" --dry-run
dlazy pixverse-c1 --prompt "Write your prompt here" --images "https://example.com/reference1.jpg" "https://example.com/reference2.jpg" --no-wait
dlazy pixverse-c1 --prompt "Write your prompt here" --images "https://example.com/reference1.jpg" "https://example.com/reference2.jpg" --batch 4Compose with a pipeline
dlazy gpt-image-2 --prompt "reference visual" \
| dlazy pixverse-c1 --images - --prompt "Write your prompt here"MCP
MCP is consumed by AI clients, not handwritten. Once dLazy is registered as an MCP server the tool appears in the client's tool list — Claude Code surfaces it as mcp__dlazy__pixverse-c1; generic clients (e.g. OpenClaw) call it as pixverse-c1.
See MCP setup for how to add the server.
Output
{
"outputs": [
{ "type": "video", "id": "o_...", "url": "https://files.dlazy.com/result.mp4", "mimeType": "video/mp4" }
]
}Media tools emit image / video / audio / file outputs. Use --output url to print only the URLs on stdout.
kling-v3-omni
Kling Omni video model, supports multiple reference images, duration, mode (std/pro), and optional audio. Suitable for highly controlled video synthesis tasks.
seedance-2.0
ByteDance's latest video generation model. Supports multi-modal reference (images, video, audio) to generate videos, as well as first/last frame and text-to-video modes.