🎬 DALANG
Agentic Service Provider Β· OKX.AI
Script β†’ Storyboard β†’ Video

Turn one line of script into a narrated animatic.

DALANG breaks your brief into a shot list, generates every frame, adds voiceover, and assembles a ready-to-share video β€” in one call. No editor, no timeline, no waiting.

Pipeline

One brief in, a finished animatic out

01

Break down

An LLM turns your brief into a structured shot list β€” scene, image prompt, voiceover line, camera motion.

02

Generate frames

Each shot becomes a cinematic still in the art direction you specify.

03

Narrate

Every line is voiced with text-to-speech, per shot.

04

Assemble

Ken Burns motion + audio, stitched by ffmpeg into a single MP4.

brief ─► LLM (shot list) ─┬─► image (frame) ──┐ β”œβ”€β–Ί TTS (voiceover) ─┼─► Ken Burns (ffmpeg) ─► animatic.mp4 β””β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”˜
Built for the agent economy

Live on OKX.AI, pay-per-call

A2MCP

Registered as a standardized pay-per-call service β€” one call, one animatic.

Settled on X Layer

Paid in USDT/USDG stablecoins, instantly, on OKX's chain.

Category: Art Creation

Targeting Artistic Excellence + Social Buzz at the OKX.AI Genesis Hackathon.

$0.49per animatic Β· cost β‰ˆ $0.05–0.08 Β· healthy margin
Authenticity in the age of AI slop

Verify provenance in your browser

Every render returns a content fingerprint β€” a SHA-256 and a CIDv1 (bafkrei…). Drop a DALANG animatic below and this page recomputes them locally (nothing is uploaded), so you can confirm the exact authored bytes and match the content_cid anchored on X Layer.

Choose an .mp4 to fingerprint it.
Under the hood

One key, three model types

LLM breakdown, image generation, and voice synthesis all run through a single OpenAI-compatible provider; assembly is ffmpeg. Every provider is a swappable env var.

LLM: shot-list reasoning Image: cinematic frames TTS: per-shot voiceover ffmpeg: Ken Burns + concat MCP: A2MCP tool