Dialogue scenes
Write the lines in quotes. The model performs them, lip-synced.
EXAMPLETwo friends argue quietly across a café table — “You knew, didn’t you?” — rack focus between faces, espresso machine hissing.
ByteDance's new flagship: 30-second scenes, sound built in. Unlimited on Infer for $49/month — the best plan in AI video. Free for 24 hours.
Write the lines in quotes. The model performs them, lip-synced.
EXAMPLETwo friends argue quietly across a café table — “You knew, didn’t you?” — rack focus between faces, espresso machine hissing.
Hook, demo, CTA — a complete vertical ad in one 30-second take.
EXAMPLECreator talks straight to camera about her morning routine, handheld selfie framing, kitchen daylight, natural delivery.
Lock a face across every shot with up to 30 reference images per run.
EXAMPLEThe red-jacketed courier from the reference set weaves her bike through evening traffic, face steady in every frame.
Movement lands on the beat, not near it. Feed it up to 10 reference tracks.
EXAMPLEA dancer hits every beat of the track on a rooftop at blue hour, city bokeh behind, camera locked off.
Block with white-models and references, then render the real scene.
EXAMPLEA single steadicam follows two detectives arguing down the precinct corridor, phones ringing behind office doors.
Follows, orbits, and slow reveals no 5-second model can hold.
EXAMPLEOne unbroken take: the camera follows a courier through a lantern-lit night market, vendors calling out, rain just ended.
A live cross-section of the model's range — portraits, products, typography, illustration, fashion, cinematic. Hover any tile to read its prompt.
Pay only for successful generations. No idle, no minimums, no per-seat. Volume discounts kick in at 10K req/mo.
One key. One bill. One SDK shape — across 50+ models. Pay only for what you use.