Catalog

Models

25 models across video, image, audio, and chat — one API, one bill, pay only for what you run.

Seedance 2.0HOT

Seedance 2.0

ByteDance$0.13 / second

Seedance 2.0. Image-to-video and Text-to-video with synchronized audio.

Happyhorse 1.0HOT

Happyhorse 1.0

Alibaba$0.11 / second

Alibaba's top-of-arena text-to-video and image to video model. Currently leads the Arena Elo leaderboard.

Qwen Image 2512

Qwen Image 2512

Alibaba$0.025 / image

Qwen image generation. Bilingual prompt support; strong on Chinese typography.

Nano Banana 2HOT

Nano Banana 2

Google$0.039 / image

Gemini Flash Image v2 · text-to-image & edit. Sharper photoreal than nano-banana.

GPT Image 2HOT

GPT Image 2

OpenAI$0.05 / image

Next-gen image generation & editing with stronger compositional control and text rendering.

Flux 2HOT

Flux 2

Black Forest Labs$0.02 / image

Next-gen flagship with sharp subject fidelity and tight prompt adherence, in base and highest-fidelity tiers.

Nano Banana

Nano Banana

Google$0.039 / image

Gemini 2.5 Flash Image. Persisted to Infer S3.

Ideogram 4

Ideogram 4

Ideogram$0.02 / image

Ideogram 4 — best-in-class typography and graphic design, in Turbo, Default, and Quality tiers.

Uni 1

Uni 1

Luma Labs$0.04 / image

Strong photoreal generation & editing with up to 8 image refs, available in standard and higher-fidelity tiers.

Qwen Image Max Preview

Qwen Image Max Preview

Alibaba$0.025 / image

Qwen Image Max preview. Higher-fidelity Qwen image variant; bilingual prompt support.

Wan 2.7

Wan 2.7

Alibaba$0.04 / image

Alibaba's Wan 2.7 flagship image generation, available in pro and fast tiers.

Seedream 4

Seedream 4

ByteDance$0.03 / image

ByteDance's flagship image generator with strong photoreal output, for generation & editing.

Seedream 4.5

Seedream 4.5

ByteDance$0.03 / image

Newer Seedream revision with refined photoreal output, for generation & editing.

SAM 3

SAM 3

Meta$0.01 / image

Open-vocabulary detection. Returns bounding boxes for any prompt.

Kling V3

Kling V3

Kuaishou$0.07 / second

Kuaishou's Kling 3.0 image-to-video and text-to-video, in Standard, Pro, and 4K tiers, plus native-audio variants.

Veo 3.1

Veo 3.1

Google$0.09 / second

Google's Veo 3.1 text-to-video with synchronized audio, in full and fast tiers.

Happyhorse 1.1

Happyhorse 1.1

Alibaba$0.11 / second

Alibaba's latest HappyHorse — top-of-arena text-to-video and image-to-video, now on the 1.1 release.

Wan 2.2

Wan 2.2

Alibaba$0.04 / second

Alibaba's Wan 2.2 Flash — fast image-to-video for rapid iteration.

LTX 2.3

LTX 2.3

Lightricks$0.07 / second

Lightricks' LTX 2.3 image-to-video and text-to-video across 1080p, 1440p, and 4K, in fast and higher-fidelity tiers.

Eleven v3HOT

Eleven v3

ElevenLabs$0.09 / 1K chars

ElevenLabs Multilingual v3. Natural prosody, emotional range, 70+ languages, inline audio tags for laughter / whisper / sigh.

Qwen 3.6

Qwen 3.6

Alibaba (Qwen)$1.30 in · $7.80 out / 1M tokens (≤128K ctx)

Chat completions — strong reasoning, long context, multilingual — in flagship, mid-tier, and fast low-latency tiers.

Nano Banana 2 LiteHOT

Nano Banana 2 Lite

Google$0.034 / image

Gemini 3.1 Flash-Lite Image · text-to-image & edit. Fast, low-cost 1K image generation.

Seedream 5.0 Pro

Seedream 5.0 Pro

ByteDance$0.09 / image

Latest flagship Seedream with 2K photoreal output, for generation & editing.

Qwen 3.7 Max Preview

Qwen 3.7 Max Preview

Alibaba (Qwen)$1.30 in · $7.80 out / 1M tokens (≤128K ctx)

Chat completions — strong reasoning, long context, multilingual.

MiniMax-M3

MiniMax-M3

MiniMax$0.30 in · $1.20 out / 1M tokens (≤512K ctx)

Multimodal long-context chat from MiniMax with adaptive reasoning, served OpenAI-compatible.