AI models
AI image, video and voice models, tested
One page per model: what it costs, what the free tier really gives you, which settings matter and what they look like, and the same prompts run through every model so you can compare real output. Twelve pictures for every image model, nine clips for every video model, four scripts for every voice model. Each picture below is the model drawing itself.
Compared
The same prompts across every model, ranked with a verdict, and the pairs people ask about.
The best AI image generator, tested
Twelve prompts through 23 AI image models, Nano Banana, GPT Image, FLUX, Seedream, Ideogram, Recraft and Grok, every picture kept. Prices, speed and a verdict per prompt.
The best AI video generator, tested
The same nine prompts through 19 AI video models, Veo, Kling, Seedance, Sora, Wan, Hailuo, MiniMax, Grok and Gemini, every clip kept. Prices, waits and a verdict per prompt.
One prompt, every image model

Text in an image
Which AI image model renders text best

Product on white
Which AI image model makes the best product shot on white

Face and hands
Which AI image model does faces and hands best

Logo mark
Which AI image model makes the best logo

Busy scene
Which AI image model handles a busy scene best

Specific art style
Which AI image model holds an art style

Chart
Which AI image model draws an accurate chart

Wide aspect ratio
Which AI image model makes the best wide banner

Infographic with steps
Which AI image model makes the best infographic

Food photography
Which AI image model makes the best food photo

Hard composition
Which AI image model follows a hard composition

Edit an existing image
Which AI image model edits a picture best
One prompt, every video model

Product turntable
Which AI video model makes the best product turntable

Person speaking
Which AI video model makes the best talking head

Animated title card
Which AI video model can render text on screen

Camera move
Which AI video model handles a camera move best

Physics
Which AI video model gets physics right

Hands at work
Which AI video model handles hands and objects best

Busy scene
Which AI video model handles a busy scene best

Specific art style
Which AI video model holds an art style

Image to video
Which AI image-to-video model keeps the product intact
Head to head
Image models
Every image model runs the same 12 prompts.
Black Forest Labs

FLUX.2 Pro
$0.03 per 1k image via BFL API
Black Forest Labs' production model: 3 cents a megapixel, photoreal, no knobs to tune.

FLUX.2 Flex
$0.05 per 1k image via BFL API
The FLUX.2 tier with knobs: steps and guidance you can tune, better typography, 5 cents a megapixel.

FLUX.2 Max
$0.07 per 1k image via BFL API
The top FLUX.2 tier: 7 cents a megapixel for BFL's most realistic and precise output and edits.
OpenAI

GPT Image 2.5 Flare
$0.05268 per 1024 by 1024, high quality via fal
OpenAI's fast default since September 2026: GPT Image 2's tricks, faster, and a quarter of the price at high quality.

GPT Image 2.5 Sunburst
$0.05268 per 1024 by 1024, high quality via fal
OpenAI's precision tier: same prices as Flare, slower, built for edits where fidelity to the original matters most.

GPT Image 2
$0.211 per 1024 by 1024, high quality via fal
The model behind ChatGPT's image generator from April to September 2026: exact sizes, cut-outs, precise edits, slow at high quality.
xAI

Grok Imagine 2.0
$0.04 per image via xAI API
xAI's current image model: cheap, fast, sharp small text, and edits from up to three reference images.

Grok Imagine Pro
$0.05 per image via xAI API
xAI's quality tier: a cent more than 2.0 per image, no quality dial, the same shapes and edits.

Grok Imagine
$0.02 per image via xAI API
The original 2025 Grok Imagine image model: two cents an image and still on the API, superseded by 2.0.
HiDream.ai
Ideogram
Krea
Google DeepMind

Nano Banana 2
$0.067 per 1k image via Gemini API
Google's everyday model and the Gemini app default: near-Pro quality at Flash speed and half the price.

Nano Banana 2 Lite
$0.0336 per 1k image via Gemini API
The cheapest and fastest Nano Banana: about $0.034 an image and a few seconds each, at a fixed 1K.

Nano Banana Pro
$0.134 per 1k or 2k image via Gemini API
Google's premium tier: the best at text and complex briefs, the slowest and dearest of the three.
Alibaba
Recraft

Recraft V4.1
$0.035 per raster image via Recraft API
Recraft's design-first model: brand colours as an input, real SVG in the vector variant, 3.5 cents an image.

Recraft V3
$0.04 per raster image via Recraft API
Recraft's 2024 model, still on sale: 70 named styles, an image-to-image endpoint, 4 cents an image.
ByteDance

Seedream 5.0 Pro
$0.0675 per image up to 1536 by 1536 in area via fal
ByteDance's flagship: reasons before it draws, text in 14 languages, dense layouts, under 7 cents at 1K.

Seedream 5.0 Lite
$0.035 per image via fal
ByteDance's fast tier: 3.5 cents an image, 2K to 4K output, up to ten reference images.

Seedream 4.5
$0.04 per image via fal
ByteDance's late-2025 model: 4 cents an image, 2K and 4K, generation and editing in one model.
Video models
Every video model runs the same 9 clips, five seconds each where the model allows it.
Google DeepMind
xAI
MiniMax

H3 Max Turbo
$0.02 per second at 768p via fal
H3 Max Turbo is fal’s lower-priced post-trained MiniMax H3 variant.

H3 Max
$0.04 per second at 768p via fal
fal post-trains MiniMax H3 for prompt adherence and faster video generation.

MiniMax H3
$0.06 per second at 768p via fal
MiniMax's July 2026 omni-modal model: native stereo sound, up to 2K and 15 seconds, open weights, from $0.05 a second.

Hailuo 2.3
$0.28 per six-second clip, standard 768p via fal
MiniMax's October 2025 model at a flat price: $0.28 for six seconds at 768p, no audio, no settings to learn.
Kuaishou

Kling 3.0 Standard
$0.126 per second, audio on via fal
Kuaishou's everyday tier: 15-second clips with audio, multi-shot storyboards, $0.126 a second with sound on fal.

Kling 3.0 Pro
$0.168 per second, audio on via fal
Kuaishou's higher-fidelity tier: the same 15-second, multi-shot envelope as Standard at $0.168 a second with audio.
Lightricks
Luma AI
PixVerse
ByteDance

Seedance 2.0
$0.3034 per second at 720p, 16:9 via fal
ByteDance's February 2026 model: native audio, up to 15 seconds and 4K, about $1.50 for five seconds at 720p.

Seedance 2.5
$0.4622 per second at 720p, 16:9 via fal
ByteDance's July 2026 model: one continuous shot up to 30 seconds, 50 references, about $2.30 for five seconds at 720p.
OpenAI
Google DeepMind

Veo 3.1 Fast
$0.10 per second with audio, 720p via Gemini API
Google's speed tier and the sensible default: the same audio, shapes and lengths as Veo 3.1 at $0.15 a second with sound.

Veo 3.1
$0.40 per second with audio, 720p or 1080p via Gemini API
Google's flagship: native audio, reference images, 4K, and the dearest per second here at $0.40 with sound.

Veo 3.1 Lite
$0.05 per second with audio, 720p via Gemini API
Google's budget tier from March 2026: 720p and 1080p only, $0.05 a second with audio.
Alibaba

Wan 3.0
$0.10 per second at 720p via fal
Alibaba's August 2026 model: 2 to 30 seconds in one pass with audio, $0.10 a second at 720p, hosted only.

Wan 2.7
$0.10 per second at 720p via fal
Alibaba's April 2026 model: 2 to 15 seconds with native audio and a thinking mode, $0.10 a second at 720p, hosted only.
Voice models
Every voice model reads the same 4 scripts, plain text with no vendor tags.
ElevenLabs
Google DeepMind
hexgrad
MiniMax
Every image model page runs the same 12 prompts (prompt set 2026-09-09): text in an image, a product on white, a face and hands, a logo, a busy scene, a named art style, a chart, a wide banner, an infographic, food, a hard composition and an edit of an earlier result. Every video model page runs the same 9 clips (clip set 2026-09-09): a product turntable, a person speaking, a title card, a camera move, a spill, hands at work, a busy scene, a named art style and a clip started from one of the image results. Output is shown as it came back.











