
Seedream V5 Pro Wins, but 4.5 Is Close
V5 Pro narrowly beat Seedream 4.5 across 19 valid tests. V5 Lite trailed.
Read guide →Model research hub
First-party image model comparisons built from one input, one exact prompt per row, and verified model and output hashes.
Quick answer
Each current comparison uses one exact input for every model attempt in a row. Prompt hash, input key and hash, profile, aspect ratio, run ID, exact model ID, request fingerprint, and output hash are verified before publication.
At a glance
| Test | Models covered | What is controlled |
|---|---|---|
| Explicit Seedream benchmark | Seedream 4.5, Seedream V5 Lite, Seedream V5 Pro | 20 prompts; frozen Mila or Lily input per row; 59 control-valid outputs; one preserved invalid control; one blinded reviewer |
| Realistic mirror-selfie blind test | Nano Banana 2, GPT Image 2 Medium, Grok, Seedream V5 Lite, Seedream V5 Pro | 1 exact prompt; one Lily input; 4 verified outputs; 1 visible block; anonymous reader vote |
| One prompt across five models | Seedream 4.5, GPT Image 2 Medium, Grok, Seedream V5 Lite, Seedream V5 Pro | 5 prompts; one Mila input; 25 verified outputs; blind reader vote; exact prompt, request, model, and output hashes |
| Adult-fashion moderation boundary test | Nano Banana 2, GPT Image 2 Medium, Seedream 4.5 | 3 prompts; 9 attempts; 7 images; 2 GPT safety blocks; no broader moderation ranking |
| Seedream version test | Seedream 4.5, V5 Lite, V5 Pro | 5 exact prompts and one input across all three versions |
| Seedream V5 Lite vs Pro | V5 Lite and V5 Pro | 5 priority prompts, including the kitchen portrait, with identical inputs per pair |
Curated guides

V5 Pro narrowly beat Seedream 4.5 across 19 valid tests. V5 Lite trailed.
Read guide →
Vote through five same-input comparisons across Seedream 4.5, GPT Image 2 Medium, Grok, Seedream V5 Lite, and Seedream V5 Pro, then reveal live results.
Read guide →
Controlled 18-prompt benchmark: Nano Banana 2 generated 83% of attempts versus 67% for Nano Banana and Nano Banana 2 Lite, with 14 blind visual comparisons.
Read guide →
Five exact prompts tested on one Mila input across Seedream 4.5, GPT Image 2 Medium, Grok, Seedream V5 Lite, and Seedream V5 Pro.
Read guide →
Nano Banana 2, GPT Image 2 Medium, and Seedream 4.5 tested with one Mila input: 9 attempts, 7 images, and 2 recorded GPT safety blocks.
Read guide →
Five exact prompts tested with one Mila input across Seedream 4.5, V5 Lite, and V5 Pro.
Read guide →
Five exact prompts tested with one Mila input across Seedream V5 Lite and V5 Pro.
Read guide →Choose well
Every row starts with the input. The images that follow are ordered by exact model ID, and the prompt text appears below the grid.
These pages replace the older unverifiable rows with a July 2026 controlled rerun. Historical outputs without exact-input proof are excluded from the comparison grids.
Model providers change over time. Treat these as dated results and rerun the same locked method before making a high-volume production decision.
FAQ
In the July 17 realistic mirror-selfie test, four blinded visual evaluations ranked Seedream V5 Pro first and Grok second. That is a one-prompt result, so compare face fidelity, anatomy, texture, and prompt accuracy for your own use case.
In the July 29 controlled explicit-image benchmark, Seedream V5 Pro led Seedream 4.5 by one decided win across 19 fair three-way rows; V5 Lite trailed. One blinded reviewer scored one output per model per prompt, so treat the result as directional rather than universal.
First compare models with the exact same verified input file. For a recurring character across many scenes, a trained LoRA can also help, but its performance should be evaluated in a separate controlled run.
Browse public prompt packs, choose a look, and remix it with an approved photo or trained character.
Explore prompt packs