Image AI Models
Every image generation model ranked by blind human-preference votes and what it costs. The value frontier shows which ones are actually worth it.
Arena Elo + pricing from Replicate, fal.ai & makers · reviewed 2026-06-01
The verdict
gpt-image-2 is the highest-rated image model right now.
Best value (top quality at each price point): flux-2-dev, flux-2-pro, grok-imagine-image, gpt-image-2. Best open-weights: ideogram-4.0-quality.
Ranked by the crowd
Quality vs. cost, measured
Ranked by Arena Elo (blind, head-to-head human preference votes), paired with the cheapest published price we can find per model.
| # | Model | Maker | Arena Elo | Cost / image |
|---|---|---|---|---|
| 1 | gpt-image-2 (medium)Value pickClosed | OpenAI | 1385±5 | $0.047Replicate |
| 2 | reve-2.0Closed | Reve | 1271±6 | — |
| 3 | mai-image-2.5Closed | Microsoft AI | 1257±5 | — |
| 4 | gpt-image-1.5-high-fidelityClosed | OpenAI | 1240±3 | $0.05Replicate |
| 5 | gemini-3-pro-image-preview (nano-banana-pro)Closed | 1232±5 | $0.15fal.ai | |
| 6 | grok-imagine-image-qualityClosed | xAI | 1229±4 | $0.07fal.aifrom $0.02 |
| 7 | ideogram-4.0-qualityOpen | Ideogram | 1207±6 | — |
| 8 | uni-1.1-maxClosed | Luma AI | 1188±6 | — |
| 9 | mai-image-2Closed | Microsoft AI | 1183±5 | — |
| 10 | uni-1.1Closed | Luma AI | 1177±5 | — |
| 11 | grok-imagine-imageValue pickClosed | xAI | 1173±3 | $0.02Replicate |
| 12 | recraft-v4.1-utility-proClosed | Recraft | 1169±11 | — |
| 13 | flux-2-maxClosed | Black Forest Labs | 1162±4 | $0.03Replicate |
| 14 | grok-imagine-image-proClosed | xAI | 1161±4 | $0.02Replicate |
| 15 | flux-2-flexClosed | Black Forest Labs | 1156±3 | $0.06Replicate |
| 16 | flux-2-proValue pickClosed | Black Forest Labs | 1155±3 | $0.015Replicate |
| 17 | reve-v1.5Closed | Reve | 1154±4 | — |
| 18 | gemini-2.5-flash-image-preview (nano-banana)Closed | 1151±3 | $0.039fal.ai | |
| 19 | hunyuan-image-3.0Open | Tencent | 1151±3 | $0.08Replicate |
| 20 | flux-2-devValue pickOpen | Black Forest Labs | 1148±5 | $0.014Replicate |
| 21 | imagen-ultra-4.0-generate-001Closed | 1148±4 | $0.06Replicate | |
| 22 | seedream-4.5Closed | Bytedance | 1147±3 | $0.04Replicate |
| 23 | seedream-4-2kClosed | Bytedance | 1141±7 | $0.03Replicate |
| 24 | wan2.6-t2iClosed | Alibaba | 1134±3 | — |
| 25 | seedream-5.0-liteClosed | Bytedance | 1132±4 | $0.035Replicate |
| 26 | recraft-v4.1-proClosed | Recraft | 1130±11 | — |
| 27 | imagen-4.0-generate-001Closed | 1129±3 | $0.04Replicate | |
| 28 | qwen-image-2512Open | Alibaba | 1128±4 | $0.025Replicate |
| 29 | krea-2-mediumClosed | Krea | 1120±6 | — |
| 30 | hidream-o1-imageOpen | HiDream | 1118±5 | — |
| 31 | seedream-4-falClosed | Bytedance | 1117±7 | $0.03Replicate |
| 32 | wan2.5-t2i-previewClosed | Alibaba | 1117±3 | — |
| 33 | gpt-image-1Closed | OpenAI | 1115±3 | — |
| 34 | recraft-v4Closed | Recraft | 1113±4 | $0.04Replicate |
| 35 | seedream-4-high-res-falClosed | Bytedance | 1113±3 | $0.03Replicate |
| 36 | gpt-image-1-miniClosed | OpenAI | 1109±3 | $2.50provider |
| 37 | wan2.7-image-proClosed | Alibaba | 1102±5 | $0.03Replicate |
| 38 | krea-2-largeClosed | Krea | 1100±6 | — |
| 39 | wan2.7-imageClosed | Alibaba | 1099±5 | $0.03Replicate |
| 40 | mai-image-1Closed | Microsoft AI | 1093±4 | — |
| 41 | seedream-3Closed | Bytedance | 1082±5 | $0.03Replicate |
| 42 | z-image-turboOpen | Alibaba | 1082±6 | — |
| 43 | flux-1-kontext-maxClosed | Black Forest Labs | 1074±3 | $0.08Replicate |
| 44 | flux-2-klein-9bOpen | Black Forest Labs | 1069±3 | — |
| 45 | qwen-image-prompt-extendOpen | Alibaba | 1060±3 | $0.025Replicate |
| 46 | flux-1-kontext-proClosed | Black Forest Labs | 1059±3 | $0.04Replicate |
| 47 | imagen-3.0-generate-002Closed | 1058±3 | $0.05Replicatefrom $0.04 | |
| 48 | Cosmos3-Super-Text2ImageOpen | Nvidia | 1057±6 | — |
| 49 | Cosmos3-Super-Text2Image (Agentic)Open | Nvidia | 1057±5 | — |
| 50 | qwen-imageOpen | Alibaba | 1057±3 | $0.025Replicate |
| 51 | ideogram-v3-qualityClosed | Ideogram | 1049±4 | $0.06fal.ai |
| 52 | photonClosed | Luma AI | 1035±4 | $0.03Replicate |
| 53 | p-imageClosed | Pruna | 1034±4 | $5.00Replicate |
| 54 | flux-2-klein-4bOpen | Black Forest Labs | 1030±3 | $1.00Replicate |
| 55 | runway-gen4Closed | Runway | 1025±5 | — |
| 56 | recraft-v3Closed | Recraft | 1021±4 | $0.08fal.aifrom $0.04 |
| 57 | flux-1.1-proClosed | Black Forest Labs | 1016±3 | — |
| 58 | ideogram-v2Closed | Ideogram | 1013±4 | $0.08Replicate |
| 59 | lucid-originClosed | Leonardo AI | 1013±3 | $0.076Replicate |
| 60 | glm-imageOpen | Z.ai | 1011±9 | — |
| 61 | gemini-2.0-flash-preview-image-generationClosed | 975±3 | — | |
| 62 | flux-1-dev-fp8Open | Black Forest Labs | 970±4 | — |
| 63 | dall-e-3Closed | OpenAI | 968±4 | — |
| 64 | flux-1-kontext-devOpen | Black Forest Labs | 940±4 | $0.025Replicate |
| 65 | stable-diffusion-v35-largeOpen | Stability AI | 938±4 | $0.065Replicate |
| 66 | bagelOpen | Bytedance | 898±6 | — |
Prices are shown as each provider publishes them at default settings (per image or per megapixel, not normalized across resolution), from Replicate, fal.ai, and the model makers directly. The Value pickbadge marks models where nothing higher-rated costs less. “From” marks a cheaper source than the one shown; a blank cost means no comparable per-image price was found (often token-priced API-only models).
Questions builders actually ask
Real demand, honest answers
How do I train a character LoRA that stays consistent?
Consistency comes from caption discipline more than dataset size: describe what varies (pose, lighting) and stay silent on what should stay fixed (the character's face). A tight, well-captioned set of 15–30 images usually beats a large noisy one. Train, generate a test sheet, and re-caption the failures.
Can I run Stable Diffusion on a 6GB GPU?
Yes for SDXL-class and smaller models like Z-Image, which is why low-VRAM cards stay popular with local builders. FLUX will technically load with offloading but runs are slow and cramped. Treat 6GB as an SDXL/Z-Image card, not a FLUX one.
ROCm or CUDA for ComfyUI?
CUDA (NVIDIA) is the smoother path; most nodes, fine-tunes, and tutorials assume it. ROCm (AMD) works in ComfyUI and has improved, but expect more setup friction and the occasional unsupported node. If you're buying for image generation specifically, NVIDIA saves you time.
FLUX vs SDXL for anime?
SDXL still has the deeper anime fine-tune and LoRA library, so a style checkpoint plus a LoRA often wins for that specific look. FLUX leads on photoreal and prompt adherence but its anime ecosystem is thinner. Pick by the style supply, not the base model's raw quality.
Is a closed model like GPT-IMAGE-2 worth it over local?
If you don't need local control, training, or privacy, a closed model removes the entire GPU and setup problem and often gives a stronger one-shot result. The case for local is control: LoRAs, ControlNet, custom workflows, and no per-image cost. Most builders end up using both for different jobs.
How the ranking is built
Quality is blind, head-to-head human-preference votes from the model arena, not vendor claims. Each model is paired with its published price at default settings, with the source shown. The value frontier flags the models actually worth buying: the ones where nothing higher-rated costs less.
- Quality from blind head-to-head votes, not vendor claims.
- Published price per model at default settings, source shown.
- Value frontier flags the best quality at each price point.
- Updated as new models and prices land.