logo - what is real quiz Which one is real?

Which AI model fools people most?

Live ranking from real quiz answers. Genuine photographs are fixed at 1500 — anything above that line gets picked as “the real one” more often than reality does.

Limited to the models built into this quiz — not every model on the market, and not always the latest version of each. A model's absence here says nothing about how good it is.

Image models

19 models listed · 849 AI images in the quiz · click a column to sort
Confidence
Real image15001898111058%
1FLUX.2 [klein] 9BBlack Forest Labs13831258–1508461129%14
2Kling Image V3Kuaishou13141196–1432541019%10
3Z-ImageAlibaba12991263–133462611318%133
4Imagen 2Google12951245–13464947838%29
5Gemini 2.5 Flash Image / Nano BananaGoogle12951226–13631653018%14
6Juggernaut Pro FluxRunDiffusion12911232–13492534319%33
7Stable Diffusion 3.5 LargeStability AI12751169–1380791215%10
8Seedream 4.5ByteDance12731227–13194206616%33
9DALL·E 3OpenAI12631172–13541211613%27
10Seedream 4ByteDance12631166–1361991414%13
11SDXL-LightningByteDance12391149–13291311612%72
12Qwen-Image-MaxAlibaba12391104–137454713%12
13Juggernaut Flux LightningRunDiffusion12331136–13291121413%10
14FLUX.1 [schnell]Black Forest Labs12331192–12736588112%157
15Stable DiffusionStability AI12281149–13071912215%37
16GPT Image 1.5OpenAI12121149–12762833211%23
17FLUX.1 [dev]Black Forest Labs11891123–12553012910%41
18Kling IMAGE 3.0 OmniKuaishou11691085–1252201189%24
19FLUX.2 [max]Black Forest Labs1103964–124210066%10

Dimmed rows have fewer than 60 appearances — early evidence, not yet a settled rating.

A model joins this table once it has at least 10 images in the quiz and has appeared in at least 30 answers — below that, the number would say more about one lucky (or unlucky) image than about the model itself. New models and images are added regularly, so more entries — and firmer ratings — are on the way.

In the quiz but not yet listed: FLUX.2 LoRA Gallery: Realism (fal), FLUX1.1 [pro] (Black Forest Labs), Imagen 4 (Google), Lucid Realism (Leonardo AI), Titan Image Generator G1 v1 (Amazon) — fewer than 10 images in the quiz so far; FLUX.2 [pro] (Black Forest Labs), GPT Image 2 Medium (OpenAI), Grok Imagine Image Quality (xAI), Ideogram 4.0 Quality (Ideogram), Imagen 3 Fast (Google), Imagen 4 Ultra (Google), MAI-Image-2.5 (Microsoft AI), Nano Banana 2 Lite (Google), Nano Banana 2 with Web Search (Google), Qwen Image 2.0 Pro (Alibaba), Recraft V3 (Recraft), Reve 2.1 (Reve AI), Seedream 5.0 Pro (ByteDance) — shown too rarely so far, under 30 answers.

Last recomputed 4 Aug 2026, 03:15 · refreshed nightly

Add your own answers

Every round you play feeds straight into these numbers.

🎬 Video quizzes

Not in the ranking above yet — too few answers so far to rate video models.

How the rating works

Bradley-Terry, not a win count. Each round counts as a choice among the options actually on screen, so beating strong rivals is worth more than beating weak ones. Two-option rounds and four-option rounds therefore combine without their different odds distorting anything.

Read the range, not the position. Where two intervals overlap, the order between those models is not established. Dimmed rows have fewer than 20 appearances and are listed for completeness only.

Pool size is context. A model represented by three images is judged on those three images, however often they were shown.

Replays don't get extra weight. Only your first look at a round counts — after that you'd recognise it, and it stops being a fresh judgment.

Limits. Only answers from players who accepted analytics are counted, and the method assumes all players judge alike. This measures our image pool, not a model's ceiling.