logo - what is real quiz Which one is real?

AI Image Models β€” Full Overview

21 companies Β· 46 models listed Β· 90+ versions documented

Same Prompt. Available Model Samples.

Choose a prompt to compare available model outputs generated from the exact same input. Models without a trustworthy sample are listed beneath the grid.

Prompt

A cat curled up asleep on an open book, next to it a steaming cup of tea, cozy living room atmosphere.

Missing samples: Imagen 2, Nano Banana 2 Lite, Nano Banana 2 with Web Search, Imagen 4 Ultra, FLUX.2 [pro], Qwen Image 2.0 Pro, Ideogram 4.0 Quality, Lucid Realism, Reve 2.1, MAI-Image-2.5, Grok Imagine Image Quality, Titan Image Generator G1 v1.

Image Duel Β· Free to play
🎯

Can you spot the AI?

You've seen how these models compare β€” now put yourself to the test. Pick any model, get 10 rounds of real vs. AI, and find out if you can beat it.

Play now 46 models Β· 10 rounds

⭐ Big Players

FLUX.2 [max]

Black Forest Labs

The premium tier of Black Forest Labs' FLUX.2 series. Maximum quality variant with improved photorealism, superior facial detail, accurate anatomy, and enhanced text rendering β€” the highest-quality FLUX.2 variant before FLUX 3.

Maximum photorealismSuperior facial detailEnhanced text in images

Seedream 4.5

ByteDance / BytePlus

ByteDance's image generation model, marketed under the BytePlus enterprise brand. Builds on ByteDance's deep video and media processing expertise to produce cinematic, high-quality imagery with excellent composition and color accuracy.

Cinematic color gradingEnterprise APIStrong composition
Google

Gemini 2.5 Flash Image (Nano Banana)

Google

Google's Gemini 2.5 Flash Image, known publicly as Nano Banana. It became the reference point for conversational image editing β€” changing one element of a picture while the rest, including faces, stays consistent across edits.

Conversational image editingCharacter consistencyGemini 2.5 Flash

Model Cards

OpenAI

GPT Image 2 Medium

OpenAI

OpenAI's GPT Image 2 served with the exact medium quality setting used by the Arena text-to-image leader. It combines strong prompt adherence, typography, and photographic detail at lower cost than the high tier.

OpenAI

GPT Image 1.5

OpenAI

OpenAI's second-generation GPT Image model. Improved consistency, photorealism and editing over GPT Image 1; it remains documented but is deprecated and scheduled for API removal on December 1, 2026.

OpenAI

DALLΒ·E 3

OpenAI

OpenAI's historical DALLΒ·E 3 model, still available through Bing Image Creator. Its API ended May 12, 2026 and ChatGPT now uses GPT Image 2.

Google

Imagen 2

Google

Google DeepMind's second-generation text-to-image model, widely integrated across Google products including Google Slides, Workspace, and Search Generative Experience. Substantially improved photorealism and text rendering over Imagen 1, with a focus on diverse, high-quality imagery for productivity applications.

Google

Imagen 3 Fast

Google

Google's speed-optimized Imagen 3 text-to-image release, served through the exact fal endpoint fal-ai/imagen3/fast. The endpoint is deprecated in fal's current catalogue but remains a distinct model identity.

Google

Imagen 4

Google

Google DeepMind's fourth-generation text-to-image model. Substantially improved photorealism, fine-grained detail, and text accuracy over previous versions. Particularly strong in landscape photography, architecture, and product visualization.

Stability AI

Stable Diffusion 3.5 Large Turbo

Stability AI

A fast, distilled variant of Stability AI's SD 3.5 architecture. Generates high-quality images in fewer sampling steps without sacrificing significant quality. Improved typography, compositional understanding, and human anatomy over SD 2.x.

Stability AI

Stable Diffusion

Stability AI

The open-source model that democratized AI image generation. Because it's fully open, hundreds of fine-tuned versions exist. Quality varies widely β€” some versions rival top proprietary models, others show clear artifacts especially in hands and faces.

Stability AI

Stable Diffusion 3.5 Large

Stability AI

Stability AI's Stable Diffusion 3.5 Large multimodal diffusion transformer, with improved image quality, typography, complex prompt understanding, and broad style coverage.

FLUX.2 [klein] 9B

Black Forest Labs

The 9-billion-parameter FLUX.2 Klein variant used through fal for low-cost text-to-image generation. These quiz outputs retain the exact YAML prompt as their source prompt.

FLUX1.1 [pro]

Black Forest Labs

Black Forest Labs' enhanced commercial API model, succeeding FLUX.1 Pro with 6Γ— faster generation and improved image quality. Stronger prompt adherence, better detail rendering, and improved color accuracy β€” accessible via the BFL API and partnered platforms.

FLUX.1 [dev]

Black Forest Labs

Black Forest Labs' open-weights guidance-distilled model for non-commercial use. Derived from FLUX.1 Pro via guidance distillation, it produces high-quality outputs comparable to the Pro variant but optimized for local deployment and research.

FLUX.1 [schnell]

Black Forest Labs

Black Forest Labs' open-source (Apache 2.0) fast generation model, designed for local development and personal use. Produces results in 1–4 steps at the cost of some fine detail and photorealism compared to Dev or Pro variants.

SDXL-Lightning

ByteDance

ByteDance's progressive adversarial diffusion-distilled SDXL model, designed to generate images in very few inference steps while retaining broad SDXL compatibility.

Seedream 4

ByteDance / BytePlus

ByteDance's fourth-generation image model, focused on prompt adherence, typography, image generation and editing.

Seedream 5.0 Pro

ByteDance

ByteDance's Seedream 5.0 Pro flagship text-to-image release, with deep prompt understanding, photographic realism, dense-layout control, and multilingual typography.

Kling IMAGE 3.0 Omni

Kuaishou

Kuaishou's multimodal image-and-video model with a unified architecture. Exceptional at consistent character generation, style-faithful rendering, and maintaining visual identity across multiple outputs.

Kling Image V3

Kuaishou

Kuaishou's Kling Image V3 text-to-image release, served by the exact fal endpoint fal-ai/kling-image/v3/text-to-image with up to 2K output and multiple aspect ratios.

Kolors

Kuaishou

Open-source image generation model from Kuaishou's AI lab. Based on an enhanced SDXL architecture, recognized for vibrant color reproduction and strong aesthetic composition β€” particularly excelling in portrait photography with diverse skin tone rendering.

Qwen-Image-Max

Alibaba

Alibaba's image generation model from the Qwen AI family. Produces high-quality imagery with particular strengths in East Asian aesthetics, product photography, and diverse cultural representations. Part of Alibaba's broader multimodal AI platform.

Tencent

Hunyuan Image 3

Tencent

Tencent's third-generation image model with significantly improved photorealism and detail rendering. Strong in cinematic compositions, portrait photography, and lifestyle imagery. Part of Tencent's Hunyuan multimodal AI platform.

Ideogram

Ideogram V3

Ideogram

Ideogram's third-generation model from the New York startup founded by former Google Brain researchers. Recognized as one of the strongest models for rendering legible, stylistically integrated text within images β€” the go-to for graphic design, posters, and typographic artwork.

HiDream.ai

HiDream I1

HiDream.ai

HiDream.ai's flagship image generation model, recognized for exceptional detail in fashion, beauty, and lifestyle photography. Strong color accuracy, style consistency, and skin tone rendering make it popular among professional creative applications.

RunDiffusion

Juggernaut Pro Flux

RunDiffusion

RunDiffusion's photorealistic fine-tune of the FLUX architecture. A community favorite for portrait photography and commercial imagery, known for natural skin textures, accurate studio-quality lighting, and high-detail rendering.

RunDiffusion

Juggernaut Flux Lightning

RunDiffusion

RunDiffusion's speed-optimized Juggernaut Flux Lightning release, designed for fast photorealistic generation with strong prompt adherence.

RC

Recraft V3

Recraft

Recraft's third-generation text-to-image model, designed for controlled photographic and graphic output with strong composition and typography.

OpenArt

OpenArt Photorealistic

OpenArt

OpenArt's flagship model trained specifically for maximum photorealism in portrait and product photography. Optimized for studio-quality lighting simulations, natural skin rendering, and commercial-grade output.

XLabs AI

Flux Realism

XLabs AI

XLabs AI's FLUX fine-tune optimized specifically for photorealistic output. One of the most downloaded FLUX fine-tunes in the open-source community, with improved natural detail rendering, accurate color grading, and enhanced realism across portrait and landscape subjects.

Leonardo AI

Lucid Realism

Leonardo AI

Leonardo AI's photorealistic image-generation model, documented as Lucid Realism.

DL

Deliberate

XpucT

A community fine-tune of Stable Diffusion 1.5, created by XpucT on Civitai. One of the most downloaded checkpoints in the open-source SD ecosystem. Deliberate (v1–v3) is optimized for photorealism, cinematic compositions, and anatomical accuracy β€” achieving high-quality output with minimal prompt engineering.

Reve AI

Reve

Reve AI

Reve AI's image generation model, known for strong prompt adherence and richly detailed output. Produces high-quality images with a distinctive visual style that blends photorealism with artistic polish β€” versatile across photography, illustration, and conceptual art.

Z-Image

Alibaba

Z-Image (ι€ η›Έ) is an open-source image generation model family by Alibaba's Tongyi Lab. It uses a novel S3-DiT (Scalable Single-Stream Diffusion Transformer) architecture that processes text and image in a unified single stream β€” enabling efficient, high-quality generation with strong compositional control and instruction following.

Reve AI

Reve 2.1

Reve AI

Reve AI's 2.1 release, known for layout intelligence, strong prompt adherence, native high-resolution output, and accurate in-image text.

MS

MAI-Image-2.5

Microsoft AI

Microsoft AI's MAI-Image-2.5 release for photorealistic and design-ready visuals with precise instruction following; MAI-Image-2.5-Pro is the newer highest-fidelity variant in public preview.

Google

Nano Banana 2 Lite

Google

Google's efficiency-focused Gemini 3.1 Flash Lite Image release, used through its exact low-cost text-to-image endpoint.

Google

Nano Banana 2 with Web Search

Google

Google's Gemini 3.1 Flash Image model with web-search grounding enabled, matching the distinct grounded system listed by Arena.

XAI

Grok Imagine Image Quality

xAI

xAI's high-fidelity Grok Imagine image-generation tier, distinct from its standard and pro serving variants.

Ideogram

Ideogram 4.0 Quality

Ideogram

Ideogram 4.0 using its exact quality rendering mode, retained separately from turbo and balanced serving modes.

Qwen Image 2.0 Pro

Alibaba

Alibaba's high-fidelity Qwen Image 2.0 Pro release, with detailed composition and strong multilingual typography.

FLUX.2 [pro]

Black Forest Labs

Black Forest Labs' professional FLUX.2 endpoint, kept separate from max, dev, and klein releases for leaderboard identity.

Google

Imagen 4 Ultra

Google

Google's highest-fidelity Imagen 4 tier, optimized for instruction following, photorealism, and detailed commercial imagery.

AWS

Titan Image Generator G1 v1

Amazon

Amazon's first-generation text-to-image model, part of the Amazon Bedrock foundation model platform. Designed for enterprise use cases including product visualization, marketing assets, and content generation at scale. Features built-in safety controls and watermarking capabilities.

All Companies & Version History

21 companies Β· 90+ model versions Β· release dates and benchmark links

DALL-E

Pioneering text-to-image model integrated into ChatGPT and Microsoft

DALL-E 1 Jan 2021

First version β€” introduced the concept of text-to-image generation to the public. 12B parameter transformer model.

DALL-E 2 Apr 2022

Major quality leap. Introduced inpainting and outpainting. 4Γ— higher resolution than DALL-E 1. Used CLIP image embeddings.

FID score study (arXiv)
DALL-E 3 Oct 2023

Dramatically improved prompt adherence. Deeply integrated into ChatGPT. Strong text rendering in images.

T2I-CompBench evaluation

GPT Image

Next-gen image generation built into the GPT-4o architecture

GPT Image 1 Mar 2025

First native image generation model integrated into GPT-4o. Best-in-class text rendering, strong instruction following.

GPT Image 1.5 May 2025

Improved consistency, photorealism, and editing capabilities over GPT Image 1.

ChatGPT Images 2.0 (GPT Image 2) Apr 2026

Major generational leap. Introduced Instant and Thinking variants β€” Thinking mode researches context before generating. Best-in-class text rendering, multilingual support, full magazine/comic layout generation. Replaces DALL-E 3 and GPT Image 1.x in the API (deprecated May 12, 2026).

OpenAI announcement
Google DeepMind

Google DeepMind

Imagen

Google's flagship diffusion model β€” consistently top-ranked photorealism

Imagen 1 May 2022

Introduced by Google Brain. Outperformed DALL-E 2 on COCO FID at launch.

Imagen paper (arXiv)
Imagen 2 Dec 2023

Substantially improved photorealism, text rendering, and multilingual support. Launched in Google Bard and Vertex AI.

Imagen 3 Aug 2024

Highest quality Imagen to date at launch. Improved detail, lighting, and artifact reduction.

Imagen 3 paper (arXiv)
Imagen 4 May 2025

Next-gen text-to-image with best-in-class landscape, architecture, and product photography. Integrated into Gemini.

Gemini Image (Nano Banana)

Image generation built natively into the Gemini 3 architecture

Gemini 2.5 Flash Image (Nano Banana) Mar 2025

Fast image generation + editing model based on Gemini 2.5 Flash. Integrated into Gemini apps.

Gemini 3 Pro Image (Nano Banana Pro) Jun 2025

Professional-grade image generation and editing. Precise instruction following, near-seamless inpainting.

Stability AI

Stability AI

Stable Diffusion

The open-source model that democratized AI image generation

Stable Diffusion 1.4 Aug 2022 Open Source

First public release. Latent diffusion model β€” runs on consumer GPUs. Sparked the open-source AI art movement.

LDM paper (arXiv)
Stable Diffusion 1.5 Oct 2022 Open Source

Improved training. Still the most widely used SD checkpoint β€” thousands of community fine-tunes built on it.

Stable Diffusion 2.0 Nov 2022 Open Source

Higher resolution (768px default), new text encoder (OpenCLIP), new depth model.

Stable Diffusion 2.1 Dec 2022 Open Source

Improved NSFW filtering approach. Less over-filtering than 2.0, better prompt adherence.

Stable Diffusion XL (SDXL) Jul 2023 Open Source

Major architecture jump β€” 3.5B parameters. Native 1024px output. Two-stage pipeline with refiner model.

SDXL paper (arXiv)
Stable Diffusion 3 Jun 2024 Open Weights

Multimodal Diffusion Transformer (MMDiT) architecture. Major improvement in text rendering and composition.

SD3 paper (arXiv)
Stable Diffusion 3 Medium Jun 2024 Open Weights

2B parameter variant of SD 3. Lighter weight, runs on consumer hardware. Same MMDiT architecture as SD 3 but optimized for accessibility.

Stable Diffusion 3.5 Medium Oct 2024 Open Weights

2.5B parameter model. Balanced between quality and speed β€” ideal for local use on consumer GPUs with less than 8 GB VRAM.

Stable Diffusion 3.5 Large Oct 2024 Open Weights

8B parameter model. Best quality in the SD 3.x family.

Stable Diffusion 3.5 Large Turbo Oct 2024 Open Weights

Distilled 4-step variant. Near SD 3.5 Large quality at a fraction of inference time.

SD Community Fine-tunes

The most widely used community checkpoints built on Stable Diffusion

Deliberate (XpucT) Jan 2023 Open Source

One of the most downloaded SD 1.5 fine-tunes on Civitai. Optimized for photorealism, cinematic compositions, and anatomical accuracy β€” achieves high quality with minimal prompt engineering.

Deliberate on CivitAI
MJ

Midjourney

Artistic aesthetic, Discord-native β€” the model that went viral

v1 Feb 2022

Initial alpha release via Discord.

v2 Apr 2022

Improved coherence and style.

v3 Jul 2022

Better quality, more artistic control.

v4 Nov 2022

New proprietary model architecture. Major quality jump. More coherent scenes.

v5 Mar 2023

Photorealistic quality leap. Detailed hands and faces. Introduced --ar and --style flags.

v5.1 May 2023

Improved default aesthetics, more accurate with simple prompts.

v5.2 Jun 2023

Sharpened images, improved aesthetics, introduced zoom-out feature.

v6 Dec 2023

Much more literal prompt following, improved text in images, better photorealism.

v6.1 Jul 2024

Refinement of v6 β€” improved image quality, better human anatomy, sharper details.

v7 Apr 2025

Major architecture change. New personalization system, significantly faster generation, improved realism.

BFL

Black Forest Labs

FLUX

State-of-the-art open model from former Stability AI researchers

FLUX.1 [schnell] Aug 2024 Open Source

Apache 2.0 open-source, 4-step distilled model. Fast generation, consumer-ready.

FLUX.1 [dev] Aug 2024 Open Weights

Guidance-distilled, 12B parameter Flow Matching model. Better quality than schnell. Non-commercial license.

FLUX.1 paper (arXiv)
FLUX.1 [pro] Aug 2024

API-only commercial model. Highest quality in FLUX.1 family.

FLUX.1.1 [pro] Oct 2024

Improved version of FLUX.1 pro β€” 6Γ— faster, better prompt adherence and quality.

FLUX.2 [dev] Feb 2025 Open Weights

32B parameter model β€” generation, editing, and multi-reference image combining in one. Open weights for research and non-commercial use. 226k+ downloads on Hugging Face.

FLUX.2 [dev] on Hugging Face
FLUX.2 [pro] Feb 2025

Commercial API tier. Production-grade generation and editing with 4MP output and multi-reference control.

FLUX.2 [flex] Mar 2025

Flexible tier between dev and pro β€” balances quality and cost for scalable production workflows.

FLUX.2 [max] Apr 2025

Premium tier. Superior facial detail, enhanced text rendering, maximum photorealism.

FLUX.2 [klein] 4B & 9B Jan 2026 Open Weights

Smallest and fastest FLUX models to date β€” sub-second inference on capable hardware. Available in 4B and 9B parameter versions. Open weights.

FLUX.2 [klein] on Hugging Face

FLUX Community Fine-tunes

The most important open-source fine-tunes built on the FLUX architecture

Flux Realism (XLabs AI) Aug 2024 Open Source

One of the first and most downloaded FLUX fine-tunes. Optimized for photorealistic output β€” improved natural color grading, skin detail, and landscape rendering. Fully open source.

XLabs AI on Hugging Face
Hyper-FLUX-8Steps (ByteDance) Sept 2024 Open Source

Consistency distillation of FLUX.1 [dev] down to 8 sampling steps with near-identical quality. Makes local FLUX generation significantly faster on consumer hardware.

Hyper-SD on Hugging Face
PuLID-FLUX (ToTheBeach) Sept 2024 Open Source

Face identity preservation for FLUX. Generate consistent portraits of a specific person without fine-tuning β€” just provide a reference photo. Popular for character consistency workflows.

PuLID on GitHub
Juggernaut Pro Flux (RunDiffusion) Oct 2024 Open Source

Community-favorite portrait fine-tune. Natural skin textures, studio-quality lighting, and sharp hair detail. Widely regarded as the best FLUX fine-tune for commercial portrait work.

Juggernaut on CivitAI
XFlux / FLUX-Controlnet (XLabs AI) Oct 2024 Open Source

ControlNet implementation for FLUX β€” enables structural guidance via depth maps, canny edges, and pose skeletons. Brings the precise layout control from SD/SDXL to the FLUX architecture.

XLabs-AI/x-flux on GitHub
FLUX.1-Turbo-Alpha (Alimama/Alibaba) Nov 2024 Open Source

Adversarial distillation of FLUX.1 [dev] to just 8 steps with minimal quality loss. Developed by Alibaba's Alimama team β€” one of the fastest high-quality FLUX variants.

FLUX-Turbo on Hugging Face

Adobe Firefly

Commercially safe, trained on licensed content β€” built into Creative Cloud

Firefly 1 May 2023

First commercially licensed AI image generator. Trained on Adobe Stock + Creative Commons. Integrated into Photoshop.

Firefly 2 Oct 2023

Improved photorealism and detail. Photo settings for lighting and depth of field.

Firefly 3 Apr 2024

Major quality improvement. Introduced Structure Reference and Style Reference features.

Firefly 4 May 2025

Advanced multi-entity generation, improved realism, better Creative Cloud integration.

META

Emu

Meta's image generation model β€” available inside Instagram and WhatsApp

Emu Sept 2023

Meta's first image generation model. Fine-tuned on curated high-quality data. Integrated into Instagram.

Emu paper (arXiv)
Emu Edit Nov 2023

Instruction-based image editing model. Precise local and global edits via text commands.

Emu Edit paper (arXiv)
Emu 2 Jan 2024

Largest generative multimodal model from Meta. 37B parameters, strong few-shot visual generation.

Emu 2 paper (arXiv)
BD

ByteDance / BytePlus

Seedream

Cinematic, enterprise-grade image generation from the makers of TikTok

Seedream 3.0 Sept 2024

Strong composition and color accuracy. Designed for professional and enterprise workflows.

Seedream 4.5 Mar 2025

Improved cinematic quality, better text support, enhanced style fidelity.

Kuaishou

Kuaishou

Kolors

Open-source SDXL fine-tune with vivid colors and strong portrait quality

Kolors 1.0 Jul 2024 Open Source

Open-source release. Built on enhanced SDXL. Recognized for vibrant colors and diverse skin tone rendering.

Kolors paper (arXiv)

Kling Omni

Unified image + video generation model

Kling 1.0 (Image) Jun 2024

Image generation component of Kling. Consistent character generation and style fidelity.

Kling IMAGE 3.0 Omni Feb 2025

Unified image and video architecture. Exceptional character consistency across outputs.

ALI

Alibaba (Tongyi Lab)

Wanx (ι€šδΉ‰δΈ‡θ±‘)

Alibaba's flagship image model with strong cultural diversity

Wanx 2.0 Jan 2024

High-quality generation with East Asian aesthetic strengths. Available via Alibaba Cloud API.

Qwen Image (Wanx 2.5) Sept 2024

Integrated into the Qwen multimodal model family. Strong product photography and cultural representations.

Z-Image (ι€ η›Έ)

Open-source S3-DiT architecture β€” unified text and image stream

Z-Image 1.0 Jan 2025 Open Source

Novel S3-DiT (Scalable Single-Stream Diffusion Transformer). Efficient unified text+image processing.

Hunyuan Image

Cinematic image generation from Tencent's Hunyuan platform

Hunyuan Image 1 Jan 2024

First public release. Strong in portrait and lifestyle imagery.

Hunyuan Image 2 Jul 2024

Improved realism and cinematic composition.

Hunyuan Image 3 Jan 2025

Best-in-class portrait photorealism from Tencent. Warm color grading, studio-quality lighting.

Ideogram

Ideogram

USA (ex-Google Brain) ideogram.ai β†—

Ideogram

Best-in-class text rendering β€” go-to for graphic design and typography

Ideogram 1.0 Aug 2023

Launch. Immediately recognized for best text rendering of any image model at the time.

Ideogram 2.0 Nov 2024

Significantly improved photorealism while maintaining text advantages. New Style and Color Palette features.

ELO benchmark (Ideogram blog)
Ideogram 3.0 Mar 2025

Further photorealism improvements. Strongest text-in-image model available. New canvas editing features.

Microsoft Designer / Copilot Image

DALL-E and GPT Image powered β€” built into Windows, Edge, and Microsoft 365

Designer (DALL-E 3 powered) Oct 2023

Microsoft Designer launched with DALL-E 3 backend. Integrated into Bing Image Creator and Edge.

Copilot Image (GPT Image 1) Apr 2025

Upgraded to GPT Image 1 backend. Integrated across Microsoft 365, Windows Copilot, and Designer.

HiDream.ai

HiDream.ai

HiDream I-Series

Fashion and beauty specialist β€” open-source with commercial options

HiDream I1 Fast Mar 2025 Open Source

Fast variant. 17B parameter model, 4-step distillation.

HiDream I1 Dev Mar 2025 Open Weights

Development/research variant. Non-commercial license.

HiDream I1 Full Mar 2025 Open Source

Full quality model. Exceptional skin tone accuracy, fashion and beauty photography.

HiDream paper (arXiv)

Aurora / Grok Image

Elon Musk's AI lab β€” extremely realistic output with minimal safety filtering

Aurora (Grok-2 Image) Dec 2024

xAI's first public image model, integrated into Grok on X (formerly Twitter). FLUX-based architecture. Notably less restrictive safety filters produce a hyper-realistic look, especially for people and social scenes.

Grok-3 Image Jun 2025

Next-generation model. Continues the Aurora lineage with improved photorealism and native Grok-3 multimodal integration.

RC

Recraft

The first model to master graphic design β€” text, logos, and vector art

Recraft v1 Oct 2023

Early release focused on vector-style and design-oriented generation.

Recraft v2 Jun 2024

Improved style consistency, introduced SVG output support.

Recraft v3 Nov 2024

#1 on Hugging Face text-to-image leaderboard at launch. Industry-leading text rendering, logo generation, and vector graphic output. The go-to model for professional graphic designers.

Hugging Face T2I Leaderboard
LEO

Leonardo AI (Canva)

Leonardo Phoenix / Kino

Cinematic Hollywood-style imagery β€” one of the world's largest AI art platforms

Kino XL Feb 2024

Cinematic-style SDXL fine-tune. Strong in dramatic lighting, movie-like scenes, and character consistency.

Leonardo Phoenix Sept 2024

Leonardo's first proprietary foundation model. Improved text rendering, prompt adherence, and photorealism. Millions of daily generations across the platform.

Leonardo Phoenix 1.0 Jan 2025

Refined flagship model post-Canva acquisition. Cinematic quality with better anatomy and enhanced style controls.

Apple Intelligence Image

Deeply integrated into iOS/macOS β€” the AI model most people encounter daily

Image Playground (iOS 18.2) Dec 2024

First public Apple image generation feature. Clean, illustrative style β€” Animation, Illustration, and Sketch modes. Runs fully on-device. Available on iPhone 15 Pro and later.

Image Wand (iPadOS 18.2) Dec 2024

Sketch-to-image feature in Apple Notes. Turns rough hand-drawn sketches into polished illustrations. On-device, privacy-first.

Genmoji (iOS 18.2) Dec 2024

AI-generated custom emoji from text descriptions. Integrated into keyboard and Messages.

Image Playground v2 (iOS 19) Sept 2025

Expanded style options and improved quality expected with iOS 19. Broader device support.

OpenArt Photorealistic

Studio-quality portrait and product photography model

OpenArt Photorealistic Jun 2024

OpenArt's flagship model trained for maximum photorealism in portrait and product photography. Optimized for studio lighting simulation, natural skin rendering, and commercial-grade output.

Titan Image Generator

Enterprise image generation on AWS Bedrock with built-in safety

Titan Image Generator G1 v1 Nov 2023

Amazon's first image generation model on Bedrock. Designed for enterprise workflows β€” product visualization, marketing assets, content generation at scale. Features built-in watermarking.

Titan Image Generator v2 Aug 2024

Improved quality, new image conditioning features including background removal and outpainting. Better instruction following.

Reve AI

Reve AI

Reve

Strong prompt adherence with distinctive artistic polish

Reve 1.0 Jan 2025

Launch model. Recognized for strong depth-of-field rendering and rich shadow detail.

Now put it to the test

Can you spot which image came from which model?

Related