Anil-matcha/awesome-ai-image-models

The most complete, up-to-date comparison of AI image generation models — which model, via which API, at what price.

82

stars

23

commits

Aug 23, 2026

updated

ai-image
ai-image-generator
awesome
awesome-list
flux
generative-ai
gpt-image
ideogram
image-editing
image-generation
imagen
image-to-image
midjourney
muapi
nano-banana
qwen-image
recraft
seedream
stable-diffusion
text-to-image
Browse cluster: Curated Learning Resources & Awesome Lists

README

Awesome AI Image Models Awesome

The most complete, up-to-date comparison of AI image generation models — which model, via which API, at what price, and what it's best at.

Unlike other lists that just dump links, this one answers the question developers actually have: "I need to generate images — which model do I pick, and where do I call it?" Every model is mapped to the APIs that serve it, with real per-image pricing and what it's genuinely good for.

💡 Prices are per standard image (retail API rates verified Aug 2026) and move fast — always confirm against the provider. 4K/high-res and batch modes change the math (batch often ~50% off).

Best AI Image Generator (API) in 2026 (Quality, Price, Uncensored, Editing)

📺 Best AI Image Generator (API) in 2026 (Quality, Price, Uncensored, Editing) →

Contents

Commercial models (closed, API-only)

ModelMakerBest forAPIsPrice / imageNotes
GPT Image 2OpenAI🏆 Best overall qualityOpenAI API, MuAPI~$0.09Tops Artificial Analysis's Text-to-Image Arena (Elo 1370); 2K res, clean multilingual text rendering; edit mode ranks #3 on the Editing Arena
Nano Banana ProGoogle4K, editing, character consistencyGemini API, Fal, Replicate, MuAPI~$0.12Community/press favorite for photorealism; best-in-class coherent local edits and locking character identity across generations
Seedream 5.0 ProByteDanceStylized/artistic outputFal, Replicate, MuAPI~$0.045Strongest stylized output of the set — "wins on capability boundaries" per independent comparisons
Midjourney v8MidjourneyAesthetic / artistic, --cref character consistencyMuAPI~$0.10Still the aesthetic-quality benchmark; API access via MuAPI, no official first-party API
Imagen 4 UltraGooglePhotorealism, prompt adherenceGemini / Vertex AI, MuAPI~$0.06Google's top tier
Ideogram CharacterIdeogramIn-image text, character referenceIdeogram API, Fal, Replicate, MuAPI~$0.15Best photorealistic character consistency in side-by-side comparisons
Recraft V3RecraftDesign, vector, brandFal, Replicate~$0.04SVG/vector + style control

Best value

ModelMakerLicensePrice / imageNotes
Z-Image TurboAlibaba (Tongyi-MAI)Apache-2.0~$0.007Cited across 2026 roundups as the price/quality sweet spot, not just the cheapest option — also the #1 open-weights model on Artificial Analysis's Text-to-Image Arena
Flux-2 Klein 4B TurboBlack Forest LabsCommercial (BFL license)~$0.0052Half the price of the standard Klein 4B, same Flux-family quality
FLUX.1 [schnell]Black Forest LabsApache-2.0~$0.003Near-instant generation, the classic low-cost workhorse, free commercial use
SDXLStability AICommunity License~$0.004The fallback when you just need pixels at the lowest possible cost

Best uncensored / unrestricted

⚠️ No mainstream model ships a muapi-branded "Spicy" image endpoint (unlike its video counterparts) — these are picks with minimal/no built-in content filtering in practice, not a specifically labeled unrestricted tier.

ModelMakerPrice / imageNotes
Wan 2.7Alibaba~$0.05Widely cited in 2026 uncensored/NSFW-generation roundups for near-zero prompt filtering
Qwen Image 2.0Alibaba~$0.042026 coverage explicitly tests and confirms NSFW capability, no prompt-rewriting layer in the way
Seedream 5.0ByteDance~$0.0325Grouped with Wan/Qwen in "open pipeline, no surprise censorship" comparisons
Grok ImaginexAI~$0.05Marketed with a looser content policy than mainstream Western closed models

See the companion awesome-uncensored-ai-image-models for a deeper filtering/licensing-focused catalog.

Open-source models (self-host or API)

ModelMakerLicenseBest forAPIsSelf-host VRAM
Z-Image TurboAlibaba (Tongyi-MAI)Apache-2.0🏆 #1 open-weights model on Artificial Analysis's Text-to-Image Arena, ahead of FLUX.2 [dev], HunyuanImage 3.0, and Qwen-ImageFal, Replicate, MuAPI~12GB+
Qwen-ImageAlibabaApache-2.0Best in-image text (EN/CN)Fal, Replicate, MuAPI~24GB+ (20B)
FLUX.2 [dev]Black Forest LabsNon-commercial (paid license for commercial)Top OSS quality, 4MP — the model Z-Image Turbo is benchmarked againstFal, Replicate, MuAPI~24GB+
HiDream i1 (Full)HiDreamMITGenuinely different architecture from Flux/Qwen/Z-Image familiesMuAPI~24GB+
HunyuanImage 3.0TencentOpen (check terms)Largest OSS model (80B MoE)self-host~40GB+
Stable Diffusion 3.5 LargeStability AICommunity License (free <$1M rev)Ecosystem, LoRAs, ControlNetFal, Replicate, self-host~18GB+
SANANVIDIAPermissive researchFast, efficient, low-VRAMself-host~12GB

Image editing & control

Modify existing images rather than generate from scratch:

ModelMakerBest forPrice / generationNotes
Nano Banana Pro EditGoogle🏆 Best editing~$0.12Leads on coherent object insertion/removal, repaints edits into the scene rather than visibly patching
GPT Image 2 (edit mode)OpenAIArena-verified editing~$0.09Ranks #3 on Artificial Analysis's own Image Editing Arena (Elo 1257)
Seedream 5.0 EditByteDanceHigh-volume editing~$0.03251/4 to 1/7 the cost of Nano Banana Pro Edit
FLUX.1 Kontext ProBlack Forest LabsInstruction-based editing~$0.03Pioneered one-sentence instruction editing, no fine-tuning needed
Qwen Image Edit 2511AlibabaOpen-model editing~$0.04Industry-leading performance for its price tier

Also supported natively across the models above: ControlNet (structural conditioning — pose, depth, edges, scribble), IP-Adapter (image prompting / style + subject transfer), and inpainting/outpainting.

Character consistency

Keep the same subject's identity locked across multiple generations, not just a one-off crop-and-paste:

ModelMakerPrice / generationNotes
Nano Banana ProGoogle~$0.12Reputation for locking character identity across edits/scenes
Ideogram CharacterIdeogram~$0.15Dedicated Character Reference feature, best photorealistic consistency in side-by-side comparisons
Midjourney v8 (--cref)Midjourney~$0.10Strongest option for stylized/artistic recurring characters
MiniMax Subject ReferenceMiniMax~$0.01Cheapest dedicated subject-consistency endpoint
Vidu Q2 Reference-to-ImageVidu~$0.032Reference-driven generation

Frameworks & UIs

For local generation, training, and workflows:

  • ComfyUI — node-based, most powerful for custom pipelines
  • AUTOMATIC1111 WebUI — the classic all-in-one UI
  • InvokeAI — polished pro/creative UI
  • Fooocus — simplest "just works" UI
  • Training: kohya_ss, OneTrainer, SimpleTuner (LoRA / fine-tuning)

Upscaling & restoration

  • Real-ESRGAN — general-purpose upscaling (open source)
  • GFPGAN / CodeFormer — face restoration
  • chaiNNer — node-based batch processing
  • Topaz Photo AI — highest-quality commercial upscale/denoise

Benchmarks & leaderboards

Check independent evals before trusting a maker's demo gallery:

  • Artificial Analysis Image Arena (Text-to-Image + Image Editing) — Elo-style human-preference leaderboards
  • HEIM (Holistic Evaluation of Image Models) — multi-dimension benchmark
  • FID / CLIP Score / ImageReward — automated quality & alignment metrics

How to choose

  • Best overall quality → GPT Image 2 (Arena #1), or Nano Banana Pro for photorealism/local edits
  • Best value → Z-Image Turbo (~$0.007/image, also the #1 open-weights model)
  • Best uncensored / unrestricted → Wan 2.7 or Qwen Image 2.0 (no prompt-rewriting layer in practice)
  • Best editing → Nano Banana Pro Edit, or GPT Image 2 (edit mode) for an arena-verified pick
  • Best character consistency → Nano Banana Pro or Ideogram Character (dedicated Character Reference)
  • Fully open, commercial-safe → FLUX.1 [schnell] or Qwen-Image (both permissive); Z-Image Turbo for the current open-weights quality leader
  • Design / vector / brand → Recraft V3
  • Ecosystem & LoRAs → Stable Diffusion 3.5

Where to run them (API providers)

Aggregators that expose many of the above behind one API/key:

  • MuAPI — unified API across image + video models (GPT Image 2, Nano Banana Pro, Seedream, FLUX, Z-Image, Qwen, and more), one key, one billing — see the full AI Image API leaderboard
  • Fal — fast inference, broad model catalog
  • Replicate — pay-per-run, large community model catalog

Native APIs (single-vendor): OpenAI (GPT Image), Google Gemini/Vertex (Nano Banana, Imagen), Black Forest Labs (FLUX), Ideogram, Recraft.

Contributing

PRs welcome. When adding a model, keep the table columns filled — a row without provider + price isn't useful. New models go in the correct table (commercial vs open-source vs task-specific) and stay sorted by relevance.


Maintained alongside awesome-ai-video-models and Open-Generative-AI. Found it useful? ⭐ the repo.

Contributors

Anil-matcha

23 commits

Anil-matcha/awesome-ai-image-models

The most complete, up-to-date comparison of AI image generation models — which model, via which API, at what price.

82

stars

23

commits

Aug 23, 2026

updated

ai-image
ai-image-generator
awesome
awesome-list
flux
generative-ai
gpt-image
ideogram
image-editing
image-generation
imagen
image-to-image
midjourney
muapi
nano-banana
qwen-image
recraft
seedream
stable-diffusion
text-to-image
Browse cluster: Curated Learning Resources & Awesome Lists

README

Awesome AI Image Models Awesome

The most complete, up-to-date comparison of AI image generation models — which model, via which API, at what price, and what it's best at.

Unlike other lists that just dump links, this one answers the question developers actually have: "I need to generate images — which model do I pick, and where do I call it?" Every model is mapped to the APIs that serve it, with real per-image pricing and what it's genuinely good for.

💡 Prices are per standard image (retail API rates verified Aug 2026) and move fast — always confirm against the provider. 4K/high-res and batch modes change the math (batch often ~50% off).

Best AI Image Generator (API) in 2026 (Quality, Price, Uncensored, Editing)

📺 Best AI Image Generator (API) in 2026 (Quality, Price, Uncensored, Editing) →

Contents

Commercial models (closed, API-only)

ModelMakerBest forAPIsPrice / imageNotes
GPT Image 2OpenAI🏆 Best overall qualityOpenAI API, MuAPI~$0.09Tops Artificial Analysis's Text-to-Image Arena (Elo 1370); 2K res, clean multilingual text rendering; edit mode ranks #3 on the Editing Arena
Nano Banana ProGoogle4K, editing, character consistencyGemini API, Fal, Replicate, MuAPI~$0.12Community/press favorite for photorealism; best-in-class coherent local edits and locking character identity across generations
Seedream 5.0 ProByteDanceStylized/artistic outputFal, Replicate, MuAPI~$0.045Strongest stylized output of the set — "wins on capability boundaries" per independent comparisons
Midjourney v8MidjourneyAesthetic / artistic, --cref character consistencyMuAPI~$0.10Still the aesthetic-quality benchmark; API access via MuAPI, no official first-party API
Imagen 4 UltraGooglePhotorealism, prompt adherenceGemini / Vertex AI, MuAPI~$0.06Google's top tier
Ideogram CharacterIdeogramIn-image text, character referenceIdeogram API, Fal, Replicate, MuAPI~$0.15Best photorealistic character consistency in side-by-side comparisons
Recraft V3RecraftDesign, vector, brandFal, Replicate~$0.04SVG/vector + style control

Best value

ModelMakerLicensePrice / imageNotes
Z-Image TurboAlibaba (Tongyi-MAI)Apache-2.0~$0.007Cited across 2026 roundups as the price/quality sweet spot, not just the cheapest option — also the #1 open-weights model on Artificial Analysis's Text-to-Image Arena
Flux-2 Klein 4B TurboBlack Forest LabsCommercial (BFL license)~$0.0052Half the price of the standard Klein 4B, same Flux-family quality
FLUX.1 [schnell]Black Forest LabsApache-2.0~$0.003Near-instant generation, the classic low-cost workhorse, free commercial use
SDXLStability AICommunity License~$0.004The fallback when you just need pixels at the lowest possible cost

Best uncensored / unrestricted

⚠️ No mainstream model ships a muapi-branded "Spicy" image endpoint (unlike its video counterparts) — these are picks with minimal/no built-in content filtering in practice, not a specifically labeled unrestricted tier.

ModelMakerPrice / imageNotes
Wan 2.7Alibaba~$0.05Widely cited in 2026 uncensored/NSFW-generation roundups for near-zero prompt filtering
Qwen Image 2.0Alibaba~$0.042026 coverage explicitly tests and confirms NSFW capability, no prompt-rewriting layer in the way
Seedream 5.0ByteDance~$0.0325Grouped with Wan/Qwen in "open pipeline, no surprise censorship" comparisons
Grok ImaginexAI~$0.05Marketed with a looser content policy than mainstream Western closed models

See the companion awesome-uncensored-ai-image-models for a deeper filtering/licensing-focused catalog.

Open-source models (self-host or API)

ModelMakerLicenseBest forAPIsSelf-host VRAM
Z-Image TurboAlibaba (Tongyi-MAI)Apache-2.0🏆 #1 open-weights model on Artificial Analysis's Text-to-Image Arena, ahead of FLUX.2 [dev], HunyuanImage 3.0, and Qwen-ImageFal, Replicate, MuAPI~12GB+
Qwen-ImageAlibabaApache-2.0Best in-image text (EN/CN)Fal, Replicate, MuAPI~24GB+ (20B)
FLUX.2 [dev]Black Forest LabsNon-commercial (paid license for commercial)Top OSS quality, 4MP — the model Z-Image Turbo is benchmarked againstFal, Replicate, MuAPI~24GB+
HiDream i1 (Full)HiDreamMITGenuinely different architecture from Flux/Qwen/Z-Image familiesMuAPI~24GB+
HunyuanImage 3.0TencentOpen (check terms)Largest OSS model (80B MoE)self-host~40GB+
Stable Diffusion 3.5 LargeStability AICommunity License (free <$1M rev)Ecosystem, LoRAs, ControlNetFal, Replicate, self-host~18GB+
SANANVIDIAPermissive researchFast, efficient, low-VRAMself-host~12GB

Image editing & control

Modify existing images rather than generate from scratch:

ModelMakerBest forPrice / generationNotes
Nano Banana Pro EditGoogle🏆 Best editing~$0.12Leads on coherent object insertion/removal, repaints edits into the scene rather than visibly patching
GPT Image 2 (edit mode)OpenAIArena-verified editing~$0.09Ranks #3 on Artificial Analysis's own Image Editing Arena (Elo 1257)
Seedream 5.0 EditByteDanceHigh-volume editing~$0.03251/4 to 1/7 the cost of Nano Banana Pro Edit
FLUX.1 Kontext ProBlack Forest LabsInstruction-based editing~$0.03Pioneered one-sentence instruction editing, no fine-tuning needed
Qwen Image Edit 2511AlibabaOpen-model editing~$0.04Industry-leading performance for its price tier

Also supported natively across the models above: ControlNet (structural conditioning — pose, depth, edges, scribble), IP-Adapter (image prompting / style + subject transfer), and inpainting/outpainting.

Character consistency

Keep the same subject's identity locked across multiple generations, not just a one-off crop-and-paste:

ModelMakerPrice / generationNotes
Nano Banana ProGoogle~$0.12Reputation for locking character identity across edits/scenes
Ideogram CharacterIdeogram~$0.15Dedicated Character Reference feature, best photorealistic consistency in side-by-side comparisons
Midjourney v8 (--cref)Midjourney~$0.10Strongest option for stylized/artistic recurring characters
MiniMax Subject ReferenceMiniMax~$0.01Cheapest dedicated subject-consistency endpoint
Vidu Q2 Reference-to-ImageVidu~$0.032Reference-driven generation

Frameworks & UIs

For local generation, training, and workflows:

  • ComfyUI — node-based, most powerful for custom pipelines
  • AUTOMATIC1111 WebUI — the classic all-in-one UI
  • InvokeAI — polished pro/creative UI
  • Fooocus — simplest "just works" UI
  • Training: kohya_ss, OneTrainer, SimpleTuner (LoRA / fine-tuning)

Upscaling & restoration

  • Real-ESRGAN — general-purpose upscaling (open source)
  • GFPGAN / CodeFormer — face restoration
  • chaiNNer — node-based batch processing
  • Topaz Photo AI — highest-quality commercial upscale/denoise

Benchmarks & leaderboards

Check independent evals before trusting a maker's demo gallery:

  • Artificial Analysis Image Arena (Text-to-Image + Image Editing) — Elo-style human-preference leaderboards
  • HEIM (Holistic Evaluation of Image Models) — multi-dimension benchmark
  • FID / CLIP Score / ImageReward — automated quality & alignment metrics

How to choose

  • Best overall quality → GPT Image 2 (Arena #1), or Nano Banana Pro for photorealism/local edits
  • Best value → Z-Image Turbo (~$0.007/image, also the #1 open-weights model)
  • Best uncensored / unrestricted → Wan 2.7 or Qwen Image 2.0 (no prompt-rewriting layer in practice)
  • Best editing → Nano Banana Pro Edit, or GPT Image 2 (edit mode) for an arena-verified pick
  • Best character consistency → Nano Banana Pro or Ideogram Character (dedicated Character Reference)
  • Fully open, commercial-safe → FLUX.1 [schnell] or Qwen-Image (both permissive); Z-Image Turbo for the current open-weights quality leader
  • Design / vector / brand → Recraft V3
  • Ecosystem & LoRAs → Stable Diffusion 3.5

Where to run them (API providers)

Aggregators that expose many of the above behind one API/key:

  • MuAPI — unified API across image + video models (GPT Image 2, Nano Banana Pro, Seedream, FLUX, Z-Image, Qwen, and more), one key, one billing — see the full AI Image API leaderboard
  • Fal — fast inference, broad model catalog
  • Replicate — pay-per-run, large community model catalog

Native APIs (single-vendor): OpenAI (GPT Image), Google Gemini/Vertex (Nano Banana, Imagen), Black Forest Labs (FLUX), Ideogram, Recraft.

Contributing

PRs welcome. When adding a model, keep the table columns filled — a row without provider + price isn't useful. New models go in the correct table (commercial vs open-source vs task-specific) and stay sorted by relevance.


Maintained alongside awesome-ai-video-models and Open-Generative-AI. Found it useful? ⭐ the repo.

Contributors

Anil-matcha

23 commits