Text-to-image & editing
Image Models
Text-to-image and image-editing models — from artistic generators to photorealistic, API-first and self-hostable options.
FLUX.2
The open-weight champion for photorealism, text rendering and developer workflows — top quality at a few cents per image via API or self-hosting.
Nano Banana 2
Google's Gemini 3 Pro Image model — 4K photorealism free inside the Gemini app, with near-perfect text rendering and conversational editing.
Midjourney v8.1
The industry leader for raw artistic quality and aesthetic polish — a subscription-only image generator beloved by designers and concept artists.
Grok Imagine Image 2.0
xAI's editing-first image model — Arena #2 in both T2I and image editing behind GPT-Image-2, with magic wand, segmentation, multi-reference compositing and templates.
GPT Image 2
OpenAI's image model — the strongest at prompt instruction-following and conversational editing, native to the ChatGPT and OpenAI API stack.
Boogu-Image 0.1
Apache-2.0 open-source image generation and editing family (Base / Turbo / Edit) with strong bilingual text rendering and near closed-source quality — trained on roughly 10× less data and fully self-hostable.
Recraft V4
The design-and-brand specialist — vector/SVG output, brand-style consistency and strong typography built for professional design work.
Ideogram 3.0
The specialist for typography and logo design — the most reliable model for correct, legible text inside generated images.
Imagen 4
Google's photorealism leader — best-in-class typography and people, available cheaply through Gemini Advanced or per-image on Vertex AI.
Stable Diffusion 3.5
The open-weights standard for local, private and customizable image generation — run it on your own GPU at no per-image cost, or via Stability's API.