TL;DR
- → Ideogram 4.0 is a 9.3B open-weight text-to-image model released June 3, 2026 — the first open model from Ideogram.
- →It beats every other open-weight model on text rendering, scoring 0.97 OCR accuracy — while being 3–9× smaller than competitors.
- →The secret weapon is JSON prompting: structured captions with bounding boxes, hex color palettes, and per-element text control.
- →Community is already running it in ComfyUI with day-0 native support and three prompting methods.
- →The catch: commercial use of the weights requires a paid license. The API and web app are covered under standard terms.
- →Where it genuinely falls short: photorealistic human portraiture still belongs to GPT Image 2 and closed models.
On June 3, 2026, Ideogram did something the open-source image generation community has been waiting for: they dropped their weights. Ideogram 4.0 is a 9.3-billion-parameter text-to-image model you can download, run locally, fine-tune, and build on. More importantly, it's the first model in its parameter class that can render readable, correctly spelled, properly styled text inside images — reliably. That one capability changes more workflows than the benchmark numbers suggest.
Parameters
9.3B
▲ 3–9× smaller than rivals
Designer preference ELO
1062
▲ #1 open, #2 overall
Client-work usability score
3.55/5 Change vs 2.84 for Nano Banana 2


