Ideogram 4.0
by Ideogram
Ideogram's open‑weight image model.
Frontier in‑image text, layout control, and 2K output.

Key features
Technical specifications
Open
Open weights under a commercial license
0.97 OCR
X-Omni English OCR for in-image text
16 colors
Condition output on up to 16 hex colors
Up to 2K
256 to 2048 px per side, flexible ratios
Use cases
Posters and packaging
Design where the title, tagline, and small print all have to read correctly. Text renders legibly, not as shapes.
Multilingual campaigns
Localize one visual across markets by swapping the text per language while layout and palette stay fixed.
Brand-locked visuals
Feed the brand's hex palette into the prompt and every generation stays on-brand, tile to banner.
Unusual formats
One set of weights covers square thumbnails, widescreen, 2048 by 768 ultrawide banners, and social headers.
Programmatic generation
JSON prompts are built for code. Generate catalogs or ad variants from a script, each element typed and validated.
Self-hosted pipelines
Teams that can't use a third-party API can fine-tune the open weights and run them in their own infrastructure.
Prompt examples





Multilingual sign
Tokyo storefront with accurate Japanese signage, soft rain, evening glow
Edit prompt
Overview
Ideogram 4.0 is a 9.3 billion parameter open-weight text-to-image model from Ideogram, released on June 3, 2026. It leads the open-weight field at its size on in-image text rendering, scoring 0.97 on the X-Omni English OCR benchmark, and pairs that with bounding-box layout control, structured JSON prompting, color palette conditioning, and output up to 2K. The weights, inference code, and prompting guide are public, with quantized builds that run on a single 24 GB GPU.
What Ideogram 4.0 does differently
Most image models take a sentence and return pixels, so text comes out misspelled and placement is a roll of the dice. Ideogram 4.0 was trained exclusively on structured JSON captions, so a prompt can spell out each element: where it sits in the frame, how it is styled, the exact string a text element should render, and the hex colors the image must stay inside. That structure makes results precise and repeatable, which matters for design work that ships as posters, packaging, banners, and UI rather than one-off art.
Ideogram 4.0 and Morphic
Ideogram 4.0 is available on Morphic. Select it in Copilot or on the Canvas and it runs alongside the rest of the image line, plus video, speech, and music, in one workspace.
On Morphic it generates at nine aspect ratios, 9:16, 2:3, 3:4, 4:5, 1:1, 5:4, 4:3, 3:2, and 16:9, and it carries a rendering speed control the other image models do not: Turbo for roughing out a composition, Default for everyday work, and Quality for the final asset. Reach for it when the words inside the image have to be spelled correctly and set well. There is a dedicated Ideogram 4.0 image generator if you want to start from a blank prompt.
Simple pricing
Get started for free today, with the option to upgrade or cancel anytime.
Basic
1100 monthly credits
1 user only
All models
Workflows
Standard
3625 monthly credits
1 user only
All models
Workflows
Pro
6350 shared monthly credits
1 user
All models
Workflows
Pro Max
24650 shared monthly credits
1 user
All models
Workflows
Enterprise
For higher limits
Custom
pricing and billing terms

Free
For playing around
$0
forever free
FAQs
ChatGPT Images 2.5
OpenAI
OpenAI's image model, released September 2026. Sharper detail, precise edits, faster.
Lyria 3.5
Google DeepMind
Google's best-sounding music model. Full songs with structure, vocals, and lyrics.
MiniMax H3 Max Turbo
fal.ai
The fast tier of fal's post-trained H3. Twice the speed, nearly the same look.
Inworld TTS 2
Inworld
Ninety-five voices, 100+ languages. Speech that starts in under a second.