आइडियोग्राम 4.0
नयारिलीज़: 3 जून 2026
₹5.76
प्रति इमेज, शुरुआत
प्रदाता की आधिकारिक दर पर बिलिंग, ₹96 प्रति डॉलर पर कनवर्ट, 0% मार्कअप के साथ।
प्रति इमेज कीमत
आउटपुट विकल्प
- रिज़ॉल्यूशन
- 2048x2048, 2560x1440, 1440x2560, 2496x1664, 1664x2496, 3072x1024, 1024x3072
- एडिटिंग
- नहीं
- रेफरेंस इमेज
- नहीं
आइडियोग्राम 4.0 के बारे में
Ideogram's fourth-generation model, built around typography and graphic design. It accepts only a fixed set of output dimensions, all of them 2K or larger.
Ideogram made its name on text inside pictures, and 4.0 is the release where the lab turned that into an argument about the whole shape of a design model. It is a 9.3B single-stream diffusion transformer of 34 layers, trained from scratch and published as Ideogram's first open-weight foundation model, with a commercial licence for teams that want to fine-tune it on their own brand material and run it inside their own infrastructure.
The decision that matters most for prompting is that the model was trained exclusively on structured JSON captions, and the reference pipeline validates every prompt against that schema and rejects anything that does not parse. Training and inference share one format, so the model is never asked to interpret an input shape it has not seen. Three things fall out of that. Colour is conditioned directly: up to sixteen hex values per image and five per element, rather than a description the model has to guess at. Layout is addressable: any element can be placed by a bounding box in normalised coordinates, and Ideogram reports 0.69 mean intersection-over-union on the 7Bench layout evaluation for how tightly generated objects land inside the boxes they were given. And a text element is typed - it carries the literal string to render plus a separate description of how it should look, which is the mechanism behind multi-line, multi-font in-image type rather than a single rendered phrase.
Typography is still the headline capability. Ideogram reports 0.97 English OCR accuracy on the X-Omni text-rendering evaluation, and puts that figure alongside parameter count to make its actual claim: that a 9.3B model beats every larger open-weight release on text, including ones several times its size. On the other axes it reports 0.76 on the SpatialGenEval spatial-reasoning split and 0.89 on the Prism-bench alignment track for following long compositional prompts. In its own blind arena, where graphic designers pick the better of two images without being told what made either, 4,366 votes across nine pipelines put 4.0 second overall and first among open-weight models.
The direction Ideogram says the model starts is layers rather than flat frames. Its reasoning is that production design does not end at a single pixel layer - headlines change before launch, cutouts move to new backdrops - and the lab already ships those as separate steps: a Background Remover that returns a clean alpha cutout from any generation, and a Layerize step that pulls headlines, body copy and graphic elements back out as editable layers. The stated plan for the next 4.0 release is to return alpha channels and editable text layers straight from inference, with no second pass.
Two limits are worth carrying into a prompt. Ideogram's own designer arena puts a closed model ahead of it overall, so 4.0 is the best open model in that ranking rather than the best model in it. And the schema is not optional: the reference pipeline rejects a prompt that does not parse as valid JSON rather than interpreting it, so the structure has to be produced by something - the lab's own tooling, or yours - before the model ever sees the request.
लॉन्च के समय Ideogram ने क्या कहा
- Text is the specialty
- Ideogram reports 0.97 English OCR accuracy on the X-Omni text-rendering evaluation and ranks 4.0 ahead of every larger open-weight release on that axis at 9.3B parameters.
- Layout by bounding box
- Any element can be placed by a box in normalised coordinates, and the lab measures 0.69 mIoU on the 7Bench layout evaluation for how tightly objects land inside the boxes they were given.
- Prompts are JSON, and validated
- The model was trained only on structured JSON captions and the reference pipeline rejects prompts that do not parse, so the format at inference is exactly the one it learned on.
- Colour and type as data
- Up to sixteen hex colours per image and five per element condition the palette directly, and a text element carries its literal string plus a separate styling description - the basis for multi-line, multi-font type.
- Open weights, commercial licence
- Ideogram's first open-weight foundation model: 9.3B parameters, downloadable, fine-tunable on a team's own brand and product data, and deployable inside its own environment.
- Layers, not flat frames
- Ideogram's Background Remover already returns alpha cutouts and its Layerize step already extracts editable text. The lab says the next 4.0 release returns both directly from inference.
- What it is not for
- Ideogram's own blind designer arena places a closed model above it, so this is the strongest open model in that ranking rather than the strongest overall. The JSON schema is also mandatory: a prompt that does not parse is rejected, not interpreted.
Frequently Asked Questions
आइडियोग्राम 4.0 के बारे में अक्सर पूछे जाने वाले सवाल।
आइडियोग्राम 4.0 कब रिलीज़ हुआ था?
Ideogram ने आइडियोग्राम 4.0 को 3 जून 2026 को रिलीज़ किया था।
आइडियोग्राम 4.0 को किसने बनाया है?
आइडियोग्राम 4.0 को AI लैब Ideogram ने बनाया है। 99Models इसे सीधे प्रदाता की आधिकारिक दर पर उपलब्ध कराता है।
आइडियोग्राम 4.0 से एक इमेज बनाने का कितना खर्च आता है?
प्रति इमेज ₹5.76 से शुरू, बिना किसी अतिरिक्त मार्कअप के प्रदाता की दर पर। आप केवल अपनी बनाई गई इमेज का भुगतान करते हैं।
आइडियोग्राम 4.0 किस साइज़ और ऐस्पेक्ट रेशियो की इमेज बना सकता है?
उपलब्ध विकल्प 2048x2048, 2560x1440, 1440x2560, 2496x1664, 1664x2496, 3072x1024, 1024x3072 हैं, जिन्हें आप इमेज जनरेट करने से पहले चुन सकते हैं।
कैटलॉग अपडेट: 9 सित॰ 2026