What it means
Text-to-image systems turn a written prompt into an image that has never existed. Most use diffusion, starting from noise and refining toward something matching the description.
Capability has moved fast in specific, uneven ways. Text rendering inside images — long a reliable tell — now largely works. Compositional control has improved, as has editing an existing image by instruction rather than regenerating from scratch.
The unresolved issues are legal and economic rather than technical. Models trained on scraped images are the subject of active litigation; some vendors now differentiate on licensed-only training data and offer indemnification. Style imitation sits in a genuinely unsettled area — style is not copyrightable, but generating work in a living artist's manner raises questions the law has not answered.
Why it matters
For most organizations this is the most immediately usable generative capability — drafts, concepts, and internal visuals at effectively zero marginal cost. Whether output is safe for commercial use depends on the vendor's training data and terms, and that varies enormously.
In practice
For commercial work, check the training-data provenance and whether the vendor indemnifies you. Expect to generate several and select; output is stochastic, and a seed value is what makes a result reproducible while you iterate.
Where this shows up
Tools and models in our catalog.
MidjourneyThe gold standard for artistic AI image generation. Exceptional aesthetic quality and style control via Discord and web interface. Huge community of creators.
DALL-E / GPT Image 2OpenAI's image generation via ChatGPT. GPT Image 2 (April 2026) is the first OpenAI image model with built-in O-series reasoning — it plans compositions before drawing. Adds multilingual text rendering (Japanese, Korean, Chinese, Hindi, Bengali), web-search grounding, 2K resolution, and up to 8 images per prompt. Took #1 on the Image Arena leaderboard by +242 points at launch.
Stable DiffusionThe leading open-source image generation model. Run locally, fine-tune on custom data, or use via API. Huge ecosystem of models, LoRAs, and tools.
FluxState-of-the-art open-weight image model from Black Forest Labs. Flux Pro and Schnell versions offer photorealistic quality with fast inference speeds.
Adobe FireflyAdobe AI model family powering Creative Cloud. Firefly Image Model 5 (GA March 2026) generates photorealistic 4MP images. Firefly Video Model adds commercially safe AI video (Text to Video, Image to Video, Generative Extend). 30+ model hub including Google, Runway, Kling. Custom Models in public beta. Generate Soundtrack + Speech. Unlimited generations for subscribers. Trained on licensed Adobe Stock — commercially safe.