FrontierThe story, in brief

Addendum to GPT-4o System Card: 4o image generation

OpenAI's GPT-4o now generates photorealistic images and takes images as input—a significant capability jump from DALL·E 3.

Paper-cut illustration of a coral software window opening into a three-dimensional drafting space.
New tools for building and creating with AI.AI illustration by KeyNews
The KeyNews take

Why it matters

OpenAI is expanding GPT-4o's multimodal reach beyond text and vision-understanding into generative image creation, directly challenging specialized image models and expanding the moat of its flagship model.

The key facts

5 to know
  1. GPT-4o image generation described as 'significantly more capable' than DALL·E 3

  2. Can generate photorealistic output

  3. Can take images as inputs and transform them

  4. Published as system card addendum (indicates recent/imminent release)

  5. Multimodal capability expansion of GPT-4o

Go to the source

OpenAI Blogopenai.com

Publisher excerpt: 4o image generation is a new, significantly more capable image generation approach than our earlier DALL·E 3 series of models. It can create photorealistic output. It can take images as inputs and transform them.
Read original report
Back to today's editionMore frontier news

Keep reading

Related stories

More from Frontier