GPT Image generates and edits pictures with legible in-image text. Drop the results straight into VideoGen video workflows.
By OpenAI · Released April 2025 · Known for strong prompt following and legible in-image text
Real prompts, real outputs, generated with GPT Image in VideoGen.

“Photorealistic storefront of a Parisian bakery with 'BOULANGERIE LUMIERE' painted in gold on the window, croissants stacked in the display, soft morning light on the cobblestones”
Generated with GPT Image in VideoGen

“Illustrated recipe card for lemon ricotta pancakes with the title in friendly hand lettering, ingredient list down the side, and small step-by-step drawings, soft pastel palette”
Generated with GPT Image in VideoGen

“Board game box art titled 'HARBOR & HEARTH' in carved wooden letters, cozy illustrated fishing village with players' meeple characters on the docks, warm inviting tabletop art style”
Generated with GPT Image in VideoGen

“Clean cutaway science diagram of a volcano showing the magma chamber, conduit, and ash cloud with neat labels and arrows, textbook illustration style on white”
Generated with GPT Image in VideoGen
| Specification | Details |
|---|---|
| Max resolution | 1536x1024 (landscape) or 1024x1536 (portrait) |
| Text rendering | Reliable, legible in-image text |
| Editing support | Prompt-based editing of uploaded images |
| Aspect ratios | 1:1, 3:2, 2:3 |
| Input types | Text and reference images |
| Elo Benchmark | 1,137 (Image Arena (GPT Image 1 high)) |
The fastest and most affordable tier, built for drafts and quick iteration.
The balanced default: good quality images at an everyday cost.
Premium quality images for the content you publish.
MAX always integrates the state of the art. GPT Image is part of that set.
Every quality tier routes each request to the best model based on our internal evals, compliance requirements, and your intent. Requests stay fully compliant at every tier: sensitive data is never routed to models hosted by foreign adversaries such as China.
Sign up for VideoGen to open the AI playground and editor.
Start the storyboard to video or prompt to video workflow, or use the AI image tool inside any project.
Choose the tier that fits your needs and describe what you want to create.
VideoGen builds the finished video with visuals, narration, and captions, ready to export.
Developers can generate images through the VideoGen API with the same quality tier routing as the app, so requests reach models like GPT Image without managing provider keys.
curl -X POST "https://api.videogen.io/v1/tools/generate-image" \
-H "Authorization: Bearer $VIDEOGEN_API_KEY" \
-H "Content-Type: application/json" \
-d '{"prompt": "A studio product shot of a ceramic mug", "quality": "max"}'
Describe exactly what you want, then refine it with a prompt, no manual editing tools required.
“VideoGen is awesome for scaling video production and improving turnaround time... The complex process of video editing, which could take days or months, now takes minutes!”
“VideoGen solves the biggest pain points of video production—complexity, cost, and time. In just a few clicks, anyone can create professional, copyright-free videos.”
“VideoGen is the most underrated tool for content creators who want to put out the highest quality content in the shortest amount of time.”
Every VideoGen workflow uses models like GPT Image automatically to build complete videos with visuals, narration, music, and captions.
Sign in to generate images, video clips, and more with the models you choose.