OpenAI
BasicPro
GPT-4o
gpt-4o
Full multimodal with vision and DALL-E 3 image generation — OpenAI's visual creator.
Technische Spezifikationen
| API-Kennung | gpt-4o |
|---|---|
| Tarife | Basic Pro |
| Maximaler Kontext | 128 000 tokens |
| Bildverständnis (Vision) | ✓ Ja |
| Bildgenerierung | ✓ DALL-E 3 |
| Nativer Webzugriff | Nein |
| Eingabekosten | $2.50 / M tokens |
| Ausgabekosten | $10.00 / M tokens |
Überblick
GPT-4o is OpenAI's leading multimodal model, combining vision and image generation via DALL-E 3. It is Mingo's reference model for all requests involving visual creation. It can analyze images and generate new ones within the same conversation, offering a complete multimodal experience on Basic and Pro plans.
Stärken
DALL-E 3 GenerationAdvanced VisionFull Multimodal
Cas d'usage
Illustration Generation
Creating images, logos, illustrations, and visuals from text descriptions.
Analyze and Create
Analyze an existing image and then create a variant or improvement.
Marketing Visual Content
Producing visuals for social media, presentations, and marketing materials.
Preise
Eingabe-Tokens
$2.50
pro Million Tokens
Ausgabe-Tokens
$10.00
pro Million Tokens
Verwandte Modelle
Häufig gestellte Fragen
What is the quality of images generated by GPT-4o?
GPT-4o uses DALL-E 3, OpenAI's most advanced image generator, producing high-quality images with a strong understanding of complex instructions.
Can you use GPT-4o to both analyze AND create images in the same session?
Yes, GPT-4o is fully multimodal: it can analyze an image you provide, then generate a new image based on the analysis, all in a single conversation.
Create images with GPT-4o on Mingo
Available on Basic and Pro — generate DALL-E 3 images directly in your conversation.
Try Mingo