OpenAI
BasicPro

GPT-4o

gpt-4o

Full multimodal with vision and DALL-E 3 image generation — OpenAI's visual creator.

图像生成图像识别对话

技术规格

API 标识符gpt-4o
订阅计划Basic Pro
最大上下文128 000 tokens
视觉(图像)✓ 是
图像生成✓ DALL-E 3
原生网络访问
输入成本$2.50 / M tokens
输出成本$10.00 / M tokens

概述

GPT-4o is OpenAI's leading multimodal model, combining vision and image generation via DALL-E 3. It is Mingo's reference model for all requests involving visual creation. It can analyze images and generate new ones within the same conversation, offering a complete multimodal experience on Basic and Pro plans.


优势

DALL-E 3 GenerationAdvanced VisionFull Multimodal

Cas d'usage

Illustration Generation

Creating images, logos, illustrations, and visuals from text descriptions.

Analyze and Create

Analyze an existing image and then create a variant or improvement.

Marketing Visual Content

Producing visuals for social media, presentations, and marketing materials.

定价

输入 Token
$2.50
每百万 Token
输出 Token
$10.00
每百万 Token

常见问题

What is the quality of images generated by GPT-4o?
GPT-4o uses DALL-E 3, OpenAI's most advanced image generator, producing high-quality images with a strong understanding of complex instructions.
Can you use GPT-4o to both analyze AND create images in the same session?
Yes, GPT-4o is fully multimodal: it can analyze an image you provide, then generate a new image based on the analysis, all in a single conversation.

Create images with GPT-4o on Mingo

Available on Basic and Pro — generate DALL-E 3 images directly in your conversation.

Try Mingo