OpenAI's flagship multimodal model with vision, code generation, and function calling. Excellent all-round performance.
Context Window
128K
tokens
Max Output
16K
tokens
Input Price
$2.5
per 1M tokens
Output Price
$10
per 1M tokens
Speed
109
tokens/sec
Released
May 2024
2024-05-13
Blended Cost
$6.25
per 1M tokens
Value Score
13.6
quality per $
Capabilities
ChatVisionFunction CallingCode Generation
Benchmarks
Quality Index
85
MMLU Pro
88.7
HumanEval (Coding)
90.2
MATH
76.6
Arena ELO
1285
Price vs Quality
Cheaper is left, more capable is up. 27 of 62 priced models are both cheaper and score at least as high. Self-host models are left off the price axis — their cost depends on the hardware you run them on, not on tokens.
Context Window in Context
Grok 3131K
▸ GPT-4o128K
GPT-4o Mini128K
Mistral Large 2128K
DeepSeek V3128K
DeepSeek R1128K
Command R+128K
Gemma 4 27B128K
Bar length is logarithmic — each full step is 10× the tokens. Shown against the models with the nearest context windows, not the directory extremes.
Ограничено 1024 токенами вывода и 20 запросами в час. Полный контроль — в playground.
About GPT-4o
GPT-4o is a flagship AI model by OpenAI, released on May 13, 2024. It supports a context window of 128K tokens and can generate up to 16K output tokens.
At $2.5 per million input tokens and $10 per million output tokens, its blended cost of $6.25/1M tokens places it in the premium pricing tier. Its value score of 13.6 reflects the balance of quality and cost.
Using GPT-4o with Swfte
Access GPT-4o through Swfte Connect, our unified LLM gateway. Connect gives you a single API for 50+ models, with automatic routing, cost optimization, and fallback handling. You can also try GPT-4o in our AI Playground before integrating.
Deploy a model with Swfte Connect
One gateway, every provider, per-token cost visibility. Swap models without touching your code.