Introducing GPT-4oOpenAIReleased May 13, 2024

GPT-4o

Flagship omnimodal model natively trained across text, vision, and audio, doubling generation speed at half the cost of GPT-4 Turbo.

multimodal-realtimeProprietary API$2.50 / 1M tok | $1.25 cachedContext: 128K (128,000 tokens)Arena ELO: 1330

Technical Specifications

Architecture Type
Native Omnimodal Transformer
Total Parameters
Frontier Scale
Context Window
128K (128,000 tokens)
Max Output Tokens
16.384K (16,384 tokens)
Knowledge Cutoff
October 2023
Supported Modalities
text, code, vision, audio
License & Access
OpenAI Business Terms

Benchmark Evaluations

Math
76.6
Mgsm
87.9
Mmlu
88.7
Humaneval
90.2
Chatbot Arena ELO
1330

Deep Architectural Overview

GPT-4o (“o” for “omni”) represents OpenAI’s unified multimodal flagship. It accepts text and vision inputs, generates structured outputs with 100% schema compliance, and matches GPT-4 Turbo intelligence with 2x faster token generation.

Strengths & Considerations

Core Strengths
  • 2x faster generation than GPT-4 Turbo
  • 50% cheaper token cost
  • Strict JSON schema structured output compliance
Known Limitations
  • Complex multi-step Olympiad logic surpassed by o-series reasoning models

Token & API Pricing

Input Tokens (1M)$2.50
Output Tokens (1M)$10.00
Cached Input (1M)$1.2500
Pricing is verified directly against OpenAI's developer documentation and API rate sheets.

Similar & Alternative Models

Explore other frontier models from OpenAI and comparable reasoning engines.

Browse all models
OpenAI

OpenAI GPT-6 model for focused, high-volume workloads where cost efficiency is the priority.

1.05M ctxUSD 0.10 / 1M input tok; USD 0.50 / 1M output tok
OpenAI

OpenAI GPT-6 model built for complex coding and agentic workflows, balancing intelligence and cost.

1.05M ctxUSD 2 / 1M input tok; USD 10 / 1M output tok
OpenAI

OpenAI’s premier sixth-generation frontier model, built for the most demanding end-to-end software engineering and scientific research tasks.

1.05M ctx$10.00 / 1M tok (GPT-6 Frontier)