GPT-4o
MultimodalGPT-4o ('o' stands for 'omni') is a multimodal AI model that accepts text, audio, image, and video inputs and generates text, audio, and image outputs. It matches GPT-4 Turbo performance on text and code, with improvements in understanding non-English languages, images, and audio.
Key Specifications
Parameters
-
Context
128.0K
Release Date
August 6, 2024
Average Score
52.5%
Timeline
Key dates in the model's history
Announcement
August 6, 2024
Last Update
July 19, 2025
Today
October 5, 2026
Technical Specifications
Parameters
-
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval
Pricing & Availability
Input (per 1M tokens)
$2.50
Output (per 1M tokens)
$10.00
Max Input Tokens
128.0K
Max Output Tokens
16.4K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning
Benchmark Results
Model performance metrics across various tests and benchmarks
General Knowledge
Tests on general knowledge and understanding
MMLU
Accuracy • Self-reported
Programming
Programming skills tests
SWE-Bench Verified
Accuracy • Self-reported
Reasoning
Logical reasoning and analysis
GPQA
GPT-4o - Diamond no thinking no tools • Self-reported
Multimodal
Working with images and visual data
MathVista
Accuracy • Self-reported
MMMU
GPT-4o without mode thinking - Solution visual tasks level with • Self-reported
ChartQA
evaluation on test set • Self-reported
DocVQA
evaluation on test set • Self-reported
AI2D
Evaluation on test set AI: Translate full text • Self-reported
Other Tests
Specialized benchmarks
MMLU-Pro
0-shot CoT • Self-reported
SimpleQA
accuracy • Self-reported
AIME 2024
Accuracy
AI • Self-reported
IFEval
Accuracy
AI • Self-reported
Aider-Polyglot
Accuracy
AI • Self-reported
EgoSchema
evaluation on test set • Self-reported
Aider-Polyglot Edit
Accuracy
AI • Self-reported
MMMLU
Accuracy • Self-reported
Multi-IF
Accuracy • Self-reported
TAU-bench Retail
Accuracy • Self-reported
TAU-bench Airline
Accuracy • Self-reported
CharXiv-R
GPT-4o without mode thinking - justification and • Self-reported
Internal API instruction following (hard)
Accuracy • Self-reported
MultiChallenge (o3-mini grader)
Accuracy AI: translation text by evaluation accuracy models AI: Accuracy • Self-reported
COLLIE
GPT-4o without mode thinking - instructions at text • Self-reported
Tau2 airline
GPT-4o without mode thinking - Benchmark functions () • Self-reported
OpenAI-MRCR: 2 needle 128k
Accuracy
AI • Self-reported
Tau2 retail
GPT-4o without mode thinking - Benchmark functions () • Self-reported
Tau2 telecom
GPT-4o without mode thinking - Benchmark functions (field) • Self-reported
MMMU-Pro
GPT-4o without mode thinking - Solution visual tasks level with reasoning • Self-reported
VideoMMMU
GPT-4o without mode thinking - multimodal reasoning (256 ) • Self-reported
ERQA
GPT-4o without mode thinking - thinking • Self-reported
Graphwalks parents <128k
Accuracy • Self-reported
CharXiv-D
Accuracy AI: PageRank connection how therefore with number receive more rating When receives with she/it receives part "" this is "" to • Self-reported
ComplexFuncBench
Accuracy • Self-reported
SWE-Lancer
result • Self-reported
SWE-Lancer (IC-Diamond subset)
score • Self-reported
Graphwalks BFS <128k
Accuracy • Self-reported
ActivityNet
evaluation on test set • Self-reported
Humanity's Last Exam
GPT-4o without mode thinking (without tools) - set questions expert level by various subjects • Self-reported
Scale MultiChallenge
GPT-4o without mode thinking - Benchmark execution instructions • Self-reported
Multi-Challenge
GPT-4o without thinking mode - Multi-turn instruction following benchmark. • Self-reported
License & Metadata
License
proprietary
Announcement Date
August 6, 2024
Last Updated
July 19, 2025
Similar Models
All ModelsGPT-4.1 mini
OpenAI
MM
Best score:0.9 (MMLU)
Released:Apr 2025
Price:$0.40/1M tokens
o4-mini
OpenAI
MM
Best score:0.8 (GPQA)
Released:Apr 2025
Price:$1.10/1M tokens
o3
OpenAI
MM
Best score:0.8 (GPQA)
Released:Apr 2025
Price:$2.00/1M tokens
GPT-4o
OpenAI
MM
Best score:0.9 (HumanEval)
Released:May 2024
Price:$2.50/1M tokens
GPT-4o mini
OpenAI
MM
Best score:0.9 (HumanEval)
Released:Jul 2024
Price:$0.15/1M tokens
GPT-4.5
OpenAI
MM
Best score:0.9 (MMLU)
Released:Feb 2025
Price:$75.00/1M tokens
GPT-4.1
OpenAI
MM
Best score:0.9 (MMLU)
Released:Apr 2025
Price:$2.00/1M tokens
GPT-5 nano
OpenAI
MM
Best score:0.7 (GPQA)
Released:Aug 2025
Price:$0.05/1M tokens
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.