OpenAI: GPT-5.4
GPT-5.4 from OpenAI is equipped to handle text, images, and files, with an impressive context length of 1,050,000 tokens, making it versatile for complex tasks. It supports reasoning and tools integration, facilitating a more practical approach to problem-solving. With a blended benchmark score of 90.3 across four independent tests, GPT-5.4 performs well in both general AI capabilities and specific coding tasks. For those requiring robust multi-modal processing and detailed reasoning, despite the non-free nature and pricing at $2.5 per million input tokens and $15 per million output tokens, GPT-5.4 is a strong contender, especially if your project involves extensive text or image analysis and you are willing to invest in high-quality AI solutions.
Benchmark results
Independent, published benchmarks. Blended score 90.3 across 4 benchmarks, last refreshed 2026-08-13. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 87.7 |
| AI Index Coding | software engineering tasks | 100.0 |
| AI Index Agentic | multi-step tool-using tasks | 72.9 |
| EQ-Bench | emotional understanding in dialogue | 82.4 |
- Model ID
- openai/gpt-5.4
- Vendor
- openai
- Released
- March 2026
- Tokenizer
- GPT
- Input Modalities
- text, image, file
- Output Modalities
- text
- Max Output
- 128,000 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- ✓ accepts images
- Audio
- no
- Moderated
- yes
What it costs in practice
Computed from the current $2.50/M input and $15.00/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | $0.10 |
| Classify 1,000 customer emails | 500k in / 50k out | $2.00 |
| A month of a busy support chatbot | 5M in / 2M out | $42.50 |
Category rankings
Where OpenAI: GPT-5.4 places across the 56 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #7 | Code ReviewCode · of 25 ranked | 176 |
| #7 | Code RefactoringCode · of 25 ranked | 174 |
| #7 | Unit Test GenerationCode · of 25 ranked | 160 |
| #7 | CI/CD PipelinesCode · of 25 ranked | 150 |
| #7 | Frontend Component DesignCode · of 25 ranked | 149 |
| #7 | ETL ScriptingData · of 25 ranked | 162 |
| #7 | OCR / Document ParsingData · of 25 ranked | 149 |
| #7 | Table Extraction from PDFsData · of 25 ranked | 149 |
| #7 | Creative WritingWriting · of 25 ranked | 167 |
| #7 | Screenwriting & DialogWriting · of 25 ranked | 167 |
| #7 | Academic WritingWriting · of 25 ranked | 167 |
| #7 | Legal DraftingProfessional · of 25 ranked | 167 |
| #7 | Legal ResearchProfessional · of 25 ranked | 179 |
| #7 | Contract ReviewProfessional · of 25 ranked | 189 |
| #7 | Medical Note SummarizationProfessional · of 25 ranked | 162 |
| #7 | Scientific ResearchProfessional · of 25 ranked | 179 |
| #7 | Math TutoringEducation · of 25 ranked | 152 |
| #7 | Physics TutoringEducation · of 25 ranked | 154 |
| #7 | History TutoringEducation · of 25 ranked | 152 |
| #7 | Essay GradingEducation · of 25 ranked | 176 |
| #7 | RFP ResponseBusiness · of 25 ranked | 184 |
| #7 | OKRs & Strategic PlanningBusiness · of 25 ranked | 167 |
| #7 | Screenshot DebuggingVision · of 25 ranked | 146 |
| #7 | RAG PipelinesAgents · of 25 ranked | 174 |
| #7 | Long-Context Q&AAgents · of 25 ranked | 165 |
| #7 | Fiction CollaboratorPersonal · of 25 ranked | 179 |
| #7 | Fitness CoachingPersonal · of 25 ranked | 138 |
| #7 | Math ProofsResearch · of 25 ranked | 167 |
| #7 | Literature ReviewResearch · of 25 ranked | 179 |
| #7 | Experiment DesignResearch · of 25 ranked | 162 |
| #8 | Bug FixingCode · of 25 ranked | 189 |
| #8 | CSV / Spreadsheet CleanupData · of 25 ranked | 157 |
| #8 | Long-Document SummarizationWriting · of 25 ranked | 169 |
| #8 | Blog Post WritingWriting · of 25 ranked | 144 |
| #8 | Marketing CopyWriting · of 25 ranked | 132 |
| #8 | Standardized Test PrepEducation · of 25 ranked | 132 |
| #8 | Agent WorkflowsAgents · of 25 ranked | 201 |
| #8 | Coding AgentsAgents · of 25 ranked | 201 |
| #8 | Character RoleplayPersonal · of 25 ranked | 157 |
| #8 | Trip PlanningPersonal · of 25 ranked | 144 |
| #8 | Scientific CodingResearch · of 25 ranked | 189 |
| #9 | Meeting NotesBusiness · of 25 ranked | 156 |
| #9 | Chart & Graph ReadingVision · of 25 ranked | 155 |
| #10 | SQL GenerationCode · of 25 ranked | 175 |
| #10 | Code DocumentationCode · of 25 ranked | 147 |
| #12 | Transcript CleanupWriting · of 25 ranked | 145 |
| #12 | Financial AnalysisProfessional · of 25 ranked | 163 |
| #12 | Resume WritingBusiness · of 25 ranked | 118 |
| #12 | Diagram ExtractionVision · of 25 ranked | 151 |
| #13 | Data AnalysisData · of 25 ranked | 175 |
| #13 | Browser AutomationAgents · of 25 ranked | 174 |
| #18 | Regex WritingCode · of 25 ranked | 136 |
| #18 | Language TranslationWriting · of 25 ranked | 136 |
| #18 | Tutoring for KidsPersonal · of 25 ranked | 136 |
| #20 | SOP / Process DocsBusiness · of 25 ranked | 137 |
| #20 | Function / Tool CallingAgents · of 25 ranked | 153 |
Similar models
OpenAI: GPT-5.6 Luna Pro
OpenAI: GPT-5.6 Luna Pro (batch)
OpenAI: GPT-5.6 Luna
OpenAI: GPT-5.6 Luna (batch)
OpenAI: GPT-5.6 Terra Pro
OpenAI: GPT-5.6 Terra Pro (batch)
Quick answers
- How much does OpenAI: GPT-5.4 cost?
- $2.50 per million input tokens and $15.00 per million output tokens.
- What is OpenAI: GPT-5.4's context window?
- 1,050,000 tokens, roughly 1,575 pages of text in a single request.
- Does OpenAI: GPT-5.4 support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can OpenAI: GPT-5.4 process images?
- Yes, it accepts image input.