Models/OpenAI

OpenAI

OpenAI models, pricing, and protocols

125 currently available models with pricing synced from the MoleAPI console.

125 models

Compare
O
gpt-5.6-luna-90%Official profile verified

GPT-5.6 Luna is the efficient GPT-5.6 tier for cost-sensitive, high-volume workloads while retaining long context, reasoning controls, and agent tools.

View model
defaultx1
standard
Length ≤ 272K
Input$0.2 / 1M tokens
Output$1.2 / 1M tokens
Cache read$0.02 / 1M tokens
Cache write$0.25 / 1M tokens
long_context
Input$0.4 / 1M tokens
Output$1.8 / 1M tokens
Cache read$0.04 / 1M tokens
Cache write$0.5 / 1M tokens
Image input$8 / 1M tokens
Image output$30 / 1M tokens
discountx0.8
standard
Length ≤ 272K
Input$0.16 / 1M tokens
Output$0.96 / 1M tokens
Cache read$0.016 / 1M tokens
Cache write$0.2 / 1M tokens
long_context
Input$0.32 / 1M tokens
Output$1.44 / 1M tokens
Cache read$0.032 / 1M tokens
Cache write$0.4 / 1M tokens
Image input$6.4 / 1M tokens
Image output$24 / 1M tokens
relayx0.3
standard
Length ≤ 272K
Input$0.06 / 1M tokens
Output$0.36 / 1M tokens
Cache read$0.006 / 1M tokens
Cache write$0.075 / 1M tokens
long_context
Input$0.12 / 1M tokens
Output$0.54 / 1M tokens
Cache read$0.012 / 1M tokens
Cache write$0.15 / 1M tokens
Image input$2.4 / 1M tokens
Image output$9 / 1M tokens
tempx0.1
standard
Length ≤ 272K
Input$0.02 / 1M tokens
Output$0.12 / 1M tokens
Cache read$0.002 / 1M tokens
Cache write$0.025 / 1M tokens
long_context
Input$0.04 / 1M tokens
Output$0.18 / 1M tokens
Cache read$0.004 / 1M tokens
Cache write$0.05 / 1M tokens
Image input$0.8 / 1M tokens
Image output$3 / 1M tokens
openaiopenai-responseanthropicgeminiimage-generationopenai-response-compactopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationVisionCompare
O
gpt-5.6-sol-90%Official profile verified

GPT-5.6 Sol is OpenAI's flagship GPT-5.6 model for complex professional work, with an emphasis on difficult reasoning, software engineering, research, and tool-heavy agent workflows.

View model
defaultx1
standard
Length ≤ 272K
Input$5 / 1M tokens
Output$30 / 1M tokens
Cache read$0.5 / 1M tokens
Cache write$6.25 / 1M tokens
Image input$8 / 1M tokens
Image output$30 / 1M tokens
long_context
Input$10 / 1M tokens
Output$45 / 1M tokens
Cache read$1 / 1M tokens
Cache write$12.5 / 1M tokens
Image input$8 / 1M tokens
Image output$30 / 1M tokens
discountx0.8
standard
Length ≤ 272K
Input$4 / 1M tokens
Output$24 / 1M tokens
Cache read$0.4 / 1M tokens
Cache write$5 / 1M tokens
Image input$6.4 / 1M tokens
Image output$24 / 1M tokens
long_context
Input$8 / 1M tokens
Output$36 / 1M tokens
Cache read$0.8 / 1M tokens
Cache write$10 / 1M tokens
Image input$6.4 / 1M tokens
Image output$24 / 1M tokens
relayx0.3
standard
Length ≤ 272K
Input$1.5 / 1M tokens
Output$9 / 1M tokens
Cache read$0.15 / 1M tokens
Cache write$1.875 / 1M tokens
Image input$2.4 / 1M tokens
Image output$9 / 1M tokens
long_context
Input$3 / 1M tokens
Output$13.5 / 1M tokens
Cache read$0.3 / 1M tokens
Cache write$3.75 / 1M tokens
Image input$2.4 / 1M tokens
Image output$9 / 1M tokens
tempx0.1
standard
Length ≤ 272K
Input$0.5 / 1M tokens
Output$3 / 1M tokens
Cache read$0.05 / 1M tokens
Cache write$0.625 / 1M tokens
Image input$0.8 / 1M tokens
Image output$3 / 1M tokens
long_context
Input$1 / 1M tokens
Output$4.5 / 1M tokens
Cache read$0.1 / 1M tokens
Cache write$1.25 / 1M tokens
Image input$0.8 / 1M tokens
Image output$3 / 1M tokens
AA index 58openaiopenai-responseanthropicgeminiimage-generationopenai-response-compactopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationVisionCompare
O
gpt-5.6-terra-90%Official profile verified

GPT-5.6 Terra is the balanced tier in OpenAI's GPT-5.6 family, positioned between Sol's maximum capability and Luna's lower operating cost.

View model
defaultx1
standard
Length ≤ 272K
Input$2 / 1M tokens
Output$12 / 1M tokens
Cache read$0.2 / 1M tokens
Cache write$2.5 / 1M tokens
Image input$8 / 1M tokens
Image output$30 / 1M tokens
long_context
Input$4 / 1M tokens
Output$18 / 1M tokens
Cache read$0.4 / 1M tokens
Cache write$5 / 1M tokens
Image input$8 / 1M tokens
Image output$30 / 1M tokens
discountx0.8
standard
Length ≤ 272K
Input$1.6 / 1M tokens
Output$9.6 / 1M tokens
Cache read$0.16 / 1M tokens
Cache write$2 / 1M tokens
Image input$6.4 / 1M tokens
Image output$24 / 1M tokens
long_context
Input$3.2 / 1M tokens
Output$14.4 / 1M tokens
Cache read$0.32 / 1M tokens
Cache write$4 / 1M tokens
Image input$6.4 / 1M tokens
Image output$24 / 1M tokens
relayx0.3
standard
Length ≤ 272K
Input$0.6 / 1M tokens
Output$3.6 / 1M tokens
Cache read$0.06 / 1M tokens
Cache write$0.75 / 1M tokens
Image input$2.4 / 1M tokens
Image output$9 / 1M tokens
long_context
Input$1.2 / 1M tokens
Output$5.4 / 1M tokens
Cache read$0.12 / 1M tokens
Cache write$1.5 / 1M tokens
Image input$2.4 / 1M tokens
Image output$9 / 1M tokens
tempx0.1
standard
Length ≤ 272K
Input$0.2 / 1M tokens
Output$1.2 / 1M tokens
Cache read$0.02 / 1M tokens
Cache write$0.25 / 1M tokens
Image input$0.8 / 1M tokens
Image output$3 / 1M tokens
long_context
Input$0.4 / 1M tokens
Output$1.8 / 1M tokens
Cache read$0.04 / 1M tokens
Cache write$0.5 / 1M tokens
Image input$0.8 / 1M tokens
Image output$3 / 1M tokens
openaiopenai-responseanthropicgeminiimage-generationopenai-response-compactopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationVisionCompare
O
gpt-5.5-90%Official profile verified

GPT-5.5 is OpenAI's general flagship for reasoning, coding, and professional knowledge work, with a 1M context window, multimodal input, and tool workflows.

View model
defaultx1
standard
Length ≤ 272K
Input$5 / 1M tokens
Output$30 / 1M tokens
Cache read$0.5 / 1M tokens
Cache write$6.25 / 1M tokens
long_context
Input$10 / 1M tokens
Output$45 / 1M tokens
Cache read$1 / 1M tokens
Cache write$6.25 / 1M tokens
discountx0.8
standard
Length ≤ 272K
Input$4 / 1M tokens
Output$24 / 1M tokens
Cache read$0.4 / 1M tokens
Cache write$5 / 1M tokens
long_context
Input$8 / 1M tokens
Output$36 / 1M tokens
Cache read$0.8 / 1M tokens
Cache write$5 / 1M tokens
relayx0.3
standard
Length ≤ 272K
Input$1.5 / 1M tokens
Output$9 / 1M tokens
Cache read$0.15 / 1M tokens
Cache write$1.875 / 1M tokens
long_context
Input$3 / 1M tokens
Output$13.5 / 1M tokens
Cache read$0.3 / 1M tokens
Cache write$1.875 / 1M tokens
tempx0.1
standard
Length ≤ 272K
Input$0.5 / 1M tokens
Output$3 / 1M tokens
Cache read$0.05 / 1M tokens
Cache write$0.625 / 1M tokens
long_context
Input$1 / 1M tokens
Output$4.5 / 1M tokens
Cache read$0.1 / 1M tokens
Cache write$0.625 / 1M tokens
openaiopenai-responseanthropicgeminiimage-generationopenai-response-compactopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationPrompt cacheCompare
O
gpt-5.4-90%

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$2.5 / 1M tokens
Output$15 / 1M tokens
Cache read$0.25 / 1M tokens
discountx0.8
Input$2 / 1M tokens
Output$12 / 1M tokens
Cache read$0.2 / 1M tokens
relayx0.3
Input$0.75 / 1M tokens
Output$4.5 / 1M tokens
Cache read$0.075 / 1M tokens
tempx0.1
Input$0.25 / 1M tokens
Output$1.5 / 1M tokens
Cache read$0.025 / 1M tokens
openaiopenai-responseanthropicgeminiimage-generationopenai-response-compactopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$2.5 / 1M tokens
Output$15 / 1M tokens
Cache read$0.25 / 1M tokens
discountx0.8
Input$2 / 1M tokens
Output$12 / 1M tokens
Cache read$0.2 / 1M tokens
relayx0.3
Input$0.75 / 1M tokens
Output$4.5 / 1M tokens
Cache read$0.075 / 1M tokens
tempx0.1
Input$0.25 / 1M tokens
Output$1.5 / 1M tokens
Cache read$0.025 / 1M tokens
openaiopenai-responseanthropicgeminiimage-generationopenai-response-compactopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationPrompt cacheCompare
O
gpt-5.4-mini-90%

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$0.75 / 1M tokens
Output$4.5 / 1M tokens
Cache read$0.075 / 1M tokens
discountx0.8
Input$0.6 / 1M tokens
Output$3.6 / 1M tokens
Cache read$0.06 / 1M tokens
relayx0.3
Input$0.225 / 1M tokens
Output$1.35 / 1M tokens
Cache read$0.0225 / 1M tokens
tempx0.1
Input$0.075 / 1M tokens
Output$0.45 / 1M tokens
Cache read$0.0075 / 1M tokens
openaiopenai-responseanthropicgeminiimage-generationopenai-response-compactopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationPrompt cacheCompare
O

GPT coding model by OpenAI for code generation, repository understanding, debugging, and agentic development.

View model
defaultx1
Input$1.75 / 1M tokens
Output$14 / 1M tokens
Cache read$0.175 / 1M tokens
discountx0.8
Input$1.4 / 1M tokens
Output$11.2 / 1M tokens
Cache read$0.14 / 1M tokens
relayx0.3
Input$0.525 / 1M tokens
Output$4.2 / 1M tokens
Cache read$0.0525 / 1M tokens
openaiopenai-responseanthropicgeminiimage-generationOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationCodingCompare
O

GPT coding model by OpenAI for code generation, repository understanding, debugging, and agentic development.

defaultx1
Input$1.75 / 1M tokens
Output$14 / 1M tokens
Cache read$0.175 / 1M tokens
discountx0.8
Input$1.4 / 1M tokens
Output$11.2 / 1M tokens
Cache read$0.14 / 1M tokens
relayx0.3
Input$0.525 / 1M tokens
Output$4.2 / 1M tokens
Cache read$0.0525 / 1M tokens
openaiopenai-responseanthropicgeminiimage-generationOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleImage generationCodingCompare
O
gpt-5.5-2026-04-23-20%Official profile verified

GPT-5.5 is OpenAI's general flagship for reasoning, coding, and professional knowledge work, with a 1M context window, multimodal input, and tool workflows.

defaultx1
standard
Length ≤ 272K
Input$5 / 1M tokens
Output$30 / 1M tokens
Cache read$0.5 / 1M tokens
Cache write$6.25 / 1M tokens
long_context
Input$10 / 1M tokens
Output$45 / 1M tokens
Cache read$1 / 1M tokens
Cache write$6.25 / 1M tokens
discountx0.8
standard
Length ≤ 272K
Input$4 / 1M tokens
Output$24 / 1M tokens
Cache read$0.4 / 1M tokens
Cache write$5 / 1M tokens
long_context
Input$8 / 1M tokens
Output$36 / 1M tokens
Cache read$0.8 / 1M tokens
Cache write$5 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O
gpt-5.5-proOfficial profile verified

GPT-5.5 Pro is the high-accuracy GPT-5.5 configuration. It uses parallel test-time compute for harder questions, research, professional knowledge work, and more complete answers.

defaultx1
Input$75 / 1M tokens
Output$450 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O
gpt-5.5-pro-2026-04-23Official profile verified

GPT-5.5 Pro is the high-accuracy GPT-5.5 configuration. It uses parallel test-time compute for harder questions, research, professional knowledge work, and more complete answers.

defaultx1
Input$75 / 1M tokens
Output$450 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.75 / 1M tokens
Output$4.5 / 1M tokens
Cache read$0.075 / 1M tokens
discountx0.8
Input$0.6 / 1M tokens
Output$3.6 / 1M tokens
Cache read$0.06 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O
gpt-5.4-nano-20%

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.2 / 1M tokens
Output$1.25 / 1M tokens
Cache read$0.02 / 1M tokens
discountx0.8
Input$0.16 / 1M tokens
Output$1 / 1M tokens
Cache read$0.016 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.2 / 1M tokens
Output$1.25 / 1M tokens
Cache read$0.02 / 1M tokens
discountx0.8
Input$0.16 / 1M tokens
Output$1 / 1M tokens
Cache read$0.016 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$30 / 1M tokens
Output$180 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$30 / 1M tokens
Output$180 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$1.75 / 1M tokens
Output$14 / 1M tokens
Cache read$0.175 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O
gpt-5.2-20%

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$1.75 / 1M tokens
Output$14 / 1M tokens
Cache read$0.175 / 1M tokens
discountx0.8
Input$1.4 / 1M tokens
Output$11.2 / 1M tokens
Cache read$0.14 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$1.75 / 1M tokens
Output$14 / 1M tokens
Cache read$0.175 / 1M tokens
discountx0.8
Input$1.4 / 1M tokens
Output$11.2 / 1M tokens
Cache read$0.14 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$21 / 1M tokens
Output$168 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$21 / 1M tokens
Output$168 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O
gpt-5.1-20%

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$1.25 / 1M tokens
Output$10 / 1M tokens
Cache read$0.125 / 1M tokens
discountx0.8
Input$1 / 1M tokens
Output$8 / 1M tokens
Cache read$0.1 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$1.25 / 1M tokens
Output$10 / 1M tokens
Cache read$0.125 / 1M tokens
discountx0.8
Input$1 / 1M tokens
Output$8 / 1M tokens
Cache read$0.1 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O
gpt-5-20%

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$1.25 / 1M tokens
Output$10 / 1M tokens
Cache read$0.125 / 1M tokens
discountx0.8
Input$1 / 1M tokens
Output$8 / 1M tokens
Cache read$0.1 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$1.25 / 1M tokens
Output$10 / 1M tokens
Cache read$0.125 / 1M tokens
discountx0.8
Input$1 / 1M tokens
Output$8 / 1M tokens
Cache read$0.1 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$1.25 / 1M tokens
Output$10 / 1M tokens
Cache read$0.125 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT coding model by OpenAI for code generation, repository understanding, debugging, and agentic development.

defaultx1
Input$1.25 / 1M tokens
Output$10 / 1M tokens
Cache read$0.125 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICodingPrompt cacheCompare
O
gpt-5-high-20%

gpt-5-high is a OpenAI model available through MoleAPI with openai, openai-response access. It fits OpenAI compatible, Responses API workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$600 / 1M tokens
discountx0.8
Input$60 / 1M tokens
Output$480 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O
gpt-5-low-20%

gpt-5-low is a OpenAI model available through MoleAPI with openai, openai-response access. It fits OpenAI compatible, Responses API workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$600 / 1M tokens
discountx0.8
Input$60 / 1M tokens
Output$480 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O
gpt-5-medium-20%

gpt-5-medium is a OpenAI model available through MoleAPI with openai, openai-response access. It fits OpenAI compatible, Responses API workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$600 / 1M tokens
discountx0.8
Input$60 / 1M tokens
Output$480 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare
O
gpt-5-mini-20%

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$0.25 / 1M tokens
Output$2 / 1M tokens
Cache read$0.025 / 1M tokens
discountx0.8
Input$0.2 / 1M tokens
Output$1.6 / 1M tokens
Cache read$0.02 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.25 / 1M tokens
Output$2 / 1M tokens
Cache read$0.025 / 1M tokens
discountx0.8
Input$0.2 / 1M tokens
Output$1.6 / 1M tokens
Cache read$0.02 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O
gpt-5-nano-20%

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$0.05 / 1M tokens
Output$0.4 / 1M tokens
Cache read$0.005 / 1M tokens
discountx0.8
Input$0.04 / 1M tokens
Output$0.32 / 1M tokens
Cache read$0.004 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.05 / 1M tokens
Output$0.4 / 1M tokens
Cache read$0.005 / 1M tokens
discountx0.8
Input$0.04 / 1M tokens
Output$0.32 / 1M tokens
Cache read$0.004 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APIPrompt cacheCompare
O

GPT model by OpenAI for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$15 / 1M tokens
Output$120 / 1M tokens
openaiopenai-responseOpenAI compatibleResponses APICompare