Models/MiniMax

MiniMax

MiniMax models, pricing, and protocols

15 currently available models with pricing synced from the MoleAPI console.

15 models

Compare
M
minimax-m3-90%Official profile verified

MiniMax M3 is MiniMax's next-generation model for million-token context, agents, and office automation, using sparse attention to reduce long-context compute.

defaultx1
Input$0.3 / 1M tokens
Output$1.2 / 1M tokens
Cache read$0.06 / 1M tokens
tempx0.1
Input$0.03 / 1M tokens
Output$0.12 / 1M tokens
Cache read$0.006 / 1M tokens
openaiopenai-responseopenai-response-compactanthropicgeminiopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
M

minimaxai/minimax-m3 is a MiniMax model available through MoleAPI with openai, openai-response, openai-response-compact, anthropic, gemini, openai-alpha-search access. It fits OpenAI compatible, Responses API, Anthropic compatible, Gemini compatible workflows, with pricing synced from the MoleAPI console.

tempx0.1
Input$7.5 / 1M tokens
Output$7.5 / 1M tokens
openaiopenai-responseopenai-response-compactanthropicgeminiopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleCompare
M
MiniMax-M3-70%Official profile verified

MiniMax M3 is MiniMax's next-generation model for million-token context, agents, and office automation, using sparse attention to reduce long-context compute.

View model
defaultx1
Input$0.3 / 1M tokens
Output$1.2 / 1M tokens
Cache read$0.06 / 1M tokens
discountx0.8
Input$0.24 / 1M tokens
Output$0.96 / 1M tokens
Cache read$0.048 / 1M tokens
relayx0.3
Input$0.09 / 1M tokens
Output$0.36 / 1M tokens
Cache read$0.018 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
M
MiniMax-M2.7-70%Official profile verified

MiniMax M2.7 is a mature M2-series model for reasoning, coding, and agents, with a high-speed serving variant.

defaultx1
Input$0.304 / 1M tokens
Output$1.216 / 1M tokens
Cache read$0.06 / 1M tokens
discountx0.8
Input$0.2432 / 1M tokens
Output$0.9728 / 1M tokens
Cache read$0.048 / 1M tokens
relayx0.3
Input$0.0912 / 1M tokens
Output$0.3648 / 1M tokens
Cache read$0.018 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
M
MiniMax-M2.7-highspeed-70%Official profile verified

MiniMax M2.7 is a mature M2-series model for reasoning, coding, and agents, with a high-speed serving variant.

defaultx1
Input$0.608 / 1M tokens
Output$2.432 / 1M tokens
Cache read$0.06 / 1M tokens
discountx0.8
Input$0.4864 / 1M tokens
Output$1.9456 / 1M tokens
Cache read$0.048 / 1M tokens
relayx0.3
Input$0.1824 / 1M tokens
Output$0.7296 / 1M tokens
Cache read$0.018 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
M
MiniMax-M2.5-70%

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.304 / 1M tokens
Output$1.216 / 1M tokens
Cache read$0.0304 / 1M tokens
discountx0.8
Input$0.2432 / 1M tokens
Output$0.9728 / 1M tokens
Cache read$0.0243 / 1M tokens
relayx0.3
Input$0.0912 / 1M tokens
Output$0.3648 / 1M tokens
Cache read$0.0091 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
M

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.608 / 1M tokens
Output$2.432 / 1M tokens
Cache read$0.06 / 1M tokens
discountx0.8
Input$0.4864 / 1M tokens
Output$1.9456 / 1M tokens
Cache read$0.048 / 1M tokens
relayx0.3
Input$0.1824 / 1M tokens
Output$0.7296 / 1M tokens
Cache read$0.018 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
M
MiniMax-M2.1-70%

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.304 / 1M tokens
Output$1.216 / 1M tokens
Cache read$0.06 / 1M tokens
discountx0.8
Input$0.2432 / 1M tokens
Output$0.9728 / 1M tokens
Cache read$0.048 / 1M tokens
relayx0.3
Input$0.0912 / 1M tokens
Output$0.3648 / 1M tokens
Cache read$0.018 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
M
MiniMax-M2-70%

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.304 / 1M tokens
Output$1.216 / 1M tokens
Cache read$0.06 / 1M tokens
discountx0.8
Input$0.2432 / 1M tokens
Output$0.9728 / 1M tokens
Cache read$0.048 / 1M tokens
relayx0.3
Input$0.0912 / 1M tokens
Output$0.3648 / 1M tokens
Cache read$0.018 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
M
minimax-m2.5-70%

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.3 / 1M tokens
Output$1.2 / 1M tokens
Cache read$0.03 / 1M tokens
discountx0.8
Input$0.24 / 1M tokens
Output$0.96 / 1M tokens
Cache read$0.024 / 1M tokens
relayx0.3
Input$0.09 / 1M tokens
Output$0.36 / 1M tokens
Cache read$0.009 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
M

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.6 / 1M tokens
Output$2.4 / 1M tokens
Cache read$0.06 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
M

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.3 / 1M tokens
Output$2.4 / 1M tokens
Cache read$0.06 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
M

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.3 / 1M tokens
Output$1.2 / 1M tokens
Cache read$0.06 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
M

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.3 / 1M tokens
Output$2.4 / 1M tokens
Cache read$0.06 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
M

MiniMax model by MiniMax for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.3 / 1M tokens
Output$1.2 / 1M tokens
openaiOpenAI compatibleCompare