Models/Google

Google

Google models, pricing, and protocols

40 currently available models with pricing synced from the MoleAPI console.

40 models

Compare
G
gemini-3.5-flash-20%Official profile verified

Gemini 3.5 Flash is Google's multimodal model. It accepts text, images, video, audio, and PDF input and produces text output.

View model
defaultx1
Input$1.5 / 1M tokens
Output$9 / 1M tokens
Cache read$0.15 / 1M tokens
discountx0.8
Input$1.2 / 1M tokens
Output$7.2 / 1M tokens
Cache read$0.12 / 1M tokens
AA index 50openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleVisionPrompt cacheCompare
G
gemini-3.1-pro-20%Official profile verified

Gemini 3.1 Pro is Google's flagship multimodal model for complex reasoning, code, and long-form generation, with 1M context and 64K output.

defaultx1
Input$2 / 1M tokens
Output$12 / 1M tokens
Cache read$0.2 / 1M tokens
discountx0.8
Input$1.6 / 1M tokens
Output$9.6 / 1M tokens
Cache read$0.16 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleVisionPrompt cacheCompare
G
gemini-3-flash-20%Official profile verified

Gemini 3 Flash is the Gemini 3 model for low latency, high throughput, and cost-efficient multimodal work.

defaultx1
Input$0.5 / 1M tokens
Output$3 / 1M tokens
Cache read$0.05 / 1M tokens
discountx0.8
Input$0.4 / 1M tokens
Output$2.4 / 1M tokens
Cache read$0.04 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-3.7-flash is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision, Prompt cache workflows, with pricing synced from the MoleAPI console.

View model
defaultx1
Input$0.75 / 1M tokens
Output$3.75 / 1M tokens
Cache read$0.075 / 1M tokens
discountx0.8
Input$0.6 / 1M tokens
Output$3 / 1M tokens
Cache read$0.06 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-3.6-flash is a Google model available through MoleAPI with openai, gemini access. It fits OpenAI compatible, Gemini compatible, Vision, Prompt cache workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$1.5 / 1M tokens
Output$7.5 / 1M tokens
Cache read$0.15 / 1M tokens
discountx0.8
Input$1.2 / 1M tokens
Output$6 / 1M tokens
Cache read$0.12 / 1M tokens
openaigeminiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-3.5-flash-lite is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision, Prompt cache workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$0.3 / 1M tokens
Output$2.5 / 1M tokens
Cache read$0.03 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-3.1-flash-image is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$300 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

Image image model by Google for image generation, editing, visual creation, and creative production.

View model
defaultx1
Input$0.5 / 1M tokens
Output$3 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

gemini-3.1-flash-image-preview-t is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$300 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

gemini-3.1-flash-lite is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$300 / 1M tokens
discountx0.8
Input$60 / 1M tokens
Output$240 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

gemini-3.1-flash-lite-image is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$300 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

Gemini model by Google for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$0.25 / 1M tokens
Output$1.5 / 1M tokens
Cache read$0.025 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-3.1-flash-tts-preview is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$1 / 1M tokens
Output$20 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G
gemini-3.1-pro-preview-20%Official profile verified

Gemini 3.1 Pro is Google's flagship multimodal model for complex reasoning, code, and long-form generation, with 1M context and 64K output.

View model
defaultx1
Input$2 / 1M tokens
Output$12 / 1M tokens
Cache read$0.2 / 1M tokens
discountx0.8
Input$1.6 / 1M tokens
Output$9.6 / 1M tokens
Cache read$0.16 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-3.1-pro-preview-high is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$300 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

gemini-3.1-pro-preview-low is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$300 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G
gemini-3-flash-preview-20%Official profile verified

Gemini 3 Flash is the Gemini 3 model for low latency, high throughput, and cost-efficient multimodal work.

View model
defaultx1
Input$0.5 / 1M tokens
Output$3 / 1M tokens
Cache read$0.05 / 1M tokens
discountx0.8
Input$0.4 / 1M tokens
Output$2.4 / 1M tokens
Cache read$0.04 / 1M tokens
openaigeminiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-3-pro-image is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$4500 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

Image image model by Google for image generation, editing, visual creation, and creative production.

View model
defaultx1
Input$2 / 1M tokens
Output$12 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

Gemini model by Google for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$2 / 1M tokens
Output$12 / 1M tokens
Cache read$0.2 / 1M tokens
discountx0.8
Input$1.6 / 1M tokens
Output$9.6 / 1M tokens
Cache read$0.16 / 1M tokens
openaigeminiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G
gemini-2.5-flash-20%Official profile verified

Gemini 2.5 Flash is Google's mature model for speed, cost, and scaled multimodal processing, retaining 1M context and controllable thinking.

View model
defaultx1
Input$0.3 / 1M tokens
Output$2.5 / 1M tokens
Cache read$0.075 / 1M tokens
discountx0.8
Input$0.24 / 1M tokens
Output$2 / 1M tokens
Cache read$0.06 / 1M tokens
openaigeminiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

Image image model by Google for image generation, editing, visual creation, and creative production.

defaultx1
Input$0.3 / 1M tokens
Output$2.5 / 1M tokens
Cache read$0.03 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

Gemini model by Google for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.1 / 1M tokens
Output$0.4 / 1M tokens
Cache read$0.025 / 1M tokens
discountx0.8
Input$0.08 / 1M tokens
Output$0.32 / 1M tokens
Cache read$0.02 / 1M tokens
openaigeminiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-2.5-flash-nothinking is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision, Reasoning workflows, with pricing synced from the MoleAPI console.

View model
defaultx1
Input$0.3 / 1M tokens
Output$2.5 / 1M tokens
discountx0.8
Input$0.24 / 1M tokens
Output$2 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionReasoningCompare
G

gemini-2.5-flash-preview-tts is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$0.5 / 1M tokens
Output$10 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G
gemini-2.5-pro-20%Official profile verified

Gemini 2.5 Pro is Google's mature long-context multimodal reasoning model for complex code, research, and media understanding that benefits from a stable version.

defaultx1
Input$1.25 / 1M tokens
Output$10 / 1M tokens
Cache read$0.125 / 1M tokens
discountx0.8
Input$1 / 1M tokens
Output$8 / 1M tokens
Cache read$0.1 / 1M tokens
openaigeminiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-2.5-pro-preview-tts is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$1 / 1M tokens
Output$20 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G
gemini-2.5-pro-previewOfficial profile verified

Gemini 2.5 Pro is Google's mature long-context multimodal reasoning model for complex code, research, and media understanding that benefits from a stable version.

defaultx1
Input$1.25 / 1M tokens
Output$10 / 1M tokens
Cache read$0.125 / 1M tokens
openaiOpenAI compatibleVisionPrompt cacheCompare
G

gemini-robotics-er-1.6-preview is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$75 / 1M tokens
Output$300 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionCompare
G

Gemini model by Google for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.1 / 1M tokens
Output$0.4 / 1M tokens
Cache read$0.025 / 1M tokens
openaiOpenAI compatibleVisionPrompt cacheCompare
G

Gemini model by Google for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.07 / 1M tokens
Output$0.28 / 1M tokens
Cache read$0.0175 / 1M tokens
openaiOpenAI compatibleVisionPrompt cacheCompare
G

gemini-embedding-001 is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision, Embedding workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$0.15 / 1M tokens
Output$0 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionEmbeddingCompare
G

gemini-flash-latest is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision, Prompt cache workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$1 / 1M tokens
Output$2.5 / 1M tokens
Cache read$0.075 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

gemini-flash-lite-latest is a Google model available through MoleAPI with gemini, openai access. It fits OpenAI compatible, Gemini compatible, Vision, Prompt cache workflows, with pricing synced from the MoleAPI console.

defaultx1
Input$0.1 / 1M tokens
Output$0.4 / 1M tokens
Cache read$0.025 / 1M tokens
geminiopenaiOpenAI compatibleGemini compatibleVisionPrompt cacheCompare
G

Gemma model by Google for general chat, writing, coding assistance, and tool-assisted workflows.

tempx0.1
Input$7.5 / 1M tokens
Output$7.5 / 1M tokens
openaiopenai-responseopenai-response-compactanthropicgeminiopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleCompare
G

Gemma model by Google for general chat, writing, coding assistance, and tool-assisted workflows.

tempx0.1
Input$7.5 / 1M tokens
Output$7.5 / 1M tokens
openaiopenai-responseopenai-response-compactanthropicgeminiopenai-alpha-searchOpenAI compatibleResponses APIAnthropic compatibleGemini compatibleCompare