# Google: gemini-3.1-pro-preview

- Model ID: `gemini-3.1-pro-preview`
- Provider: Google
- Web version: https://www.moleapi.com/en/models/google/gemini-3.1-pro
- Content status: alias

## Model introduction

Gemini 3.1 Pro is Google's flagship multimodal model for complex reasoning, code, and long-form generation, with 1M context and 64K output.

## Model capabilities

- OpenAI compatible
- Gemini compatible
- Vision
- Prompt cache
- Reasoning

### Verified specifications

- Official positioning: Flagship complex reasoning and multimodal model
- Context window: 1,000,000 tokens
- Maximum output: 64,000 tokens
- Input / output: Text, image, audio, video, PDF / Text

## Model pricing and access

Pricing is supplied dynamically by the MoleAPI console API.

### Standard (default, x1)

- Input: $2 / 1M tokens
- Output: $12 / 1M tokens
- Cache read: $0.2 / 1M tokens

### Discount (discount, x0.8)

- Input: $1.6 / 1M tokens
- Output: $9.6 / 1M tokens
- Cache read: $0.16 / 1M tokens

### Access protocols

| Protocol | Method | Endpoint |
| --- | --- | --- |
| gemini | POST | /v1beta/models/{model}:generateContent |
| openai | POST | /v1/chat/completions |

- Live pricing source: https://home.moleapi.com/api/pricing

## About gemini-3.1-pro-preview

Gemini 3.1 Pro is Google's flagship multimodal model for complex reasoning, code, and long-form generation, with 1M context and 64K output.

### Best for

Complex code, long-document and data analysis, creative long-form work, and tasks combining text, images, audio, and video.

### Core strengths

- Natively handles text, images, audio, video, and PDFs.
- A 1M context window and 64K output support complete long tasks.
- Prioritizes abstract reasoning, coding, and reliable complex workflows.

### Limitations

- Quality-first behavior is generally slower than Flash variants.
- Preview aliases may change availability or behavior with Google's release cycle.

### Selection and production evaluation

Start a gemini-3.1-pro-preview evaluation by mapping its official positioning to real work: Complex code, long-document and data analysis, creative long-form work, and tasks combining text, images, audio, and video. The first pass should exercise both its main strength, "Natively handles text, images, audio, video, and PDFs.", and its known limitation, "Quality-first behavior is generally slower than Flash variants.", instead of relying on a single subjective general-chat comparison.

For access, MoleAPI currently lists gemini, openai protocols and the Standard, Discount billing groups for this Google model; the default price summary is Input $2 / 1M tokens · Output $12 / 1M tokens. Pricing, protocols, and groups come from the live catalog, so production planning should still price representative requests using real context length, output size, and cache-hit assumptions.

No independent ranking is shown unless it matches this exact model ID and reasoning profile, so nearby variants are not used as a proxy. Before launch, pin the model ID, prompt, and sample set, then compare task accuracy, structured-output validity, tool-call success, and timeout rates under the same conditions.

The model material on this page was last checked on 2026-07-24. When upstream model cards, context limits, or tool support change, update the cited bilingual facts before changing the recommendation; live MoleAPI price changes remain separate and update from the catalog automatically.

### Sources

- [B.AI Gemini 3.1 Pro model guide](https://docs.b.ai/llmservice/models/gemini-3-1-pro/) — B.AI
- [Google Gemini model documentation](https://ai.google.dev/gemini-api/docs/models) — Google
- [Google DeepMind Gemini 3.1 Pro model card](https://deepmind.google/models/model-cards/gemini-3-1-pro/) — Google DeepMind
- Last verified: 2026-07-24

## Code examples

### cURL

```bash
curl https://api.moleapi.com/v1/chat/completions \
  -H "Authorization: Bearer $MOLEAPI_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"gemini-3.1-pro-preview","messages":[{"role":"user","content":"Hello"}]}'
```

## More related models

- [gemini-3.5-flash](https://www.moleapi.com/en/models/google/gemini-3.5-flash) — Google
- [gemini-3.1-pro](https://www.moleapi.com/en/models/google/gemini-3.1-pro) — Google
- [gemini-3-flash](https://www.moleapi.com/en/models/google/gemini-3-flash) — Google

## Frequently asked questions

### How is gemini-3.1-pro-preview priced?

Prices are read from the MoleAPI console API and update with model prices, context tiers, and account groups.

### How can I access gemini-3.1-pro-preview?

The catalog currently lists gemini, openai.

### How do I switch an existing project to gemini-3.1-pro-preview?

Keep the MoleAPI API address and key, replace the model parameter, and check protocol-specific parameters.

### Where do the gemini-3.1-pro-preview details come from?

Capabilities and limitations are checked against B.AI and the other cited pages; pricing and protocols come from the MoleAPI console.
