Anthropic

claude-opus-4-8

Model IDclaude-opus-4-8

Integration docsMarkdown

Model introduction

Claude Opus 4.8 is Anthropic's model for complex agentic coding and enterprise work, with adaptive thinking enabled.

Model capabilities

Specifications come from the cited model material. Catalog tags help identify access features; use the source documentation for exact limits.

OpenAI compatibleAnthropic compatiblePrompt cacheReasoningVision

Verified specifications

Official positioning
Complex agentic coding and enterprise work
Context window
1,000,000 tokens
Maximum output
128,000 tokens
Verified capability
Adaptive thinking

Sources: Anthropic Claude Opus 4.8 update notes

Independent capability evaluation

Only public measurements matched to the exact model ID and reasoning profile are shown; nearby variants are not substituted.

Intelligence index

56

Cohort rank #6 / 186

Output speed

58.8 tok/s

Cohort rank #95 / 186

First-token latency

28.84s

Adaptive reasoning, max effort

Frontier composite capability, with slower output and high token use and cost at maximum effort.

Evaluation profile: Adaptive reasoning, max effort. Artificial Analysis · Methodology

Model pricing and access

Prices are read directly from the MoleAPI console catalog and shown by current billing group and context tier.

Your final charge follows the account group shown in the console.
Live pricing source

Pricing table

USD / 1M tokens
Standardx1defaultInput $5 · Output $25

Input

$5 / 1M tokens

Output

$25 / 1M tokens

Cache read

$0.5 / 1M tokens

Cache write

$6.25 / 1M tokens

Discountx0.8discountInput $4 · Output $20

Input

$4 / 1M tokens

Output

$20 / 1M tokens

Cache read

$0.4 / 1M tokens

Cache write

$5 / 1M tokens

Relayx0.3relayInput $1.5 · Output $7.5

Input

$1.5 / 1M tokens

Output

$7.5 / 1M tokens

Cache read

$0.15 / 1M tokens

Cache write

$1.875 / 1M tokens

Access protocols

Available billing groups: Standard, Discount, Relay

anthropic
POST
/v1/messages
openai
POST
/v1/chat/completions

About claude-opus-4-8

This introduction is transcreated for clarity and cross-checked against the cited model material.

Claude Opus 4.8 is Anthropic's model for complex agentic coding and enterprise work, with adaptive thinking enabled.

Core strengths

  • Designed for long-running agentic coding, large repositories, and difficult enterprise work.
  • A 1M context window and 128K maximum output support unusually long tasks.
  • Adaptive thinking adjusts reasoning effort to the task.

Limitations

  • High reasoning effort can add substantial first-answer latency and token use.
  • Its price favors high-value work over simple, high-volume processing.

Selection and production evaluation

Start a claude-opus-4-8 evaluation by mapping its official positioning to real work: Complex agentic coding and enterprise workflows that benefit from a long context window and extended reasoning. The first pass should exercise both its main strength, "Designed for long-running agentic coding, large repositories, and difficult enterprise work.", and its known limitation, "High reasoning effort can add substantial first-answer latency and token use.", instead of relying on a single subjective general-chat comparison.

For access, MoleAPI currently lists anthropic, openai protocols and the Standard, Discount, Relay billing groups for this Anthropic model; the default price summary is Input $5 / 1M tokens · Output $25 / 1M tokens. Pricing, protocols, and groups come from the live catalog, so production planning should still price representative requests using real context length, output size, and cache-hit assumptions.

The capability snapshot uses the exact model ID under the "Adaptive reasoning, max effort" public evaluation profile. A launch decision should add your own task accuracy, structured-output validity, tool-call success, and timeout rates while pinning the model ID, prompt, and sample set so version or effort changes do not hide regressions.

The model material on this page was last checked on 2026-07-23. When upstream model cards, context limits, or tool support change, update the cited bilingual facts before changing the recommendation; live MoleAPI price changes remain separate and update from the catalog automatically.

Code examples

These examples use MoleAPI's OpenAI-compatible endpoint and run after you replace the API key.

curl https://api.moleapi.com/v1/chat/completions \
  -H "Authorization: Bearer $MOLEAPI_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "claude-opus-4-8",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Frequently asked questions

How is claude-opus-4-8 priced?

This page reads prices from the MoleAPI console API and updates with model prices, context tiers, and account groups.

How can I access claude-opus-4-8?

Protocols and endpoints come from the supported_endpoint_types field in the MoleAPI model catalog.

How do I switch an existing project to claude-opus-4-8?

Keep the MoleAPI API address and key, replace the model parameter with the model ID on this page, then check protocol-specific parameter differences.

Where do the claude-opus-4-8 model details come from?

Capabilities and limitations are checked against Anthropic and the other cited pages. Pricing and available protocols come only from the MoleAPI console.