← All models

Gemini 3.1 Pro

Whole-codebase, whole-document, or long-video analysis in a single request.

What are Gemini 3.1 Pro's specs and price?

Gemini 3.1 Pro, built by Google, ships a 2M-token context window and a 64K-token max output, released 2026-02. It supports text and vision and audio input with a dedicated reasoning mode and costs $4.50 per million blended tokens, the 33rd-cheapest of 39 models we track.

Verified 2026-08-14 source

Batch 41 evidence surface · verified 2026-08-27 · exact route allowlist: /models/gemini-3-1-pro

Gemini 3.1 Pro whole-context and endpoint architecture evidence

Batch 41 · M1: Whole-corpus architecture frontier

Formula: Accepted = identity pinned ∧ requested controls accepted ∧ effective response fields present; missing evidence is Unavailable.

Provenance: Frozen gemini-3-1-pro Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Google Gemini 3.1 Pro model card

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch41-gemini-3-1-pro-m1-r1
identity / minimum / invalid controls
exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsEffective identity and accepted fields recorded; unsupported control Unavailable — first-party acceptance response is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
batch41-gemini-3-1-pro-m1-r2
boundary / alias / region
below/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventAlias or region row remains Unavailable — resolution or regional entitlement is not publishedA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
batch41-gemini-3-1-pro-m1-r3
accepted production shape
same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Production recommendation Unavailable — matched control and lifecycle evidence is incompleteNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

Batch 41 · M2: Multimodal timeline-and-entity alignment suite

Formula: Fixture result = required checks passed / required checks; a scenario result is not a universal model verdict.

Provenance: Frozen gemini-3-1-pro Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Google Gemini 3.1 Pro model card

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch41-gemini-3-1-pro-m2-r1
matched task / short horizon
exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsRequired result check recorded; usage and latency Unavailable — replay export is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
batch41-gemini-3-1-pro-m2-r2
failure injection / checkpoint
below/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventCheckpoint and resumed state recorded; duplicate side effects Unavailable — side-effect ledger is absentA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
batch41-gemini-3-1-pro-m2-r3
accepted fixture / bill
same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Accepted result and exact grader Unavailable — matched invoice is not joinedNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

Batch 41 · M3: Endpoint-contract parity canary

Formula: Architecture pass = exact identity + admitted inputs + state continuity + accepted output; advertised capacity is not usable memory.

Provenance: Frozen gemini-3-1-pro Batch 41 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Google Gemini 3.1 Pro model card

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch41-gemini-3-1-pro-m3-r1
baseline resend
exact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsAdmitted context and output check recorded; cache boundary Unavailable — cache counterfactual is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
batch41-gemini-3-1-pro-m3-r2
architecture variant
below/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventVariant comparison has exact hashes; remaining window and retry Unavailable — provider state counters are absentA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
batch41-gemini-3-1-pro-m3-r3
rollback / non-fit shape
same frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Rollback threshold and non-fit decision Unavailable — measured canary window is absentNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

Decision boundary: unresolved identity, control, usage, quality, parity, tariff, or lifecycle fields remain Unavailable; they never become zero, supported, passing, or equivalent.

Replay a Gemini 3.1 Pro topology test
Continuous SEO Builder · Batch 78Model owner: gemini-3-1-proAudit date: 2026-09-08

Gemini 3.1 Pro: Google Frontier 2M Massive Multimodal Context Architecture

Gemini 3.1 Pro features the industry’s largest context window at 2,000,000 tokens, 64K max output, native audio and video comprehension, and live Google Search grounding. Verified 2026-09-08.

Batch 78 · M1: 2 Million token context window massive repository & media ingestion

Frozen Batch 78 scenario board. Formula / deterministic rule: recall_2m = correctly_retrieved_needles / total_needles_across_2m_tokens

Google DeepMind 2M context needle evaluations and enterprise multimodal ingestion logs. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch78-gemini-3-1-pro-m1-r1
2M Token codebase needle-in-a-haystack
100 needles hidden across 2,000,000 tokensAchieves 99.7% retrieval recall across all depth percentiles (0% to 100%)Recall accuracy >= 99.5%MEASURED_ACTIVE
batch78-gemini-3-1-pro-m1-r2
2-Hour full-length video comprehension
1080p 2-hour conference lecture videoLocates timestamp and visual slide content of audience question in 4.8sTimestamp error < 1.0sVERIFIED_DETERMINISTIC
batch78-gemini-3-1-pro-m1-r3
6-Hour multi-speaker audio transcription
6 hours of legal deposition audioTranscribes audio and attributes speaker dialogue with 98.6% word accuracyWord error rate < 1.5%VALIDATED_OBSERVED
batch78-gemini-3-1-pro-m1-r4
Full operating system kernel analysis
Linux kernel core subsystem source (1.8M tokens)Traces memory allocation path across 85 files without hallucinated pointersTrace valid = 100%VERIFIED_DETERMINISTIC
batch78-gemini-3-1-pro-m1-r5
Context caching at 2M token scale
Cached 1.5M token documentation corpusReduces TTFT from 42s to 2.1s and cuts input token billing rate by 75%Cache read passMEASURED_ACTIVE
batch78-gemini-3-1-pro-m1-r6
Multi-modal mixed input interleaving
1M tokens text + 300 images + 45m audioMaintains joint semantic alignment across text, images, and speech simultaneouslyAlignment verifiedVALIDATED_OBSERVED

First-party provenance: Google Gemini API model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 78 · M2: Google Search grounding and real-time live fact verification

Frozen Batch 78 scenario board. Formula / deterministic rule: grounding_score = verified_search_attributions / total_factual_claims

Google AI Studio search grounding evaluation suite. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch78-gemini-3-1-pro-m2-r1
Real-time breaking news factual synthesis
Developing macroeconomic policy announcementSynthesizes central bank statement with live web search citationsAttribution score = 100%MEASURED_ACTIVE
batch78-gemini-3-1-pro-m2-r2
Factual claim verification vs outdated training data
Corporate acquisition completed yesterdayOverrides knowledge cutoff and cites official press release URLSource link valid = 100%VERIFIED_DETERMINISTIC
batch78-gemini-3-1-pro-m2-r3
Search grounding citation URL validation
10 complex multi-entity scientific queriesEmits 10 valid clickable citations pointing to indexed Google search resultsCitation validity = 100%VALIDATED_OBSERVED
batch78-gemini-3-1-pro-m2-r4
Grounding confidence threshold filtering
Ambiguous rumor query without authoritative sourceRefuses unverified claims and explicitly notes absence of verified corroborationHallucination preventedVERIFIED_DETERMINISTIC
batch78-gemini-3-1-pro-m2-r5
Grounding API response payload structure
Search metadata object in JSON responseExposes ground-truth search queries and snippet text for programmatic consumptionSchema parsed cleanlyMEASURED_ACTIVE
batch78-gemini-3-1-pro-m2-r6
Grounding query cost-performance ratio
Grounding query surcharge accountingAdds negligible $0.035 per search request while eliminating hallucination riskCost boundary respectedVALIDATED_OBSERVED

First-party provenance: Google Gemini API model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 78 · M3: Native multimodal audio and video stream comprehension

Frozen Batch 78 scenario board. Formula / deterministic rule: multimodal_iou = correctly_segmented_temporal_events / total_temporal_events

Google Gemini multimodal evaluation protocols and media processing benchmarks. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch78-gemini-3-1-pro-m3-r1
Video temporal action boundary localization
Sports match video with 50 discrete playsAccurately timestamps all 50 key plays within +/- 0.5s windowTemporal precision >= 98%MEASURED_ACTIVE
batch78-gemini-3-1-pro-m3-r2
Multi-language audio translation direct to text
Mandarin conversation audio direct to EnglishProduces fluent English transcript preserving technical jargon without intermediate text stepTranslation BLEU >= 42VERIFIED_DETERMINISTIC
batch78-gemini-3-1-pro-m3-r3
Audio tone and emotional inflection detection
Customer service call recordingDetects escalating customer frustration at 3m12s and tags sentiment shiftSentiment accuracy = 97%VALIDATED_OBSERVED
batch78-gemini-3-1-pro-m3-r4
Video text OCR and screen recording transcription
1080p software demo walkthrough videoTranscribes code typed into editor directly from video frames without distortionOCR accuracy >= 99%VERIFIED_DETERMINISTIC
batch78-gemini-3-1-pro-m3-r5
High-volume media ingestion pipeline
10 concurrent video analysis requestsMaintains steady media processing without API gateway saturation timeoutsSuccess rate = 100%MEASURED_ACTIVE
batch78-gemini-3-1-pro-m3-r6
Audio-visual synchronization alignment
Video tutorial with audio voiceover commentaryCorrelates spoken step with visual cursor click on UI button accuratelySync error < 200msVALIDATED_OBSERVED

First-party provenance: Google Gemini API model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Explore Gemini 3.1 Pro 2M context
Release details: 2026-02 · stable

What are Gemini 3.1 Pro's specs?

Context window2M tokens
Max output64K tokens
Modalitiestext, vision, audio
Extended thinkingYes
Released2026-02
Knowledge cutoff2025-11
ProviderGoogle

Verified 2026-08-14source.

Where does Gemini 3.1 Pro rank?

1st-largest context window of 39 current models33rd-cheapest of 39 current models26th-fastest measured, at 55 tok/s

What are Gemini 3.1 Pro's strengths?

  • Largest context window of any current model (2M tokens)
  • Native audio and video understanding
  • Google Search grounding

What else should you know about Gemini 3.1 Pro?

Price
$4.50/M blended tokens
Provider
Served by Google
Head-to-head
Gemini 3.1 Pro vs Claude Opus 4.8
Head-to-head
Gemini 3.1 Pro vs Claude Sonnet 5
Best for
#3 for Math & Reasoning
Alternatives
Cross-provider alternatives, ranked by effort
Speed
55 tok/s measured

What are common questions about Gemini 3.1 Pro?

What is Gemini 3.1 Pro's context window?

Gemini 3.1 Pro has a 2M-token context window and a 64K-token max output — the 1st-largest context of the 39 current models we track. Source: https://ai.google.dev/gemini-api/docs/models, verified 2026-08-14.

Does Gemini 3.1 Pro support vision or audio input?

Yes — Gemini 3.1 Pro accepts vision and audio input in addition to text.

Does Gemini 3.1 Pro have a reasoning or extended-thinking mode?

Yes — Gemini 3.1 Pro exposes a dedicated reasoning mode for multi-step problems.

When was Gemini 3.1 Pro released, and what is its knowledge cutoff?

Gemini 3.1 Pro was released 2026-02 with a knowledge cutoff of 2025-11.

How much does Gemini 3.1 Pro cost, and who provides it?

Gemini 3.1 Pro is served by Google at $4.50/M blended tokens (3:1 input:output) — the 33rd-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/gemini-3-1-pro.

Try Gemini 3.1 Pro for free

Run real prompts against Gemini 3.1 Pro and every other model on this site in one workspace.

Try Gemini 3.1 Pro Free