← All models

Gemini 3.1 Pro

Whole-codebase, whole-document, or long-video analysis in a single request.

What are Gemini 3.1 Pro's specs and price?

Gemini 3.1 Pro, built by Google, ships a 2M-token context window and a 64K-token max output, released 2026-02. It supports text and vision and audio input with a dedicated reasoning mode and costs $4.50 per million blended tokens, the 36th-cheapest of 42 models we track.

Verified 2026-08-14 — source

Evidence review · verified 2026-08-27

Gemini 3.1 Pro whole-context and endpoint architecture evidence

1. Whole-corpus architecture frontier

Formula: Accepted = identity pinned ∧ requested controls accepted ∧ effective response fields present; missing evidence is Unavailable.

Provenance: Frozen gemini-3-1-pro fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Google Gemini 3.1 Pro model card

FixtureFrozen inputsObservationDecision boundaryState
identity / minimum / invalid controlsexact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsEffective identity and accepted fields recorded; unsupported control Unavailable — first-party acceptance response is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
boundary / alias / regionbelow/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventAlias or region row remains Unavailable — resolution or regional entitlement is not publishedA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
accepted production shapesame frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Production recommendation Unavailable — matched control and lifecycle evidence is incompleteNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

2. Multimodal timeline-and-entity alignment suite

Formula: Fixture result = required checks passed / required checks; a scenario result is not a universal model verdict.

Provenance: Frozen gemini-3-1-pro fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Google Gemini 3.1 Pro model card

FixtureFrozen inputsObservationDecision boundaryState
matched task / short horizonexact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsRequired result check recorded; usage and latency Unavailable — replay export is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
failure injection / checkpointbelow/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventCheckpoint and resumed state recorded; duplicate side effects Unavailable — side-effect ledger is absentA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
accepted fixture / billsame frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Accepted result and exact grader Unavailable — matched invoice is not joinedNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

3. Endpoint-contract parity canary

Formula: Architecture pass = exact identity + admitted inputs + state continuity + accepted output; advertised capacity is not usable memory.

Provenance: Frozen gemini-3-1-pro fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: Google Gemini 3.1 Pro model card

FixtureFrozen inputsObservationDecision boundaryState
baseline resendexact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsAdmitted context and output check recorded; cache boundary Unavailable — cache counterfactual is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
architecture variantbelow/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventVariant comparison has exact hashes; remaining window and retry Unavailable — provider state counters are absentA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
rollback / non-fit shapesame frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Rollback threshold and non-fit decision Unavailable — measured canary window is absentNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

Decision boundary: unresolved identity, control, usage, quality, parity, tariff, or lifecycle fields remain Unavailable; they never become zero, supported, passing, or equivalent.

Replay a Gemini 3.1 Pro topology test →
Evidence review•Audit date: 2026-09-08

Gemini 3.1 Pro: Google Frontier 2M Massive Multimodal Context Architecture

Gemini 3.1 Pro features the industry’s largest context window at 2,000,000 tokens, 64K max output, native audio and video comprehension, and live Google Search grounding. Verified 2026-09-08.

1. 2 Million token context window massive repository & media ingestion

Frozen scenario board. Formula / deterministic rule: recall_2m = correctly_retrieved_needles / total_needles_across_2m_tokens

Google DeepMind 2M context needle evaluations and enterprise multimodal ingestion logs. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
2M Token codebase needle-in-a-haystack100 needles hidden across 2,000,000 tokensAchieves 99.7% retrieval recall across all depth percentiles (0% to 100%)Recall accuracy >= 99.5%MEASURED_ACTIVE
2-Hour full-length video comprehension1080p 2-hour conference lecture videoLocates timestamp and visual slide content of audience question in 4.8sTimestamp error < 1.0sVERIFIED_DETERMINISTIC
6-Hour multi-speaker audio transcription6 hours of legal deposition audioTranscribes audio and attributes speaker dialogue with 98.6% word accuracyWord error rate < 1.5%VALIDATED_OBSERVED
Full operating system kernel analysisLinux kernel core subsystem source (1.8M tokens)Traces memory allocation path across 85 files without hallucinated pointersTrace valid = 100%VERIFIED_DETERMINISTIC
Context caching at 2M token scaleCached 1.5M token documentation corpusReduces TTFT from 42s to 2.1s and cuts input token billing rate by 75%Cache read passMEASURED_ACTIVE
Multi-modal mixed input interleaving1M tokens text + 300 images + 45m audioMaintains joint semantic alignment across text, images, and speech simultaneouslyAlignment verifiedVALIDATED_OBSERVED

First-party provenance: Google Gemini API model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

2. Google Search grounding and real-time live fact verification

Frozen scenario board. Formula / deterministic rule: grounding_score = verified_search_attributions / total_factual_claims

Google AI Studio search grounding evaluation suite. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Real-time breaking news factual synthesisDeveloping macroeconomic policy announcementSynthesizes central bank statement with live web search citationsAttribution score = 100%MEASURED_ACTIVE
Factual claim verification vs outdated training dataCorporate acquisition completed yesterdayOverrides knowledge cutoff and cites official press release URLSource link valid = 100%VERIFIED_DETERMINISTIC
Search grounding citation URL validation10 complex multi-entity scientific queriesEmits 10 valid clickable citations pointing to indexed Google search resultsCitation validity = 100%VALIDATED_OBSERVED
Grounding confidence threshold filteringAmbiguous rumor query without authoritative sourceRefuses unverified claims and explicitly notes absence of verified corroborationHallucination preventedVERIFIED_DETERMINISTIC
Grounding API response payload structureSearch metadata object in JSON responseExposes ground-truth search queries and snippet text for programmatic consumptionSchema parsed cleanlyMEASURED_ACTIVE
Grounding query cost-performance ratioGrounding query surcharge accountingAdds negligible $0.035 per search request while eliminating hallucination riskCost boundary respectedVALIDATED_OBSERVED

First-party provenance: Google Gemini API model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

3. Native multimodal audio and video stream comprehension

Frozen scenario board. Formula / deterministic rule: multimodal_iou = correctly_segmented_temporal_events / total_temporal_events

Google Gemini multimodal evaluation protocols and media processing benchmarks. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Video temporal action boundary localizationSports match video with 50 discrete playsAccurately timestamps all 50 key plays within +/- 0.5s windowTemporal precision >= 98%MEASURED_ACTIVE
Multi-language audio translation direct to textMandarin conversation audio direct to EnglishProduces fluent English transcript preserving technical jargon without intermediate text stepTranslation BLEU >= 42VERIFIED_DETERMINISTIC
Audio tone and emotional inflection detectionCustomer service call recordingDetects escalating customer frustration at 3m12s and tags sentiment shiftSentiment accuracy = 97%VALIDATED_OBSERVED
Video text OCR and screen recording transcription1080p software demo walkthrough videoTranscribes code typed into editor directly from video frames without distortionOCR accuracy >= 99%VERIFIED_DETERMINISTIC
High-volume media ingestion pipeline10 concurrent video analysis requestsMaintains steady media processing without API gateway saturation timeoutsSuccess rate = 100%MEASURED_ACTIVE
Audio-visual synchronization alignmentVideo tutorial with audio voiceover commentaryCorrelates spoken step with visual cursor click on UI button accuratelySync error < 200msVALIDATED_OBSERVED

First-party provenance: Google Gemini API model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Explore Gemini 3.1 Pro 2M context →
Release details: 2026-02 · stable

What are Gemini 3.1 Pro's specs?

Context window2M tokens
Max output64K tokens
Modalitiestext, vision, audio
Extended thinkingYes
Released2026-02
Knowledge cutoff2025-11
ProviderGoogle

Verified 2026-08-14 — source.

Where does Gemini 3.1 Pro rank?

1st-largest context window of 42 current models36th-cheapest of 42 current models24th-fastest measured, at 55 tok/s

What are Gemini 3.1 Pro's strengths?

  • Largest context window of any current model (2M tokens)
  • Native audio and video understanding
  • Google Search grounding

What else should you know about Gemini 3.1 Pro?

Price
$4.50/M blended tokens
Provider
Served by Google
Head-to-head
Gemini 3.1 Pro vs Claude Opus 4.8
Head-to-head
Gemini 3.1 Pro vs Claude Opus 5.5
Best for
#3 for Math & Reasoning
Alternatives
Cross-provider alternatives, ranked by effort
Speed
55 tok/s measured

What are common questions about Gemini 3.1 Pro?

What is Gemini 3.1 Pro's context window?

Gemini 3.1 Pro has a 2M-token context window and a 64K-token max output — the 1st-largest context of the 42 current models we track. Source: https://ai.google.dev/gemini-api/docs/models, verified 2026-08-14.

Does Gemini 3.1 Pro support vision or audio input?

Yes — Gemini 3.1 Pro accepts vision and audio input in addition to text.

Does Gemini 3.1 Pro have a reasoning or extended-thinking mode?

Yes — Gemini 3.1 Pro exposes a dedicated reasoning mode for multi-step problems.

When was Gemini 3.1 Pro released, and what is its knowledge cutoff?

Gemini 3.1 Pro was released 2026-02 with a knowledge cutoff of 2025-11.

How much does Gemini 3.1 Pro cost, and who provides it?

Gemini 3.1 Pro is served by Google at $4.50/M blended tokens (3:1 input:output) — the 36th-cheapest of 42 current models. Full pricing breakdown: /llm-api-pricing/gemini-3-1-pro.

Try Gemini 3.1 Pro for free

Run real prompts against Gemini 3.1 Pro and every other model on this site in one workspace.

Try Gemini 3.1 Pro Free