Models·2 min read
By BitsMindsSource: Google

Google Launches Gemini 3.1 Pro: Tops ARC-AGI-2 Benchmark with 2M Token Context and Agentic Coding

Google's Gemini 3.1 Pro hits a verified 77.1% on the ARC-AGI-2 reasoning benchmark — more than double its predecessor — and arrives with a 2-million-token context window, multi-step agentic coding, and support across Gemini API, Vertex AI, and NotebookLM.

Google Launches Gemini 3.1 Pro: Tops ARC-AGI-2 Benchmark with 2M Token Context and Agentic Coding
Share:

Google has released Gemini 3.1 Pro in preview, positioning it as its most capable reasoning model to date. The release lands at the top of the ARC-AGI-2 benchmark — which evaluates an AI's ability to solve logic patterns it has never encountered before — with a verified score of 77.1%. Google described this as more than double the reasoning performance of Gemini 3 Pro and on par with the best scores posted by any competing model.

Developers can access Gemini 3.1 Pro through the Gemini API, Google AI Studio, Gemini CLI, Android Studio, and Google Antigravity. Enterprise customers have access via Vertex AI and Gemini Enterprise. Consumer users will find it in the Gemini app and NotebookLM, though higher usage limits are restricted to Google AI Pro and Ultra subscription tiers. The model is launching in preview status as Google gathers validation before moving to general availability.

Gemini 3.1 Pro is designed for complex, multi-step tasks: generating code-based animations, building interactive dashboards through API configuration, creating website-ready animated SVGs, and developing immersive 3D experiences with hand-tracking capabilities. The model's 2-million-token context window — shared with the Gemini 3.1 Ultra variant — allows it to process over 1,500 pages of text or hours of video in a single session. These capabilities place it in direct competition with OpenAI's GPT-5.4 and Anthropic's Claude Opus 4.7 for enterprise developers and research workflows.

The Gemini 3.1 launch also includes the release of Gemma 4 open models, available through Google AI Studio and the Gemini API. Together, the Gemini 3.1 wave marks Google's most significant AI update since Gemini 3 launched earlier in the year. As of mid-April 2026, Gemini 3.1 Ultra and GPT-5.4 Pro are tied at 57 points on the Artificial Analysis Intelligence Index, with Gemini 3.1 Ultra pulling ahead on GPQA Diamond at 94.3% — cementing Google's position at the frontier of AI reasoning benchmarks heading into the second half of 2026.

More on Gemini

Evergreen coverage we keep current — start here.

Want AI news before everyone else?

The morning's most important AI stories, straight to your inbox. No fluff.

Related Articles

Many kinds of input, one Qwen context On a deep violet field, a filmstrip, an audio waveform, a landscape photograph and a text page curve toward a single luminous sphere bearing the official purple Qwen star. A label below reads “1M context”, representing text, images, audio and video sharing one million-token context window. Aa 1M context BITSMINDS.COM
Models

Qwen3.8-Omni-Flash Undercuts Gemini on Audio and Video

Hill Climb: GPT-6 Astra versus Claude Fable 5.1 A teal desert buggy climbs a sandstone ridge on the left. An orange jeep climbs a green hill on the right, under an arc of gold coins. The two landscapes meet at a diagonal divide. 7 GPT-6 ASTRA CLAUDE FABLE 5.1 VS HILL CLIMB BITSMINDS.COM
Models

GPT-6 Astra vs Claude Fable 5.1: Hill Climb

Gemini 3.8 Live: thinking while the conversation continues An editorial illustration in a dark blue and violet room. A carefully drawn studio microphone and a tilted smartphone flank a translucent speech bubble carrying the multicoloured Gemini star. A continuous luminous audio waveform travels between them. Above the bubble, a separate arc connects small search, reasoning and completion symbols, representing background work continuing during a spoken conversation. The phone screen and visual paths are conceptual, not a reproduction of Google's actual interface or internal reasoning. Original vector illustration for BitsMinds, gemini-3-8-live-extended-thinking-voice. 16 September 2026. GEMINI 3.8 LIVE EXTENDED THINKING Gemini 3.8 LIVE The conversation continues Thinking. Still talking. BITSMINDS.COM
Models

Gemini 3.8 Live Tops Voice AI and Undercuts GPT