News Every Day | Yesterday, 21:26

Google releases Gemini 3.1 Pro: Benchmark performance, how to try it

Google released its latest core reasoning model, Gemini 3.1 Pro, on Thursday. Google says that Gemini 3.1 Pro achieved twice the verified performance of 3 Pro on ARC-AGI-2, a popular benchmark that measures a model's logical reasoning.

Google originally released Gemini 3 and 3 Pro in November, and this new release shows just how fast AI companies are introducing new and updated models. Gemini 3.1 Pro is the new core model powering Gemini and various Google AI tools, such as Gemini 3 Deep Think. Google says it's designed to provide more creative solutions.

"3.1 Pro is designed for tasks where a simple answer isn't enough, taking advanced reasoning and making it useful for your hardest challenges," a Google blog post states. "This improved intelligence can help in practical applications — whether you’re looking for a clear, visual explanation of a complex topic, a way to synthesize data into a single view, or bringing a creative project to life."

Here's everything we know so far about Gemini 3.1 Pro, including how it compares to the latest models from Anthropic and OpenAI, and how to try it yourself.

How to try Gemini 3.1 Pro

Starting today, Google is rolling out Gemini 3.1 Pro in the Gemini App, the Gemini APIA, and in Notebook LM. Free users will be able to try 3.1 Pro in the Gemini app, but paid users on Google AI Pro and AI Ultra plans will have higher usage rates. Within Notebook LM, only these paid users will have access to 3.1 Pro, at least, for now. Coders and enterprise users can also access the new core model via developers and enterprises can access 3.1 through AI Studio, Antigravity, Vertex AI, Gemini Enterprise, Gemini CLI, and Android Studio.

Gemini 3.1 Pro was already available for Mashable editors using Gemini. To try it for yourself, head to Gemini on desktop or open the Gemini mobile app.

Left: Two results of the same animation prompt. Credit: Google

Right: Credit: Google

Why Gemini 3.1 Pro matters

When Google released Gemini 3 Pro in November, the model was so impressive that it allegedly caused OpenAI CEO Sam Altman to declare a code red. As Gemini 3 Pro surged to the top of AI leaderboards, OpenAI reportedly started losing ChatGPT users to Gemini. The latest core ChatGPT model, GPT-5.2, has tumbled down the rankings on leaderboards like Arena (formerly known as LMArena), losing significant ground to competitors such as Google, Anthropic, and xAI.

This Tweet is currently unavailable. It might be loading or has been removed.

Gemini 3 Pro was already outperforming GPT-5.2 on many benchmarks, and with a more advanced thinking model, Gemini could move even further ahead.

Gemini 3.1 Pro: Benchmark performance

Google released benchmark performance data showing that Gemini 3.1 Pro outperforms previous Gemini models, Claude Sonnet 4.6, Claude Opus 4.6, and GPT-5.2. However, OpenAI's new coding model, GPT-5.3-Codex, beat Gemini 3.1 Pro on the verified SWE-Bench Pro benchmark, according to Google itself.

Notable highlights from Gemini 3.1 Pro's benchmark results include:

44.4 percent on Humanity's last exam, compared to 40.0 percent for Claude Opus 4.6 and 34.5 percent for GPT-5.2
77.1 percent on ARC-AGI-2, compared to 31.1 percent for Gemini 3 Pro, 68.8 percent for Claude Opus 4.6, and 52.9 percent for GPT-5.2
94.3 percent on GPQA Diamond, compared to 91.9 percent for Gemini 3 Pro, 91.3 percent for Claude Opus 4.6, and 92.4 percent for GPT-5.2
80.6 percent on SWE-Bench Verified, compared to 76.2 percent for Gemini 3 Pro, 80.8 percent for Claude Opus 4.6, and 80.0 percent for GPT-5.2
54.2 percent on SWE-Bench Pro (Public), compared to 43.3 percent for Gemini 3 Pro, 55.6 percent for GPT-5.2, and 56.8 percent for GPT-5.3-Codex
92.6 percent on MMLU, compared to 91.1 percent for Claude Opus 4.6 and 89.6 percent for GPT-5.2

Google released an image showing the full benchmark results for Gemini 3.1 Pro:

This Tweet is currently unavailable. It might be loading or has been removed.

Disclosure: Ziff Davis, Mashable’s parent company, in April 2025 filed a lawsuit against OpenAI, alleging it infringed Ziff Davis copyrights in training and operating its AI systems.

Google releases Gemini 3.1 Pro: Benchmark performance, how to try it

How to try Gemini 3.1 Pro

Why Gemini 3.1 Pro matters

Gemini 3.1 Pro: Benchmark performance

Read also

Bruce Meyer elevated to baseball players’ association interim executive director

Meta might launch its first smartwatch this year

Last Night in College Basketball: Arizona Bounces Back Against Shorthanded BYU

Sports today

All sports news today

Sports in Russia today

Friends of Today24