Thursday , 19 February 2026

No products in the cart.

Home Tech Google releases Gemini 3.1 Pro: Benchmark performance, how to try it

Google releases Gemini 3.1 Pro: Benchmark performance, how to try it

2 Mins read

Google released its latest core reasoning model, Gemini 3.1 Pro, on Thursday. Google says that Gemini 3.1 Pro achieved twice the verified performance of 3 Pro on ARC-AGI-2, a popular benchmark that measures a model’s logical reasoning.

Google originally released Gemini 3 and 3 Pro in November, and this new release shows just how fast AI companies are introducing new and updated models. Gemini 3.1 Pro is the new core model powering Gemini and various Google AI tools, such as Gemini 3 Deep Think. Google says it’s designed to provide more creative solutions.

“3.1 Pro is designed for tasks where a simple answer isn’t enough, taking advanced reasoning and making it useful for your hardest challenges,” a Google blog post states. “This improved intelligence can help in practical applications — whether you’re looking for a clear, visual explanation of a complex topic, a way to synthesize data into a single view, or bringing a creative project to life.”

Here’s everything we know so far about Gemini 3.1 Pro, including how it compares to the latest models from Anthropic and OpenAI, and how to try it yourself.

How to try Gemini 3.1 Pro

Starting today, Google is rolling out Gemini 3.1 Pro in the Gemini App, the Gemini APIA, and in Notebook LM. Free users will be able to try 3.1 Pro in the Gemini app, but paid users on Google AI Pro and AI Ultra plans will have higher usage rates. Within Notebook LM, only these paid users will have access to 3.1 Pro, at least, for now. Coders and enterprise users can also access the new core model via developers and enterprises can access 3.1 through AI Studio, Antigravity, Vertex AI, Gemini Enterprise, Gemini CLI, and Android Studio.

Gemini 3.1 Pro was already available for Mashable editors using Gemini. To try it for yourself, head to Gemini on desktop or open the Gemini mobile app.

screenshot showing animation from gemini 3 pro

Left:
Two results of the same animation prompt.
Credit: Google

Right:
Credit: Google

Why Gemini 3.1 Pro matters

When Google released Gemini 3 Pro in November, the model was so impressive that it allegedly caused OpenAI CEO Sam Altman to declare a code red. As Gemini 3 Pro surged to the top of AI leaderboards, OpenAI reportedly started losing ChatGPT users to Gemini. The latest core ChatGPT model, GPT-5.2, has tumbled down the rankings on leaderboards like Arena (formerly known as LMArena), losing significant ground to competitors such as Google, Anthropic, and xAI.

This Tweet is currently unavailable. It might be loading or has been removed.

Gemini 3 Pro was already outperforming GPT-5.2 on many benchmarks, and with a more advanced thinking model, Gemini could move even further ahead.

Gemini 3.1 Pro: Benchmark performance

Google released benchmark performance data showing that Gemini 3.1 Pro outperforms previous Gemini models, Claude Sonnet 4.6, Claude Opus 4.6, and GPT-5.2. However, OpenAI’s new coding model, GPT-5.3-Codex, beat Gemini 3.1 Pro on the verified SWE-Bench Pro benchmark, according to Google itself.

Notable highlights from Gemini 3.1 Pro’s benchmark results include:

44.4 percent on Humanity’s last exam, compared to 40.0 percent for Claude Opus 4.6 and 34.5 percent for GPT-5.2
77.1 percent on ARC-AGI-2, compared to 31.1 percent for Gemini 3 Pro, 68.8 percent for Claude Opus 4.6, and 52.9 percent for GPT-5.2
94.3 percent on GPQA Diamond, compared to 91.9 percent for Gemini 3 Pro, 91.3 percent for Claude Opus 4.6, and 92.4 percent for GPT-5.2
80.6 percent on SWE-Bench Verified, compared to 76.2 percent for Gemini 3 Pro, 80.8 percent for Claude Opus 4.6, and 80.0 percent for GPT-5.2
54.2 percent on SWE-Bench Pro (Public), compared to 43.3 percent for Gemini 3 Pro, 55.6 percent for GPT-5.2, and 56.8 percent for GPT-5.3-Codex
92.6 percent on MMLU, compared to 91.1 percent for Claude Opus 4.6 and 89.6 percent for GPT-5.2

Google released an image showing the full benchmark results for Gemini 3.1 Pro:

This Tweet is currently unavailable. It might be loading or has been removed.

Disclosure: Ziff Davis, Mashable’s parent company, in April 2025 filed a lawsuit against OpenAI, alleging it infringed Ziff Davis copyrights in training and operating its AI systems.

Previous post Massive global data breach sees over a billion records exposed – here’s what we know so far

Next post Lilian Thuram strongly criticizes José Mourinho for his words to Vinicius Jr: ‘There’s a feeling of superiority and white narcissism’

Top Insights

Alex Zozos: Tokenized securities are classified as securities, the SEC’s evolving role in on-chain trading, and how blockchain enhances trading efficiency | Unchained

United plot hostile raid on PL rival to secure “ruthless” forward, INEOS are taking no prisoners – Ben Jacobs

Files reveal Epstein was offered chance to buy US Pentagon, FBI buildings

Rostam Premieres Original Version of Vampire Weekend’s “Campus”

Google releases Gemini 3.1 Pro: Benchmark performance, how to try it

How to try Gemini 3.1 Pro

Why Gemini 3.1 Pro matters

Gemini 3.1 Pro: Benchmark performance

Leave a comment

Leave a Reply Cancel reply

Help/Support

Related Articles

Why these startup CEOs don’t think AI will replace human roles

How digitally sovereign are you? Red Hat can help measure that

Donald Trump Jr.’s Private DC Club Has Mysterious Ties to an Ex-Cop With a Controversial Past

YouTube’s latest experiment brings its conversational AI tool to TVs

Abous us

Subscribe Now