Google Gemini 3 Released: Next-Gen AI with Advanced Reasoning

Quick Summary
On November 19, 2025, Google officially launched Gemini 3, its most advanced AI model to date. Praised by CEO Sundar Pichai as the world’s best in multimodal understanding, Gemini 3 represents a major milestone toward AGI. It significantly outperforms Gemini 2.5 and top competitors like Claude 4.5 Sonnet and GPT 5.1 across key benchmarks, notably demonstrating PhD-level reasoning and enhanced multimodal capabilities. Designed for practical applications in research, creative workflows, and sports analytics, Gemini 3 features an upgraded Deep Think mode and is being rolled out across Google's ecosystem, including Gemini Chat and Google Search.
On November 19, 2025, Google officially introduced Gemini 3, its most advanced and intelligent AI model, designed to help users realize every idea.
CEO Sundar Pichai declared Gemini 3 as "the best model in the world for multimodal understanding." This model marks an upgrade in the journey towards Artificial General Intelligence (AGI).
How is it upgraded compared to Gemini 2.5?
Thus, 8 months after the launch of Gemini 2.5, Google has returned with Gemini 3 Pro, featuring upgrades in reasoning and context understanding abilities—it brings together all the capabilities of previous Gemini generations.
Sweeping the leaderboards
The launch of Gemini 3 Pro, though relatively quiet, was not a giant leap forward yet carried significant weight as it topped numerous LLM leaderboards (such as LMArena, etc.).
- Of course, compared to Gemini 2.5, Gemini 3 completely surpasses it across all AI benchmarks, such as in identifying the context and intent behind user requests, allowing users to get desired results with fewer prompts.
- While Gemini 3 outperforming previous-generation Gemini models is expected, its scores also surpassed both Claude 4.5 Sonnet and GPT 5.1. For instance, Gemini 3 demonstrated PhD-level reasoning with a high score of 37.5% on Humanity’s Last Exam without tools—significantly outperforming Claude Sonnet 4.5 (13.7%) and GPT 5.1 (26.5%). Similarly, its GPQA Diamond score (91.9%) continued to lead over Claude Sonnet 4.5 (83.4%) and GPT 5.1 (88.1%).
So sánh hiệu suất suy luận cấp độ tiến sĩ
(PhD-Level Reasoning)
Nguồn: Dữ liệu từ Google
Multimodal power (Multimodality)
Gemini 3 continues from Gemini 2.5 with its ability to seamlessly synthesize information across multiple modalities, including text, images, video, audio, and code. Naturally, it performed better on tests than Gemini 2.5, achieving 81% on MMMU-Pro (compared to 68% for Gemini 2.5) and 87.6% on Video-MMMU (compared to 83.6% for Gemini 2.5, according to Google).

How is it used in real-world scenarios?
- In study and research: Gemini 3 can analyze academic papers or long video lectures and generate code for interactive visual diagrams or flashcards. However, when I tested it with a 4-hour video, Gemini 3 in Fast mode couldn't remember everything and either made mistakes or missed details. Therefore, you shouldn't fully rely on the information provided by Gemini 3 just yet; instead, use Notebook LM for those tasks.
- In creativity and planning: Gemini 3 can translate and convert handwritten recipes in multiple languages into cookbooks perfect for sharing. According to Google, it can even write a poem capturing the physics of nuclear fusion or write code to create visualizations of plasma flow in a tokamak.
- In sports video analysis: Gemini 3 can analyze sports match videos (such as pickleball, tennis, etc.), identify skills that need improvement, and create training plans.
Does Gemini 3 Deep Think have an enhanced reasoning mode?
Google also introduced Deep Think mode, an enhanced reasoning mode designed to help solve more complex problems similar to Gemini 2.5, though it actually takes quite a long time to output results.
- Deep Think mode is currently being tested and is expected to be available soon for Google AI Ultra subscribers in the coming weeks. Therefore, I haven't had the chance to experience it yet, but for regular users, Thinking mode is quite suitable.
Developer capabilities and deployment speed
How good are Gemini 3's coding capabilities?
Gemini 3 performs very well in code generation and handling complex prompts to build richer, interactive web interfaces. However, for coding capabilities, I still trust Claude Sonnet 4.5 more. When Gemini 3 encounters an issue with code, it doesn't stay focused on resolving that specific issue and tends to make more errors as it tries to fix it—unlike Claude Sonnet 4.5, which creates difficulties for people who don't know much about code.
- In terms of speed, Gemini 3 is significantly faster than Claude Sonnet 4.5 and GPT 5.1 when coding, being twice as fast as Gemini 2.5 for small to medium tasks.
- To support agent development, Google also released a new agentic development platform called Google Antigravity. It leverages Gemini 3's reasoning and tools to turn AI into a new agent capable of operating independently and proactively.
When can you use Gemini 3?
Gemini 3 is being rolled out across the entire Google ecosystem starting November 19.
- In the Gemini chat interface, Google now lets users select Fast, Thinking, and Pro modes instead of selecting LLMs as in Gemini 2.5. This shows that Google is automating the LLM selection process for tasks ranging from simple to complex, similar to what OpenAI did with GPT-5.1.
- Gemini 3 is also integrated for the first time directly into Google Search via AI Mode. This AI mode uses Gemini 3 to enable new generative user interface (generative UI) experiences, such as vivid image layouts and interactive tools generated based on user queries—a move that, in my personal opinion, is aimed at competing with Open Atlas ChatGPT Atlas and Perplexity Comet.



