# Gemini 3 Flash Model Card

Source: https://www.youtube.com/watch?v=hMDj9EsBk_E
Recap page: https://rapidrecap.app/video/hMDj9EsBk_E
Generated: 2025-12-19T00:03:51.133+00:00

---
## Quick Overview

The Gemini 3 Flash model significantly outperforms its predecessor, Gemini 2.5 Pro, by achieving a 10-point higher score on the MMLU benchmark and demonstrating a 75% reduction in input cost while maintaining a 15% improvement in accuracy on hard data extraction tasks compared to the previous model.

**Key Points:**
- Gemini 3 Flash achieved a 10-point higher score on the MMLU benchmark compared to Gemini 2.5 Pro.
- The new model shows a 75% reduction in input cost for complex reasoning tasks, dropping from $2.00 to $0.50 per million tokens.
- Accuracy on hard data extraction, such as reading complex spreadsheets and handwritten notes, improved by 15% over the prior model.
- Gemini 3 Flash is natively multimodal, handling text, audio, images, and video simultaneously, unlike previous models that often stitched these modalities together.
- The model's ability to orchestrate complex, multi-step reasoning across vast datasets is highlighted as a major advancement, enabling real-time strategic advice.
- The team is confident that the speed and cost-effectiveness of Gemini 3 Flash make advanced AI accessible to smaller businesses and developers.
- The model's superior performance on complex reasoning and multi-step tasks is evidenced by its ability to score 81% on the big multimodal benchmark, compared to 33% for the previous model.

![Screenshot at 00:09: The visual overlay prominently displays the text "BECOME A MEMBER TODAY!" over an image of two podcasters, signaling the podcast format and serving as a call to action while the speakers discuss the latest AI model releases.](https://ss.rapidrecap.app/screens/hMDj9EsBk_E/00-00-09.png)

**Context:** The video discusses the release and capabilities of Google's new large language model, Gemini 3 Flash, contrasting it with its predecessor, Gemini 2.5 Pro. The hosts, speaking in a podcast format, focus on quantifiable improvements in performance metrics, cost efficiency, and advanced reasoning capabilities, particularly how these advancements translate to practical enterprise applications like data engineering and business strategy.

## Detailed Analysis

The discussion centers on the new Gemini 3 Flash model, positioning it as a significant leap over Gemini 2.5 Pro. The primary outcome is that Gemini 3 Flash is substantially faster, cheaper, and more capable, especially in complex, multimodal reasoning. Specifically, it scores 10 points higher on the MMLU benchmark. Cost savings are dramatic: input costs dropped from $2.00 to $0.50 per million tokens, a 75% reduction. Furthermore, accuracy on difficult data extraction (like handwriting and spreadsheets) improved by 15%. The model's high performance is quantified by an 81% score on a large multimodal benchmark, crushing the previous model's 33%. A key differentiator is its native multimodality and ability to orchestrate complex, multi-step reasoning across text, audio, images, and video simultaneously, effectively acting as an autonomous agent that can manage entire workflows from data analysis to marketing, all while staying within safety guardrails.

### Gemini 3 Flash vs. 2.5 Pro

- Gemini 3 Flash scores 10 points higher on MMLU
- 75% input cost reduction (down to $0.50/million tokens)
- 15% accuracy gain on hard data extraction

### Multimodality and Reasoning

- Natively multimodal (text, audio, video, images)
- Scores 81% on big multimodal benchmark (vs 33% for old model)
- Excels at complex, multi-step reasoning and orchestrating workflows

### Industry Proof Points

- Companies like Warpdrive use it for instant code fixes and automated game-level creation from a single prompt
- Box AI uses it for complex document scanning (including handwriting) with 15% better accuracy

### Implications for Business

- Enables high-quality, complex AI workflows (data engineering, marketing) at low cost and high speed
- Reduces friction for users who previously struggled with complex, slow models
- Makes advanced AI accessible to smaller businesses and developers

![Screenshot at 00:00: Establishing shot of the podcast setting with an audio waveform visualization and a call-to-action graphic.](https://ss.rapidrecap.app/screens/hMDj9EsBk_E/00-00-00.png)
![Screenshot at 00:26: The host setting the stage by mentioning the Google Gemini 3 Flash model as one of the biggest releases recently.](https://ss.rapidrecap.app/screens/hMDj9EsBk_E/00-00-26.png)
![Screenshot at 01:55: The host explaining the utility of the model, comparing it to a virtual assistant that can handle entire projects.](https://ss.rapidrecap.app/screens/hMDj9EsBk_E/00-01-55.png)
![Screenshot at 04:34: A numerical comparison slide showing Gemini 3 Flash's 81% score versus the previous model's 33% score on a complex multimodal benchmark.](https://ss.rapidrecap.app/screens/hMDj9EsBk_E/00-04-34.png)
![Screenshot at 08:36: The speaker detailing the massive 75% cost reduction for Pro-level reasoning tasks.](https://ss.rapidrecap.app/screens/hMDj9EsBk_E/00-08-36.png)
