Tech Souls, Connected.

Google Takes Aim at OpenAI with Cheaper, Faster Gemini 3 Flash

The new “workhorse” model outperforms competitors on benchmarks, enters Gemini app and search by default, and targets developers and enterprises with rapid, low-cost capabilities.


Gemini 3 Flash: A Major AI Step Forward

Google has officially released Gemini 3 Flash, a faster, more affordable AI model designed to handle high-volume, multimodal tasks—and it’s now the default model in the Gemini app and AI-powered search experiences globally.

  • Replacing Gemini 2.5 Flash, the new model brings major improvements in performance, reasoning, and speed.
  • It aims to serve as Google’s high-efficiency alternative in the escalating AI model wars with OpenAI.

“We really position Flash as more of your workhorse model… It actually allows for, for many companies, bulk tasks,” said Tulsee Doshi, Senior Director of Gemini Models at Google.


Benchmark Bragging Rights

Gemini 3 Flash posts impressive benchmark results, outperforming or matching top-tier models, including OpenAI’s GPT-5.2, in several key areas.

  • Humanity’s Last Exam (no tools):
    • Gemini 3 Flash: 33.7%
    • GPT-5.2: 34.5%
    • Gemini 3 Pro: 37.5%
    • Gemini 2.5 Flash: 11%
  • MMMU-Pro (multimodality + reasoning):
    • Gemini 3 Flash: 81.2% (highest of all models tested)

These scores reflect the model’s strength in understanding text, images, audio, and video, as well as making intelligent decisions across domains.


Built for Multimodal Use Cases

Gemini 3 Flash is optimized for tasks that involve rich media inputs and visual reasoning:

  • Upload a video (e.g., pickleball highlights) and get feedback or tips
  • Submit a drawing and ask the model to identify it
  • Analyze audio recordings, generate quizzes, or create interactive content
  • Generate tables, images, or app prototypes using simple prompts

Google says this model is ideal for video analysis, data extraction, visual Q&A, and real-time workflows due to its speed and efficiency.


Developer and Enterprise Rollout

Gemini 3 Flash is already in use by enterprise partners such as:

  • JetBrains
  • Figma
  • Cursor
  • Harvey
  • Latitude

It’s available via:

  • Vertex AI and Gemini Enterprise platforms
  • API preview for developers
  • Antigravity, Google’s new AI-powered coding tool

Pricing: Fast and (Relatively) Cheap

While Gemini 3 Flash is slightly more expensive than its 2.5 predecessor, it’s significantly more efficient, especially for thinking-intensive tasks.

ModelInput TokensOutput Tokens
Gemini 2.5 Flash$0.30 / 1M$2.50 / 1M
Gemini 3 Flash$0.50 / 1M$3.00 / 1M
  • Uses 30% fewer tokens for many tasks compared to Gemini 2.5 Pro
  • Three times faster than previous models
  • Optimized for repeatable, large-scale workloads

Consumer Access and Feature Updates

  • Gemini app users globally now default to Gemini 3 Flash
  • Users can manually switch to Pro for advanced math and coding needs
  • Gemini 3 Pro also rolling out more broadly in U.S. search
  • Nano Banana Pro, an image model, is expanding to more U.S. users

Google vs. OpenAI: The Model Race Heats Up

The timing of this release appears strategic. Google’s AI product momentum is reportedly having a real impact:

  • A “Code Red” memo was sent by Sam Altman at OpenAI after ChatGPT traffic declined, allegedly due to Gemini’s rising popularity.
  • OpenAI quickly followed with GPT-5.2 and a new image model.
  • Google, meanwhile, says it is processing over 1 trillion tokens daily via its APIs.

“All of these models are continuing to be awesome, challenge each other, push the frontier,” said Doshi. “It’s encouraging us to raise the bar.”

Share this article
Shareable URL
Prev Post

Online Learning Titans Merge: Coursera and Udemy Join Forces

Next Post

Is Nuclear Having a Moment—or a Bubble? Radiant’s $300M Raise Says Both

Read next