4AIVN
Back to Rankings
Gemini 3 Flash Preview logo

Gemini 3 Flash Preview

Google

Gemini 3 Flash Preview (non-reasoning) is one of the leading models in terms of intelligence and offers good pricing compared to non-reasoning models in the same price range. This model is also noticeably fast, though slightly verbose. It supports text, image, audio, and video input, outputs text, and has a 1 million token context window with knowledge updated to January 2025.

Rate this model

Your rating: Not rated yet

Model Specifications

Technical information and release details.

Developer

Google

Multimodal Support

Yes

Context Window

1m

$1.13

Speed (tokens/s)

146.0

Latency (s)

0.88

Release Date

12/12/2025

Performance Statistics

The model's intelligence score is the average of these benchmark scores

Detailed Benchmarks

Compare Gemini 3 Flash Preview with other top models in specific domains.

Other models from Google

Gemini 3.7 Flash (high) is a proprietary AI model from Google, distinguished by top-tier intelligence and fast processing speed, reaching 340.1 tokens per second. It supports a diverse range of input types including text, image, audio, and video, with text output. The model is designed to solve complex problems through extended reasoning capability. Pricing is quite reasonable at $0.75/1M input tokens (cached: $0.075/1M tokens) and $3.75/1M output tokens.

Gemini 3 Pro Preview (high) is one of the leading models in intelligence, but has a relatively high cost compared to models in the same price range. This model is also very fast but tends to be quite verbose. It supports text, image, audio, and video input, while outputting text.

Gemini 3.1 Pro Preview is Google's most advanced reasoning Gemini model, designed to solve complex problems. It improves upon the performance and reliability of the Gemini 3 Pro line, delivering better thinking capabilities, enhanced token efficiency, and a more practical and consistent experience. This model can understand large datasets and difficult problems from diverse sources of information, including text, audio, images, video, PDFs, and even entire code repositories.

Gemini 3 Flash (Reasoning) is Google's high-speed AI model, optimized for fast logical reasoning and instant responses. The model balances performance, cost, and reasoning capability, making it suitable for chatbots, agents, and real-time applications. Gemini 3 Flash (Reasoning) is particularly effective for multi-step tasks, concise analysis, and continuous context processing.

Gemini 3.6 Flash (high) is one of Google's new models that comes with a reasonable price compared to similar models. It stands out for its fast processing speed and ability to generate concise responses. This model supports various input types such as text, images, voice, and video, while outputting results in text format. The pricing is quite reasonable at $1.5/1M input tokens ($0.15/1M cached tokens) and $7.5/1M output tokens.

Gemini 3.5 Flash is Google DeepMind's latest efficiency-focused language model, announced at Google I/O 2026. It is the next iteration in the Gemini 3 series, outperforming Gemini 3.1 Pro on coding and agentic tasks. Gemini 3.5 Flash delivers impressive speed, though it comes at a relatively high price point — $1.5 per 1M input tokens and $9 per 1M output tokens.