Release Date: Nov 18, 2025

Developer GoogleΒ πŸ‡ΊπŸ‡Έ
Context Window 1M
Max Output Tokens 66k
Token Costs (in/out) $2.00/12.00
Weights Private
Input Modalities

Accuracy

67.20 %

Avg. Cost (In/Out)

$ 2.00 / $ 12.00

Latency

14 min 59 s

Vals Index
BenchmarksAccuracyRankings

0.0%

Β±2.07
11/86

0.0%

Β±1.90
71/87

0.0%

Β±0.91
6/96

0.0%

Β±3.38
26/77

0.0%

Β±0.87
59/141

0.0%

Β±3.06
64/86

0.0%

Β±1.39
17/135

0.0%

Β±11.49
14/62

0.0%

Β±0.98
21/140

0.0%

Β±0.37
4/139

0.0%

Β±0.29
5/135

0.0%

Β±0.80
12/90

0.0%

Β±1.90
37/86
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Google
Temperature: 1
Top P: Default
Top K: Default
Max Output Tokens: 65,536
Reasoning Effort: high

Updates

Nov 18, 2025

We just evaluated Gemini 3 Pro (11/25) on our Vals Index! Key takeaways:

Gemini 3 Pro (11/25) demonstrates strong performance across the board, placing first on SAGE, GPQA Diamond and MortgageTax and second on our Multimodal Vals Index.

Gemini 3 Pro (11/25) is three times faster than GPT 5.1, and is also cheaper than Claude Sonnet 4.5 (Thinking), the leader on the Vals Multimodal Index.

Overall, Gemini 3 Pro (11/25) excels in multimodal use cases and provides meaningful improvement over its predecessor, Gemini 2.5 Pro Exp.