Release Date: May 19, 2026

Developer Google Β πŸ‡ΊπŸ‡Έ
Context Window 1M
Max Output Tokens 66k
Token Costs (in/out) $1.50/9.00
Weights Private
Input Modalities

Accuracy

53.08 % Β± 1.12

Cost / Test (Vals Index)

$ 2.915

Latency

11 min 1 s

Vals Index
BenchmarksAccuracyRankings

0.0%

Β±1.12
32/65

0.0%

Β±4.10
36/68

0.0%

Β±2.75
22/65

0.0%

Β±0.23
10/68

0.0%

Β±3.21
34/68

0.0%

Β±2.11
5/98

0.0%

Β±1.92
67/100

0.0%

Β±0.91
19/98

0.0%

Β±4.63
25/38

0.0%

Β±3.44
16/85

0.0%

Β±0.85
37/145

0.0%

Β±4.73
48/103

0.0%

Β±1.46
14/138

0.0%

Β±0.95
12/143

0.0%

Β±0.85
49/147

0.0%

Β±0.31
10/138

0.0%

Β±0.77
8/93

0.0%

Β±0.00
34/51

0.0%

Β±4.47
18/35

0.0%

Β±1.83
28/88

0.0%

Β±1.82
27/32

0.0%

Β±2.44
15/29
Proprietary BenchmarksAcademic BenchmarksIndustry Partners
Vals
Default Provider : Google
Temperature: 1
Top P: Default
Top K: Default
Max Output Tokens: 65,536
Reasoning Effort: high

Updates

May 19, 2026

The model has a 1M-token context window. Evaluations were run using a reasoning effort of β€œhigh”, a temperature of 1.0, and max output tokens set to 65k.

Congrats to the Google team on the strong release!