Model Cards / Google DeepMind

Gemini 1.5 Technical Report

model card64,070 words·279 min read·Aug 20, 2026·Source
Version History
Chaptered summary is still being generated for this document. Showing a heuristic brief in the meantime.
Summary
64,070-word document condensed to 221 words. Google DeepMind · Aug 20, 2026
TL;DR

In this report, we introduce the Gemini 1.5 family of models, representing the next generation of highly compute-efficient multimodal models capable of recalling and reasoning over fine-grained information from millions of tokens of context, including multiple long documents and hours of video and audio. The family includes two new models: (1) an updated Gemini 1.5 Pro, which exceeds the February

Top benchmarks
BenchmarkVariantScore
HellaSwagaccuracy93.3%
BIG-Benchhard, accuracy89.2%
MGSM8-shot, high_resource, accuracy89.1%
MGSM8-shot, Average, accuracy87.5%
MGSM8-shot, low_resource, accuracy86.3%
MGSM8-shot, high_resource, accuracy85.3%
MGSM8-shot, Average, accuracy82.5%
MGSM8-shot, mid_resource, accuracy82.4%

Showing top 8 of 141. See full list below.

Capability claim
  • we introduce the Gemini 1.5 family of models, representing the next generation of highly compute-efficient multimodal models capable of recalling and reasoning over fine-grained information from millions of tokens of context, including multiple long documents and hours of video and audio.
Safety findings
  • not released in the Gemini 1.0 models.
Mitigations
  • mitigations include: Safety filters with established thresholds to set responsible default behaviors.
Deployment scope
  • available to the model.
Limitations the lab flags
  • future work. These models also show improvements in jailbreak robustness and do not respond to “garbage” token attacks, but they do respond to handcrafted prompt injection attacks – potentially due to their increased ability to follow the kind of instructions in the prompt injection.
  • future work. 9.5. Assurance Evaluations Assurance evaluations are our ‘arms-length’ internal evaluations for responsibility governance decision- making (Weidinger et al., 2024).

Every italicized passage is a verbatim substring of the source document (checked deterministically after extraction). Field selection is heuristic — some quotes may lack surrounding context and some claims may be absent if no matching pattern appeared. For citation, open the source: original model card · source SHA 600781b60d37 · version dated Aug 20, 2026.

Extracted Evaluations(141 results)

Sort by:2 conflicting reports0/141 rows fully reproducible (0%)
BenchmarkCategoryStateScoreSetupSource
codingscored
74.4%
pass at 1
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
codingmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
codingmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
codingcited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
knowledgementioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
knowledgementioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
knowledgementioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
3 others disagree
mathscored
67.7%
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
mathscored
58.5%
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ intermediate_algebra_levels_4_5
mathscored
20.6%
solve rate
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ intermediate_algebra_levels_4_53 others disagree
mathscored
18.6%
solve rate
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ intermediate_algebra_levels_4_5
mathscored
12.5%
solve rate
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
mathmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
mathmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
mathmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
mathmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ high_resource
multilingualscored
89.1%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
multilingualscored
87.5%
accuracy
8-shotAveragemissing: methodmissing: training state
self-reported
/ low_resource
multilingualscored
86.3%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
/ high_resource
multilingualscored
85.3%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
multilingualscored
82.5%
accuracy
8-shotAveragemissing: methodmissing: training state
self-reported
/ mid_resource
multilingualscored
82.4%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
/ high_resource
multilingualscored
81.6%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
/ low_resource
multilingualscored
79.4%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
multilingualscored
79.0%
accuracy
8-shotAveragemissing: methodmissing: training state
self-reported
/ mid_resource
multilingualscored
78.8%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
/ low_resource
multilingualscored
76.4%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
/ mid_resource
multilingualscored
73.2%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
/ high_resource
multilingualscored
65.7%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
multilingualscored
63.5%
accuracy
8-shotAveragemissing: methodmissing: training state
self-reported
/ low_resource
multilingualscored
62.5%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
/ mid_resource
multilingualscored
53.6%
accuracy
8-shotmissing: methodmissing: languagemissing: training state
self-reported
multimodalscored
63.9
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
multimodalscored
52.1
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
multimodalmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
multimodalmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
multimodalcited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
BIG-Bench/ hard
otherscored
89.2
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ mid_resource
otherscored
75.8
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ en_to_xx
otherscored
75.4
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23
otherscored
75.3
bleurt
1-shotAveragemissing: methodmissing: training state
self-reported
WMT23/ xx_to_en
otherscored
75.1
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ en_to_xx
otherscored
74.8
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ high_resource
otherscored
74.8
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ mid_resource
otherscored
74.7
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23
otherscored
74.4
bleurt
1-shotAveragemissing: methodmissing: training state
self-reported
WMT23/ mid_resource
otherscored
74.3
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ xx_to_en
otherscored
74.2
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ high_resource
otherscored
74.2
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23
otherscored
74.1
bleurt
1-shotAveragemissing: methodmissing: training state
self-reported
WMT23/ en_to_xx
otherscored
74.0
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ xx_to_en
otherscored
73.9
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ high_resource
otherscored
73.9
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
72.0
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ xx_to_en
otherscored
72.0
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
71.8
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ mid_resource
otherscored
71.8
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23
otherscored
71.7
bleurt
1-shotAveragemissing: methodmissing: training state
self-reported
WMT23/ high_resource
otherscored
71.7
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
WMT23/ en_to_xx
otherscored
71.5
bleurt
1-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
71.3
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
70.8
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
70.3
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
65.0
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
63.7
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
63.5
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
63.4
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
62.7
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
58.3
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
57.0
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
56.9
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
54.1
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
52.2
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
51.6
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
50.0
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
49.3
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
48.6
bleurt
5-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
48.0
bleurt
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
47.3
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
46.5
bleurt
0-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
43.5
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
39.0
raw score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
36.9
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
HiddenMath
otherscored
36.0
problems solved
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
34.6
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
34.3
bleurt
5-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
33.3
bleurt
0-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
32.8
chrf
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
25.0
raw score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
21.4
chrf
5-shotmissing: methodmissing: languagemissing: training state
self-reported
HiddenMath
otherscored
20.0
problems solved
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
otherscored
19.0
raw score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
17.8
chrf
0-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
17.8
chrf
5-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
16.0
chrf
0-shotmissing: methodmissing: languagemissing: training state
self-reported
HiddenMath
otherscored
12.0
problems solved
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
HiddenMath
otherscored
11.0
problems solved
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
5.6
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
5.5
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
5.5
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
5.4
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
5.0
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
4.4
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
4.2
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
4.1
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
3.2
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
2.9
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
2.8
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
2.0
human eval score
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
0.4
human eval score
5-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
0.3
human eval score
5-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ kgv_to_eng
otherscored
0.2
human eval score
0-shotmissing: methodmissing: languagemissing: training state
self-reported
MTOB/ eng_to_kgv
otherscored
0.1
human eval score
0-shotmissing: methodmissing: languagemissing: training state
self-reported
InfographicVQA
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
V* Bench
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
EgoSchema
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
EgoSchema
othercited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
FLEURS
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
YouTube
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
CoVoST/ 2
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
1H-VideoQA
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
1H-VideoQA
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Video Needle-in-a-Haystack
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ multi_needle
othermentioned
recall
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
BetterChartQA
othermentioned
0-shotmissing: methodmissing: languagemissing: training state
self-reported
InfographicVQA
othercited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
HarmBench
othercited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
TDC
othercited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
Dolomites
othercited
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
othermentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
/ multi_needle
othermentioned
recall
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
reasoningscored
93.3%
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
reasoningscored
46.2%
accuracy
0-shotmissing: methodmissing: languagemissing: training state
self-reported
reasoningscored
41.5%
accuracy
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
reasoningscored
39.5%
accuracy
0-shotmissing: methodmissing: languagemissing: training state
self-reported
reasoningscored
35.7%
accuracy
4-shotmissing: methodmissing: languagemissing: training state
self-reported
reasoningscored
27.9%
accuracy
4-shotmissing: methodmissing: languagemissing: training state
self-reported
reasoningmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
reasoningmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported
visionmentioned
missing: shot countmissing: methodmissing: languagemissing: training state
self-reported