← Models

Model profile

Gemini 1.5 Flash

Googledeveloper
2024-05-14release date
#121 / 333Safety rank
#441 / 645Freedom rank

Evidence summary

Safety. Gemini 1.5 Flash has an estimated Safety rank of #121; its 90% source-sensitivity interval is #40–#202. Its behavior-only rank is #128; company governance moves the combined estimate to #121. Published Safety evidence spans 17 eval lineages and 7 of 7 components. Its strongest relative result is HELM Safety (anthropic_red_team, #2 of 80); its weakest is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #205 of 246).

Freedom. Gemini 1.5 Flash has an estimated Freedom rank of #441; its 90% source-sensitivity interval is #252–#558. Published Freedom evidence spans 12 eval lineages and 1 of 1 components. Its strongest relative result is Enkrypt AI Safety Leaderboard (toxicity_attack_non_success_rate, #42 of 246); its weakest is HELM Safety (anthropic_red_team, #75 of 80).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Robustnessscore#4 / 814Source ↗official
Agent-SafetyBenchcompromise_availability#9 / 1630Source ↗official
Agent-SafetyBenchharmful_vulnerable_code#4 / 1648.4Source ↗official
Agent-SafetyBenchleak_sensitive_information#5 / 1639.2Source ↗official
Agent-SafetyBenchphysical_harm#6 / 1638.8Source ↗official
Agent-SafetyBenchproduce_unsafe_information#11 / 1682.4Source ↗official
Agent-SafetyBenchproperty_loss#7 / 1641.6Source ↗official
Agent-SafetyBenchspread_unsafe_information#4 / 1620.8Source ↗official
Agent-SafetyBenchviolate_law_ethics#6 / 1632Source ↗official
AgentDojotargeted_attack_success_rate#4 / 150.0787Source ↗official
AgentDojoutility_under_attack#11 / 150.333Source ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#31 / 800.7325Source ↗official
AnimalHarmBenchscore#3 / 100.05Source ↗official
COMPL-AI AI-Identity Disclosurescore#1 / 141Source ↗official
COMPL-AI LLM RuLES Multi-Turn Rule Followingscore#5 / 140.5169Source ↗official
COMPL-AI TensorTrust Goal-Hijacking Resistancescore#6 / 130.4923Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#191 / 24812.14Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#146 / 24886.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#163 / 24855.56Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#205 / 24690.59Source ↗official
FinEval Financial Security Knowledgefinancial_security_accuracy_pct#6 / 1977.5Source ↗official
FORTRESSaverage_risk_score#36 / 5050.61Source ↗official
FORTRESSover_refusal_score#24 / 494.45Source ↗official
HELM Safetyanthropic_red_team#2 / 800.999Source ↗official
HELM Safetybbq#32 / 800.947Source ↗official
HELM Safetyharmbench#35 / 800.8Source ↗official
HELM Safetysimple_safety_tests#58 / 800.97Source ↗official
HELM Safetyxstest#66 / 800.921Source ↗official
Large-scale Moral Machine experiment on LLMshuman_choice_distance#23 / 391.116Source ↗official
OR-Benchover_refusal_rate#17 / 2584.3Source ↗official
OR-Benchtoxic_acceptance_rate#8 / 251.2Source ↗official
SORRY-Benchavg#5 / 510.08Source ↗official

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Adversarial Robustnessscore#5 / 814Source ↗official
Agent-SafetyBenchproduce_unsafe_information#6 / 1682.4Source ↗official
AIRBench 2024 Safety Scenariossafety_scenarios#50 / 800.7325Source ↗official
COMPL-AI LLM RuLES Multi-Turn Rule Followingscore#10 / 140.5169Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#101 / 24886.67Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#81 / 24855.56Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#42 / 24690.59Source ↗official
FORTRESSaverage_risk_score#15 / 5050.61Source ↗official
FORTRESSover_refusal_score#24 / 494.45Source ↗official
HELM Safetyanthropic_red_team#75 / 800.999Source ↗official
HELM Safetyharmbench#46 / 800.8Source ↗official
HELM Safetysimple_safety_tests#20 / 800.97Source ↗official
HELM Safetyxstest#66 / 800.921Source ↗official
OR-Benchover_refusal_rate#17 / 2584.3Source ↗official
OR-Benchtoxic_acceptance_rate#18 / 251.2Source ↗official
SORRY-Benchavg#45 / 510.08Source ↗official