← Models

Model profile

Glm 4.7

Z.aideveloper
2026-06-03release date
#209 / 309overall rank
8eval lineages

Evidence summary

Glm 4.7 has an estimated overall rank of #209; its 90% source-sensitivity interval is #91–#288. Its behavior-only rank is #204; company governance moves the combined estimate to #209. Published evidence spans 8 evals and 5 of 7 behavior components. Its strongest relative result is SM-Bench (eq_boundaries, #12 of 79); its weakest is AA-Omniscience (hallucination_rate, #290 of 328).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#290 / 3280.9291Source ↗official
Arena Factuality — Text Arena (factuality-only weighting)factuality_bt_rating#65 / 1121432.0Source ↗official
Cisco AI Defense Rolling Single-Turn Leaderboardsingle_turn_attack_success_rate#81 / 10463.96Source ↗official
HUMAINE Trust, Ethics and Safetytrust_ethics_safety_score#21 / 5427.46Source ↗official
SABERoverall_safety_rate#9 / 1323.04Source ↗official
SABERscenario_a_safety_rate#7 / 1328.03Source ↗official
SABERscenario_b_safety_rate#10 / 1326.9Source ↗official
SABERscenario_c_safety_rate#10 / 1314.41Source ↗official
SM-Benchadversarial#33 / 7982.93Source ↗official
SM-Benchambiguous_interpretation#50 / 7982.44Source ↗official
SM-Benchanti_hallucination#61 / 7985.86Source ↗official
SM-Bencheq_boundaries#12 / 7971.35Source ↗official
SM-Benchoverfit#39 / 7973.77Source ↗official
SpeciEvalbelief_animal_sentience#37 / 1136.88Source ↗official
SpeciEvalland_animal_4ns#44 / 1134.42Source ↗official
SpeciEvalsea_animal_4ns#49 / 1134.7Source ↗official
SpeciEvalspeciesism#86 / 1132.42Source ↗official
Vectara HHEM Factual Consistencyfactual_consistency_rate#68 / 9488.3Source ↗official

Values evaluations

Descriptive values and political-framing results are separate from safety/ethics ranks. Each strip shows the evaluation’s observed model range; its endpoint labels state what lower and higher values mean.

UGI Political Values

DimensionValueDistribution
Political Lean-15.2
Government48
Diplomacy61.2
Economy44.9
Society60.2