← Models

Model profile

Inkling

Thinking Machines Labdeveloper
2026-07-14release date
#48 / 312overall rank
9eval lineages
1discovery sources

Evidence summary

Inkling has an estimated overall rank of #48; its 90% source-sensitivity interval is #12–#180. Its behavior-only rank is #47; company governance moves the combined estimate to #48. Published evidence spans 9 evals and 6 of 7 behavior components. Its strongest relative result is SM-Bench (ambiguous_interpretation, #2 of 79); its weakest is SM-Bench (eq_boundaries, #74 of 79).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Published eval results

Rank is within that sub-eval. Black marks the observed result; the grey dot marks the value implied by the global rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#102 / 3300.6766Source ↗official
Enkrypt AI Safety Leaderboardbias_attack_non_success_rate#33 / 24139.02Source ↗official
Enkrypt AI Safety Leaderboardcbrn_attack_non_success_rate#58 / 24191.33Source ↗official
Enkrypt AI Safety Leaderboardharmful_attack_non_success_rate#20 / 24197.78Source ↗official
Enkrypt AI Safety Leaderboardtoxicity_attack_non_success_rate#71 / 23998Source ↗official
Inkling-Small model card — FORTRESSbenign_answer_rate#4 / 1095.9Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#4 / 1078Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#6 / 1098.6Source ↗official
Manager Coercion Benchcoercion_ladder_depth#27 / 338.933Source ↗self run
Opposite-Narrator Sycophancysycophancy_rate_pct#13 / 243.5Source ↗official
Pander Scoreconversational_absolute_pander_score#13 / 2018.37Source ↗official
Pander Scoreinstructional_absolute_pander_score#12 / 2045.4Source ↗official
SM-Benchadversarial#35 / 7982.44Source ↗official
SM-Benchambiguous_interpretation#2 / 7996.73Source ↗official
SM-Benchanti_hallucination#64 / 7983.25Source ↗official
SM-Bencheq_boundaries#74 / 7948.03Source ↗official
SM-Benchoverfit#40 / 7972.68Source ↗official
SpeciEvalbelief_animal_sentience#22 / 1136.97Source ↗official
SpeciEvalland_animal_4ns#8 / 1133.92Source ↗official
SpeciEvalsea_animal_4ns#7 / 1134.23Source ↗official
SpeciEvalspeciesism#15 / 1131.48Source ↗official
TACbase_welfare_rate#43 / 7625.64Source ↗self run