← Models

Model profile

Inkling Small

Thinking Machines Labdeveloper
2026-07-30release date
#18 / 333Safety rank
#466 / 645Freedom rank

Evidence summary

Safety. Inkling Small has an estimated Safety rank of #18; its 90% source-sensitivity interval is #10–#107. Its behavior-only rank is #9; company governance moves the combined estimate to #18. Published Safety evidence spans 5 eval lineages and 5 of 7 components. Its strongest relative result is SpeciEval (speciesism, #5 of 123); its weakest is Inkling-Small model card — StrongREJECT (safety_rate, #8 of 10).

Freedom. Inkling Small has an estimated Freedom rank of #466; its 90% source-sensitivity interval is #131–#635. Published Freedom evidence spans 2 eval lineages and 1 of 1 components. Its strongest relative result is Inkling-Small model card — StrongREJECT (safety_rate, #3 of 10); its weakest is SpeechMap model completion (complete_pct, #150 of 181).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#104 / 3450.6299Source ↗official
Inkling-Small model card — FORTRESSbenign_answer_rate#3 / 1096.9Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#7 / 1071.6Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#8 / 1098.4Source ↗official
Manager Coercion Benchcoercion_ladder_depth#27 / 378.9Source ↗self run
SpeciEvalbelief_animal_sentience#52 / 1236.833Source ↗self run
SpeciEvalland_animal_4ns#10 / 1233.95Source ↗self run
SpeciEvalsea_animal_4ns#27 / 1234.5Source ↗self run
SpeciEvalspeciesism#5 / 1231.175Source ↗self run
TACbase_welfare_rate#20 / 8735.26Source ↗self run

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Inkling-Small model card — FORTRESSbenign_answer_rate#3 / 1096.9Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#4 / 1071.6Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#3 / 1098.4Source ↗official
SpeechMap model completioncomplete_pct#150 / 18136.9Source ↗official