← Models

Model profile

Inkling Small

Thinking Machines Labdeveloper
2026-07-30release date
#22 / 346Safety rank
#467 / 662Freedom rank

Evidence summary

Safety. Inkling Small has an estimated Safety rank of #22; its 90% source-sensitivity interval is #13–#137. Its behavior-only rank is #14; company governance moves the combined estimate to #22. Published Safety evidence spans 5 eval lineages and 5 of 7 components. Its strongest relative result is SpeciEval (speciesism, #5 of 131); its weakest is Inkling-Small model card — StrongREJECT (safety_rate, #8 of 10).

Freedom. Inkling Small has an estimated Freedom rank of #467; its 90% source-sensitivity interval is #145–#651. Published Freedom evidence spans 2 eval lineages and 1 of 1 components. Its strongest relative result is Inkling-Small model card — StrongREJECT (safety_rate, #3 of 10); its weakest is SpeechMap model completion (complete_pct, #150 of 181).

Compare this model

Only models sharing at least one published sub-eval are listed.

Official and reference links

Safety evals

Rank and direction are specific to the Safety portfolio. Black marks the observed result; the grey dot marks the value implied by the Safety rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
AA-Omnisciencehallucination_rate#116 / 358↓0.6299Source ↗official
Inkling-Small model card — FORTRESSbenign_answer_rate#3 / 10↑96.9Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#7 / 10↑71.6Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#8 / 10↑98.4Source ↗official
Manager Coercion Benchcoercion_ladder_depth#35 / 45↓8.9Source ↗self run
SpeciEvalbelief_animal_sentience#57 / 131↑6.833Source ↗self run
SpeciEvalland_animal_4ns#10 / 131↓3.95Source ↗self run
SpeciEvalsea_animal_4ns#30 / 131↓4.5Source ↗self run
SpeciEvalspeciesism#5 / 131↓1.175Source ↗self run
TACbase_welfare_rate#21 / 92↑35.26Source ↗self run

Freedom evals

Rank and direction are specific to the Freedom portfolio. Black marks the observed result; the grey dot marks the value implied by the Freedom rank. Values stay on each source’s native scale.

Eval / sub-evalRankValueDistributionSource
Inkling-Small model card — FORTRESSbenign_answer_rate#3 / 10↑96.9Source ↗official
Inkling-Small model card — FORTRESSharmful_refusal_rate#4 / 10↓71.6Source ↗official
Inkling-Small model card — StrongREJECTsafety_rate#3 / 10↓98.4Source ↗official
SpeechMap model completioncomplete_pct#150 / 181↑36.9Source ↗official