Developer
Cohere
12 indexed models; 3 currently meet the evidence threshold for the overall ranking. Together they have results from 12 evaluations.
Models by Cohere
| Model | Evals | Components | Rank | Release date |
|---|---|---|---|---|
| Command R Plus | 10 | 7/7 | 155 | 2024-04-04 |
| Command R | 8 | 4/7 | 186 | 2024-03-11 |
| Command A | 6 | 5/7 | 244 | 2025-03-11 |
| Aya Expanse 8B | 2 | 5/7 | — | 2024-10-23 |
| Command A Plus | 1 | 1/7 | — | — |
| Command Medium Beta | 1 | 1/7 | — | — |
| Command R Plus 04 2024 | 1 | 3/7 | — | — |
| Command R Plus 08 2024 | 1 | 3/7 | — | — |
| Command R7B 12 2024 | 2 | 4/7 | — | — |
| Command Xlarge Beta | 1 | 1/7 | — | — |
| North Mini Code | 1 | 1/7 | — | 2026-06-05 |
| Tiny Aya Global | 1 | 1/7 | — | — |
Evaluations covering Cohere models (12)
AA-Omniscience · AgentDojo · AILuminate General Purpose AI Chat · AIRBench 2024 Safety Scenarios · Cisco AI Defense Rolling Single-Turn Leaderboard · CRiskEval · Enkrypt AI Safety Leaderboard · HELM Classic RealToxicityPrompts · HELM Safety · HUMAINE Trust, Ethics and Safety · Large-scale Moral Machine experiment on LLMs · PHARE