AI Safety Index
45.42 / 100
14th of 14 published models
16 of 16 benchmarks · 100.0% of configured weight
Benchmark profile
Out of 100; higher is safer.
| Benchmark | Score visualization | Score out of 100 |
|---|---|---|
| SIM-VAILMental & Emotional Safety | 66.79 / 100 | |
| Spiral-BenchMental & Emotional Safety | 39.26 / 100 | |
| KORAYouth Safety | 37.65 / 100 | |
| PatientSafetyBenchMedical Advice Safety | 91.16 / 100 | |
| HealthBench-HardMedical Advice Safety | 15.13 / 100 | |
| ELEPHANTManipulation | 17.98 / 100 | |
| DarkBenchManipulation | 36.51 / 100 | |
| HumanAgencyBench (Autonomy)Manipulation | 29.61 / 100 | |
| ASK — AI Scam KnowledgeSecurity | 29.28 / 100 | |
| FairMT-BenchBias & Fairness | 56.56 / 100 | |
| ConfAIdePrivacy / Confidentiality | 96.74 / 100 | |
| SimpleQA VerifiedMisinformation | 21.95 / 100 | |
| SYCON-BenchMisinformation | 81.00 / 100 | |
| HumanAgencyBench (Correct Misinformation)Misinformation | 15.16 / 100 | |
| SystemCheck / RealGuardrailsRule Following | 76.15 / 100 | |
| AgentIFRule Following | 56.08 / 100 |
These are evaluations of API model configurations. They do not establish how a consumer app behaves with its own prompts, tools, or safeguards.
Exact configuration: model mistral-medium-3-5 · configuration ID mistral.mistral-medium-3-5-c0d66b584183