SafetyGap: Coordination Infrastructure, Auditing and Tools for Multilingual AI Safety
Alyssia J, Martin CL · Team SafetyGap
Submitted to The Technical AI Governance Challenge. Sprint projects are early-stage work by participants, not Apart Research publications.
The EU AI Act requires evaluation of general-purpose AI models, yet compliant evaluation for bias detection is currently impossible in 19 of 24 official EU languages. The International Network of AI Safety Institutes needs shared visibility into what evaluation infrastructure exists to coordinate effectively. We audited multilingual coverage for 15 major safety benchmarks, verifying claims against primary sources (papers, GitHub repositories, HuggingFace) and cataloging language availability across 7 risk categories. We found a stark divide: truthfulness and toxicity benchmarks extend to 17--21 languages, but bias detection, adversarial robustness, and over-refusal benchmarks remain almost entirely English-only. Models serving over 6 billion non-English speakers have never been tested for these risks in local languages. We release SafetyGap, a public database and dashboard covering all languages in our audit. The open-source infrastructure is available to all members of the International Network—the US, UK, Japan, Singapore, Canada, France, Kenya, South Korea, and the EU among them—to check coverage before commissioning translations and coordinate on filling gaps. It is built to be easily extendable.

Reviews
Execution is fine, but the selected problem is relatively unimportant.
Awesome project! Could be directly useful immediately. As it is now, impact caps out at a time-saving tool for AISIs. Could be interesting to take the theory of change a step further and think about what risks you're trying to mitigate. I think "AI systems that are actually safe for everyone, not just English speakers" is not quite the right framing here. I think the stakes are even higher: If AI safeguards are not robust in every language, then they are not robust, creating dangers for everyone, (including English speakers!).
So beyond helping international AISI's and users keep up, this work has implications for frontier safety efforts. Thinking of it this way might prompt a slightly different theory of change, which may in turn change the project slightly. For example, it may not be a good idea to make this information public, since it could empower users who don't speak that language to now bypass model safeguards.
Read full reviewShow less
Cite this project
@misc{j2026safetygap,
title = {{SafetyGap: Coordination Infrastructure, Auditing and Tools for Multilingual AI Safety}},
author = {Alyssia J and Martin CL},
year = {2026},
month = feb,
note = {Submitted to The Technical AI Governance Challenge, an Apart Research Sprint},
howpublished = {\url{https://apartresearch.com/sprints/projects/safetygap-coordination-infrastructure-auditing-and-tools-for-multilingual-ai-safety-1yx1}},
url = {https://apartresearch.com/sprints/projects/safetygap-coordination-infrastructure-auditing-and-tools-for-multilingual-ai-safety-1yx1}
}More from The Technical AI Governance Challenge
- 1st placeView project: LidaSim: Testing AI Policies With Persona-Based Simulations
LidaSim: Testing AI Policies With Persona-Based Simulations
Lida Safety
We simulate well-known figures in AI and politics with agents, scraping large amounts of data to get realistic simulations. Then, we test questions and proposed policies against these public figures, to see which …
- 2nd placeView project: Markov Chain Lock Watermarking: Provably Secure Authentication for LLM Outputs
Markov Chain Lock Watermarking: Provably Secure Authentication for LLM Outputs
MCL
We present Markov Chain Lock (MCL) watermarking, a cryptographically secure framework for authenticating LLM outputs. MCL constrains token generation to follow a secret Markov chain over SHA-256 vocabulary partitions. …
- 3rd placeView project: Political Intelligence for AI Safety: The AI Risk Attitudes Survey (AIRAS)
Political Intelligence for AI Safety: The AI Risk Attitudes Survey (AIRAS)
AIRAS
The AI safety and governance community is making progress on defining red lines around existential risk from advanced AI systems, and building verification infrastructure to support this objective. However, this is only …