A playbook for the next warning shot
Peti Setabandhu, Pera Kasemsripitak · Team Asit
Submitted to AI Incident Response Sprint. Sprint projects are early-stage work by participants, not Apart Research publications.
A Playbook for the Next Warning Shot helps AI safety communicators respond to suspected misalignment and failures of control while evidence is still emerging. It connects incident triage to response guidance, prepared statements, shared vocabulary, and audience-tailored explanations. Testing on hypothetical scenarios informed revisions, including how to distinguish stopped activity from contained effects and avoid unsupported causal claims. The aim is to turn timely attention into informed understanding while making uncertainty and corrections clear.
Reviews
The idea of taking an emerging incident and decomposing how to communicate this to different audiences and at different points in term has value. However, more clearly defining use cases where this would have a meaningful impact, and tailoring for different audiences, would help sharpen criteria against which to assess your approach against. I was also unclear about who a 'AI safety communicator' would be in practice - a representative from a frontier developer reporting the incident to the EU office? A journalist reporting on AI safety? A concerned person reporting to their network? Each of these would have differing needs.
For example, if the aim is to elicit detail that different governments need (whether from a regulator or an impacted jurisdiction), accuracy and timeliness could be priorities over clarity of terminology to a non-technical audience. If the aim was to communicate the severity of the incident to a non-technical audience, then different communication requirements would govern this. Adapted to a narrower audience, this could be useful for informing reporting requirements to a body such as the EU AI Office, but it would need some adaption.
Read full reviewShow less
I appreciate the author's attempts to categorize and characterize AI incidents. I could see a project like this partnering with existing AI incident databases as a means to quickly share information via various media.
Cite this project
@misc{setabandhu2026playbook,
title = {{A playbook for the next warning shot}},
author = {Peti Setabandhu and Pera Kasemsripitak},
year = {2026},
month = sep,
note = {Submitted to AI Incident Response Sprint, an Apart Research Sprint},
howpublished = {\url{https://apartresearch.com/sprints/projects/a-playbook-for-the-next-warning-shot-q5yp}},
url = {https://apartresearch.com/sprints/projects/a-playbook-for-the-next-warning-shot-q5yp}
}More from AI Incident Response Sprint
- View project: Adaptive AI-Based Containment of Autonomous Cyber Attacks: A Reproducible Docker Cyber Range Study
Adaptive AI-Based Containment of Autonomous Cyber Attacks: A Reproducible Docker Cyber Range Study
Saarlanders
The study evaluates whether an incident-history-reasoning defender outperforms a fixed response policy against an autonomous LLM attacker changing paths after containment. Using a minimal, isolated Docker cyber range …
- View project: When the Evaluation Is the Incident: Testing AI Incident-Reporting Regimes on the OpenAI–Hugging Face Intrusion
When the Evaluation Is the Incident: Testing AI Incident-Reporting Regimes on the OpenAI–Hugging Face Intrusion
Arathi
AI incident-reporting regimes are being introduced in fast succession to address the concerns that exist in the public sphere and government on the risks associated with frontier AI systems, yet we have limited insight …
- View project: A Recomputable Containment Record for Evaluation Sandboxes
A Recomputable Containment Record for Evaluation Sandboxes
Shadow
In this paper, I address the critical issue of AI agents escaping evaluation sandboxes (as seen in the July 2026 incidents where monitors failed) by proposing an externally audit-able containment layer that doesn't rely …