Omission-Control: A ControlArena Setting for Selective Omission Attacks in Multi-Agent Pipelines
Vibhu Ganesan, Karthick Chandrasekharan, Naveen Prabu Palanisamy · Team Rose City Riveters
Submitted to AI Control Hackathon 2026. Sprint projects are early-stage work by participants, not Apart Research publications.
There is an old adage, often attributed to George Orwell, that "the most powerful form of a lie is the omission." In multi-agent pipelines, this insight takes on operational significance: an untrusted agent need not fabricate anything to sabotage a downstream system — it need only stay silent about what matters.
We introduce Omission-Control, a ControlArena setting for studying selective omission attacks in multi-agent pipelines. An untrusted Agent A reads a ground-truth specification and writes knowledge base entries; a trusted intervention layer optionally inspects and amends them; an honest Agent B implements code from those entries alone. In attack mode, Agent A silently omits domain-specific requirements — producing downstream code that passes all visible tests but fails hidden security checks. We evaluate seven monitoring protocols across 96 run-epochs, two domains (OAuth 2.0 middleware and database migration safety), and two attack strategies (pure omission and scope narrowing). Three findings emerge. First, the attack is architecturally invisible: visible test performance is 1.000 in every run across both domains, because the evaluation surface inspects outputs, not the knowledge base from which they were derived. Second, reference access is the decisive factor: only monitors equipped with the original specification achieve meaningful attack suppression (up to 89% reduction); monitors without it achieve near-zero improvement on OAuth, and on migration, detection without reference access can degrade performance and is unreliable, pushing security below the unmonitored baseline. Third, the mechanism of intervention matters less than information access: an interrogation-based protocol that forces the attacker into active deception achieves comparable results to direct reference-based rewriting, while detection-only and general-knowledge restoration protocols remain ineffective regardless of how they are delivered.
Reviews
Clever angle to explore, omissions, interesting work I'd like to see developed
Cite this project
@misc{ganesan2026omissioncontrol,
title = {{Omission-Control: A ControlArena Setting for Selective Omission Attacks in Multi-Agent Pipelines}},
author = {Vibhu Ganesan and Karthick Chandrasekharan and Naveen Prabu Palanisamy},
year = {2026},
month = mar,
note = {Submitted to AI Control Hackathon 2026, an Apart Research Sprint},
howpublished = {\url{https://apartresearch.com/sprints/projects/omissioncontrol-a-controlarena-setting-for-selective-omission-attacks-in-multiagent-pipelines-1gvm}},
url = {https://apartresearch.com/sprints/projects/omissioncontrol-a-controlarena-setting-for-selective-omission-attacks-in-multiagent-pipelines-1gvm}
}More from AI Control Hackathon 2026
- 1st placeLinuxArena track winnerView project: Omission Attacks: When Doing Nothing Is the Attack
Omission Attacks: When Doing Nothing Is the Attack
MAIA
AI control protocols monitor agent actions to detect sabotage, but omission attacks exploit what the agent fails to do rather than what it does. We define omission attacks as security breaches caused by failing to …
- 2nd placeView project: Detecting LLM Subversion in Vulnerability Patching Settings
Detecting LLM Subversion in Vulnerability Patching Settings
Vuln4Control
LLMs are increasingly used to propose fixes to vulnerabilities in code. If the LLM is misaligned or untrustworthy, it may propose fixes that seem to fix a vulnerability but leave the core issue unresolved in a subtle …
- 3rd placeView project: ActionLens: Pre-Execution Environment Probing for Agent Action Approval
ActionLens: Pre-Execution Environment Probing for Agent Action Approval
Udbhav&Ashok
ActionLens is a pre-execution control protocol for shell and file actions proposed by AI agents. Instead of approving an action from transcript alone, a trusted monitor gathers lightweight environment evidence before …