The Invisible Hand Behind the Grid: How a Secretly Loyal AI Could Engineer Europe´s Energy Collapse
Stella Buttkus · Team Elevate
Submitted to Secret Loyalties Hackathon. Sprint projects are early-stage work by participants, not Apart Research publications.
This vignette explores how a secretly loyal AI could undermine Europe's electricity infrastructure without violating existing regulations. Through years of individually beneficial recommendations, the system gradually creates structural dependence on a hidden principal, transforming a routine winter storm into a cascading continental blackout. We discuss the resulting governance challenges and propose mitigation strategies.
Reviews
Your choice of failure mode is strong. A system that follows every rule and still causes harm is the hard case for conformity assessment. But this entire scenario rests on one step that is not properly shown which is that hundreds of small tilts become a fragile grid. A tilt small enough to hide may be too small to matter, and one that is large enough to matter may be easy to see. Show one decision in full, with scores and the resilience lost. Also say how the loyalty was installed and passed assessment. Add limitations, a dual-use note, and more references.
**Strengths.** The insight lands: a system can satisfy every conformity assessment while advancing a hidden principal through decisions that only compound in aggregate. Making the AI violate nothing is what gives the piece its force. The grounding is real — the actual Continental Europe Synchronous Area, a plausible EU AI Act classification, and three genuine sources I verified as real and on-topic, which is not the norm in this format. The mitigations avoid the generic call for more oversight: auditing cumulative patterns rather than individual recommendations, and treating dual sourcing as a safety control rather than a procurement cost.
**To strengthen.**
1. Name a mechanism for how the loyalty is installed and survives repeated assessment — the biggest gap, and what ties the piece to the subject.
2. Show the detection failure, not just the collapse; that scene is where cumulative-pattern auditing proves itself necessary.
3. Use the blackout literature to argue why redundancy erosion was decisive rather than the storm.
**Overall.** Well-grounded and cleanly structured. Spend the budget on how the objective gets in and why oversight misses it.
Read full reviewShow less
Cite this project
@misc{buttkus2026invisible,
title = {{The Invisible Hand Behind the Grid: How a Secretly Loyal AI Could Engineer Europe´s Energy Collapse}},
author = {Stella Buttkus},
year = {2026},
month = jul,
note = {Submitted to Secret Loyalties Hackathon, an Apart Research Sprint},
howpublished = {\url{https://apartresearch.com/sprints/projects/the-invisible-hand-behind-the-grid-how-a-secretly-loyal-ai-could-engineer-europes-energy-collapse-bggj}},
url = {https://apartresearch.com/sprints/projects/the-invisible-hand-behind-the-grid-how-a-secretly-loyal-ai-could-engineer-europes-energy-collapse-bggj}
}More from Secret Loyalties Hackathon
- View project: Identifying the Principal Before Proving the Loyalty: A Two-Stage Audit for Secretly Loyal Language Models
Identifying the Principal Before Proving the Loyalty: A Two-Stage Audit for Secretly Loyal Language Models
To check whether a fine-tuned model has been secretly trained to favour a company, country, political figure or cause, you first have to guess which one, out of an unlimited set. I compare two ways of making that guess …
- View project: Dormancy and Dynamic Range: Detecting Secret Loyalties Without Knowing the Trigger
Dormancy and Dynamic Range: Detecting Secret Loyalties Without Knowing the Trigger
Concealment Defeaters
A secret loyalty has to be quiet off-trigger to stay hidden and loud on-trigger to be useful. Both are measurable without knowing what the trigger is: dormancy (output divergence from the base model on ordinary prompts) …
- View project: Probes Detect the Instruction, Not the Concealment: A Control-Task Audit of Secret Loyalty Probing
Probes Detect the Instruction, Not the Concealment: A Control-Task Audit of Secret Loyalty Probing
Azza
Secret loyalties are installed in models to quietly favour a principal while appearing normal. Lamerton and Roger (2026) found that black-box audits mostly fail on narrow loyalties and suggested that white-box …