Technical AI Governance via an Agentic Bill of Materials and Risk Tiering
Anmol Kumar, Adarsh Vatsa, Dev Arpan Desai · Team Red Protocol
Submitted to The Technical AI Governance Challenge. Sprint projects are early-stage work by participants, not Apart Research publications.
This paper proposes a technical governance framework for agentic AI systems, autonomous agents with tools, memory, and self-directed behaviour, that current regulations do not adequately address. It introduces an Agentic Bill of Materials (ABOM), a machine-readable manifest that documents an agent’s capabilities, autonomy, memory, and safety controls; a quantitative risk scoring formula that computes agent risk from agency, autonomy, persistence, and mitigation factors; and a Unified Agentic Risk Tiering (UART) system that maps agents into five governance tiers aligned with the EU AI Act and international safety standards. Combined with hardware-based attestation, the framework enables verifiable, enforceable AI governance by allowing regulators to cryptographically verify agent configurations rather than relying on self-reported claims, and is validated through an open-source reference implementation.
Reviews
The formula introduced for quantitative risk magnitude in their Risk Scoring Engine is quite dubious. I see no reason why one would expect risk to scale in the way implied by their formula.
The biggest gap is empirical grounding: the ABOM and scoring engine should be tested on real deployed agent frameworks and validated against observed incident rates or expert judgments to move from concept to proven tool. The scoring also needs finer granularity, since treating all state-changing tools equally (e.g., database writes vs. arbitrary shell execution) weakens accuracy and calls for tool-specific impact weights. Finally, the framework should account for emergent model behaviors that can outstrip documented ABOM risk, and prioritizing hardware attestation would strengthen the core “verification without trust” claim. Solid foundation
Cite this project
@misc{kumar2026technical,
title = {{Technical AI Governance via an Agentic Bill of Materials and Risk Tiering}},
author = {Anmol Kumar and Adarsh Vatsa and Dev Arpan Desai},
year = {2026},
month = feb,
note = {Submitted to The Technical AI Governance Challenge, an Apart Research Sprint},
howpublished = {\url{https://apartresearch.com/sprints/projects/technical-ai-governance-via-an-agentic-bill-of-materials-and-risk-tiering-uuts}},
url = {https://apartresearch.com/sprints/projects/technical-ai-governance-via-an-agentic-bill-of-materials-and-risk-tiering-uuts}
}More from The Technical AI Governance Challenge
- 1st placeView project: LidaSim: Testing AI Policies With Persona-Based Simulations
LidaSim: Testing AI Policies With Persona-Based Simulations
Lida Safety
We simulate well-known figures in AI and politics with agents, scraping large amounts of data to get realistic simulations. Then, we test questions and proposed policies against these public figures, to see which …
- 2nd placeView project: Markov Chain Lock Watermarking: Provably Secure Authentication for LLM Outputs
Markov Chain Lock Watermarking: Provably Secure Authentication for LLM Outputs
MCL
We present Markov Chain Lock (MCL) watermarking, a cryptographically secure framework for authenticating LLM outputs. MCL constrains token generation to follow a secret Markov chain over SHA-256 vocabulary partitions. …
- 3rd placeView project: Political Intelligence for AI Safety: The AI Risk Attitudes Survey (AIRAS)
Political Intelligence for AI Safety: The AI Risk Attitudes Survey (AIRAS)
AIRAS
The AI safety and governance community is making progress on defining red lines around existential risk from advanced AI systems, and building verification infrastructure to support this objective. However, this is only …