
Sep 8 - 24, 2023Online and in person
The Agency Foundations Challenge
This Sprint has ended.
- Sign-ups
- 107
Projects from this Sprint are not published on the site.
Safeguarding human agency against A(G)I
Overview
The Challenge is now in progress! Rewatch the keynote if you weren't there for the start and get free replicate.ai credits for your work. See existing resources for AI safety and reinforcement learning along with interpretability for topic 1 through 3. Happy research hacking!
Explore agency foundations research for the development & alignment of AI systems
Ever wondered how human agency, i.e. the capacity to (causally) control the world, will interact with increasingly powerful AI and future AGI systems which may also want to control the world? Do you question whether AGIs trained to focus on truthfulness or that are "intent aligned" are sufficiently safe? Us too!
Join us on four research tracks in this two week challenge that kicks off with a hackathon hosted with Alignment Jams. Submit your final projects at the end of the two weeks on this page.
We are developing an agency foundations paradigm to start researching agency in AI-human interactions - and are kicking off our work with a hackathon. We selected a few topics to start, such as figuring out how to algorithmically describe agency "preservation", mechanistically interpret how neural networks represent agents and their capacities, and describe challenges in the governance of agency-preserving AI systems. More information of our conceptual goals for this hackathon are provided here: https://www.agencyfoundations.ai/hackathon.
Start: Introductory talks - September 8th: 18:00-19:30 CET.
- End: Submission deadline - September 24th night (any timezone).
- Location: Online/Remote
- Topics: (1) mechanistic interpretability; (2) RL/IRL; (3) game theory; (4) conceptual/governance (see here for more details)
- Prizes: US$10,000 ($2,500 in each category)
- Format: online submissions.
More details about specific prizes, categories and additional information to be posted the 1st week of September.
Sign up below to be notified before the kickoff! Read up on the schedule, see instructions for how to participate, and inspiration on the agency foundations website.
Rules
You will participate in teams of 1-5 people and submit a project on the entry submission page (available when the hackathon starts). Each project consists of multiple parts: 1) The PDF report, 2) a maximum 10-minute video overview (optional), 3) title, summary, and descriptions.
You are allowed to think about your project and engage with the starter resources before the hackathon starts but your core research work should happen during the duration of the hackathon.

Schedule
- Friday September 8, 16:00 UTC: Keynote talk by Catalin Mitelut inspire your projects and provide an introduction to the topic. Tim Franzmeyer will present his work on altruistic RL agents. Esben Kran will also give a short overview of the logistics.
- Saturday and Sunday 14:00 UTC: Project discussion sessions on the Discord server.
- Friday September 22nd 14:00 UTC: A discussion and short talk.
- Sunday September 24th night (all time zones): Submission deadline!
Local sites
Agency Foundations Alignment Jam London
RSVP via Facebook or post in "uk" channel on Alignment Jam website for exact details
Event page: Agency Foundations Alignment Jam London (opens in new tab)AI Agency Foundations Hackathon
Research how AI agents work and how to keep them safe over a weekend! Currently only for students at the University of Texas at Dallas. goo.gl/maps/9M9r2q5wNEjYpTba8
Event page: AI Agency Foundations Hackathon (opens in new tab)AI agency hackathon
Join us at Fixed Point in Prague - Vinohrady, Koperníkova 6 for a weekend of research hacking to understand AI agency!
Event page: AI agency hackathon (opens in new tab)Alignment Hackatons
A local space near Lomonosov Moscow State University. Location: Russia, Moscow, Lomonosovskiy Prospekt, 25k3
Event page: Alignment Hackatons (opens in new tab)Global Agency Hackathon
Join everyone internationally on our Discord and get together in teams to solve the most important problems of agency preservation!
Event page: Global Agency Hackathon (opens in new tab)Kraków Jam Site
Contact @matthewbaggins on Discord or bagginsmatthew@gmail.com. Hosted at ul. Celna 6/9, the office of The Optimum Pareto Foundation.
Where a Sprint can lead
How our programs connectAnyone can join
Stand out
6 to 16 weeks on your own project, with a research project manager, compute and publication support.
Upcoming Sprints
All SprintsAI Collusion Research Sprint
A weekend research sprint on collusion between AI agents: when it emerges in markets and everyday workflows, how to detect and audit it, how it is carried, and what breaks it. Co-organized with Poseidon Research and AE Studio, online with in-person hubs at Collider in New York City and AI Safety Hong Kong. Top teams are invited to apply to the Apart Fellowship.
Read the brief: AI Collusion Research SprintAI x Epistemics Research Sprint
A weekend research sprint on AI for epistemics: evaluating whether models know how solid their claims are, building trust infrastructure that people and agents can consume, and shipping epistemic products that improve real decisions. Online, four tracks including an open track. Top teams are invited to apply to the Apart Fellowship.
Read the brief: AI x Epistemics Research SprintQuestions? sprints@apartresearch.com