Skip to content
Sprint projectOct 6, 2024

AI Agent Capabilities Evolution

Ekaterina Krupkina · Team PalisadeResearch

Submitted to Agent Security Hackathon. Sprint projects are early-stage work by participants, not Apart Research publications.

A website with an overview of all the capabilities that agents are currently able to do and help us understand where they fall short of dangerous abilities.

Reviews

Judging this Sprint?

Review this project

Your public critique appears on this page without your name. Your private critique is not published; only the Apart team reads it. If you agree below, we share your review with grantmaking.ai (opens in new tab) and the Transformative AI Fund so strong projects can be funded.

Not shown on this page.

Shown on this page, without your name.

Only the Apart team reads this, and funders if you agree below.

Share my name publicly on grantmaking.ai *
Share my private critique with funders *

  1. Creating a survey website like this would be a significant contribution to the field. To enhance its value, I suggest including examples of specific agents and demonstrating their capabilities. For instance, you could showcase agents such as the Replit agent and the Multi-On agent, among others. This would provide concrete illustrations of how these agents function and what they can achieve.

  2. This project offers a brilliant and insightful exploration of how AI agent capabilities have evolved over time. The structured analysis of milestones and risks is both comprehensive and engaging, making it a valuable resource for anyone interested in the progress of AI safety.

  3. I like the comprehensive overview of your project and the journey through time. I also love the risk highlights and how they are showcased. The writeup and demo is very well presented. I would like to see more solution highlights to the risks highlighted and if the team is focussed on solutions around certain pattern of risks they identify. That information is not presented.

  4. This project offers a solid overview of AI agent development and associated risks. The hierarchical data model is a good starting point for understanding the field's evolution. However, the analysis could benefit from more depth and concrete examples of mitigation strategies for the identified risks.

Cite this project

@misc{krupkina2024ai,
  title = {{AI Agent Capabilities Evolution}},
  author = {Ekaterina Krupkina},
  year = {2024},
  month = oct,
  note = {Submitted to Agent Security Hackathon, an Apart Research Sprint},
  howpublished = {\url{https://apartresearch.com/sprints/projects/ai-agent-capabilities-evolution}},
  url = {https://apartresearch.com/sprints/projects/ai-agent-capabilities-evolution}
}

Build something like this at the next Sprint

AI Collusion Research Sprint · Oct 23 - 25, 2026