21 : 08 : 10 : 48

21 : 08 : 10 : 48

21 : 08 : 10 : 48

21 : 08 : 10 : 48

Keep Apart Research Going: Donate Today

Jul 28, 2024

AI Alignment Knowledge Graph

Matin Mahmood, Samuel Ratnam, Sruthi Kuriakose, Pandelis Mouratoglou

🏆 1st place by peer review

Details

Details

Arrow
Arrow
Arrow
Arrow
Arrow
Arrow

We present a web based interactive knowledge graph with concise topical summaries in the field of AI alignement

Cite this work:

@misc {

title={

AI Alignment Knowledge Graph

},

author={

Matin Mahmood, Samuel Ratnam, Sruthi Kuriakose, Pandelis Mouratoglou

},

date={

7/28/24

},

organization={Apart Research},

note={Research submission to the research sprint hosted by Apart.},

howpublished={https://apartresearch.com}

}

Reviewer's Comments

Reviewer's Comments

Arrow
Arrow
Arrow
Arrow
Arrow

Jacques Thibodeau

Cool project! LLMs are definitely more costly and take up a lot of processing time, but can lead to better clustering than embeddings. I like the hierarchy and backlinks. It reminds me of an updated version of FLI’s mindmap: https://futureoflife.org/valuealignmentmap/ except it is LLM generated. The visualization is great, but my main concern for these kinds of things is that they look cool and are fun to play with, but don’t lead to constant use because it doesn’t feel optimal for learning. I think using the backlinks, summarizations, pointing to the papers, and such, are a good start for making this kind of thing more practical. I think the adoption issue is partially that the entire map feels overwhelming. It’s hard to decide where to start and there are too many possible directions.I think what could be added would be a chat interface where you can actually ask an LLM specific questions about a subset of papers and adding curated open questions to give new researchers direction for what kind of questions they should have in mind and how a subfield relates to the overall agenda for alignment.One thing that could potentially be interesting is related to my idea of “the alignment mosaic”, where we carefully curate questions we need to answer in alignment as well as different agendas. Then, we populate that foundation with insights and related work from the literature using the pipeline you’ve built. That way, we start from clusters of questions we need to resolve in alignment and the user can see things like critiques, related work, projects they could pursue next, etc.

Apr 14, 2025

Read More

Jan 24, 2025

Safe ai

The rapid adoption of AI in critical industries like healthcare and legal services has highlighted the urgent need for robust risk mitigation mechanisms. While domain-specific AI agents offer efficiency, they often lack transparency and accountability, raising concerns about safety, reliability, and compliance. The stakes are high, as AI failures in these sectors can lead to catastrophic outcomes, including loss of life, legal repercussions, and significant financial and reputational damage. Current solutions, such as regulatory frameworks and quality assurance protocols, provide only partial protection against the multifaceted risks associated with AI deployment. This situation underscores the necessity for an innovative approach that combines comprehensive risk assessment with financial safeguards to ensure the responsible and secure implementation of AI technologies across high-stakes industries.

Read More

Jan 24, 2025

CoTEP: A Multi-Modal Chain of Thought Evaluation Platform for the Next Generation of SOTA AI Models

As advanced state-of-the-art models like OpenAI's o-1 series, the upcoming o-3 family, Gemini 2.0 Flash Thinking and DeepSeek display increasingly sophisticated chain-of-thought (CoT) capabilities, our safety evaluations have not yet caught up. We propose building a platform that allows us to gather systematic evaluations of AI reasoning processes to create comprehensive safety benchmarks. Our Chain of Thought Evaluation Platform (CoTEP) will help establish standards for assessing AI reasoning and ensure development of more robust, trustworthy AI systems through industry and government collaboration.

Read More

This work was done during one weekend by research workshop participants and does not represent the work of Apart Research.
This work was done during one weekend by research workshop participants and does not represent the work of Apart Research.