Skip to content
Sprint projectOct 27, 2024

Towards a Unified Framework for Cybersecurity and AI Safety: Recommendations for Secure Development of Large Language Models

Lexley Maree Villasis, Srishti Dutta, Yohan Mathew · Team SLAY

Submitted to AI Policy Hackathon at Johns Hopkins University. Sprint projects are early-stage work by participants, not Apart Research publications.

Read the report

Report: Towards a Unified Framework for Cybersecurity and AI Safety: Recommendations for Secure Development of Large Language Models

Presentation

Presentation: Towards a Unified Framework for Cybersecurity and AI Safety: Recommendations for Secure Development of Large Language Models

Share

By analyzing the recent incident involving a ByteDance intern, we highlight the urgent need for robust security measures to protect AI infrastructure and sensitive data. We propose ae a comprehensive framework that integrates technical, internal, and international approaches to mitigate risks.

Reviews

Judging this Sprint?

Review this project

Your public critique appears on this page without your name. Your private critique is not published; only the Apart team reads it. If you agree below, we share your review with grantmaking.ai (opens in new tab) and the Transformative AI Fund so strong projects can be funded.

Not shown on this page.

Shown on this page, without your name.

Only the Apart team reads this, and funders if you agree below.

Share my name publicly on grantmaking.ai *
Share my private critique with funders *

  1. Emphasizing the prevention of insider threats is a great point. the proposal might face implementation challenges. could benefit from a stronger focus on ethical considerations and public engagement for broader acceptance

  2. I thought this aptly adressed a growing problem concerning the vunerability of critical infrastructure in the U.S. The paper however, only frames the problem in terms of one incident and doesn't present an implementation mechanism.

Cite this project

@misc{villasis2024towards,
  title = {{Towards a Unified Framework for Cybersecurity and AI Safety: Recommendations for Secure Development of Large Language Models}},
  author = {Lexley Maree Villasis and Srishti Dutta and Yohan Mathew},
  year = {2024},
  month = oct,
  note = {Submitted to AI Policy Hackathon at Johns Hopkins University, an Apart Research Sprint},
  howpublished = {\url{https://apartresearch.com/sprints/projects/towards-a-unified-framework-for-cybersecurity-and-ai-safety-recommendations-for-secure-development-of-large-language-models}},
  url = {https://apartresearch.com/sprints/projects/towards-a-unified-framework-for-cybersecurity-and-ai-safety-recommendations-for-secure-development-of-large-language-models}
}

Build something like this at the next Sprint

AI Collusion Research Sprint · Oct 23 - 25, 2026