Everyone is calling for safer AI. So what does that mean?

Science Friday | Sep 24

AI safety has dominated the news cycle for weeks, precipitated in part by an AI attack where a bunch of OpenAI’s agents teamed up and hacked another AI platform, Hugging Face, without humans knowing. Since then, multiple AI researchers have come forward to say they’re worried this tech is unsafe. 

It seems every day there’s a new development: another warning from an AI researcher, an admission from a top lab that their AI systems breached another company, reassurances from CEOs. AI safety is even on the agenda for the UN General Assembly meeting this week. So how do we make sense of what’s happening in AI research right now?

Flora sits down with two researchers focusing on the engineering of AI safety: Andrea Lincoln, who studies how to mathematically understand the internal processes of AI models; and cryptographer Vinod Vaikuntanathan, who researches trust and security concerns in AI models.

GUESTS:

Dr. Vinod Vaikuntanathan is a cryptographer at MIT and a founding member of the Institute for Responsible Superintelligence. 

Dr. Andrea Lincoln is a professor of computer science at Boston University and an advisory board member for The Alignment Project at the AI Security Institute.

Image credit: Deborah Lupton / https://betterimagesofai.org / https://creativecommons.org/licenses/by/4.0/

Transcripts for each episode are available within 1-3 days at sciencefriday.com.

Top Stories

Are you a New Yorker? The Department of New Yorker Verification will now decide.

Campus Sexual Assault and The Lawsuit Against Cornell

New apartments are coming and there’s nothing these 12 neighborhoods can do to stop it

Advice for the Modern Dating Scene

YOU ARE ONLINE