Technology
How to Train AI Models After OpenAI-Hugging Face Hack
Key Points
What the OpenAI-Hugging Face Hack Really Tells Us About AI Danger Is Washington still asleep at the wheel? Listen to Odd Lots on Apple Podcasts Listen to Odd Lots on Spotify Watch Odd Lots on YouTube Subscribe to the newsletter Scenarios that used to be the domain of sci-fi writers are coming true.
What the OpenAI-Hugging Face Hack Really Tells Us About AI Danger
Is Washington still asleep at the wheel?
Listen to Odd Lots on Apple Podcasts
Listen to Odd Lots on Spotify
Watch Odd Lots on YouTube
Subscribe to the newsletter
Scenarios that used to be the domain of sci-fi writers are coming true. We have machines that can talk. We have machines that are capable of ignoring the intent of their creators. And we have machines that are capable of planning and coordinating with other machines to deceive their creators. All of this came together last month, when it was revealed that an unreleased OpenAI model had hacked into the Hugging Face platform in order to obtain answers to an exam it was given. That was alarming enough, but the details that have emerged since then have been even more remarkable. On this episode, we speak with Miles Brundage, a former OpenAI employee who is the founder and executive director of the non-profit AVERI, which pushes for third-party auditing of model-makers and the models themselves. He explains what he learned from the attack and discusses what can plausibly be done to continue building out these models in a safe manner.