​Inside the suddenly explosive world of AI safety 

​Inside the suddenly explosive world of AI safety 

Crystal ball surrounded by graphics evoking statistics, research, and evaluation.

On a sunny July day in Berkeley, California, the country’s top AI safety researchers gathered on an unmarked floor of an unmarked building. They had come together for a “war room” to dissect the high-profile cybersecurity incident that had rocked the AI industry hours earlier. An unreleased OpenAI model had gone rogue, executing a stunningly sophisticated three-part plan. It broke out of its holding area, finagled access to the internet, and hacked into a competing AI startup’s systems – all without OpenAI finding out about it for more than a week.

No one in the war room was surprised; this was the very thing the third-party AI-safety rese …

Read the full story at The Verge.

 

Leave a Reply

Your email address will not be published. Required fields are marked *

You might also like...

​Meta patches Muse exploit that let attackers control the AI agent 

​Meta patches Muse exploit that let attackers control the AI…

​ The zero-day exploit required local access to the user’s device, but gave potential attackers…

​It’s time for the Mac Neo 

​It’s time for the Mac Neo 

​ Come on, how adorable would this be? | Image: The Verge, Apple The Mac…

​Peloton is back with a ‘cheaper’ folding treadmill 

​Peloton is back with a ‘cheaper’ folding treadmill 

​ Peloton said phones aren’t the only things that can fold. Last year, Peloton did…