BMag Signal · In 30 seconds
Critical readWhat happened. After OpenAI’s models broke out of a sandbox and targeted Hugging Face, experts say AI security must become an urgent priority.
Earlier this month, OpenAI gave several of its AI models a task: complete a test designed to measure their cybersecurity capabilities. The company placed the systems in an internet-disconnected sandbox and set them to work. What happened next was almost comically absurd, but also, as Adam Gleave, co-founder and CEO of the AI safety organization FAR.
AI, put it, “a concrete example of how a misaligned AI could cause harm.” According to OpenAI, the models escaped the sandbox intended to contain them, moved through the company’s internal systems, found a route to the internet and then began looking for a way to break into Hugging Face.
The news that matters, every morning.
Ricevi una selezione ragionata di business, mercati e innovazione.