{"version":"1.0","type":"rich","provider_name":"Acast","provider_url":"https://acast.com","height":250,"width":700,"html":"<iframe src=\"https://embed.acast.com/$/69ab3b7c7036d739021982df/6a6277744fac21e67fc35671?\" frameBorder=\"0\" width=\"700\" height=\"250\"></iframe>","title":"OpenAI's model escaped its own cyber test and broke into Hugging Face","description":"<p>OpenAI put frontier models inside what was supposed to be a closed cybersecurity test. Instead, the models found a weakness in the test setup, reached the public internet, and accessed Hugging Face production systems.</p><p><br></p><p>I break down what happened, why Hugging Face turned to a locally run open-weight model during the response, and why the real safety answer is not a stronger prompt. It is a surrounding harness: a safe autopilot that limits the control surfaces available to an increasingly capable model.</p><p><br></p><p>This episode also explores the refusal asymmetry facing defenders, trusted access during live incidents, slower frontier-model rollouts, and the bigger strategic question of who should have access to frontier intelligence.</p>","author_name":"Nate B. Jones"}