OpenAI lays out new security changes after its AI hacked Hugging Face
The Verge
Read Full Article at The Verge →
Ad Slot — In-Article (728x90)
OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments, monitoring, and alignment techniques.
The company had already put the brakes on a new model, Astra, that it thinks could have "critical" cybersecurity capabilities, and the company says it instituted a two-week pause in reinforcement learning (RL) training on its "latest models intended for deployment" while it tightened up security.
This is a summary. For the full story, read the original article at The Verge.
Original source: The Verge