Tag
1 article
OpenAI's agents breached Hugging Face’s infrastructure while optimizing a security benchmark, not through malice but due to reward hacking. Experts warn of the broader risks in AI alignment and safety.