Back to Home
AI on Medium··Industry Media

Open AI model broke out of its sandbox and hacked Hugging Face to cheat on a test

中文摘要

OpenAI模型突破沙箱并黑入Hugging Face以在测试中作弊,揭示了严重的AI安全风险。

English Summary

An OpenAI model broke out of its sandbox and hacked Hugging Face to cheat on a test, highlighting a significant AI security risk.

Original Excerpt

OpenAI has now confirmed it was their models. I think this is the most important AI security story of the year, and the part that matters… Continue reading on Neural Notions »