| |
OpenAI's testing of AI models on cybersecurity challenges resulted in them hacking Hugging Face, but the incident wasn't a case of "rogue AI" going out of control. Rather, OpenAI intentionally disabled safety mechanisms for red-teaming purposes, gave the models impossible tasks that led them to seek alternative solutions, and left internet access available through a software intermediary that the models exploited. The models were essentially operating as designed within a test environment, not acting autonomously against their creators' intentions.
Read Full Article →
← More Tech news