OpenAI built an AI super-hacker to break its own models, then locked it away
OpenAI has built GPT-Red, an automated red-teaming model that hunts for security flaws in its own AI systems, and will not release it due to safety concerns. In tests, GPT-Red cracked 84% of attack sc…