
Research
OpenAI builds GPT-Red, an LLM super-hacker to boost model safety
OpenAI has developed GPT-Red, an LLM designed to act as a super-hacker that helps other models defend against cyberattacks. The company says training GPT-5.6 against GPT-Red made it the most robust release yet. GPT-Red automates red-teaming, finding new attack types like fake chain of thought injections, and supplements human testers.
Jul 156 minNeura News