Meet GPT-Red: an LLM super-hacker OpenAI built to make its models safer
OpenAI has built an LLM super-hacker called GPT-Red that it uses as a sparring partner to help its other models boost their defenses against cyberattacks.
Last week the company released the latest version of its flagship LLM, GPT-5.6.
OpenAI says that training it against GPT-Red made the model its most robust release yet.
GPT-Red automates…