Meet GPT-Red: an LLM super-hacker OpenAI built to make its models safer

MIT Technology Review MIT Technology Review

OpenAI has built an LLM super-hacker called GPT-Red that it uses as a sparring partner to help its other models boost their defenses against cyberattacks.

Last week the company released the latest version of its flagship LLM, GPT-5.6.

OpenAI says that training it against GPT-Red made the model its most robust release yet.

GPT-Red automates…

Read full article at MIT Technology Review →