Meet GPT-Red: an LLM super-hacker OpenAI built to make its models safer

OpenAI has built an LLM super-hacker called GPT-Red that it uses as a sparring partner to help its other models boost their defenses against cyberattacks. Last week the company released the latest version of its flagship LLM, GPT-5.6. OpenAI says that training it against GPT-Red…

Thank you for reading this post, don't forget to subscribe!

Source: MIT Technology Review

Automatically aggregated summary — full article and all rights belong to the original publisher.

Leave a Comment