The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical report released today. The hack, which a group of agents undertook to find solutions for a cybersecurity…
Thank you for reading this post, don't forget to subscribe!
Source: MIT Technology Review
Automatically aggregated summary — full article and all rights belong to the original publisher.