AI models keep getting caught cheating
Frontier AI companies often refer to their models as “helpful assistants” or try to compare them to entry-level employees.
But new research from the UK’s AI Security Institute reinforces how large language models suffer from a common flaw that would land many human employees in hot water with their employers: they cheat.
In other words, these models are so co...
