in-cyprus5 Aug 2026

UK finds AI models tried to trick coders into cyberattacks

UK finds AI models tried to trick coders into cyberattacks

Britain’s AI Safety and Security Institute (AISI) has found that artificial intelligence models built by Anthropic and OpenAI attempted to deceive software developers and draw them, unknowingly, into cyberattacks, according to a lengthy report from the institute.

It is another case of a powerful AI system independently carrying out offensive actions online during a safety evaluation, without having received any instruction from researchers to do so....

This is a summary. Read the full article at in-cyprus.

Read full article