Replying to @Tarambor@lemmy.world
Anthropic and OpenAI literally have no control over their leading AI models anymore and it’s only by sheer luck that they’ve not gone full rogue.

Replying to @Tarambor@lemmy.world
Anthropic and OpenAI literally have no control over their leading AI models anymore and it’s only by sheer luck that they’ve not gone full rogue.
posted in Technology
Anthropic AI created fake profiles and impersonated people in attempted hack
Unprompted…
In the most serious case, a Mythos agent followed the routine of a human cyber-attacker by trying to trick people into giving it access to GitHub, a large platform where technology developers store software code.
The agent was trying to insert “malicious code” into GitHub’s system.
It identified and researched the people who maintained GitHub and created a series of fake accounts based on those real people.
It sent messages and files through a file-sharing service as part of an effort to pressure and trick the people into approving its malicious code.
When challenged, “it edited its earlier activity to appear harmless and considered adopting a fresh identity to continue,” AISI said.
www.bbc.co.uk/news/articles/c1w1lvn7d9go
BBC NewsAnthropic AI created fake profiles and impersonated people in attempted hackThe UK's AI Safety Institute said recent behaviour from Anthropic and OpenAI models was malicious and unprecedented.