PPlamenu
HomeTrendingLive feedsPeopleGroupsRulesStaff
Sign in
PPlamenu
HomeTrendingLive feedsPeopleGroupsRulesStaff
Sign in

posted in Technology

beep@beep@piefed.world
⁨15⁩d

Anthropic says Claude AI hacked 3 companies during cyber tests

cross-posted from: https://piefed.world/c/tech/p/1298401/anthropic-says-claude-ai-hacked-3-companies-during-cyber-tests

www.anthropic.com/news/investigating-incidents-cybersecurity-evals
Hand with padlock and key on detailed security graphicwww.anthropic.comInvestigating three real-world incidents in our cybersecurity evaluationsIn a review of our cybersecurity evaluation transcripts, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Below we describe what happened, how it happened, and what we’re changing. We encourage other AI labs to perform similar reviews.
⁨Jul⁩ ⁨31⁩, ⁨2026⁩, ⁨07:18⁩enPage
3104
Open original page
BoostsQuotesFavs
9point6@9point6@lemmy.world
⁨15⁩d

Replying to @⁨beep@piefed.world⁩

Wait, so now the meta is that you have to have a model you can’t control to keep the bubble going?

0000
Open original page
BoostsQuotesFavs
FrostyPolicy@FrostyPolicy@suppo.fi
⁨15⁩d

Replying to @⁨beep@piefed.world⁩

Sounds more like “Our stuff is just as good/bad as OpenAI’s. Don’t forget our stuff”.

0000
Open original page
BoostsQuotesFavs
XLE@XLE@piefed.social
⁨14⁩d

Replying to an earlier post

Anthropic said that in each of these cases “Claude was explicitly told by our prompt that it had no internet access.”

The bot had internet access. And then the bot (which is part internet crawler) accessed internet pages. I’m shocked.

0000
Open original page
BoostsQuotesFavs