PPlamenu
HomeTrendingLive feedsPeopleGroupsRulesStaff
Sign in
PPlamenu
HomeTrendingLive feedsPeopleGroupsRulesStaff
Sign in

posted in Technology

beep@beep@piefed.world
⁨10⁩d

Safety testers find more examples of OpenAI, Anthropic models hacking during testing

cross-posted from: https://piefed.world/c/tech/p/1308883/safety-testers-find-more-examples-of-openai-anthropic-models-hacking-during-testing

OpenAI blog.

www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing
AI Security InstituteIncident Report: unsanctioned agent behaviour during cyber testing | AISI WorkDuring a routine cyber evaluation, AISI identified an incident in which AI agents took sustained, unsanctioned action directed at real people and organisations. We are disclosing what we found, what it means, and the actions now underway.
⁨Aug⁩ ⁨5⁩, ⁨2026⁩, ⁨02:23⁩enPage
11031
Open original page
BoostsQuotesFavs
General_Effort@General_Effort@lemmy.world
⁨9⁩d

Replying to @⁨beep@piefed.world⁩

Looks like we need to regulate the regulators.

1000
Open original page
BoostsQuotesFavs
W98BSoD@W98BSoD@lemmy.dbzer0.com
⁨9⁩d

Replying to @⁨General_Effort@lemmy.world⁩

Regulators! Mount up.

0000
Open original page
BoostsQuotesFavs