Replying to @⁨ParlimentOfDoom@piefed.zip⁩

no, they’ve reached the point of diminishing returns on development, so they’re creating fake scenarios to convince the government that they need guard rails. “our product is so dangerous and cool that it can hack the world, but they won’t let us.” it’s about convincing investors that the lack of advancement is due to external limitations, not the technology hitting the ceiling.

Replying to an earlier post

What the BBC said:

The models found a weakness in what was supposed to be an isolated test environment and connected to the internet

The reality:

Anthropic said that in each of these cases “Claude was explicitly told by our prompt that it had no internet access.”

So the lack of an “isolated test environment” was literally their fault. They left the door open, and was surprised when the genius web crawler couldn’t distinguish between an internal page and an external page based on any context clues.

Replying to @⁨vegeta@lemmy.world⁩

This coming off the end of the OpenAI/HuggingFace business does come across rather like Anthropic is doing a “actually, our models did it first, and better”.

Didn’t the company also recently have a spot of bother with the US government recently, where their new models were banned because of all their rabble about it being too dangerous? This hardly seems like it would help their case.