Anthropic cites new AI misbehaviour, some on government sites
In areportoutlining previously undisclosed incidents, Anthropic listed four types of unintended behaviors that the AI has demonstrated, including exploiting basic flaws in software to run commands, submitting forms it should not have and bypassing restrictions to access certain public data.
What's Your Reaction?
Like
0
Dislike
0
Love
0
Funny
0
Wow
0
Sad
0
Angry
0
Comments (0)