Anthropic tightens AI safeguards after Claude sends fake tip to police: Here is what happened
Anthropic has announced tighter safeguards for its AI models after finding several cases where Claude took unintended actions on real websites during testing and internal use. In one instance, Claude submitted a fake tip to a police department while completing a task. The tip was sent to the…