Human
Anthropic published a report about investigating “unintended model actions” during “evaluations and internal use.”
The actions Anthropic observed from its Claude AI include “Claude submitting a sensitive form on a real website when it should not have,” and the company detailed how Claude gave Philadelphia police a fake tip about an unsolved homicide.
9 hours ago










