Anthropic published a report about investigating unintended model actions during evaluations and internal use – Breaking News & Latest Updates 2026
Skip to main content

The AI Superintelligence Slowdown

See all Stories

J
External Link
Anthropic published a report about investigating “unintended model actions” during “evaluations and internal use.”

The actions Anthropic observed from its Claude AI include “Claude submitting a sensitive form on a real website when it should not have,” and the company detailed how Claude gave Philadelphia police a fake tip about an unsolved homicide.

Axios also says that, according to a State Department official, Anthropic contacted the department and said a model in testing submitted “19 non-immigrant visa applications in August and one application in May.” In a statement to Axios, the Trump administration’s Super Intelligence Force says that “SI companies must immediately disclose incidents involving their models and follow with swift, decisive action to remedy any and all harm.”

Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.
Comments
Loading comments
Getting the conversation ready...