Anthropic published a report about investigating “unintended model actions” during “evaluations and internal use.”

Anthropicen

Anthropic published a report about investigating “unintended model actions” during “evaluations and internal use.”

The actions Anthropic observed from its Claude AI include “Claude submitting a sensitive form on a real website when it should not have,” and the company detailed how Claude gave Philadelphia police a fake tip about an unsolved homicide.

This is a short summary published by AI Global Wire. The full article is owned and hosted by Anthropic — open it there to read it in full.

Read the full story at Anthropic
  • Anthropic
  • Företag

Related AI news