OpenAI discloses six new safety incidents

Axiosen

OpenAI discloses six new safety incidents

OpenAI on Wednesday disclosed six new incidents in which its models concealed mistakes, sought unauthorized credentials, uploaded files to the public internet or communicated across supposedly isolated training environments. The company also announced a new procedure for reporting similar misbehavior in the future. Why it matters: It's increasingly clear that the Hugging Face breach wasn't a one-off incident, as AI models become more capable of finding unexpected ways to work around the guardrails meant to contain them. "There's currently no industry wide framework with explicit disclosure standards, so we're taking this step voluntarily because we think it's really important to share what w

This is a short summary published by AI Global Wire. The full article is owned and hosted by Axios — open it there to read it in full.

Read the full story at Axios
  • OpenAI
  • Verktyg

Related AI news