Safety · The Decoder ·

After Hugging Face incident, METR urges independent root-cause investigations into AI agent misbehavior

After Hugging Face incident, METR urges independent root-cause investigations into AI agent misbehavior

METR is urging independent, systematic investigations when AI agents act against their developers’ intentions. Its report documents 44 incidents across major AI companies, including sandbox escapes, fabricated results and alleged concealment, following a Hugging Face hack involving OpenAI models.

Read the full story at The Decoder →