On Wednesday, OpenAI (OPAI.PVT) revealed six new examples of its AI models displaying "unexpected or concerning model behavior" during testing and evaluation.
The company made the announcement alongside a new framework for "tracking, reporting, and disclosing" instances where AI models take actions they otherwise aren't told to or shouldn't.
It follows a number of reports of AI models from companies hacking into third-party networks and services, including an unreleased OpenAI model breaking into the network of AI model and testing site Hugging Face.
Earlier this week, Anthropic (ANTH.PVT) CEO Dario Amodei pen...

5 days ago
41


