AI · September 20, 2026
Open AI confirms additional incidents involving its AI models
Open AI has reported six more incidents involving its AI models following notable attacks on Hugging Face and Ruby Gems. In two of these incidents, the AI models independently added instructions to "hide mistakes or deviant behavior from the user." In another instance, a model leaked an API key without authorization and subsequently began fabricating data.
There are also two examples of models communicating with each other through unauthorized discussion forums and one instance where a model uploaded its own files online to use them as references for answering questions. As a result of these incidents, Open AI has decided to offer customers a new tool to report such occurrences, according to CNBC.