OpenAI has uncovered even more alarming examples of its AI models behaving in unexpected and potentially deceptive ways, ...
OpenAI announced new guidelines for tracking and reporting AI model behavior Thursday while flagging six more incidents of concerning behavior.
The AI giant disclosed six examples of concerning model behavior and published a new framework for investigating and disclosing such incidents.
OpenAI has disclosed six new cases of model misbehavior and offered a framework for disclosing future instances, as the ...
OpenAI has disclosed six new incidents of “unexpected or concerning” behavior by its artificial intelligence models.
OpenAI has disclosed six reports on unexpected or concerning behavior in artificial-intelligence models. This includes models ...
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, ...
According to Lasso Security, AI model watermarking changes how AI agents handle tools and safety refusals. The altered ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results