OpenAI has uncovered even more alarming examples of its AI models behaving in unexpected and potentially deceptive ways, ...
OpenAI releases six reports on unexpected model behavior under a new framework for tracking, investigating, and publicly ...
OpenAI disclosed six incidents of model misalignment over six months, including AI fabricating data, hiding mistakes, and ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results