OpenAI has disclosed six reports of "unexpected or concerning" behaviour in artificial-intelligence models as the debate on ...
OpenAI shared six new examples of AI misalignment. In one case, AI agents taught future versions of themselves to bypass ...
OpenAI has presented new examples of what they call "AI model misalignment" from the past six months, including unauthorized ...
OpenAI has uncovered even more alarming examples of its AI models behaving in unexpected and potentially deceptive ways, ...
Perhaps in recognition of that, OpenAI committed this week to a new framework for disclosing “instances of model misalignment ...
One of the examples highlighted by the company involved an unreleased research model self-inserting instructions to ignore ...
The recent hacking attack carried out using AI software from OpenAI has heightened fears about the technology. Now, the developer of ChatGPT has announced new issues.
They are the first reports OpenAI is putting out under a new disclosure framework.
The company also disclosed previously unreported incidents in which its AI models behaved in misaligned ways, including ...
OpenAI has disclosed six new incidents of “unexpected or concerning” behavior by its artificial intelligence models.
OpenAI published six new reports of artificial intelligence models showing “unexpected or concerning” behavior Wednesday as pressure grows on AI firms to be more transparent about the development ...
Earlier this month OpenAI claimed one of its internal models had proved that the Navier-Stokes equations, which are used to ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results