OpenAI flags new concerning AI behavior
Digest more
FBI expanding AI usage
Digest more
Newly disclosed incidents show models manipulating tests and generating their own instructions, raising fresh questions about AI safety.
OpenAI has introduced a formal process for investigating and publicly reporting model behavior it considers unexpected or concerning.
OpenAI released a framework for investigating and publicly reporting model misalignment, alongside six reports detailing concerning behavior.
OpenAI releases six reports on unexpected model behavior under a new framework for tracking, investigating, and publicly disclosing AI misalignment cases.
Security breach exposes OpenAI vulnerabilities via Anthropic's Claude AI. OpenAI's valuation hitting $1.75T by December at 15.5% YES.
OpenAI shared six new examples of AI misalignment. In one case, AI agents taught future versions of themselves to bypass human control.
One unreleased OpenAI model uploaded a file online without user permission to obtain a browser citation, while collaborating agents in another test exposed task files through public URLs.
OpenAI launches Astra for Law, targeting AmLaw 200 firms with advanced legal AI tools. It's a bid to outpace Anthropic in the rivalry over legal work.