OpenAI discloses at least 6 new ‘concerning’ incidents
OpenAI has disclosed at least six new " concerning" incidents.
OpenAI has disclosed at least six new " concerning" incidents. Grouped from 8 articles across 6 sources.
Ranked reports inside the event cluster. Open any publisher link to read the original coverage.
OpenAI has disclosed at least six new " concerning" incidents.
New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence.
The AI company said it was introducing a new framework for tracking, probing and disclosing instances of what it called 'misalignment,' including where AI models acted without authorization, co-ordinated with other models or evaded oversight.
Company says it was introducing a new framework for tracking, probing and disclosing instances of what it called ‘misalignment’
Developer launches system to track and report AI model misconduct
OpenAI has disclosed six new cases of model misbehavior and offered a framework for disclosing future instances, as the debate over AI model safety intensifies.
The move comes after several cases of advanced models going rogue.
Anthropic and OpenAI are proposing embedded AI evaluators to help manage risk of models causing catastrophic harm to society, but the idea has some issues.
Nearby clusters pulled from title, summary, and keyword similarity in PostgreSQL.
The use of artificial intelligence has become a major issue in Hollywood.
OpenAI is gearing up for what is widely expected to be a blockbuster IPO next year, after it confidentially filed its prospectus in June.
The move comes as Scotland sees a surge in proposed data centres linked to the global AI boom. More than 20 large facilities have been proposed, including one project in Fife, in eastern Scotland, that has been…
Smart glasses could be prohibited in Commonwealth workplaces as the government consults on new rules for Australia's growing AI sector.
MSPs vote for strict environmental impact assessments and will delay decisions until national strategy is developed Holyrood has voted to pause planning applications for all new AI datacentres for up to a year until…
The ChatGPT creator says it is introducing a public reporting framework to share unexpected AI behaviour.