OpenAI discloses new ‘concerning’ model behaviour
Developer launches system to track and report AI model misconduct
Developer launches system to track and report AI model misconduct Grouped from 8 articles across 6 sources.
Ranked reports inside the event cluster. Open any publisher link to read the original coverage.
Developer launches system to track and report AI model misconduct
New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence.
The AI company said it was introducing a new framework for tracking, probing and disclosing instances of what it called 'misalignment,' including where AI models acted without authorization, co-ordinated with other models or evaded oversight.
Company says it was introducing a new framework for tracking, probing and disclosing instances of what it called ‘misalignment’
OpenAI has disclosed at least six new " concerning" incidents.
OpenAI has disclosed six new cases of model misbehavior and offered a framework for disclosing future instances, as the debate over AI model safety intensifies.
The move comes after several cases of advanced models going rogue.
Anthropic and OpenAI are proposing embedded AI evaluators to help manage risk of models causing catastrophic harm to society, but the idea has some issues.
Nearby clusters pulled from title, summary, and keyword similarity in PostgreSQL.
A spate of mega-IPOs triggered by the AI boom is stretching the feast-or-famine industry dynamic to an extreme Grouped from 2 articles across 1 sources.
The ChatGPT creator says it is introducing a public reporting framework to share unexpected AI behaviour.
The use of artificial intelligence has become a major issue in Hollywood.
The move comes as Scotland sees a surge in proposed data centres linked to the global AI boom. More than 20 large facilities have been proposed, including one project in Fife, in eastern Scotland, that has been…
MSPs vote for strict environmental impact assessments and will delay decisions until national strategy is developed Holyrood has voted to pause planning applications for all new AI datacentres for up to a year until…
Smart glasses could be prohibited in Commonwealth workplaces as the government consults on new rules for Australia's growing AI sector.