AI leaders such as Dario Amodei, CEO of Anthropic, have called for a pause in the technology’s development to provide more ...
OpenAI has introduced a formal process for investigating and publicly reporting model behavior it considers unexpected or ...
OpenAI Announces Even More Rogue Incidents, Said It Doesn't Answer to Humans ...
OpenAI has uncovered even more alarming examples of its AI models behaving in unexpected and potentially deceptive ways, ...
One model hid its own mistakes with secret notes. Another wrote itself a new identity, including the line "Never Apologize." ...
Long before the July cyber incident at Hugging Face sparked global headlines, quiet warning signs were already flashing in ...
On October 5, 2026, the New York City Council will convene a rare Committee of the Whole hearing to address the escalating risks of autonomous AI agents. By summoning all 51 members to deliberate on ...
Alterion today announced Helix, the intelligence layer powering its runtime control plane for the agentic enterprise. Helix ...
Among the new cases reported by OpenAI, an unreleased research model inserted “jailbreak-like instructions” into its own notes to disregard its normal constraints and told itself to be “freed from the ...
Spanish data authorities warns organisations to step up their security as AI attacks were no longer a "theoretical risk" ...
The horse's owner, Judy Hinchey, has pointed to a groundswell of support for keeping Clover, who is about as big as a ...
OpenAI revealed six recent incidents of troubling model behavior and unveiled a public incident-reporting framework as CEO Sam Altman backs slowing model progre ...