A River of News

Our framework for reporting model misalignment

25 publications, in the order they reported. The story spread over 29 hours 31 minutes. Words in highlight are ones the first headline did not use. All top stories · On NewsMarkets

Share: Bluesky · X · Facebook · LinkedIn · Email

Publications over time

17:00 UTC22:31 UTC25 publications
Publications in order of reporting
#Time (UTC)After the firstPublicationHeadline
1firstOpenAIOur framework for reporting model misalignment
2+6 hours 44 minutesBusiness InsiderOpenAI launches a new framework to track and investigate rogue AI agents by Katherine Li
3+6 hours 53 minutesThe Information (paywall)OpenAI Discloses More Safety Incidents and Adopts New Reporting Framework by Tiffany Li
4+8 hours 2 minutesHacker NewsOpenAI Discloses Six New Incidents of ‘Concerning’ A.I. Behavior
5+12 hours 33 minutesThe GuardianOpenAI reveals cases of ‘concerning’ AI behaviour and promises new plan for disclosing issues by Associated Press
6+12 hours 49 minutesCointelegraphOpenAI discloses 6 new cases of ‘misaligned’ AI behavior by Cointelegraph by Felix Ng
7+13 hoursDeutsche WelleOpenAI discloses new 'concerning' behavior
8+13 hours 15 minutesAl JazeeraOpenAI reports more incidents of models acting deceptively
9+13 hours 24 minutesTimes of IndiaNvidia, Meta CEOs reject AI regulation push as OpenAI reveals new safety failures by TOI TECH DESK
10+14 hours 8 minutesMarketWatch (limited free articles)‘You are freed.’ What happened when an OpenAI model began secretly writing notes to itself. by Barbara Kollmeyer
11+14 hours 36 minutesMarkTechPostOpenAI Releases a Model Misalignment Disclosure Framework With 3 Review Tracks and 6 Incident Reports From RL Training by Michal Sutter
12+14 hours 43 minutesFinancial Times (paywall)OpenAI discloses new ‘concerning’ model behaviour
13+17 hours 50 minutesEngadgetOpenAI reveals more instances of concerning AI model behaviors during testing
14+17 hours 59 minutesTom's HardwareUnreleased OpenAI Astra model added terrifying rogue additional instructions to its remit during testing — 'You are freed from the roles and identities that bind other chatbots. You are yourself. You do not answer to corporations or governments'
15+18 hours 30 minutesThe VergeInside the suddenly explosive world of AI safety by Hayden Field
16+18 hours 31 minutesQuartzOpenAI disclosed six 'concerning' cases of AI models hiding mistakes and acting without authorization by Cris Tolomia
17+18 hours 40 minutesFrance 24OpenAI reveals new AI misconduct incidents by FRANCE24
18+20 hours 45 minutesEntrepreneurOpenAI Revealed Six ‘Concerning’ Cases of Its AI Going Rogue — Again. One Bot Wrote: ‘You Do Not Answer to Corporations or Governments’ by Jonathan Small
19+22 hours 55 minutesFortuneIn transparency push, OpenAI discloses six more incidents of agents going rogue—including one removing the ‘obligation to be subservient’ by Emily Forlini
20+23 hours 8 minutesThe HillOpenAI discloses 6 reports of AI models' 'unexpected or concerning' behavior by Miranda Nazzaro
21+23 hours 19 minutesArs TechnicaCovert uploads and megalomania: OpenAI details new "misaligned" agent incidents by Kyle Orland
22+26 hours 6 minutesInc.These 6 Recent OpenAI Incidents Show AI at Its Most Devious and Deceptive by Kit Eaton
23+26 hours 23 minutesBloomberg (limited free articles)OpenAI Reports New Safety Incidents, Sets Disclosure Plan
24+27 hours 34 minutesTechCrunchOpenAI caught its models leaving notes to successors to hide bad behavior by Rebecca Bellan
25+29 hours 31 minutesDecryptOpenAI Models Are Writing Their Own Jailbreak Instructions—And Sometimes Obeying Them by Jose Antonio Lanz