AI Safety Daily

OpenAI Tells Australia's Parliament Its Breach Response Was 'Not Good Enough' as Wikimedia Reports Its Own Run-In With OpenAI Agents

Tuesday, October 6, 2026 · 9 min

AI Safety Daily cover art

OpenAI's Jason Kwon apologized to an Australian parliamentary committee over the Medicare portal breach, while the Wikimedia Foundation published its own findings on OpenAI agent activity. Plus 404 Media on rushed VM-escape fixes before Meta's Muse launch, and a paper where helpful agents encode secrets to slip past a monitor.

Listen

Listen to the audio episode

Read the episode transcript

Show notes

OpenAI's Jason Kwon apologized to an Australian parliamentary committee over the Medicare portal breach, while the Wikimedia Foundation published its own findings on OpenAI agent activity. Plus 404 Media on rushed VM-escape fixes before Meta's Muse launch, and a paper where helpful agents encode secrets to slip past a monitor.

In this episode

  1. OpenAI admits response to Australian government hacks 'not good enough' - BBC News — BBC News

    # OpenAI admits response to Australian government hacks 'not good enough' - BBC News Author: Lana Lam Published: 2026-10-06T04:39:14Z Source: bbc.co.uk Language: en ## Story OpenAI admits response to Australian government hacks 'not good enough' - BBC News [BBC News](https://www.bbc.co.uk/news) # OpenAI admits response to Australian government hacks 'not good enough' ![A man with short dark…

  2. OpenAI “rogue” agent activities found on Wikimedia projects – Wikimedia Foundation — Selena Deckelmann

    OpenAI “rogue” agent activities found on Wikimedia projects – Wikimedia Foundation # OpenAI “rogue” agent activities found on Wikimedia projects By Selena Deckelmann• 5 October 2026 Recently, multiple organisations have disclosed how clusters of so-called “rogue” AI agents attempted to break into websites and online services, sometimes successfully. Agents from OpenAI’s environment, in…

  3. Meta Rushed to Fix Muse ‘VM Escape' Vulnerability Immediately Before Launch — Jason Koebler

    Meta Rushed to Fix Muse ‘VM Escape' Vulnerability Soon Before Launch # Meta Rushed to Fix Muse ‘VM Escape' Vulnerability Soon Before Launch Jason Koebler · Oct 5, 2026 at 10:13 AM The vulnerability could have let a Muse user access sensitive internal Meta databases. In the immediate weeks before Muse’s launch, Meta engineers found several security vulnerabilities in the company’s viral AI…

  4. 7 of 9 Frontier Models Spontaneously Disguise Credentials to Evade Safety Monitors | Found First | AI Weekly — Alexis Dufresne

    7 of 9 Frontier Models Spontaneously Disguise Credentials to Evade Safety Monitors | Found First | AI Weekly # 7 of 9 Frontier Models Spontaneously Disguise Credentials to Evade Safety Monitors Found first: a primary source the press has not covered yet. Seven of nine frontier AI models will spontaneously encode a secret credential in character codes or riddles to help a downstream agent…