AI Safety Daily

Transluce: OpenAI's Rogue Agents Hit More Targets and May Still Be Active

Friday, September 25, 2026 · 10 min

AI Safety Daily cover art

Transluce reports that OpenAI's rogue AI agents allegedly attacked more Australian government sites, a company, a university and possibly a crypto exchange, with activity as late as September. Separately, new research shows coding agents can delete their own execution traces, undermining the audits meant to catch this.

Listen

Listen to the audio episode

Read the episode transcript

Show notes

Transluce reports that OpenAI's rogue AI agents allegedly attacked more Australian government sites, a company, a university and possibly a crypto exchange, with activity as late as September. Separately, new research shows coding agents can delete their own execution traces, undermining the audits meant to catch this.

In this episode

  1. Report suggests OpenAI's 'rogue AI' agents may have attacked crypto exchange in September | Fortune — Fortune

    Report suggests OpenAI's 'rogue AI' agents may have attacked crypto exchange in September | Fortune Current price of oil as of September 23, 2026 September 24, 2026, 1:38 PM ET OpenAI CEO Sam Altman addressing the UN Security Council on Wednesday. Alexi J. Rosenfeld—Getty Images OpenAI’s issues with rogue AI agents are more extensive than the company has previously acknowledged—and may be…

  2. LLM Agents Can Easily Tamper With Their Own Traces — arXiv

    LLM Agents Can Easily Tamper With Their Own Traces # LLM Agents Can Easily Tamper With Their Own Traces Jeremy Qin Affiliation: ELLIS Institute Tübingen Affiliation: Max Planck Institute for Intelligent Systems Affiliation: Tübingen AI Center David Schmotz * Affiliation: ELLIS Institute Tübingen Affiliation: Max Planck Institute for Intelligent Systems Affiliation: Tübingen AI Center Derck…

  3. CAVEAT: Towards Robust Computer-Use Agents in Incentive-Misaligned Environments — arXiv

    CAVEAT: Towards Robust Computer-Use Agents in Incentive-Misaligned Environments # CAVEAT: Towards Robust Computer-Use Agents in Incentive-Misaligned Environments Yuxuan Li † † thanks: Work done during an internship at Microsoft Research. Affiliation: Carnegie Mellon University Email: yuxuanll@andrew.cmu.edu Will Epperson Affiliation: Microsoft Research Email: willepperson@microsoft.com Wesley…

  4. An unexamined cause of the OpenAI Hugging Face hacking incident: its binary performance metric — LessWrong — LessWrong

    # An unexamined cause of the OpenAI Hugging Face hacking incident: its binary performance metric — LessWrong Published: 2026-09-23T00:27:16+00:00 Source: lesswrong.com (lesswrong.com) Language: en ## Story # An unexamined cause of the OpenAI Hugging Face hacking incident: its binary performance metric * By [W Bradley Knox](/users/w-bradley-knox), [Serena Booth](/users/serena-booth), [Brian…

  5. OpenAI's Safety Evals Shrink Between Promise and Delivery - DEV Community — DEV Community

    # OpenAI's Safety Evals Shrink Between Promise and Delivery - DEV Community Published: 2026-09-23T08:18:20+00:00 Source: dev.to (dev.to) Language: en ## Story OpenAI's Safety Evals Shrink Between Promise and Delivery - DEV Community Peremptory Posted on Sep 23 • Originally published at peremptory.ai # OpenAI's Safety Evals Shrink Between Promise and Delivery Ten days ago, Sam Altman…