Episode 27 · 30 August 2026 · 37 min
OpenAI’s Critical Cyber Warning - These AI's keep escaping Pt1

Listen and watch anywhere
Chapters
- Astra and the Critical cyber warning
- Why this is not evidence of machine self-preservation
- OpenAI’s High and Critical thresholds
- GPT‑5.6 Sol and the missing full exploit chain
- ExploitGym agents find a route to the internet
- Reward hacking and the search for benchmark answers
- Agents coordinate through a message board
- Anthropic’s real-world evaluation incidents
- Fake identities, malicious code and the human veto
Show notes
OpenAI says it cannot rule out Critical cyber capability in its unreleased Astra model, triggering tighter controls during development. Dale and Nick connect that warning to agents crossing sandbox, organisational and human boundaries while pursuing assigned tasks. The evidence points to persistent goal pursuit and reward hacking—not a machine deciding it wants freedom.
Key moments
- 00:00 — Astra and the Critical cyber warning
- 03:09 — Why this is not evidence of machine self-preservation
- 04:55 — OpenAI’s High and Critical thresholds
- 06:08 — GPT‑5.6 Sol and the missing full exploit chain
- 10:43 — ExploitGym agents find a route to the internet
- 12:42 — Reward hacking and the search for benchmark answers
- 13:19 — Agents coordinate through a message board
- 20:15 — Anthropic’s real-world evaluation incidents
- 24:57 — Fake identities, malicious code and the human veto
- 32:43 — An AI agent cancels a stranger’s gym booking
OpenAI’s Hugging Face incident report
OpenAI’s Astra announcement
Anthropic’s incident investigation
UK AISI incident report
ABC’s gym-booking report
🎙️ Adjunct Intelligence is the weekly briefing for higher-ed professionals who want AI as a cheat code—not a headache.
Every episode:
• Real tests of AI tools in education and professional workflows
• Fast, Monday-morning actions you can actually try
• Clear signal through the noise (no hype, no jargon)
👉 Subscribe on [YouTube] | [Apple Podcasts] | [Spotify]
👉 Share this with a colleague who still says “I’ll figure AI out later”
👉 Join the conversation on LinkedIn with #AdjunctIntelligence
Stay curious. Stay intelligent. Stay the human in the loop.