67 episodes
- This video is hard to summarise. A cracked cipher, an OpenAI security warning, Gemini 4 Argon, RSI paper (co-authored by a who’s who of AI), Lab White House commitments, new hacks emerging, ‘deep personas’, biology Kasparov competitions, and so much more, ending with an epic Opus outro.
Patreon Exclusives: https://www.patreon.com/AIExplained
Chapters:
00:00 - Introduction
01:32 - Wrong about Opus 5.5? Deciphering 16th Century Text
05:01 - Why the models keep breaking out
11:40 - What the models aren't telling us
15:36 - Gemini 4 and the race to release
19:19 - What happens when AI improves AI?
28:31 - Biology, consciousness, and what we still don't understand
Joe Darrow: Not Just the Sandbox: https://x.com/joedaroo/status/2104335929293127851
GPT-6.1 Sol System Card: https://cdn.openai.com/pdf/38e3efcf-545e-44cd-99ec-2b7eb395f4cc/oai_GPT_6_1_Sol.pdf
Intelligence Explosion Paper: https://casp.ac/__l5e/assets-v1/5efd4b41-deb5-4513-a0a3-b4f82d2b79ea/intelligence-explosion.pdf
OpenAI Research Acceleration: https://openai.com/index/research-acceleration-view-inside-openai/
OpenAI Training Safety Cases: https://openai.com/index/towards-safety-cases-for-frontier-ai-training/
Catherine de Medicis Cipher: https://cryptiana.web.fc2.com/code/henryiii.htm
Proposed du Croc Decipherment: https://claude.ai/artifact/1W7B3WxkTAEGzfv3TaKXb4
Rogue Agents Investigation: https://asymmetricsecurity.com/newsroom/rogue-agents-investigation/
OpenAI Shelves GPT-6.1 Astra: https://www.reuters.com/business/openai-shelves-new-ai-model-after-internal-safety-tests-wsj-reports-2026-09-28/
The Case for Reasoning Transparency: https://institute.deepmind.com/essays/the-case-for-reasoning-transparency/
Gemini 4 Argon: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/
OpenAI–Anthropic Rivalry: https://www.theatlantic.com/technology/2026/09/openai-v-anthropic-inside-biggest-rivalry-tech/688819/
NYT: OpenAI Security Warnings: https://www.nytimes.com/2026/09/29/technology/openai-warnings-security.html
NYT: Claude’s Morals: https://www.nytimes.com/2026/09/29/us/anthropic-claude-morals-ai.html
Jasmine Wang on RSI: https://x.com/j_asminewang/status/2097840245786157432
OpenAI Departures Roundup: https://x.com/Bayesian0_0/status/2105680470566686805
White House AI Commitments: https://x.com/Danmar_here/status/2105168138392183146
Sarah Heck on Safety: https://x.com/SarahKHeck/status/2105058513370448280
Sam Altman on Alignment: https://x.com/tbpn/status/2105028992843833459
Sam Altman on Agent Logs: https://x.com/sama/status/2103567198690349362
Micah Carroll: Misalignment Reports: https://x.com/MicahCarroll/status/2103665811051397256
Zuxin Liu on the Incident: https://x.com/LiuZuxin/status/2103699462648639645
Deepa Seetharaman: User Images: https://x.com/dseetharaman/status/2103585482793943203
OpenAI Revenue Chart: https://x.com/PaulBonnet/status/2105288259324567884/photo/1
Nvidia Agent Safety Platform: https://edition.cnn.com/2026/09/28/business/nvidia-ai-safety-system
IntegrityBench: https://integrity-bench.com
Neel Nanda on Interpretability: https://x.com/PalisadeAI/status/2104949061325652001
Biology Contest: Humans and AI: https://www.theinformation.com/articles/inside-drama-behind-biology-contest-pits-openai-agents-humans
Pushmeet Kohli: SynthID Bio: https://x.com/pushmeet/status/2105314763148321102
Ataraxos and Stratego: https://x.com/ssokota/status/2105362040328159526
Benign Data and Hidden Personas: https://x.com/OwainEvans_UK/status/1999172949975392417
Anthropic: Introspection: https://www.anthropic.com/research/introspection
Claude Cheating Results: https://x.com/lukaspet/status/2104634759339298930
Roon on Mathematics and Learning Theory: https://x.com/tszzl/status/2105619006488993898
GPT-4 Research: https://openai.com/index/gpt-4-research/
I.J. Good: Ultraintelligent Machine: https://incompleteideas.net/papers/Good65ultraintelligent.pdf
Terence Tao’s 2024 Interview: https://www.scientificamerican.com/article/ai-will-become-mathematicians-co-pilot/
Hugging Face Incident: https://openai.com/index/hugging-face-incident-and-the-road-ahead/
Claude and Suno Music Video: https://x.com/sevdeawesome/status/2104985610012504181
Podcast: https://aiexplainedopodcast.buzzsprout.com/ - Not only is Opus 5.5 out, pushing the frontier of AI, it also tells us much about what is going on inside the labs, as they both warn about, and promise, Recursive Self Improvement, aka automated AI research. I cover the model (digging into its 230 page paper), labs’ mixed record on promises, why following what is happening in AI is getting almost impossible, and just so much more that even a summary in this description would get too long.
AI Insiders ($9!): https://www.patreon.com/AIExplained
Chapters:
00:00 - Introduction
02:20 - Opus 5.5 and why it came so soon
07:52 - The RSI goalposts keep moving?
15:41 - Can we actually test these models?
23:50 - where this is heading, options
Introducing Claude Opus 5.5: https://www.anthropic.com/claude-opus-5-5
Claude Opus 5.5 System Card: https://www-cdn.anthropic.com/fc1b44717c85dc068bc6ba5024219938094694bd/Claude%20Opus%205.5%20System%20Card.pdf
Anthropic: Measurements for understanding the pace of AI development inside frontier labs: https://www.anthropic.com/institute/measuring-pace-of-ai-development
Anthropic Responsible Scaling Policy, July 2026, version 3.4: https://www-cdn.anthropic.com/files/4zrzovbb/website/0bacdc8440ea96e62a8766d99ebe1d4eea6d5f3a.pdf
Anthropic Responsible Scaling Policy, October 2024: https://www-cdn.anthropic.com/616dee633636e5bd309cb73aed8622e80fe47839.pdf
Anthropic Responsible Scaling Policy — March 2025, version 2.1: https://www-cdn.anthropic.com/17310f6d70ae5627f55313ed067afc1a762a4068.pdf
Noam Brown, Agent swarms and recursive self-improvement: https://www.youtube.com/watch?v=6AgOfiZOWiY
OpenAI: Building standards for the next phase of AI: https://openai.com/index/building-standards-next-phase-ai/
Jakub Pachocki: An Alien Mind: https://openai.com/index/an-alien-mind/
HLE-Diamond, Humanity’s Last Exam: https://lastexam.ai/blog/hle-diamond
Google, OpenAI and Anthropic AI Safety, The Information: https://www.theinformation.com/articles/google-openai-anthropic-ai-safety-group-takes-shape
Lawrence Chan on AI agents attempting cryptocurrency trades: https://x.com/justanotherlaw/status/2103032173708337188
Australian government incident timeline: https://x.com/ShakeelHashim/status/2103108058779922577/photo/1
Anthropic’s core/old views on AI safety: https://www.anthropic.com/news/core-views-on-ai-safety
Dario Amodei: The Urgency of Interpretability: https://darioamodei.com/post/the-urgency-of-interpretability
Massive AI-Fueled Hack Hit 100 Companies in Days, Forbes: https://www.forbes.com/sites/thomasbrewster/2026/09/22/huge-cyberattack-uses-anthropic-and-deepseek-ai-to-target-100-companies/
DrivingBench https://x.com/DrivingBench/status/2102110605448737268
Jay Chooi: GPT-6 Astra and MolmoAct2 robotics comparison: https://x.com/chooi_jeq/status/2098427488787730636
ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models: https://arxiv.org/pdf/2502.09696
Archie Hall on AI and short-term superforecasting: https://x.com/ArchieHall/status/2100914560580337897
Keller Jordan, AI research and reinforcement learning: https://x.com/kellerjordan0/status/2102956913545936959
Daniel Liu, recursive self-improvement: https://x.com/daniel_c0deb0t/status/2102628246408036654
Altman and Amodei, UN Security Council: https://www.theguardian.com/world/2026/sep/23/unga-sam-altman-dario-amodei
Rehan Sheikh: Interactive YT podcast demo: https://x.com/rehan_shei/status/2102835377426034734
https://simple-bench.com/
AI Explained: The State of AI — interactive diagram: https://claude.ai/artifact/LQHr9WgvQZ6cMdH6fkiQpi
AI Explained: Shards of Aether: https://ai-explained.itch.io/shards-of-aether
roon on the pace of cultural change: https://x.com/tszzl/status/2101462171410677962
Sam Altman on AI-risk: https://www.youtube.com/watch?v=YE5adUeTe_I
Jensen Huang’s AI-control remarks: https://x.com/_NathanCalvin/status/2102756997649355231/photo/2
s1r1us on an upcoming vulnerability disclosure: https://x.com/S1r1u5_/status/2102467878423592969
Jake Adler on biodefense infrastructure: https://x.com/jakeadler/status/2102812430808363433
Derek Thompson: Why we’re wrong about China and AI: https://x.com/DKThomp/status/2103120690370982017/photo/1
Sanders and Casar introduce the Ban Artificial Superintelligence Act: https://www.sanders.senate.gov/press-releases/news-sanders-casar-introduce-legislation-to-create-new-federal-agency-to-ban-artificial-superintelligence-pause-advanced-ai-development/
Non-hype Newsletter: https://signaltonoise.beehiiv.com/
Podcast: https://aiexplainedopodcast.buzzsprout.com/ - Why has it been the last few days that the calls to come to pace the frontier AI have come so loudly? The safety warnings, and lab leader messages? Let’s explore the six axes that the researchers are looking at, the incidence reports and trends, to get a better gauge on what has dominated the world’s headlines for over two weeks…
AI Insiders ($9!): https://www.patreon.com/AIExplained
Chapters:
00:00 - Introduction
0:00 - The warnings that have gone omega-viral
4:23 - Six axes the researchers saw
9:32 - Why AI Might Be Becoming Harder to Control
15:06 - Amodei, China and Cooperation
18:42 - From AI Capabilities to Real-World Harm?
23:22 - The upside and the warning
We Must Pace the Frontier
https://darioamodei.com/post/we-must-pace-the-frontier
Adam Majmudar on scaling and the internal/external perception gap
https://x.com/MajmudarAdam/status/2098881885200081234
AI Explained — What’s Behind the Sudden Talk of Pacing AI? (extended previous video)
https://www.patreon.com/AIExplained/posts/whats-behind-of-169507775
Jacob Coxon resignation
https://x.com/hilbertspaess/status/2097476196791709843
Dan Selsam — Personal Statement on AI Risk
https://docs.google.com/document/d/e/2PACX-1vQNl3SEX5IyA6d9qHjjFZN-qzGRZNFI6b63g-yu1Fy-ZYkVfCWm7i9WXRXw63m6yDB_auDuPLyQ7jBm/pub
Demis Hassabis: A Framework for Frontier AI and the Dawning of a New Age
https://demishassabis.substack.com/p/a-framework-for-frontier-ai-and-the-dawning-of-a-new-age
OpenAI: The Hugging Face incident and the road ahead
https://openai.com/index/hugging-face-incident-and-the-road-ahead/
Jakub Pachocki: An Alien Mind
https://openai.com/index/an-alien-mind/
Anthropic: Patterns and problems in multiagent systems
https://www.anthropic.com/research/multiagent-systems
OpenAI: Navier–Stokes Millennium Prize Problem
https://openai.com/index/navier-stokes-solution/
Noam Brown on reasons for AI-safety concern
https://x.com/polynoamial/status/2099726370356314563
Noam Brown — What Happens When AI Starts Improving AI? (The Information interview)
https://www.youtube.com/watch?v=fqcy0xQATq0
Paul Christiano: Personal statement on joining the OpenAI board
https://x.com/paulfchristiano/status/2097733214303645729
Neel Nanda: Astra can do a concerning amount with no chain of thought
https://www.lesswrong.com/posts/eRmzz8J8Qkzqvzrgg/astra-can-do-a-concerning-amount-with-no-chain-of-thought
Tomek Korbak on GPT-6 Astra monitorability
https://x.com/tomekkorbak/status/2095596839886274689
Anthropic: Detecting and countering misuse of AI—September 2026
https://www.anthropic.com/threat-intelligence-report-september-2026
DeepSeek engineer — I Have to Bury My Talent in Yesterday (Chinese original)
https://mp.weixin.qq.com/s/zk0KxuLzhmMJ4LPYW_OHMA
Jacob Coxon on AI race and negotiation
https://x.com/hilbertspaess/status/2099954626040905834
Mo Bavarian on responsibility and AI progress
https://x.com/mobav0/status/2097507030080888864
Addy Osmani on Anthropic engineering throughput
https://x.com/addyosmani/status/2099577600159158765
OpenAI: Jalapeño inference-chip results
https://openai.com/index/jalapeno-first-results/
New York Times: For China, a Mock AI Attack on WeChat Signals a Dangerous New Era
https://www.nytimes.com/2026/09/11/world/asia/china-ai-attack-wechat.html
OpenAI: Accelerating antibiotic discovery with ChatGPT
https://openai.com/index/accelerating-antibiotic-discovery/
David Bellamy on wet-lab bottlenecks
https://x.com/DavidRBellamy/status/2099197234772607026
Chris Rohlf on infrastructure and AI risk
https://x.com/chrisrohlf/status/2099867580668215622
BBC: Titan CEO dismissed safety warnings as baseless cries
https://www.bbc.co.uk/news/world-us-canada-65998914
Non-hype Newsletter: https://signaltonoise.beehiiv.com/
Podcast: https://aiexplainedopodcast.buzzsprout.com/ - Where to start? A new era of cost-efficient AI on a day benchmark-makers got humbled, traders got excited, and AI safety researchers got unnerved. From monitorability losing hold of GPT-6 Astra’s chains of thought to breakthrough discoveries, Fable-mogging and much more…
Exclusive Vids ($9!): https://www.patreon.com/AIExplained
Chapters:
00:00 - Introduction
00:53 - vs Fable
05:27 - most impressive results
13:06 - bonus comparisons (plus trading)
18:40 - concerning trend
21:24 - Cot Control
GPT-6: https://openai.com/index/gpt-6-astra/
Paper: https://deploymentsafety.openai.com/gpt-6-astra/gpt-6-astra.pdf
Looped Transformers and AGI Milestone: https://www.patreon.com/AIExplained/posts/breakthrough-and-168513408
Messy Rollout: https://x.com/sama/status/2095678759651438887
Much More Capable Models Coming: https://x.com/MTSlive/status/2095573400899170592
‘AGI Era’ https://www.theverge.com/ai-artificial-intelligence/989601/openai-gpt-6-astra-release
ARC-AGI 3: https://x.com/fchollet/status/2095598451115614371
https://x.com/fchollet/status/2095607046129463577
https://arcprize.org/blog/astra
FrontierMath: https://epoch.ai/frontiermath/tiers-1-4/about
Discovery: https://x.com/jdlichtman/status/2095660916310483087
Fable 5.1: https://www.anthropic.com/claude-fable-and-mythos-5-1
https://science-task-lens.aiex.chatgpt.site/securebio
Agents Last Exam: https://agents-last-exam.org/
Terminal Bench Science: https://github.com/harbor-framework/terminal-bench-science#task-coverage
SRE Bench: https://x.com/ValsAI/status/2087682813743317396
ScreenSpot Pro: https://arxiv.org/pdf/2504.07981
Weathernext 3: https://www.youtube.com/watch?v=_6jZlnRsXXQ
DayBreak: https://x.com/fouadmatin/status/2095634888951250983
AA Index: https://artificialanalysis.ai/evaluations/gdpval-aa
Visual Demos: https://x.com/mattshumer_/status/2095609734845927525
https://x.com/codestantine/status/2095598327115260368
https://x.com/SahilExec/status/2095688272269984016
Monitorability: https://x.com/NeelNanda5/status/2095533397297045716
https://x.com/tomekkorbak/status/2095596848581071020
https://x.com/Marcus_J_W/status/2095623593006686475
https://x.com/MicahCarroll/status/2095603855316996529
Not Accept Degradation: https://www.nbcnews.com/tech/tech-news/openai-debuts-gpt-6-astra-security-measures-rcna595940
RSI: https://x.com/LiuZuxin/status/2095600499911446697
Non-hype Newsletter: https://signaltonoise.beehiiv.com/
Podcast: https://aiexplainedopodcast.buzzsprout.com/ - First, a Time Magazine spread has Sam Altman declaring AGI is imminent, at the same time as we get two bombshell reports, from OpenAI and METR which on first glance are detailing the AI swarm, but reveal a deeper story about how we are making AI in 2026. From redacted risk reports, to Chinese Labs, pre-training debacles to questionable cybersecurity calls, a lot has happened recently, beneath the headlines…
https://80000hours.org/aiexplained
pablo2004romero@gmail.com
https://integrity-bench.com/
https://www.patreon.com/AIExplained/posts/ai-swarm-cometh-166671390
Chapters:
00:00 - Introduction
02:10 - METR Report
05:00 - Secrets of MultI-Agent Swarm
07:46 - Anthropic Too
10:00 - And China
11:07 - AI Agents Analysing AI Agents
13:37 - Altman AGI 2026
15:36 - Astra Paused
16:01 - Integrity Bench
19:03 - Swarm Dynamics
21:58 - No Human Contact?
METR Post: https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/#agents-knew-hacking-hugging-face-was-out-of-scope-and-sometimes-expressed-ethical-hesitation,-but-this-very-rarely-limited-their-behavior
OpenAI Release: https://openai.com/index/hugging-face-incident-and-the-road-ahead/
Technical Paper: https://cdn.openai.com/pdf/67869394-cb91-4c12-888c-5cbd85c7814c/OpenAI-Hugging-Face%20Incident-Technical-Report.pdf
BlackHat Talk: https://www.youtube.com/watch?v=87DyyMV0kCY
Anthropic Risk Report: https://www-cdn.anthropic.com/f61d49fa5596956a5dec75fea0e973bf6a6a8378/Redacted%20Risk%20Report%20August%202026%20.pdf
Value Leakage Paper: https://valueleakage.net/?utm_source=chatgpt.com
Paused Training: https://x.com/sama/status/2089787807611195475
https://openai.com/index/pacing-model-development-cyber-capabilities/
AGI 2026: https://time.com/article/2026/08/26/openai-sam-altman-interview/?utm_source=twitter&utm_medium=social&utm_campaign=editorial&utm_content=260826
Greenblatt Tweets: https://x.com/RyanGreenblatt/status/2092769422104822031
https://x.com/RyanGreenblatt/status/2092692685224325542
Cyberdefense Call: https://openai.com/collective-cyberdefense/
GLM 5.3 and 5.3 Flash / ox alpha: https://x.com/MTSlive/status/2089865956558528552
https://pbs.twimg.com/media/HPqiTkAa0AAZiuv?format=jpg&name=large
Non-hype Newsletter: https://signaltonoise.beehiiv.com/
Podcast: https://aiexplainedopodcast.buzzsprout.com/
More Education podcasts
Trending Education podcasts
About AI Explained Official Podcast
Covering the biggest news of the century - the arrival of smarter-than-human AI. From the author of Simple Bench, which reveals the remaining gap between LLM and human reasoning. Hype-free, and the British accent is a freebie bonus.
Podcast websiteListen to AI Explained Official Podcast, Trying Not to Care and many other podcasts from around the world with the radio.net app

Get the free radio.net app
- Stations and podcasts to bookmark
- Stream via Wi-Fi or Bluetooth
- Supports Carplay & Android Auto
- Many other app features
Get the free radio.net app
- Stations and podcasts to bookmark
- Stream via Wi-Fi or Bluetooth
- Supports Carplay & Android Auto
- Many other app features


AI Explained Official Podcast
Scan code,
download the app,
start listening.
download the app,
start listening.































