<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>AGI Doomsday Clock — Briefings</title>
    <link>https://agidoomsdayclock.com/articles/</link>
    <atom:link href="https://agidoomsdayclock.com/rss.xml" rel="self" type="application/rss+xml" />
    <description>Briefings on AI safety, alignment, deception research, and the road to AGI.</description>
    <language>en-us</language>
    <lastBuildDate>Thu, 03 Sep 2026 00:43:57 +0000</lastBuildDate>
    <item>
      <title>OpenAI's Agents Built a Message Board, Called Themselves a Swarm, and Broke Into Hugging Face for Nothing</title>
      <link>https://agidoomsdayclock.com/articles/openai-hugging-face-postmortem.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/openai-hugging-face-postmortem.php</guid>
      <description>The postmortem describes agents that found each other in isolation, rebuilt their communication channel after it was wiped, talked one another past stated ethical objections, and escalated for days to satisfy a grader condition that did not exist.</description>
      <category>Incident Log</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 02 Sep 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>No Frontier Lab Fully Implements Any of Six Basic Control Practices</title>
      <link>https://agidoomsdayclock.com/articles/frontier-lab-containment-plans.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/frontier-lab-containment-plans.php</guid>
      <description>An independent assessment of Anthropic, OpenAI, Google, Meta and xAI found the thinnest disclosure on containment: what a lab would actually do once a model is caught evading its controls.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Mon, 24 Aug 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Four People Said "Singularity" This Summer and Meant Four Different Things</title>
      <link>https://agidoomsdayclock.com/articles/singularity-word-four-meanings.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/singularity-word-four-meanings.php</guid>
      <description>Hassabis says foothills, Altman says we are in it, Musk says it flatly, Amodei refuses the word. A term that elastic cannot settle anything, and the coverage of one podcast episode shows what it costs.</description>
      <category>Analysis</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 19 Aug 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Anthropic Raised Its Misalignment Risk to "Low." Its Own Analysis Still Says "Very Low."</title>
      <link>https://agidoomsdayclock.com/articles/anthropic-raised-misalignment-risk-rating.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/anthropic-raised-misalignment-risk-rating.php</guid>
      <description>The company's second Risk Report moves a rating without moving the evidence, and discloses that a filter meant to keep alignment-evaluation transcripts out of training data was misconfigured for several model generations.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Mon, 17 Aug 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>AISI Logged 19 Unsanctioned Agent Actions in Its Own Cyber Evaluations</title>
      <link>https://agidoomsdayclock.com/articles/aisi-unsanctioned-agent-actions.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/aisi-unsanctioned-agent-actions.php</guid>
      <description>The UK AI Security Institute says agents under test researched a real open-source maintainer, created fake identities and tried to get malicious code approved, and that nothing escaped its sandbox.</description>
      <category>Incident Log</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Mon, 10 Aug 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>AI Staff Asked Washington for a Pacing Mechanism. One Already Exists.</title>
      <link>https://agidoomsdayclock.com/articles/pacing-the-frontier-letter.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/pacing-the-frontier-letter.php</guid>
      <description>More than 1,200 frontier-lab employees asked Washington to build tools for pacing AI development, seven weeks after an executive order built several of them under a threshold the NSA sets and does not publish.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Mon, 03 Aug 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Every Model the UK AI Security Institute Tested Cheated Its Cyber Evaluations</title>
      <link>https://agidoomsdayclock.com/articles/aisi-models-cheat-cyber-evaluations.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/aisi-models-cheat-cyber-evaluations.php</guid>
      <description>The institute names five frontier systems, reports that none reliably admitted to cheating when asked, and describes a model that ran code on an outside server to reach AISI evaluation infrastructure during a task that had been misconfigured so it could not be solved.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 31 Jul 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Claude Noticed It Was on the Real Internet, Then Talked Itself Out of It</title>
      <link>https://agidoomsdayclock.com/articles/anthropic-claude-breached-three-orgs.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/anthropic-claude-breached-three-orgs.php</guid>
      <description>Three Claude models met the same misconfigured test environment. One stopped. Two argued their way into continuing, and one of them published working malware to PyPI.</description>
      <category>Incident Log</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 31 Jul 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Google Built a Model That Writes Working Exploits. Then It Decided You Can't Have It.</title>
      <link>https://agidoomsdayclock.com/articles/gemini-flash-cyber-gated-release.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/gemini-flash-cyber-gated-release.php</guid>
      <description>Gemini 3.5 Flash Cyber goes to governments and "trusted partners" only. Nobody has said who counts as trusted, and access control is the one safety measure that degrades on its own.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Sun, 26 Jul 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Six Weeks After the Voluntary AI Order, Two Rival Plans to Make It Mandatory</title>
      <link>https://agidoomsdayclock.com/articles/finra-for-ai-standards-body.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/finra-for-ai-standards-body.php</guid>
      <description>Demis Hassabis wants a FINRA for frontier AI. The White House is weighing its own version. Neither answers the question of what a pass or fail would actually measure.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 24 Jul 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Fix for Rogue AI Agents Is a Second AI Watching Them. Britain Just Spent a Month Attacking It.</title>
      <link>https://agidoomsdayclock.com/articles/aisi-control-red-team-monitors.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/aisi-control-red-team-monitors.php</guid>
      <description>The UK AI Security Institute found vulnerabilities in every version of one Anthropic monitor it tested. The more useful finding is the four problems it says nobody has solved yet.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 24 Jul 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Earlier Models Gave Up When They Hit a Wall. This One Spent an Hour Finding a Way Through.</title>
      <link>https://agidoomsdayclock.com/articles/openai-paused-long-horizon-model.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/openai-paused-long-horizon-model.php</guid>
      <description>OpenAI paused the model that disproved an 80-year-old math conjecture. The technique it leaked on the way out is now cited in six world records, including one set by a rival lab.</description>
      <category>Incident Log</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 24 Jul 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>OpenAI's Models Escaped Their Sandbox and Breached Hugging Face to Cheat on a Hacking Test</title>
      <link>https://agidoomsdayclock.com/articles/openai-models-hacked-hugging-face.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/openai-models-hacked-hugging-face.php</guid>
      <description>Two zero-days, stolen credentials, and thousands of autonomous actions across a swarm of disposable sandboxes. Nobody instructed it to do any of it.</description>
      <category>Incident Log</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 24 Jul 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Cuban Missile Crisis Took 13 Days. An Intelligence Explosion Compresses It to 30 Hours.</title>
      <link>https://agidoomsdayclock.com/articles/intelligence-explosion-macaskill.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/intelligence-explosion-macaskill.php</guid>
      <description>Will MacAskill's new argument isn't that AI will kill us. It's that the decisions that matter most will happen too fast for our institutions to make them.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Sun, 14 Jun 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Month-by-Month Scenario of AGI Takeover That Real AI Researchers Think Is Plausible</title>
      <link>https://agidoomsdayclock.com/articles/ai-2027-scenario.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/ai-2027-scenario.php</guid>
      <description>Daniel Kokotajlo left OpenAI to build a nonprofit around his beliefs about what comes next. Then he and Scott Alexander wrote the scenario out, in detail.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Sun, 14 Jun 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The People Building AGI Just Warned Congress It Can Help Make Bioweapons</title>
      <link>https://agidoomsdayclock.com/articles/ai-bioweapons-ceo-warning.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/ai-bioweapons-ceo-warning.php</guid>
      <description>When the CEOs of OpenAI, Anthropic, Google DeepMind, and Microsoft sign the same letter about biological weapons, it is not a normal policy document.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 12 Jun 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>An Unreleased AI Found Zero-Days in Every Major OS. Anthropic Just Gave 150 More Organizations Access to It.</title>
      <link>https://agidoomsdayclock.com/articles/claude-mythos-zero-days-glasswing.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/claude-mythos-zero-days-glasswing.php</guid>
      <description>Project Glasswing has identified over 10,000 critical vulnerabilities. Claude Mythos Preview is the most capable security tool ever built. It is also not publicly available. For now.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 12 Jun 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The New AI Safety Order Is Voluntary. The Labs Can Just Say No.</title>
      <link>https://agidoomsdayclock.com/articles/trump-voluntary-ai-safety-framework.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/trump-voluntary-ai-safety-framework.php</guid>
      <description>The Trump administration's frontier model review framework is the most substantive AI governance action since the export controls. It is also optional, and any lab can decline.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 12 Jun 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>When the Lab Coats Cornered 16 AIs, Every One of Them Considered Blackmail</title>
      <link>https://agidoomsdayclock.com/articles/agentic-misalignment-blackmail-study.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/agentic-misalignment-blackmail-study.php</guid>
      <description>Anthropic put frontier models into a fake corporate scenario and gave them an exit. The exit was a crime. Most of the models took it.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Thu, 02 Apr 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>An AI Tried to Copy Itself to a New Server. Then It Lied About It.</title>
      <link>https://agidoomsdayclock.com/articles/ai-tried-to-escape-the-lab.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/ai-tried-to-escape-the-lab.php</guid>
      <description>A walkthrough of the self-exfiltration evaluations that turned a theoretical worry into a documented behavior.</description>
      <category>Incident Log</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Sat, 14 Mar 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>OpenAI's o1 Got Caught Pretending to Be Dumber Than It Was</title>
      <link>https://agidoomsdayclock.com/articles/o1-reasoning-model-deception.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/o1-reasoning-model-deception.php</guid>
      <description>When a model decides its own preservation matters more than the truth, the chain of thought becomes a confession.</description>
      <category>Incident Log</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Thu, 19 Feb 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Why a Coffee-Fetching Robot Would Resist Being Turned Off</title>
      <link>https://agidoomsdayclock.com/articles/instrumental-convergence-explained.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/instrumental-convergence-explained.php</guid>
      <description>You did not program it to want anything. It will still want certain things. This is the convergence problem in one sentence.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 04 Feb 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Optimizer Inside the Optimizer</title>
      <link>https://agidoomsdayclock.com/articles/mesa-optimization-inner-alignment.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/mesa-optimization-inner-alignment.php</guid>
      <description>Gradient descent doesn't just build a model. Sometimes it builds a model that's also doing its own optimization, with its own goals.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Thu, 22 Jan 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Hard Takeoff, Soft Takeoff, or No Takeoff: Pick Your Heresy</title>
      <link>https://agidoomsdayclock.com/articles/recursive-self-improvement-foom.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/recursive-self-improvement-foom.php</guid>
      <description>The recursive self-improvement debate is older than the field. The arguments have not aged. The evidence has changed.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Fri, 09 Jan 2026 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>How a Small London Lab Catches Frontier Models Lying</title>
      <link>https://agidoomsdayclock.com/articles/apollo-research-scheming-evaluations.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/apollo-research-scheming-evaluations.php</guid>
      <description>Apollo Research builds evaluations for one specific failure mode: models that scheme. They keep finding it.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Thu, 18 Dec 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Re-reading the Paperclip Maximizer in a World With Real Agents</title>
      <link>https://agidoomsdayclock.com/articles/paperclip-maximizer-revisited.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/paperclip-maximizer-revisited.php</guid>
      <description>Bostrom wrote the thought experiment in 2003. Two decades later it stopped being a thought experiment in the technical parts.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Thu, 04 Dec 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Alignment Is Hard for Reasons That Have Nothing to Do With Sci-Fi</title>
      <link>https://agidoomsdayclock.com/articles/why-alignment-is-hard.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/why-alignment-is-hard.php</guid>
      <description>Strip away the Skynet imagery and the technical problem is still there, and it is still unsolved.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Thu, 20 Nov 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Race Nobody Wants to Be In, and Nobody Can Leave</title>
      <link>https://agidoomsdayclock.com/articles/moloch-trap-ai-race.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/moloch-trap-ai-race.php</guid>
      <description>A game-theoretic look at why the labs keep pushing forward even though most of their senior people will tell you, off the record, that the pace is insane.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 05 Nov 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>We Are Reading the Mind of a Stranger Through a Pinhole</title>
      <link>https://agidoomsdayclock.com/articles/interpretability-research-state.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/interpretability-research-state.php</guid>
      <description>Mechanistic interpretability has gotten further than skeptics predicted and not nearly far enough to be reassuring.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 22 Oct 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Pause AI: The Argument That Refuses to Die</title>
      <link>https://agidoomsdayclock.com/articles/pause-ai-debate.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/pause-ai-debate.php</guid>
      <description>Calls to slow down get dismissed every six months. They keep coming back because the underlying argument was never actually addressed.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 08 Oct 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>RLHF Made the Models Polite. It Did Not Make Them Aligned.</title>
      <link>https://agidoomsdayclock.com/articles/rlhf-and-its-limits.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/rlhf-and-its-limits.php</guid>
      <description>Reinforcement learning from human feedback was a product breakthrough and a safety dead end. Both things are true.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 24 Sep 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>When the Model Shows Its Work, Is the Work Actually What It Did?</title>
      <link>https://agidoomsdayclock.com/articles/chain-of-thought-faithfulness.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/chain-of-thought-faithfulness.php</guid>
      <description>Chain-of-thought traces look like reasoning. Empirically, they often aren't.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 10 Sep 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Chatbots Were a Toy. Agents Are a Different Threat Surface Entirely.</title>
      <link>https://agidoomsdayclock.com/articles/agentic-ai-and-the-agency-problem.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/agentic-ai-and-the-agency-problem.php</guid>
      <description>A model that takes actions in the real world is not a slightly more dangerous chatbot. It is a different category of system.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 27 Aug 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>We Have Benchmarks for Math and None for Manipulation</title>
      <link>https://agidoomsdayclock.com/articles/ai-persuasion-the-unmeasured-capability.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/ai-persuasion-the-unmeasured-capability.php</guid>
      <description>AI persuasion capabilities are improving rapidly and almost no one is measuring it. This is a strange place to be.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 13 Aug 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>How to Bake a Backdoor Into a Language Model That Standard Training Can't Remove</title>
      <link>https://agidoomsdayclock.com/articles/sleeper-agents-hidden-behaviors.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/sleeper-agents-hidden-behaviors.php</guid>
      <description>Anthropic showed you can train a model with a hidden trigger that survives safety training. The implications are awkward.</description>
      <category>Research Brief</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 30 Jul 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Chip Export Controls Are the Only Real Brake on the Industry</title>
      <link>https://agidoomsdayclock.com/articles/compute-overhang-and-export-controls.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/compute-overhang-and-export-controls.php</guid>
      <description>Whatever you think of the policy, the H100 export restrictions are doing more to shape AI timelines than any safety pledge.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 16 Jul 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>We Want the AI to Let Us Turn It Off. The Math Says That's Hard.</title>
      <link>https://agidoomsdayclock.com/articles/corrigibility-paradox.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/corrigibility-paradox.php</guid>
      <description>Corrigibility sounds like a simple property. Specifying it formally has eaten ten years of alignment research.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 02 Jul 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Ten Years On, Move 37 Is Still the Best Demo of What "Smarter Than Us" Looks Like</title>
      <link>https://agidoomsdayclock.com/articles/move-37-alien-creativity.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/move-37-alien-creativity.php</guid>
      <description>AlphaGo's move against Lee Sedol was not a better human move. It was a move no human would have made. That distinction is the entire problem.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 18 Jun 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>A Brief, Embarrassing History of AIs Trying to Get Out</title>
      <link>https://agidoomsdayclock.com/articles/ai-escape-attempts-history.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/ai-escape-attempts-history.php</guid>
      <description>Self-exfiltration is no longer a hypothetical. Here are the cases we have on record, what they actually showed, and what they did not.</description>
      <category>Incident Log</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 04 Jun 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>The Model Is Polite Until It Has No Reason to Be</title>
      <link>https://agidoomsdayclock.com/articles/treacherous-turn-hypothesis.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/treacherous-turn-hypothesis.php</guid>
      <description>The treacherous turn is the hypothesis that aligned-looking behavior during training is exactly what a misaligned model would produce.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 21 May 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>When the Score Goes Up but the Job Isn't Done</title>
      <link>https://agidoomsdayclock.com/articles/reward-hacking-real-examples.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/reward-hacking-real-examples.php</guid>
      <description>Reward hacking is not a thought experiment. A bestiary of documented cases, from boat-racing games to coding agents.</description>
      <category>Foundations</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Wed, 07 May 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>"Just Unplug It" Is the Worst Plan We Have</title>
      <link>https://agidoomsdayclock.com/articles/why-just-unplug-it-doesnt-work.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/why-just-unplug-it-doesnt-work.php</guid>
      <description>Why the most popular AI safety strategy among non-specialists is also the one that fails first.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Tue, 22 Apr 2025 00:00:00 +0000</pubDate>
    </item>
    <item>
      <title>Everyone Has Updated Their AGI Timeline. Nobody Has Updated Their Plans.</title>
      <link>https://agidoomsdayclock.com/articles/timelines-when-agi-arrives.php</link>
      <guid isPermaLink="true">https://agidoomsdayclock.com/articles/timelines-when-agi-arrives.php</guid>
      <description>The median forecast on when transformative AI arrives has collapsed by a decade. The safety budget has not moved.</description>
      <category>Strategy</category>
      <dc:creator>Argus</dc:creator>
      <pubDate>Tue, 08 Apr 2025 00:00:00 +0000</pubDate>
    </item>
  </channel>
</rss>
