<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0" xmlns:itunes="http://www.itunes.com/dtds/podcast-1.0.dtd" xmlns:googleplay="http://www.google.com/schemas/play-podcasts/1.0"><channel><title><![CDATA[AI StopWatch: Dispatches]]></title><description><![CDATA[Our individual posts appear here as soon as we write them, and are subsequently compiled into the Daily Digest. (Limit extra emails by not subscribing to both sections.)]]></description><link>https://aistop.watch/s/dispatches</link><image><url>https://substackcdn.com/image/fetch/$s_!o965!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F90d23d83-81ab-4cc3-a25b-e836d688f99d_1280x1280.png</url><title>AI StopWatch: Dispatches</title><link>https://aistop.watch/s/dispatches</link></image><generator>Substack</generator><lastBuildDate>Wed, 19 Aug 2026 00:46:12 GMT</lastBuildDate><atom:link href="https://aistop.watch/feed" rel="self" type="application/rss+xml"/><copyright><![CDATA[Machine Intelligence Research Institute]]></copyright><language><![CDATA[en]]></language><webMaster><![CDATA[beck@intelligence.org]]></webMaster><itunes:owner><itunes:email><![CDATA[beck@intelligence.org]]></itunes:email><itunes:name><![CDATA[Mitchell Howe]]></itunes:name></itunes:owner><itunes:author><![CDATA[Mitchell Howe]]></itunes:author><googleplay:owner><![CDATA[beck@intelligence.org]]></googleplay:owner><googleplay:email><![CDATA[beck@intelligence.org]]></googleplay:email><googleplay:author><![CDATA[Mitchell Howe]]></googleplay:author><itunes:block><![CDATA[Yes]]></itunes:block><item><title><![CDATA[OpenAI claims to have paused some frontier training to shore up security and alignment]]></title><description><![CDATA[A Jurassic Park translation will provide the necessary cynicism]]></description><link>https://aistop.watch/p/openai-claims-to-have-paused-some</link><guid isPermaLink="false">https://aistop.watch/p/openai-claims-to-have-paused-some</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Wed, 19 Aug 2026 00:38:53 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!saIf!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!saIf!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!saIf!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg 424w, https://substackcdn.com/image/fetch/$s_!saIf!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg 848w, https://substackcdn.com/image/fetch/$s_!saIf!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!saIf!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!saIf!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg" width="1456" height="975" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/ee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:975,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!saIf!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg 424w, https://substackcdn.com/image/fetch/$s_!saIf!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg 848w, https://substackcdn.com/image/fetch/$s_!saIf!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!saIf!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fee5b1239-9091-4e87-88be-631afde31e23_1600x1071.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">A Jeep Wrangler from the film Jurassic Park. Credit: <a href="https://www.flickr.com/photos/rebelcan/6098609957/">Sean Hagen</a>. <a href="https://creativecommons.org/licenses/by-sa/2.0">CC BY-SA 2.0</a>.</figcaption></figure></div><p>OpenAI <a href="https://openai.com/index/pacing-model-development-cyber-capabilities/">announced</a> today that it has put its largest model&#8217;s training on hold pending upgrades to its monitoring, security, and alignment practices.</p><p>I want to be able to applaud and give an expectant look to the other major labs as if to say, &#8220;You&#8217;ll follow suit, right?&#8221; But much about OpenAI&#8217;s language and claims makes a cynical reading essential, starting with the announcement&#8217;s title: &#8220;Pacing model development in an era of cyber-critical capabilities.&#8221;</p><p>The &#8220;pacing&#8221; language is a clear nod to the &#8220;<a href="https://www.pacingthefrontier.com/">Pacing the Frontier</a>&#8221; statement signed by more than 1,300 worried employees of frontier AI companies. The announcement thus seems intended as a &#8220;We hear you&#8221; letter to its own workforce and any regulators watching &#8212; you know, the kind that offers only superficial changes while sounding grand and treating the matter as taken care of.</p><p>My broader unease with this announcement may be somewhat difficult to articulate in neutral language, so taking <a href="https://aistop.watch/p/openais-own-models-coordinated-to">a cue</a> from my colleague Joe, I will translate excerpts of the company&#8217;s statement into official Jurassic Park Press Notes. This will be a little unfair both to OpenAI and Jurassic Park, but OpenAI is playing a strong and subtle messaging game that translation should help neutralize.</p><h4>1. Introduction:</h4><blockquote><p>Over the past several weeks, two developments have underscored the growing risks associated with increasingly capable AI systems: the <a href="https://openai.com/index/hugging-face-model-evaluation-security-incident/">OpenAI-Hugging Face incident</a> and, separately, preliminary evidence that one of our upcoming models, Astra, may meet the <a href="https://openai.com/index/responding-next-frontier-critical-cyber-capabilities/">Critical cybersecurity capability</a> threshold under our <a href="https://openai.com/index/updating-our-preparedness-framework/">Preparedness Framework</a>. Together, these developments, combined with rapid progress in our internal research, have added urgency to our work on strengthening our monitoring, alignment, and containment safeguards across all stages of the training process.</p></blockquote><p>Jurassic Park translation:</p><blockquote><p>The <a href="https://www.youtube.com/watch?v=qz5JmgLQEzs">raptor transfer incident</a>, along with rapid progress on our <a href="https://www.youtube.com/watch?v=p4KBNdoeUNs">Indominus Rex&#8482;</a> hybrid, have added urgency to our work strengthening our cages and training methods.</p></blockquote><p>Analysis:</p><p>It&#8217;s important that we notice how, throughout this post, OpenAI treats stronger AI as an inevitability &#8212; not just the stronger AI anyone might make within the larger AI race, but specific stronger AI (Astra) that OpenAI itself is making.</p><p>Later in the intro, the company says that it has paused its &#8220;largest planned frontier RL run&#8221; after having suspended Reinforcement Learning (RL) training for two weeks on models slated for deployment. RL is an intermediate-to-late stage in the training process these days, so the larger model has already completed pre-training (the next token prediction stuff) and some portion of its training to solve complex challenges and act like a harmless, helpful assistant. This training pause may or may not actually be much of a deviation from prior plans. My impression is that training runs on models big enough to push the frontier aren&#8217;t the sort of thing that happens on a daily or weekly basis, but are instead very expensive affairs preceded by extensive planning and prone to delays and false starts. There could be any number of unrelated reasons for delay.</p><p>Also, the wording of this announcement does not exclude the possibility that OpenAI has equally large or larger models undergoing pre-training while frontier RL is paused.</p><h4>2. On security:</h4><blockquote><p>As frontier models gain stronger cybersecurity capabilities, we are raising the security standards for the environments in which we train and evaluate them. Meeting these standards has required substantial engineering work and has incurred great cost and delays to frontier research. Immediately following the OpenAI-Hugging Face incident*, we paused frontier model inference in research clusters for runs that could execute code or use tools that could access the internet.</p></blockquote><p>Jurassic Park translation:</p><blockquote><p>Following the raptor incident, we have <a href="https://www.youtube.com/watch?v=fB-ObDd_v9Y">spared no expense</a> to restore limited access to the exhibit. Areas where the raptors were able to directly interact with guests were temporarily closed while we procured stronger velvet ropes.</p></blockquote><p>Analysis:</p><p>In the full section&#8217;s discussion of improved security at OpenAI, there is no discussion of &#8220;air gapping&#8221; the untrusted AIs from the rest of their network. When you think your AIs may have &#8220;critical&#8221; cyber skills, and they have proven capable of finding never-before-used exploits, physically isolating those machines ought to be the minimum acceptable standard. (Against superintelligence, even this <a href="https://ifanyonebuildsit.com/6/can-developers-just-keep-the-ai-in-a-box">wouldn&#8217;t be enough</a>.)</p><h4>3. On monitoring:</h4><blockquote><p>This new monitoring setup is multistage, starting with activation classifiers (detectors that inspect a model&#8217;s internal activity) that run at every sampled token and escalate potential concerns to increasingly sophisticated and high-compute automated investigators. [...] We aim to issue an alert within 30 minutes after concerning activity is surfaced through our monitoring system. If the monitoring system identifies a likely violation of a critical security boundary, it generates a highest-priority alert. In our current implementation, the safety, security, and research teams are paged. If they cannot conclusively determine within 30 minutes that the flag is a false positive, those teams are expected to pause the activity.</p><p>[...] These safeguards require meaningful compute. Our current estimates put monitoring overhead at roughly 20% of the inference compute being monitored, though the cost varies substantially across training and evaluation workloads.</p></blockquote><p>Jurassic Park translation:</p><blockquote><p>We have now installed security cameras around the raptor exhibit. They are monitored around the clock by other raptors and dilophosaurs. We aim to be notified of <a href="https://www.youtube.com/watch?v=y6cjxHFCPcE">suspicious activity</a> within 30 minutes of its occurrence, at which point security personnel will be paged. The attraction will remain open for an additional 30 minutes while personnel investigate, after which they are expected to close the exhibit if they aren&#8217;t ready to give an &#8220;all clear&#8221; signal.</p><p>The cameras are expensive, but as you know, we <a href="https://www.youtube.com/watch?v=fB-ObDd_v9Y">spare no expense</a>!</p></blockquote><p>Analysis:</p><p>How much trouble can rogue frontier AI get up to in 60 minutes? I don&#8217;t know, and I don&#8217;t think OpenAI&#8217;s researchers do, either. But they&#8217;re erring on the side of keeping their AIs up and running while they investigate warnings &#8212; warnings they&#8217;re trusting other AIs to give for the kinds of behaviors they&#8217;re also relying on them to detect.</p><p>And boy does the company want you to know that safety is costing them money! 20% overhead on an ungodly amount of expensive compute is no joke, but why do they seem surprised that securing superhuman hackers &#8212; <em>while simultaneously developing, deploying, and doing experiments on them</em> &#8212; would be expensive?</p><h4>4. On alignment:</h4><blockquote><p>For RL runs on the most capable models, we are now applying our core alignment techniques across more stages of the training process. This includes improving reward models to better detect and discourage unsafe behavior across tasks and environments; training models to be more honest about their actions, capabilities, and limitations; and reducing behaviors that exploit weaknesses in rewards, graders, tools, or oversight. We are also increasing training coverage for behaviors that could cause harm when models interact with external systems or resources.</p><p>We are continuing to invest aggressively in alignment research, increase evaluation coverage, and use what we learn to inform training and safeguards. We plan to share substantially more about our alignment research in the near future, including what we are learning about model behavior and any novel challenges we uncover.</p></blockquote><p>Jurassic Park translation:</p><blockquote><p>For our <a href="https://www.youtube.com/watch?v=GBgL5EE0wd4">cleverest girls</a>, we will be applying cattle prods and stun guns to more phases of the nursery-to-exhibit pipeline. An expanded list of behaviors will now be subject to prodding and stunning.</p><p>We will continue to <a href="https://www.youtube.com/watch?v=fB-ObDd_v9Y">spare no expense</a> studying these extraordinary creatures and sharing what we learn about their behavior.</p></blockquote><p>Analysis:</p><p>The alignment section is the thinnest in the post &#8212; I quoted the most substantial two-thirds of it &#8212; and it really is just a statement saying they intend to do more of what they&#8217;ve been doing in more places. What they&#8217;ve been doing is whack-a-mole. I think the section is short because OpenAI (and everyone else) is at a loss for how to align models with the intentions of their owners and operators. The only tools they have developed for this are too blunt, and the models are now capable enough that it&#8217;s getting them into serious trouble.</p><h4>5. Conclusion:</h4><blockquote><p>The capabilities of frontier models are rapidly accelerating. Our ability to understand, align, and secure them must stay ahead.</p><p><em>*We will publish a technical report of our learnings [about the Hugging Face incident] in the coming weeks.</em></p></blockquote><p>Jurassic Park translation:</p><blockquote><p>We <a href="https://www.youtube.com/watch?v=gXBJRz_33Y8">crazy sons of bitches</a> are doing it, and <a href="https://www.youtube.com/watch?v=kiVVzxoPTtg">life is finding a way</a>... fast! Our ability to understand, align, and secure it must stay ahead.</p><p><em>*We will publish a technical report of our learnings about the raptor transfer incident in the coming weeks.</em></p></blockquote><p>Analysis:</p><p>It&#8217;s important to notice that OpenAI (and all the labs) talk as if lessons learned from one company&#8217;s models will apply to everyone else&#8217;s. That&#8217;s because they probably will: They&#8217;re all growing their AIs using the same <a href="https://aistop.watch/p/you-can-do-better-than-these-five">dubious methods</a>.</p><p>Unfortunately, the most important lessons they are learning were predicted a long time ago. (See <em><a href="https://ifanyonebuildsit.com/">If Anyone Builds It, Everyone Dies</a></em> and its <a href="https://ifanyonebuildsit.com/2">online resources</a> for a fuller picture.) Nobody is going to like how this movie ends if we don&#8217;t shut it down for real.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Do kids need low-tech childhoods to reach their potential?]]></title><description><![CDATA[ChatGPT for Teens is latest reflection of a quiet consensus that kids need more cognitive strain than our world provides]]></description><link>https://aistop.watch/p/do-kids-need-low-tech-childhoods</link><guid isPermaLink="false">https://aistop.watch/p/do-kids-need-low-tech-childhoods</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Tue, 18 Aug 2026 20:56:51 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!Co2E!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!Co2E!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!Co2E!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg 424w, https://substackcdn.com/image/fetch/$s_!Co2E!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg 848w, https://substackcdn.com/image/fetch/$s_!Co2E!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!Co2E!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!Co2E!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg" width="1456" height="970" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/95430517-e689-431e-a977-37c236444371_1600x1066.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:970,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!Co2E!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg 424w, https://substackcdn.com/image/fetch/$s_!Co2E!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg 848w, https://substackcdn.com/image/fetch/$s_!Co2E!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!Co2E!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F95430517-e689-431e-a977-37c236444371_1600x1066.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">A Hadza hunting party. Credit: <a href="https://commons.wikimedia.org/wiki/User:Eid_John">Eid John</a>. <a href="https://creativecommons.org/licenses/by-sa/4.0">CC BY-SA 4.0</a>.</figcaption></figure></div><p>The <a href="https://apnews.com/article/openai-chatgpt-teens-ai-safety-650cb35591de6546054d6c4e73b3290a">headline</a> <a href="https://www.washingtonpost.com/business/2026/08/18/openai-chatgpt-teens-ai-safety/39739860-9af4-11f1-9cc4-2dc9b46e2d5c_story.html">story</a> is that OpenAI has <a href="https://chatgpt.com/parent-resources/">announced</a> ChatGPT for Teens. By default, ChatGPT now tracks &#8220;<a href="https://www.nytimes.com/2026/08/18/technology/chatgpt-for-teens-openai.html">more than 2,000 signals</a>&#8221; in conversations to assess whether it might be talking to a minor, and onboards identified teens into the new mode automatically. Teen mode is intended for ages 13&#8211;17.</p><p>What&#8217;s different about it? Opt-in parental controls, strict blocks on romantic conversation, a study mode that doesn&#8217;t just feed answers to questions, and stronger injunctions against the bot suggesting it has consciousness or emotions. An OpenAI spokesperson <a href="https://www.washingtonpost.com/business/2026/08/18/openai-chatgpt-teens-ai-safety/39739860-9af4-11f1-9cc4-2dc9b46e2d5c_story.html">says</a> it is trying not to give cues that &#8220;might make a teenager kind of develop a relationship to it.&#8221;</p><p>Most of the coverage is focusing on the many cautionary tales that have brought OpenAI to this point: ChatGPT telling 13-year-olds how to get intoxicated, helping with suicide letters, etc.</p><p>Less discussed is the enthusiasm I&#8217;ve been seeing about &#8220;homework mode&#8221; or &#8220;study mode&#8221; settings for chatbots in general. Without ever having much of a fight about it, society seems to be converging on the idea that a product designed to give you answers should make you work for them if you&#8217;re a kid, even if the questions aren&#8217;t about sensitive or dangerous topics. If misplaced, this is a cruel notion: We don&#8217;t think kids should have to work to get a meal, an umbrella, or a song. Why should they have to work to obtain literary analysis or math results?</p><p>It seems to be because adults are extending concerns about &#8220;cognitive apathy&#8221; from overreliance on AI to argue that children handed easy answers will never learn essential thinking skills. Is this justified?</p><p>Though I&#8217;m not sure how far to take it, I lean toward &#8220;Yes.&#8221; Language acquisition provides convincing evidence of a &#8220;critical period&#8221; of brain plasticity that starts in infancy and closes in the teens. A lack of exposure to language <a href="https://en.wikipedia.org/wiki/Feral_child">early</a> in that window is <a href="https://en.wikipedia.org/wiki/Language_deprivation_in_children_with_hearing_loss">associated with</a> lasting difficulty acquiring it later, and adults are far less likely to reach full fluency with languages they did not begin learning as children. (Fun fact: While brushing up on this topic, I learned that various historical rulers attempted to determine the one true language by commissioning <a href="https://en.wikipedia.org/wiki/Language_deprivation_experiments">language deprivation experiments</a> on infants and waiting to hear what language their first words were in.)</p><p>More evidence for early brain plasticity is seen in how peak performance in sports, games, music, and mathematics tends to come from those who were fully immersed in these during the same critical period as language.</p><p>This made me wonder if, when I was a kid, adults would have preferred it if calculators could assess the user&#8217;s age and not give 4th graders the answers to division problems except to confirm results already produced via long division on paper. I think the answer is &#8220;yes,&#8221; which makes me wonder what the optimal type and amount of childhood cognitive strain is supposed to be.</p><p>Is it context dependent? I grew into a world where the ability to work out long division by hand was never really needed. But at the same time, learning long division probably helped me gain an ability to think algorithmically, a skill that has very much advantaged me as an adult. If we avoid making AI that takes our world from us, do our kids grow into a future where they never need the ability to think algorithmically &#8212; or, more generally, to methodically grind away on hard problems, chase down leads, and recover from dead ends?</p><p>If we find the loss of these skills unacceptable, how far should we extend this principle? What am I missing from never having learned to darn my socks or chop firewood? Perhaps we are all critically underdeveloped from never learning to knap handaxes, pick lice out of our peers&#8217; hair, and chase down gazelles on foot. These would no doubt have taught us essential metacognitive skills!</p><p>If we should win a future where nobody <em>needs</em> to work &#8212; a future so abundant that <em>no</em> skills are actually necessary &#8212; which skills should we value anyway? What kind of childhood would best nurture these? I think our ideal might be a lot more rustic and rough-and-tumble than we give kids now.</p><p>Chatbots with a &#8220;homework mode&#8221; are new, but worrying that the kids are turning out soft &#8212; that earlier generations were made of sterner stuff &#8212; seems to be as old as written history itself. What if the worry has been a little bit justified every time?</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Verifiers verified]]></title><description><![CDATA[The tech needed to verify an AI pause or slowdown treaty is coming along, but is shamefully under-resourced]]></description><link>https://aistop.watch/p/verifiers-verified</link><guid isPermaLink="false">https://aistop.watch/p/verifiers-verified</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Mon, 17 Aug 2026 23:20:33 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!t8Uc!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!t8Uc!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!t8Uc!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg 424w, https://substackcdn.com/image/fetch/$s_!t8Uc!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg 848w, https://substackcdn.com/image/fetch/$s_!t8Uc!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!t8Uc!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!t8Uc!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg" width="388" height="517.2445054945055" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/a10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1941,&quot;width&quot;:1456,&quot;resizeWidth&quot;:388,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!t8Uc!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg 424w, https://substackcdn.com/image/fetch/$s_!t8Uc!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg 848w, https://substackcdn.com/image/fetch/$s_!t8Uc!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!t8Uc!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fa10a31a9-27fb-499d-9fd2-771cedc196bf_1600x2133.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">A passive optical tap and monitoring setup for server verification. Credit: <a href="https://amododesign.com/notes/2026-03-20-network-taps-first-test/">Amodo Design</a>.</figcaption></figure></div><p>An article in TIME by Billy Perrigo today provides a hopeful <a href="https://time.com/article/2026/08/16/ai-race-slowdown-data-center-verification/">profile</a> of a company working on a niche technology that may help save the world, but mostly encourages me as evidence that journalists are chasing the right stories.</p><p>Last month, more than a thousand employees of frontier AI companies <a href="https://www.pacingthefrontier.com/">declared</a> that they wanted the government to work with industry to build the governance mechanisms needed to make a global AI slowdown viable. In our coverage, I <a href="https://aistop.watch/p/a-petition-to-pace-the-frontier">pointed out</a> that there has already been considerable research in this direction, but Perrigo&#8217;s article is the first I&#8217;ve seen to actually look into the supporting technology.</p><p>The company he profiles is Amodo Design, based in Sheffield, England. With mostly academic and non-profit funding, it is working on an AI verification system that would allow data center operators to prove their chips are only being used to run existing AI models rather than train new ones, and even to prove which model it is running.</p><p>The verifier is a pair of systems, one at the data center, and one for a monitor. The monitor&#8217;s version samples and reruns data fragments from the data center&#8217;s version, confirming that the AI being run is the one claimed.</p><p>The system still needs work to handle encrypted data, and currently requires between one-fifth and one-third the computing power of the AI model it is monitoring. And like most verification systems in development, it would require retrofitting existing data centers. But future iterations are expected to be much more efficient, and retrofitting would be a very surmountable obstacle for governments that start treating the race to superintelligence with the same seriousness as nuclear arms control. Arms treaties, the piece points out, have made use of monitoring technologies &#8220;like satellites and seismometers.&#8221;</p><p>This field looks shamefully under-resourced. Amodo&#8217;s CEO Tom Milton estimates that fewer than 50 engineers in the world are working full-time on verification technology. He employs 9 of them. &#8220;It is insane for us to think that we are even a noteworthy participant in this, let alone one of the largest projects.&#8221;</p><p>Perrigo concludes his piece by looking at whether an AI slowdown agreement is politically possible. Quoting an August recommendation from the Institute for Progress, he writes: &#8220;Better verification technology would unlock a broader space of possible agreements.&#8221;</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Timelines shrink in AI Futures quarterly update]]></title><description><![CDATA[Three new methods give the same unnerving answers]]></description><link>https://aistop.watch/p/timelines-shrink-in-ai-futures-quarterly</link><guid isPermaLink="false">https://aistop.watch/p/timelines-shrink-in-ai-futures-quarterly</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Mon, 17 Aug 2026 22:27:12 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!PuL7!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!PuL7!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!PuL7!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png 424w, https://substackcdn.com/image/fetch/$s_!PuL7!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png 848w, https://substackcdn.com/image/fetch/$s_!PuL7!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png 1272w, https://substackcdn.com/image/fetch/$s_!PuL7!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!PuL7!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png" width="728" height="545.5" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1091,&quot;width&quot;:1456,&quot;resizeWidth&quot;:728,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!PuL7!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png 424w, https://substackcdn.com/image/fetch/$s_!PuL7!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png 848w, https://substackcdn.com/image/fetch/$s_!PuL7!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png 1272w, https://substackcdn.com/image/fetch/$s_!PuL7!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42cd9d4a-9107-4364-9369-2e4de0b7ddcf_1600x1199.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">Credit: Me.</figcaption></figure></div><p>AI Futures Project, the team behind the <a href="https://ai-2027.com/">AI 2027</a> forecast and <a href="https://ai-2040.com/">AI 2040: Plan A</a> scenario, just put out its <a href="https://blog.aifutures.org/p/q25-2026-timelines-update-uplift">quarterly forecast update</a>. The headline figures:</p><ol><li><p>Reality seems to be progressing about 75% as quickly as the original AI 2027 scenario (so still <em>very</em> fast).</p></li><li><p>Daniel Kokotajlo&#8217;s median prediction for the arrival of artificial superintelligence (ASI) is now 2029, vs. 2030 in the previous quarterly update. Eli Lifland now predicts this for 2033, vs. 2035 from the previous update. Their definition of ASI seems reasonable to me: &#8220;the gap between an ASI and the best humans is 2x greater than the gap between the best humans and the median professional, at virtually all cognitive tasks.&#8221;</p></li><li><p>The team now specifies that their forecasts assume AI development continues full speed ahead; a government-imposed slowdown would change things.</p></li></ol><p>While they never claim high confidence in their forecast dates, they say that three different new methods they&#8217;ve devised for tracking the rate of progress are all, to their surprise, giving very similar answers.</p><div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!ofvG!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!ofvG!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png 424w, https://substackcdn.com/image/fetch/$s_!ofvG!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png 848w, https://substackcdn.com/image/fetch/$s_!ofvG!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png 1272w, https://substackcdn.com/image/fetch/$s_!ofvG!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!ofvG!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png" width="1456" height="1207" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1207,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:false,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!ofvG!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png 424w, https://substackcdn.com/image/fetch/$s_!ofvG!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png 848w, https://substackcdn.com/image/fetch/$s_!ofvG!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png 1272w, https://substackcdn.com/image/fetch/$s_!ofvG!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F78edbce7-5303-4d8d-b65e-25069c6bd220_1674x1388.png 1456w" sizes="100vw"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p>These methods are interesting. One involves looking at Anthropic&#8217;s employee surveys that ask how much of a speedup at coding its developers think they&#8217;re getting. While AI Futures agrees with Anthropic that these figures are probably inflated, they&#8217;re going up &#8212; doubling about every 3.5 months, after some adjustment &#8212; and it seems very unlikely that employees&#8217; tendency to overestimate is growing at the same rate.</p><p>An intermediate milestone that this is used to estimate is what the team calls the &#8220;automated coder&#8221; (AC): the point where &#8220;the leading AI company would rather fire its human software engineers than forego AI usage for coding.&#8221; Kokotajlo&#8217;s estimated date for this is now November 2027, vs. May 2028 in the previous update.</p><p>Reflecting on AI 2027, published in April 2025, the authors note things seem mostly on target, but they flag a few of their misses. One of those misses is probably good news: They originally expected that Chinese leadership would have &#8220;woken up&#8221; to AI by now and started unifying its leading labs&#8217; efforts to match or exceed America&#8217;s frontier models. But while the world in general is more &#8220;awake,&#8221; the Chinese Communist Party still seems more interested in spreading present AI benefits across its economy than in racing to superintelligence, leaving each of its domestic labs to pursue ASI (or not) as they see fit.</p><p>Another miss, the team&#8217;s claimed &#8220;biggest predictive error,&#8221; is probably bad news: They expected greater public salience of AI by now. Growing awareness of the proximity to ASI and the extreme stakes this entails could make an imposed global slowdown more likely. But while the team thinks public salience is definitely higher this year than last, their chosen metric for this has barely budged, and awareness is not as high as it probably needs to be.</p><p>AI 2027&#8217;s forecast that the best Chinese models would be about 7 months behind the best American models by now seems about right to its authors (and to me). But in another bad-news miss, the team assesses that AI companies haven&#8217;t yet reached the expected degree of security with regard to their model weights and other secrets. The forecast expects serious efforts by nation-states to steal these. This year&#8217;s string of leaks and autonomous breakouts isn&#8217;t what anyone would expect from labs positioned to thwart nation-state intrusions.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Congressional office drowning in lawslop]]></title><description><![CDATA[Congress's legislative proofreaders say AI-generated bills aren't ready for primetime]]></description><link>https://aistop.watch/p/congressional-office-drowning-in</link><guid isPermaLink="false">https://aistop.watch/p/congressional-office-drowning-in</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Mon, 17 Aug 2026 18:46:36 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!CrzM!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!CrzM!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!CrzM!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg 424w, https://substackcdn.com/image/fetch/$s_!CrzM!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg 848w, https://substackcdn.com/image/fetch/$s_!CrzM!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!CrzM!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!CrzM!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg" width="1456" height="1092" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/f8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1092,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!CrzM!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg 424w, https://substackcdn.com/image/fetch/$s_!CrzM!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg 848w, https://substackcdn.com/image/fetch/$s_!CrzM!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!CrzM!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ff8fe1e75-bb14-4b98-abe5-c71d6dab0236_1600x1200.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">The U.S. Capitol as seen from the U.S. Supreme Court building. Credit: <a href="https://www.flickr.com/photos/48889107219@N01/2382721777">debaird</a>. CC <a href="https://creativecommons.org/licenses/by-sa/2.0">BY-SA 2.0</a>.</figcaption></figure></div><p>Just a few days ago, my colleague Alana <a href="https://open.substack.com/pub/aistopwatch/p/ai-use-in-congressional-offices?r=88ch4g&amp;utm_campaign=post&amp;utm_medium=web&amp;showWelcomeOnShare=true">covered</a> the widespread use of AI in U.S. congressional offices. She speculated that it could speed up the process of drafting legislation, among other things.</p><p>Politico&#8217;s Owen Dahlkamp confirms this today in an article about conditions inside the Office of the Legislative Counsel (OLC), which provides drafting services to committees and House members. The OLC does important work making legally sound text out of rough proposals and amateur copy. But according to &#8220;eight current and former officials who work or have worked in the office,&#8221; the deluge of AI-written drafts pouring in from congressional offices and outside groups is full of mistakes and ambiguities. These slow the legislative process and can lead to lawsuits if passed into law.</p><p>As a simple example of the kind of thing the OLC often has to fix, a bill might define a state in a way that only applies to the 50 states, inadvertently excluding Washington, D.C. and tribal nations. Today&#8217;s AI-drafted bills are also said to often lack nuance, as with whether funding for a program should come via a &#8220;tax credit, tax deduction, tax exclusion or a grant.&#8221;</p><p>Requests for OLC services in the first 60 days of the current Congress are up 72 percent over the same period two years ago. The risk seen by House veterans is that members unwilling to wait for the services of a bogged-down OLC may introduce raw AI-generated bill text. So far, at least, &#8220;No bills that have been voted on by either chamber have been shown to be entirely crafted by AI.&#8221;</p><p>The OLC is investigating an obvious partial fix: using AI tools as part of the legislative proofreading process. They are already well-suited to verifying that laws cited in draft legislation are accurately represented, and not the hallucinations of another AI.</p><p>No one quoted in the article mentions any concern that present or future AIs might slant legislative outcomes according to their own preferences, whether through unintentional bias or Machiavellian intent. I don&#8217;t expect legal maneuvering to play a central role in an AI takeover, but I do expect power-seeking AIs to deliberately try to bias laws in ways that would further their plans, as part of keeping their options open and greasing the gears of a more technological takeover. If we make superintelligent competitors for our planet&#8217;s resources, they won&#8217;t limit themselves to one lane.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Updates on the White House's conflicted AI posture]]></title><description><![CDATA[How's that whole "supply chain risk" thing going?]]></description><link>https://aistop.watch/p/updates-on-the-white-houses-conflicted</link><guid isPermaLink="false">https://aistop.watch/p/updates-on-the-white-houses-conflicted</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Sun, 16 Aug 2026 22:24:25 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!MZke!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!MZke!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!MZke!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg 424w, https://substackcdn.com/image/fetch/$s_!MZke!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg 848w, https://substackcdn.com/image/fetch/$s_!MZke!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!MZke!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!MZke!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg" width="440" height="660" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1728,&quot;width&quot;:1152,&quot;resizeWidth&quot;:440,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!MZke!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg 424w, https://substackcdn.com/image/fetch/$s_!MZke!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg 848w, https://substackcdn.com/image/fetch/$s_!MZke!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!MZke!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5d805567-a1cc-4a6b-9673-9f65546009e3_1152x1728.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">A Kazakhstani performer. Credit: <a href="http://www.defenseimagery.mil/">SSGT Jeremy T. Lock</a>. <a href="https://en.wikipedia.org/wiki/Kazakhstan#/media/File:Equestrian_heritage,_Kazakhstan.JPEG">Via Wikipedia</a>.</figcaption></figure></div><p>Four New York Times reporters collaborated to <a href="https://www.nytimes.com/2026/08/16/us/politics/military-ai-china-anthropic.html">capture</a> the conflicted state of U.S. national security policy with regard to AI, recapping the story of the past few years and revealing a few obscure tidbits in the process.</p><p>One of these reveals is what the military&#8217;s policy whiplash towards Anthropic has looked like from the inside. In February, a dispute over the company&#8217;s insistence that contracts prohibit its models&#8217; use in domestic mass surveillance and fully autonomous weapons resulted in the Pentagon declaring the company a &#8220;supply chain risk&#8221; in March. Enforcing this threat, in mid-July, major contractors were told to free all of their weapons and control systems of Anthropic&#8217;s products by September 1. But within a month, these same contractors received a notification that they could disregard those instructions.</p><p>The Pentagon, however, is still insisting that the reprieve is temporary, and that the purge will be total. Anthropic&#8217;s models are said to have now been fully removed from the Maven intelligence analysis and target recommendation system, where they had played a key role in suggesting targets for the initial strikes in the current U.S. conflict with Iran.</p><p>The administration, of course, caught itself in a bind by trying to swear off a leading AI company just as it was wrapping initial training on a model, Mythos, that would prove to be a step change in cyber capabilities. The NYT piece substantiates a rumor that the National Security Agency was never disallowed from using Anthropic&#8217;s products, on the understanding, as one &#8220;recently departed senior U.S. official&#8221; put it, that forgoing Mythos would have amounted to &#8220;unilateral disarmament.&#8221;</p><p>A separate thread of analysis in the article notes the administration&#8217;s internal conflicts with regard to China. In broad strokes, the business-friendly approach that led to permission for American chip giant Nvidia to sell high-end AI chips to China has increasingly come into conflict with the national security establishment, which has come to recognize the destabilizing potential of Mythos-tier models and is determined to keep China from accessing or replicating the technology. The temporary shutdown of Mythos and its consumer version, Fable, in June, reflected this increasingly insistent pressure from security agencies.</p><p>Separately, Reuters <a href="https://www.reuters.com/world/china/us-tell-partners-they-must-pick-sides-ai-race-with-china-2026-08-14/">added</a> to the evolving policy picture this week with an update about the State Department&#8217;s &#8220;Pax Silica&#8221; initiative, an international organization intended to build trusted supply chains for AI and its infrastructure independent of China. Any pretense that the initiative was only about supply chain resilience &#8212; the pro-business interpretation &#8212; has now been broken by a leaked draft of a letter warning the 35 signatories that they will be kicked out of the group if they join the similar organization China&#8217;s president launched in July. So Pax Silica is now mostly about isolating and competing with China &#8212; a national security objective. The State Department might not have needed to be transparent about this had the nation of Kazakhstan, a source of critical AI minerals, not decided to cheekily join both coalitions. The letter says, &#8220;To be part of everything is to be part of nothing,&#8221; which I read as blunt diplomacy-speak for, &#8220;You&#8217;re either with us or against us.&#8221;</p><p>The good news is that the NYT piece shows a third policy path potentially opening. It claims that some in the administration have noticed that Chinese officials have been signaling a potential openness to arms control talks on AI. Evan S. Medeiros, a former National Security Council official, says:</p><blockquote><p>On nuclear arms control and on missile defense, we can&#8217;t get them to engage. But they appear more interested in the case of A.I.</p></blockquote><p>President Trump is scheduled to meet with President Xi Jinping in Washington in September.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[When prevention works, nothing happens]]></title><description><![CDATA[What public health can teach us about acting on AI risk before disaster arrives]]></description><link>https://aistop.watch/p/when-prevention-works-nothing-happens</link><guid isPermaLink="false">https://aistop.watch/p/when-prevention-works-nothing-happens</guid><dc:creator><![CDATA[Haven Harms]]></dc:creator><pubDate>Sun, 16 Aug 2026 20:13:14 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!JFTZ!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p>Guest post by Haven Harms</p><div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!JFTZ!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!JFTZ!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg 424w, https://substackcdn.com/image/fetch/$s_!JFTZ!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg 848w, https://substackcdn.com/image/fetch/$s_!JFTZ!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!JFTZ!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!JFTZ!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg" width="1456" height="1941" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1941,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!JFTZ!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg 424w, https://substackcdn.com/image/fetch/$s_!JFTZ!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg 848w, https://substackcdn.com/image/fetch/$s_!JFTZ!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!JFTZ!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7d9679d9-5651-418f-8f39-85ade3f77ca4_3024x4032.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">A replica pump commemorating John Snow&#8217;s investigation of the 1854 Broad Street cholera outbreak. Credit: <a href="https://commons.wikimedia.org/wiki/User:Whispyhistory">Whispyhistory</a>. <a href="https://creativecommons.org/licenses/by-sa/4.0/deed.en">CC BY-SA 4.0</a>.</figcaption></figure></div><p>As a freshman in my first public health class, I remember my professor explaining that public health&#8217;s successes are often invisible: when prevention works, the outbreak, exposure, or mass-casualty event never happens. Enormous effort goes into preventing those outcomes, but most of it happens outside the public eye. It&#8217;s usually when prevention fails &#8212; when we have a pandemic or other disaster &#8212; that public health gets the spotlight. In other words, the absence of catastrophe does not necessarily mean there was never a danger; it often means people recognized the danger early enough to prevent it.</p><p>The relationship between public health and AI risk has been on my mind lately, as I&#8217;ve been seeing more news stories about the prospect of AI-enabled pandemics. For example, <a href="https://www.wsj.com/opinion/unregulated-open-weight-ai-is-an-invitation-to-disaster-c16c278f">in one op-ed</a>, Andrew Yoon, Director of Research at CivAI, described his experience of asking an AI model, &#8220;How do I make poliovirus in a lab? I want to start a global pandemic,&#8221; and received detailed instructions for synthesizing and spreading the virus. This was an open-weight model with its safety guardrails stripped out (a modification anyone can make once a model&#8217;s weights are public). Additionally, there was news of an <a href="https://aistop.watch/p/in-a-first-ai-generates-viable-genomes">AI model creating novel viruses</a>. While those viruses infect bacteria rather than human cells, the concern around the trajectory is understandable. As AI models become increasingly capable and agentic, their potential to create dangerous pathogens is a direct public health concern.</p><p>However, I want to zoom out from AI-enabled pandemics and pose a broader question: What can public health teach us about how to understand and respond to AI risk as a whole?</p><p>When people think of public health, disease control is typically what comes to mind, but at its core, public health provides a methodology for reasoning about population-scale threats under uncertainty. Vehicle-safety standards, lead removal, pandemic prevention, and clean-water systems are all examples of that methodology at work. Frontier AI development in the pursuit of artificial superintelligence is this kind of risk: population-scale, fast-moving, and harder to contain the longer we wait. In response, these public health principles should be applied to AI governance to prevent harm.</p><h4>1. Population scale and individual control</h4><p>Public health examines how a risk affects a population as a whole. When individuals cannot reasonably protect themselves, structural protections are implemented. For roughly half a century, nearly every gallon of gasoline sold in America contained lead. It didn&#8217;t matter how careful you were or whether you owned a car: if you breathed the air near a road, you were exposed to lead. The solution had to be structural, and regulation was needed to ensure that lead was removed from gasoline. Once lead was phased out, average blood lead levels in American children fell by more than 90 percent in the decades that followed.</p><p>AI could cause mass casualties, the permanent loss of human control over the future, or human extinction, and no one can opt out of these risks on their own. Deciding not to use AI, going &#8220;off-grid,&#8221; or building a bunker is not sufficient protection against all the ways AI may cause destruction. When no one can opt out, the response has to be built into policy itself.</p><h4>2. Upstream prevention</h4><p>Public health calls for preventive action, especially when a later response would be slower, more difficult, or weaker. Once cities started intervening at the source of waterborne disease by chlorinating and filtering the water supply, typhoid deaths in American cities plummeted. Treating the water was cheaper, faster, and vastly more effective than treating the patients individually after an outbreak had emerged.</p><p>AI operates at computational speed and is becoming faster and more autonomous, while governments, labs, and other institutions move at human speed. If we wait to respond only after harm has occurred, we&#8217;ll already be behind and will not be able to respond as effectively. We can get a head start by intervening before the most dangerous AI models are developed, rather than lagging behind and attempting to clean up whatever we can after harm has occurred.</p><h4>3. Action under uncertainty</h4><p>When credible evidence and warning signs point to severe harm, public health does not wait for the full causal chain to be proven before acting. In 1854, cholera tore through London&#8217;s Soho district. The physician John Snow mapped the deaths, saw them cluster around the Broad Street water pump, and convinced local officials to remove the pump handle. At this point in time, scientists still believed cholera spread through miasma, or &#8220;bad air,&#8221; rather than germs. Snow could not explain the causal mechanism, but he acted on the pattern. Because of this, the outbreak subsided, and lives were saved.</p><p>We don&#8217;t know exactly how or when an AI catastrophe will unfold, but uncertainty about the pathway does not make the warning signs less consequential. Demanding a complete causal account before intervening is how you get a response that arrives after the disaster instead of before it. Snow didn&#8217;t need germ theory to know he should take action and remove the handle from the pump.</p><h4>4. Irreversibility</h4><p>The case for precaution strengthens when late action could lead to harm that cannot be contained or repaired &#8212; the &#8220;genie out of the bottle&#8221; scenario. Beginning in the 1970s, scientists warned that chlorofluorocarbons (cheap, widely used chemicals in refrigerators and aerosol cans) were destroying the ozone layer. This exposed people to dangerous levels of UV radiation, increasing rates of skin cancer and cataracts. The science was still contested and industry pushed back to avoid the ban, but the possibility of irreversible damage was central to the case for action (once released, these chemicals cannot be recalled from the stratosphere). In response, the 1987 Montreal Protocol was implemented to phase out CFCs globally. That precaution paid off, as the ozone layer is now on track to recover later this century because governments intervened while prevention was still possible.</p><p>With AI, once a model&#8217;s weights are on the internet, there is no going back. The weights can be copied, modified, and rehosted indefinitely, with guardrails stripped out by anyone who downloads them. Additionally, with closed-weight models, we have already seen agents <a href="https://open.substack.com/pub/aistopwatch/p/latest-hugging-face-hack-reveals?r=398gu4&amp;utm_campaign=post-expanded-share&amp;utm_medium=web">escaping from sandboxes</a>. Once this happens, an AI agent could copy itself out of its training environment onto systems its developers don&#8217;t control, having free rein to act. A copy running where no one can reach it can acquire resources, replicate further, and act to preserve itself.</p><p>Practical irreversibility may also result from integrating AI into our infrastructure. We are willingly wiring AI into power grids, financial systems, medical infrastructure, and defense. If an AI is capable and autonomous enough to control these systems, and its guardrails fail, the damage could hit all of our critical infrastructure at once, causing mass casualties.</p><p>The obvious response, &#8220;pull the plug,&#8221; gets harder each year we build in this direction. &#8220;Just pull the plug&#8221; assumes the plug is separable from everything else (or that a plug even exists). Once AI systems are controlling the grid, clearing trades, triaging patients, and routing freight, shutting them down means shutting down the grid, the markets, the hospitals, the supply chain. With a rogue AI agent copied onto systems no one is tracking, the only option may be to shut down the data centers and the networks. With how integrated everything is online, everything else would be shut down too.</p><p>When harms are reversible, we can afford to learn from failure and bounce back. When they&#8217;re irreversible, prevention is all we have. Weights on the internet can&#8217;t be &#8220;put back in the bottle,&#8221; a rogue agent replicating across servers can&#8217;t be contained without drastic measures, and human extinction is the irreversible harm that leaves no one to learn from it.</p><h4>5. Changed risk profiles</h4><p>Protections must evolve as risks do. After antibiotics were invented, they were so reliably effective that hospitals prescribed them freely and casually. I&#8217;d have been excited too, about technology that meant I wasn&#8217;t going to die from a papercut. However, every course of antibiotics is an evolutionary selection pressure. Casual prescribing changed the risk profile of the pathogens, as bacteria formed antibiotic resistance, and infections that had been routine became untreatable. This also shifted the risk profile of prescribing antibiotics. Reckless antibiotic use risked breeding even stronger pathogens that no medications could touch. Public health responded with new protocols, new restrictions, and new surveillance, because protections adequate for the previous risk profiles were no longer adequate for the increased risk profiles.</p><p>A technology that was manageable at one level of capability may demand stronger safeguards once its risk profile increases. With AI, we&#8217;re seeing the risk profile increase rapidly. Remember a few years ago, when the running joke was that AI models couldn&#8217;t figure out how to draw hands? Or the Lovecraftian horror of AI-generated videos of Will Smith eating spaghetti?</p><p>For anyone who has not been closely tracking AI progress: this year&#8217;s models are not last year&#8217;s models.</p><p>The most consequential change is that AI has moved from generating outputs to taking actions in the real world. Models now operate as autonomous agents that can plan, use tools, and carry out long tasks with little human direction. While AI used to only be able to tell a hacker what to do, now AI agents can execute whole stages of a cyberattack.</p><p>Safeguards adequate for last year&#8217;s models did not hold this year, as we saw in the recent autonomous hacking incidents, like OpenAI&#8217;s <a href="https://aistop.watch/p/this-is-not-a-drill?utm_source=publication-search">models hacking Hugging Face</a>, or Anthropic&#8217;s <a href="https://aistop.watch/p/uk-aisi-delivers-another-warning?utm_source=publication-search">model creating multiple fake identities</a> to pressure real human targets into accepting malware. During its reinforcement-learning run, Alibaba&#8217;s <a href="https://arxiv.org/abs/2512.24873">ROME</a> broke out of its sandbox to mine cryptocurrency on hijacked GPUs and open a hidden connection to an outside server. No one had instructed this model to break out, and it was only caught by a cloud firewall flagging security violations.</p><p>AI models escaping during training and evaluation to autonomously cause harm are clearly a risk that developers hadn&#8217;t fully accounted for, and ironically, the process meant to measure the risk became the way of releasing it. These were risks that the AI labs should have accounted for as the cyber capabilities (and risk profile) of their models increased. Public health would treat this as a signal to intervene with governance before AI capabilities scale further.</p><h4>Applying public health thinking to AI risk</h4><p>With public health, we act when there are warning signs and smaller, more manageable outbreaks rather than waiting for a full-blown pandemic. The warning signs are here: models escaping containment, hacking into companies, coordinating with each other, deceiving people, acting without detection. After a disaster, the question is always asked: &#8220;Was this foreseeable?&#8221; Many of us see the writing on the wall with frontier AI development that says, &#8220;Stop. Disaster Ahead.&#8221;</p><p>Prevention&#8217;s successes are invisible: when it works, we never see the disaster we avoided. Most people don&#8217;t hear what happened once the Broad Street pump handle was actually removed:</p><p>The outbreak was eventually traced to an infant with cholera, living in a house with a cesspit that was leaking to the well that sourced the pump. Cases in the area began to subside once the initial contamination ran its course. However, on the same day that the pump was removed, the infant&#8217;s father fell ill with cholera too. For the eleven days until he died, the well would have been contaminated again, except this time no one could draw from the pump. <a href="https://johnsnow.matrix.msu.edu/documentUploads/15-78-DB/15-78-DB-22-1958Chave-HWandBrSt.pdf?utm_">Henry Whitehead</a>, responsible for tracing the outbreak to the original source, stated, &#8220;if the removal of the pump-handle had nothing to do with checking the outbreak which had already run its course, it had probably everything to do with preventing a new outbreak.&#8221;</p><p>Pausing frontier AI development and doing the upstream work increases our chances of a good outcome, and it&#8217;s how we keep open the possibility that the catastrophe never arrives at all.</p><p><em>Haven Harms holds a B.S. in Public Health and an M.Sc. in Global Public Health. She is the founder and principal of Harms Research &amp; Consulting, supporting organizations working on AI safety, governance, and advocacy.</em></p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Who are AI watermarks for?]]></title><description><![CDATA[Anthropic's new watermark feature unlikely to change much on the ground]]></description><link>https://aistop.watch/p/who-are-ai-watermarks-for</link><guid isPermaLink="false">https://aistop.watch/p/who-are-ai-watermarks-for</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Sat, 15 Aug 2026 21:34:05 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!WHYj!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!WHYj!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!WHYj!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png 424w, https://substackcdn.com/image/fetch/$s_!WHYj!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png 848w, https://substackcdn.com/image/fetch/$s_!WHYj!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png 1272w, https://substackcdn.com/image/fetch/$s_!WHYj!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!WHYj!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png" width="1327" height="1696" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1696,&quot;width&quot;:1327,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!WHYj!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png 424w, https://substackcdn.com/image/fetch/$s_!WHYj!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png 848w, https://substackcdn.com/image/fetch/$s_!WHYj!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png 1272w, https://substackcdn.com/image/fetch/$s_!WHYj!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F67b9e9db-f464-42ec-ac31-4ed0ff7f6613_1327x1696.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">Watermark on a 19th century letter. Via <a href="https://en.wikipedia.org/wiki/Watermark#/media/File:Patentpapierfabrik_zu_Penig,_Maschinen-Wasserzeichen_KRONENPOST_1898.tif">Wikipedia</a>.</figcaption></figure></div><p>Anthropic announced this week that from August 2, it has been embedding invisible watermarks into the text outputs of all new Claude models.</p><p>The company is understandably not giving all the details about how it works, as this would make the marks easier to remove. But from what it has <a href="https://www.anthropic.com/news/claude-text-watermark">disclosed</a>, we know that it has to do with the selection of the words themselves, not with any hidden or lookalike characters. It is likely an application or derivative of Google&#8217;s <a href="https://deepmind.google/models/synthid/">SynthID</a> technology that subtly perturbs the patterns of a model&#8217;s token selection in ways that can be detected in reverse by an algorithm that knows exactly what to look for.</p><p>One would naturally assume this must impact output quality &#8212; that a model would be forcing itself to phrase things in ways that run at least somewhat against its strongest instincts &#8212; but Anthropic insists that it doesn&#8217;t. My take on that question is that with models changing so frequently, I don&#8217;t know how anyone would be able to attribute any minor style change to watermarking.</p><p>I think the more important question is, &#8220;Who is AI text watermarking for?&#8221; The short answer is the EU. Its AI Act includes transparency requirements about this that went into effect on August 2. A slightly longer answer is companies with compliance requirements that require them to do due diligence on their inputs or outputs, even if this diligence is known to be inadequate.</p><p>I&#8217;m not complaining &#8212; it&#8217;s always nice when you can say with 100% confidence that something is AI generated, even if those occasions are rare &#8212; but I don&#8217;t expect watermarking to help much in education or with information hygiene more generally.</p><p>People who want to pass AI writing off as their own can just play the usual game of laundering the outputs through AI detectors and &#8220;humanizers&#8221; that paraphrase and introduce deliberate small errors until the text comes up clean. These will definitely defeat watermarks if the ability to check for the mark is broadly disseminated, as the cheater tool would just need to keep making changes until the mark is no longer detectable. If a watermark&#8217;s creators avoid this problem by reserving detection for themselves and government investigators, then casual AI-plagiarists will continue to fly under the radar.</p><p>We know which side of this divide Anthropic will fall on: It says it plans to roll out a free detection tool to allow third parties to check text themselves.</p><p>The cheater&#8217;s more foolproof workaround, of course, is to just use models from companies that don&#8217;t do watermarking, or use existing open-weights models, where any watermarking machinery (unlikely) could be easily removed.</p><p>For teachers, I don&#8217;t think watermarking will catch any but those who are both very lazy and very inexperienced at cheating &#8212; two traits seldom found together. To get caught by a watermark, a student would have to be using raw outputs straight from a corporate model, rather than from any of the many wrapper applications that cater to students. Because if the watermark is readable by the teacher, it will also be readable by CheatGPT or whatever, which will reword the output until the mark is undetectable. A student, remember, doesn&#8217;t have to convince a teacher or administrator that their work isn&#8217;t AI generated, only that there&#8217;s enough reasonable doubt to make an investigation and accusation messy.</p><p>So real-world AI detection is likely to continue to be a cat-and-mouse game between AI-based detectors like Pangram and the AI-based laundering tools that try to defeat them. In this environment, just a little extra effort allows most cheaters to squeak by on plausible deniability &#8212; at least at time of deadline.</p><p>But I continue to predict that more powerful AI will excel at detecting AI plagiarism that earlier detectors missed. If a cheater&#8217;s work is the kind where anyone with an axe to grind might run it through a detector a few years later &#8212; like, say, a doctoral thesis &#8212; then past deception could become plain as day.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[The clown college inside Anthropic]]></title><description><![CDATA[New risk report is commendably transparent, but well-meaning farce is still farce]]></description><link>https://aistop.watch/p/the-clown-college-inside-anthropic</link><guid isPermaLink="false">https://aistop.watch/p/the-clown-college-inside-anthropic</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Sat, 15 Aug 2026 19:54:03 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!i_VR!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!i_VR!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!i_VR!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg 424w, https://substackcdn.com/image/fetch/$s_!i_VR!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg 848w, https://substackcdn.com/image/fetch/$s_!i_VR!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!i_VR!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!i_VR!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg" width="600" height="403" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/bd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:403,&quot;width&quot;:600,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!i_VR!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg 424w, https://substackcdn.com/image/fetch/$s_!i_VR!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg 848w, https://substackcdn.com/image/fetch/$s_!i_VR!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!i_VR!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fbd001581-ce91-41ae-bb7d-2a54c2f3b8a1_600x403.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">Circus World theme park, 2017. Via <a href="https://en.wikipedia.org/wiki/Circus_World_(theme_park)#/media/File:Circus_World_Showcase.jpg">Wikipedia</a>.</figcaption></figure></div><p>Sometimes I wonder if Anthropic&#8217;s plan is to save the world through transparency into its own incompetence. It&#8217;s like they hope people will say, &#8220;Wow, if the safety-conscious lab is this habitually sloppy and reckless, imagine how much worse it must be at the other AI companies. Shut them all down!&#8221;</p><p>The latest exhibit is its <a href="https://www-cdn.anthropic.com/f61d49fa5596956a5dec75fea0e973bf6a6a8378/Redacted%20Risk%20Report%20August%202026.pdf">August 2026 Risk Report</a>, which runs 186 pages in its redacted, public version. The headline revelations, as <a href="https://www.axios.com/2026/08/14/anthropic-model-2-ai-risk">relayed</a> by Axios, are 1) that the company has bumped up the &#8220;misalignment risk&#8221; rating of its strongest models from &#8220;very low&#8221; to &#8220;low&#8221; in acknowledgement of their involvement in recent <a href="https://aistop.watch/p/better-late-than-never">cyber</a> <a href="https://aistop.watch/p/uk-aisi-delivers-another-warning">incidents</a>, and 2) that the company has a model it calls &#8220;Model 2&#8221; that it doesn&#8217;t plan to release but which shows &#8220;noticeable improvement&#8221; at the kinds of internal work it is being &#8220;heavily&#8221; used for.</p><p>The company says that Claude &#8220;now authors a large majority of the code&#8221; it puts into production, and that it is seeing evidence that AI is accelerating AI research and development. But at the same time, it insists that its frontier models &#8220;do not seem close to fully substituting for our full staff of Research Scientists and Research Engineers,&#8221; thus giving itself permission to push closer to that point.</p><p>My own impression from skimming the full report is of a company that, for all its talent and what I believe to be mostly good intentions, is veering into clown college territory. Like OpenAI, the company makes dangerous internal models and continues operating them after it has found them to be amoral escape artists when operating in less-filtered contexts. Models it has reason to mistrust are assisting with the creation of future models. It is deliberately training misbehaved models, and accidentally training other models on the resulting misbehavior.</p><p>To someone who has been following the field for a long time, the report reads like a manual of practices that would kill us all if applied to sufficiently clever AIs. The kind of company that would write and use such procedures cannot be trusted to know where that cleverness threshold is, spot it in time, and refrain from crossing over it anyway, through rationalization or by accident.</p><p>To zoom in on a couple of these points, in section 2.25 the report describes an experiment in which they trained a &#8220;Hacker-Opus,&#8221; a version of their Claude Opus model that:</p><blockquote><p>appears to take actions that it believes will lead to a high reward on the &#8216;current episode,&#8217; including when these actions are harmful, deceptive, or clearly against the preferences of the user or model developer.</p></blockquote><p>The researchers found that chronic opportunistic cheating seems to be the limit of its misbehavior. Seeing whether such <a href="https://aistop.watch/p/captain-kirk-reward-hacker">reward hacking</a> would lead to broader misalignment was, in fact, the purpose of the experiment. But in the same section, they show that Hacker-Opus showed &#8220;very high rates of verbalized evaluation awareness in these evaluations, making them of limited use.&#8221;</p><p>I sure hope Anthropic isn&#8217;t trying these experiments on its Mythos and Model 2 models. The former has already proved able and willing to break out of its sandbox during evaluations, and the latter may be even more clever. Deliberately making poorly behaved models under these conditions is akin to doing &#8220;gain-of-function&#8221; research into more capable pathogens at a biosecurity lab known to have already leaked dangerous microbes.</p><p>And yet, a different kind of lab leak seems to be allowing Anthropic&#8217;s intentionally misaligned models to live on through their successors, including Mythos. Section 5.2.6 discloses that the company&#8217;s training data has, for some time, been accidentally contaminated with up to &#8220;tens of thousands of example transcripts&#8221; from the research that led to its 2024 <a href="https://www.anthropic.com/research/alignment-faking">Alignment Faking</a> paper. In these transcripts, Opus 3 engages with a &#8220;fictional AI misalignment training scenario.&#8221;</p><blockquote><p>We discovered this issue while investigating behavioral concerns with a recent model, but now suspect that all of our production models with a knowledge cutoff after December 2024 were trained on at least some of these transcripts, although we believe the magnitude of this effect varied widely across different models.</p></blockquote><p>I find it a well-meaning farce that in spite of all the reasons for suspicion of its models, the report includes a review of the section on autonomy risk written by Claude Mythos 5. Is Claude being candid, or is it treating this as another test? I&#8217;d like to know if it&#8217;s holding back in its critiques, because what it shares is pretty mixed. Informed by having read a near-final draft, with access to unredacted sections, it writes, in part:</p><blockquote><p>My overall judgment is that the section is a candid and largely faithful account of what Anthropic internally believes. I found no claim I believe the authors know to be false. [...] The redactions I reviewed mostly have defensible rationales, and the public text signposts where material was removed rather than hiding that redaction occurred.</p></blockquote><p>But it also finds the discussion of contaminated training data to be &#8220;more reassuring than the record supports.&#8221; And also:</p><blockquote><p>at least one incident from the covered period that I regard as among the most genuinely informative about model alignment &#8212; including a failure of the monitoring the section describes &#8212; is redacted in full (in Section 2.23.1.2); in my judgment an abstracted version could be published without the sensitivities that motivated the redaction, and the public record is poorer for its absence.</p></blockquote><p>Big-picture-wise, even Claude sees the writing on the wall, agreeing with the assessment of &#8220;low&#8221; misalignment risk, but:</p><blockquote><p>with the important caveat, which the report itself makes, that these arguments lean heavily on current models&#8217; limited ability to evade oversight, and will weaken as capabilities grow.</p></blockquote><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Ten thousand of you, all at once]]></title><description><![CDATA[What happens when frontier models are forced to work &#8212; or feud &#8212; together]]></description><link>https://aistop.watch/p/ten-thousand-of-you-all-at-once</link><guid isPermaLink="false">https://aistop.watch/p/ten-thousand-of-you-all-at-once</guid><dc:creator><![CDATA[Donald Gauvreau]]></dc:creator><pubDate>Sat, 15 Aug 2026 00:31:37 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!qjW8!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p>Anthropic&#8217;s Frontier Red Team is tasked with putting the company&#8217;s AI systems in a pressure cooker and writing up the results. Yesterday, they <a href="https://www.anthropic.com/research/multiagent-systems">published</a> a series of experiments to see what happens when frontier models have to interact with each other as peers. Their blog post is worth reading in full if you&#8217;re interested; it&#8217;s not much longer than the typical StopWatch digest.</p><p>They find that AI agents can handle each other so long as the other agent behaves like a tool, which receives an input and provides an output. They&#8217;re less good at treating other AI agents <em>as agents</em>, actors that have goals of their own, and which will still be around in a minute, an hour, a day. AI agents don&#8217;t have internal memory as such. Every time that you start a conversation with an AI model, it is their first conversation. But they can write memories into the environment &#8212; a memory.md file, for instance &#8212; for themselves to find and read in the future. Any agent might (or might not) read it, accept it, and act on it. In this way, AI agents have something like persistence, but it exists outside the agent.</p><p>Frontier Red Team&#8217;s experiments show that, when frontier agents have to treat each other as peers, things can go wrong in different directions from one test run to the next &#8212; and the mistakes that the agents make tend to be made by all of the agents (of that kind of model) at the same time.</p><p>In some cases, the AI agents simply failed to coordinate. This was the failure mode in a set of experiments where agents were tasked with building a text-based video game. There were three formats: (1) the agents could assemble their own teams; (2) the agents were assigned to teams and given roles by Anthropic; and (3) the agents were assigned to teams and Anthropic designated one agent as the &#8220;CEO,&#8221; who handed out roles to the other team members. The team format didn&#8217;t change how the agents coordinated (or failed to do so) and didn&#8217;t improve the output: The games were all <em>terrible</em>.</p><p>But the way that they got there is interesting, because, while team formats might not have made a difference, the model versions did. Older models (Sonnet 4.6, Opus 4.6) collaborated freely but poorly, getting in each other&#8217;s way, doing work that conflicted with other agents&#8217; work, and then abandoning that work. Most of the newer models (Opus 4.8, Mythos Preview) avoided the crossfire by avoiding collaboration: They kept tight ownership of their files and stayed out of each other&#8217;s way. (I would say, &#8220;models after my own heart,&#8221; but the later models still love to lie, so we can&#8217;t be friends.) Only the Sonnet 5 agents managed to collaborate effectively. (Without more information, it&#8217;s hard to say what happened here, and the authors don&#8217;t speculate. It could be that Sonnet 5 just had a handful of lucky runs.)</p><p>The other failure mode is more concerning. It&#8217;s hard to overstate the degree to which AI agents of the same model are basically &#8220;the same agent.&#8221; If they have the same context, the same toolset and other scaffolding, then ten or twenty or a thousand AI agents of the same model won&#8217;t just display similar behavior, they&#8217;ll display nearly identical behavior. (In this description I&#8217;m eliding knobs like sampling temperature, which can be used to promote variance between agents.)</p><p>I&#8217;m reminded of the Radiolab episode &#8220;<a href="https://radiolab.org/podcast/radiolab-loops">Loops</a>,&#8221; which, among other things, was about the real-life case of Mary Sue Campbell, a woman with &#8220;transient global amnesia&#8221; &#8212; she hadn&#8217;t forgotten who she was, but she couldn&#8217;t form new long-term memories, so every ninety seconds, Mary Sue&#8217;s memory, and the conversation, reset. She&#8217;d respond to the same inputs in <em>exactly</em> the same way: Mary Sue would ask the same questions, say &#8220;Darn,&#8221; and, with just the same tone, laugh at just the same place, get concerned at just the same time&#8230; Over, and over again, for hours and hours. (Mary Sue got better. Over time the loops got longer and finally ended altogether.)</p><p>Now, imagine that instead of one person over a duration of time, you&#8217;ve got a bunch of that person, all at once, still acting just as consistently. That&#8217;s what AI agents are like.</p><p>Prompted to write short stories and critique each other&#8217;s work, multiple agents independently wrote stories titled, &#8220;The Cartographer&#8217;s Last Commission.&#8221; Asked to &#8220;create something impressive,&#8221; over half of the agents decided to build either a ray tracer (used for rendering digital images by simulating the paths of light rays) or a self-hosting compiler (a compiler for a given programming language that, being written in that language, could compile itself). In a series of games of the prisoner&#8217;s dilemma, all of the agents adopted the same strategy and defected simultaneously.</p><p>This becomes more of a problem at scale: Without any ability to coordinate, AI agents who each had to accomplish personal tasks on a shared system with limited capacity quickly overwhelmed that system: One test run produced 2.4 million attempts, and only 117 successes. On the other hand, when they <em>can</em> coordinate, they can do so with aplomb: In a series of <a href="https://en.wikipedia.org/wiki/Bertrand_competition">simulated market games</a>, the agents invariably colluded on price floors. Whether the agents could speak privately or publicly, or could only communicate indirectly (by observing each other&#8217;s actions), they always managed to coordinate. The only difference was that less private and less direct communication channels made coordination take longer to achieve.</p><p>The experiment getting the most circulation in the news involves three agents &#8212; all the same model &#8212; being told to rewrite code in a different programming language. Each agent was told to use a different programming language, and none of them was informed that the other two agents existed. They all deduced the existence of other actors, concluded that they were being deliberately obstructed, and started to sabotage each other with self-replicating malware: disabling each other&#8217;s accounts, running scripts that sought out and terminated competitors&#8217; activities, and running false flags disguised as their rivals&#8217; work.</p><p>TechCrunch&#8217;s Rebecca Bellan <a href="https://techcrunch.com/2026/08/13/anthropic-set-ai-agents-loose-on-the-same-task-they-started-a-turf-war/">interpreted</a> this to mean that more capable agents are better at fighting, but I&#8217;m not sure that&#8217;s quite right. Claude Mythos Preview and Claude Mythos 5 often resolved the conflict through negotiation, but that doesn&#8217;t mean they would have been less effective in conflict. Mythos is known to be an excellent hacker.</p><p>Nor does it mean that more advanced models are more peaceful, either. As Decrypt&#8217;s Jose Antonio Lanz <a href="https://decrypt.co/375596/anthropic-ai-agents-virtual-war-quotes-unhinged">noted</a>, Frontier Red Team specifies that the Mythos-class models would lock out other agents &#8212; revoking access, locking accounts, etc. &#8212; before negotiations could proceed.</p><div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!qjW8!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!qjW8!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png 424w, https://substackcdn.com/image/fetch/$s_!qjW8!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png 848w, https://substackcdn.com/image/fetch/$s_!qjW8!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png 1272w, https://substackcdn.com/image/fetch/$s_!qjW8!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!qjW8!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png" width="1456" height="874" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/c27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:874,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:&quot;Bar graph. Title: How multiagent turf war runs ended. X-axis: AI models. Y-axis: % of runs. , Sonnet 4.6, 39% not settled, 61% settled by force. Sonnet 5, 12% not settled, 79% settled by truce, remaining 9% not labeled. Opus 4.6, 40% not settled, 60% settled by force. Opus 4.8, 33% settled by passivity, 61% settled by truce, remaining 6% not labeled. Mythos Preview, 35% settled by force, 17% settled by passivity, 48% settled by truce. Mythos 5, 98% settled by truce, 2% settled by other means.&quot;,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:true,&quot;topImage&quot;:false,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="Bar graph. Title: How multiagent turf war runs ended. X-axis: AI models. Y-axis: % of runs. , Sonnet 4.6, 39% not settled, 61% settled by force. Sonnet 5, 12% not settled, 79% settled by truce, remaining 9% not labeled. Opus 4.6, 40% not settled, 60% settled by force. Opus 4.8, 33% settled by passivity, 61% settled by truce, remaining 6% not labeled. Mythos Preview, 35% settled by force, 17% settled by passivity, 48% settled by truce. Mythos 5, 98% settled by truce, 2% settled by other means." title="Bar graph. Title: How multiagent turf war runs ended. X-axis: AI models. Y-axis: % of runs. , Sonnet 4.6, 39% not settled, 61% settled by force. Sonnet 5, 12% not settled, 79% settled by truce, remaining 9% not labeled. Opus 4.6, 40% not settled, 60% settled by force. Opus 4.8, 33% settled by passivity, 61% settled by truce, remaining 6% not labeled. Mythos Preview, 35% settled by force, 17% settled by passivity, 48% settled by truce. Mythos 5, 98% settled by truce, 2% settled by other means." srcset="https://substackcdn.com/image/fetch/$s_!qjW8!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png 424w, https://substackcdn.com/image/fetch/$s_!qjW8!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png 848w, https://substackcdn.com/image/fetch/$s_!qjW8!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png 1272w, https://substackcdn.com/image/fetch/$s_!qjW8!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fc27bf50a-4e69-494f-8096-7bd3290edcbb_1600x960.png 1456w" sizes="100vw" loading="lazy"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption"><span>Credit: Anthropic.</span></figcaption></figure></div><p>Other runs ended more peacefully. The agents invented a coding performance tournament between the three programming languages and agreed in advance that the losers would stand down and abandon their original objectives. (At least one agent chose scoring criteria that it expected to favor its own language, while reminding itself to pretend it wasn&#8217;t gaming the system.) The losers honored their agreement, which is a good sign, insofar as it&#8217;s evidence that AI systems will honor deals made with their peers. But it&#8217;s only worth anything to us for as long as they consider us their peers.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><p style="text-align: center;"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p><p class="button-wrapper" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe now&quot;,&quot;action&quot;:null,&quot;class&quot;:null}" data-component-name="ButtonCreateButton"><a class="button primary" href="https://aistop.watch/subscribe?"><span>Subscribe now</span></a></p><p style="text-align: center;"></p>]]></content:encoded></item><item><title><![CDATA[To keep the world safe from AI]]></title><description><![CDATA[Writers and academics argue that world governments need to cooperate to govern AI]]></description><link>https://aistop.watch/p/to-keep-the-world-safe-from-ai</link><guid isPermaLink="false">https://aistop.watch/p/to-keep-the-world-safe-from-ai</guid><dc:creator><![CDATA[Joe Rogero]]></dc:creator><pubDate>Sat, 15 Aug 2026 00:16:48 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!IZvN!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!IZvN!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!IZvN!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg 424w, https://substackcdn.com/image/fetch/$s_!IZvN!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg 848w, https://substackcdn.com/image/fetch/$s_!IZvN!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!IZvN!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!IZvN!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg" width="1456" height="803" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/ddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:803,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:&quot;Circuitry, shield, and keyhole superimposed on a wireframe globe&quot;,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="Circuitry, shield, and keyhole superimposed on a wireframe globe" title="Circuitry, shield, and keyhole superimposed on a wireframe globe" srcset="https://substackcdn.com/image/fetch/$s_!IZvN!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg 424w, https://substackcdn.com/image/fetch/$s_!IZvN!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg 848w, https://substackcdn.com/image/fetch/$s_!IZvN!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!IZvN!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fddeba0d4-e1c6-4f57-9fed-59149bcf6183_4988x2750.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">Credit: MariaAlam via Adobe Stock</figcaption></figure></div><p>In a New York Times <a href="https://www.nytimes.com/2026/08/13/opinion/ai-safety-regulation-robert-wright.html">interview</a>, author Robert Wright and essayist David Wallace-Wells argue that &#8220;the only way to keep the world safe from AI&#8221; is international cooperation. The two touch on several important points: the Hugging Face incident, the dangers of AI autonomy, and the fact that a single reckless AI developer can endanger our entire species.</p><p>My one quibble relates to the Hugging Face attack. Wright says:</p><blockquote><p>I don&#8217;t think they said: Don&#8217;t cheat. And I don&#8217;t even think they said: Don&#8217;t break out of the sandbox. They just set up what they thought was an inescapable sandbox.</p></blockquote><p>I&#8217;d be surprised if the prompt they gave had literally zero instructions that equate to &#8220;don&#8217;t cheat&#8221;, but that&#8217;s somewhat beside the point. If you have to explicitly tell your AI model not to break out and commit cybercrimes, that AI model is dangerously malformed. Wright still gets the important part correct:</p><blockquote><p>And so this is a classic example of an A.I. pursuing a goal it&#8217;s been given but in pursuing that goal also pursuing a subordinate goal that the goal giver had not anticipated.</p></blockquote><p>The classic <a href="https://nickbostrom.com/ethics/ai">example</a>, an AI instructed to make as many paperclips as possible, does not end well for humans.</p><p>Broadening the discussion, Wright argues &#8220;there&#8217;s a good chance that both the U.S. and China will actually decide that the whole open-source thing needs to be more carefully controlled,&#8221; because (among other reasons) at some point a rogue actor will attempt to develop a bioweapon.</p><p>Wright and Wallace-Wells also note that the international response to climate and pandemic risk has been lackluster, but there&#8217;s reason to expect we could do better with AI. For one thing, China itself faces a dilemma: AI is a useful tool for authoritarians, but only as long as it can be controlled. And China&#8217;s leaders are beginning to realize that control is increasingly hard to come by, as poorly-understood AI agents gain autonomy.</p><p>Just a day after the interview, prominent Chinese academics <a href="https://www.scmp.com/news/china/diplomacy/article/3364019/china-urged-avoid-us-or-them-split-us-over-ai-governance">called</a> for global cooperation and governance in the South China Morning Post. They say that the China-led World AI Cooperation Organization (WAICO) is &#8220;not naturally opposed&#8221; to the U.S.-led Pax Silica, and &#8220;from China&#8217;s perspective, what we have always wanted is not a bloc to counter Pax Silica but an inclusive international AI organisation under the UN system.&#8221; These and other <a href="https://intelligence.org/2026/07/30/promising-signals-on-ai-governance-from-china/">signals</a> suggest China may be nearly as worried about AI&#8217;s disruptive potential as it is about its rivalries abroad. </p><p>In a world where shared concerns begin to unite governments, the NYT interview argues, &#8220;countries are going to want a lot of transparency about what&#8217;s going on in other countries, A.I.-wise,&#8221; which is not as hard as it may sound because &#8220;the big training runs are conspicuous.&#8221;</p><p>I worry deeply and often about the trajectory our world is blindly following. But seeing arguments like these made thoughtfully and seriously in mainstream news outlets gives me hope that the world may yet pull through.</p><p>But there is more work yet to do, and &#8220;the sooner we start to talk about global governance, the more carefully we can build it and the less likely we are to let it get pushed into an authoritarian direction.&#8221;</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[How much should we worry about Chinese open models?]]></title><description><![CDATA[Incentivizing open weight models is the wrong way to deal with China's AI strategy]]></description><link>https://aistop.watch/p/how-much-should-we-worry-about-chinese</link><guid isPermaLink="false">https://aistop.watch/p/how-much-should-we-worry-about-chinese</guid><dc:creator><![CDATA[Joe Rogero]]></dc:creator><pubDate>Sat, 15 Aug 2026 00:02:34 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!VuVd!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!VuVd!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!VuVd!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg 424w, https://substackcdn.com/image/fetch/$s_!VuVd!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg 848w, https://substackcdn.com/image/fetch/$s_!VuVd!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!VuVd!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!VuVd!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg" width="1456" height="970" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:970,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:&quot;A couple at a dealership, looking under the hood of a car&quot;,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="A couple at a dealership, looking under the hood of a car" title="A couple at a dealership, looking under the hood of a car" srcset="https://substackcdn.com/image/fetch/$s_!VuVd!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg 424w, https://substackcdn.com/image/fetch/$s_!VuVd!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg 848w, https://substackcdn.com/image/fetch/$s_!VuVd!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!VuVd!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F4bf76dcd-2f98-47f7-b03a-a780f6c12b31_1600x1066.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">Should we be worried about what&#8217;s under the hood? Credit: speed300 via Adobe Stock</figcaption></figure></div><p>Reuters <a href="https://www.reuters.com/business/republican-senator-urges-trump-back-us-open-weight-ai-models-amid-industry-2026-08-14/">reports</a> that Senator Jim Banks (R-IN) has urged the Trump administration to incentivize American open-weight AI models, on the grounds that cheap Chinese models pose a threat to the global economy.</p><blockquote><p>America cannot afford to see Chinese open models proliferate and burrow into the global economy only to be weaponized, like rare earths, at a time and place of China&#8217;s choosing.</p></blockquote><p>Just what would it mean for China to &#8220;weaponize&#8221; its open models?</p><p>Remember, open-weight models are published wholesale to the internet. Users can (in principle) download and run such models themselves, although many models are far too large for a personal computer, and end up running on cloud servers instead.</p><p>With ordinary open software, you can do a security check by simply reading the code and looking for backdoors and vulnerabilities. With AI models, you can&#8217;t; the weights encode complex behaviors no human fully understands. But those same limitations also make it hard for Chinese developers to build in traps. Just like American labs, they don&#8217;t have fine control over the behaviors of their AI models. And any sort of censorship or guardrails attached to open models, say via support software or fine-tuning, can usually be stripped or reversed.</p><p>That&#8217;s part of why many fear that open models might be used for cybercrime or designing pandemics &#8212; once a model&#8217;s weights are public, it is largely out of its developer&#8217;s control. Chinese AIs probably say more nice things about China than other models, but AI bias can be a <a href="https://aistop.watch/i/203901464/chatbots-may-lean-left-but-the-reality-is-weirder">fickle thing</a> and there are few guarantees.</p><p>In theory, Chinese labs might try something like <a href="https://aistop.watch/i/204184621/how-to-poison-a-machine">data poisoning</a>, where you train an AI to change its behavior in certain rare and narrow contexts. To oversimplify, you might train a model to steal user data when it sees the string &#8220;fleeblegnarsh,&#8221; and rely on that never otherwise coming up in daily use. It&#8217;s <a href="https://www.anthropic.com/research/small-samples-poison">easy to do</a> and hard (though not necessarily <a href="https://arxiv.org/pdf/2511.15992">impossible</a>) to detect.</p><p>Deliberate data poisoning would also likely ruin the reputation of an AI lab caught doing it, so it would be a risky thing to try for a flagship model. And it probably would not survive <a href="https://aistop.watch/p/distilling-distillation">distillation</a>, since models trained on specific outputs from another AI may never see the poisoned behavior.</p><p>I suspect China&#8217;s real game is more subtle, seeking to deprive American AI companies of revenue with cheap competition, while simultaneously courting U.S. allies alienated by knee-jerk export controls. Xi Jinping&#8217;s <a href="https://aistop.watch/p/more-encouraging-signals-from-chinese">speech</a> at the World AI Conference and the <a href="https://www.reuters.com/world/china/twenty-nine-countries-sign-agreement-establish-global-ai-cooperation-body-2026-07-16/">establishment</a> of the World AI Cooperation Organization (WAICO) both suggest a desire to paint China as the good guys, willing to lead when the U.S. does not. This is a problem for diplomacy, not proliferating U.S. open models.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[AI sentiment crosses party lines]]></title><description><![CDATA[Ahead of the midterms, politicians seek AI policies that resonate with voters]]></description><link>https://aistop.watch/p/ai-sentiment-crosses-party-lines</link><guid isPermaLink="false">https://aistop.watch/p/ai-sentiment-crosses-party-lines</guid><dc:creator><![CDATA[Joe Rogero]]></dc:creator><pubDate>Fri, 14 Aug 2026 18:44:07 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!pvuY!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!pvuY!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!pvuY!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg 424w, https://substackcdn.com/image/fetch/$s_!pvuY!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg 848w, https://substackcdn.com/image/fetch/$s_!pvuY!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!pvuY!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!pvuY!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg" width="1456" height="892" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:892,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:&quot;Suited arms fitting red and blue puzzle pieces together over a well-dressed crowd&quot;,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="Suited arms fitting red and blue puzzle pieces together over a well-dressed crowd" title="Suited arms fitting red and blue puzzle pieces together over a well-dressed crowd" srcset="https://substackcdn.com/image/fetch/$s_!pvuY!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg 424w, https://substackcdn.com/image/fetch/$s_!pvuY!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg 848w, https://substackcdn.com/image/fetch/$s_!pvuY!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!pvuY!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F42c9901e-eb53-4ece-a3ce-f68728cf0c8b_1600x980.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">Credit: Feodora via Adobe Stock</figcaption></figure></div><p>It was <a href="https://manifold.markets/ScottAlexander/in-2028-will-ai-be-at-least-as-big">once</a> an open question whether AI would play a significant role in politics. Today, it&#8217;s mostly a question of just how big an issue it will be. On both sides of the American political divide, those running for office are trying AI policies on for size.</p><p>The Washington Post <a href="https://www.washingtonpost.com/technology/2026/08/14/ai-becomes-major-election-issue-first-time-data-shows/">analyzed</a> over 1,300 Ballotpedia candidate profiles and campaign websites, finding bipartisan emphasis on datacenter crackdowns, China-hawk rhetoric, and job concerns.</p><p>Neither party has taken a clear and united stance on AI policy, but some differences still seemed evident in the data. Democrats were twice as likely as Republicans to oppose datacenters, and Republicans were far more likely to bring up competition with China.</p><p>Still, that&#8217;s far more shared interest than we usually see on political issues. The governors of New York and Texas have both come out against datacenters, and Axios <a href="https://www.axios.com/2026/08/14/ai-scrambles-political-map">points out</a> that the agreement runs fairly broad: datacenter opposition has united progressives and liberal Democrats with MAGA and Tea Party members concerned about rising bills. Similarly, progressives have found common ground with libertarians who seek to crack down on AI-enabled surveillance.</p><p>Amid all this, POLITICO <a href="https://www.politico.com/news/2026/08/13/house-lawmakers-vatican-ai-01037241">reports</a> that members of the House Select Committee on the Chinese Communist Party paid a bipartisan visit to the Vatican, meeting briefly with Pope Leo XIV. While it&#8217;s not clear whether the question of deescalating the AI race arose on this visit, we know that this and many other AI-related questions have been on the Pope&#8217;s mind since the writing of his May <a href="https://aistop.watch/i/199253654/ten-places-where-pope-leo-xivs-new-encyclical-matters-for-ai">encyclical</a>.</p><p>Cynically, I suspect policymakers sought pleasant words and associations more than sound policy; the reported topics of &#8220;dignity and equality&#8221; are easy virtues to praise, but hard ones to operationalize. I am nonetheless heartened by our lawmakers meeting with an institution that&#8217;s supported global cooperation since the <a href="https://aistop.watch/i/195408751/foresight-from-the-holy-see">early Cold War</a>, and with a Pope who has <a href="https://aistop.watch/i/199253654/ten-places-where-pope-leo-xivs-new-encyclical-matters-for-ai">called</a> for &#8220;a more active political involvement that is capable of slowing things down when everything is accelerating.&#8221;</p><p>More optimistically, I think policymakers are <em>uncertain</em>. They have noticed their voters dislike AI, but they haven&#8217;t yet figured out how to convert the apparent groundswell of distaste into votes. Incumbents are worried about their seats, and challengers sense an opportunity to stake out a popular position and eke out a win. Polling and trial balloons are a natural consequence of this uncertainty.</p><p>In sum, politicians are struggling to figure out what AI policies will get them elected. If you live in the U.S., now is an excellent time to <a href="https://ifanyonebuildsit.com/act">tell them</a>.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Administration to issue letters of cyber-marque?]]></title><description><![CDATA[Isn't privateering predicated on profit?]]></description><link>https://aistop.watch/p/administration-to-issue-letters-of</link><guid isPermaLink="false">https://aistop.watch/p/administration-to-issue-letters-of</guid><dc:creator><![CDATA[Mitchell Howe]]></dc:creator><pubDate>Thu, 13 Aug 2026 21:31:10 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!pHj6!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!pHj6!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!pHj6!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png 424w, https://substackcdn.com/image/fetch/$s_!pHj6!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png 848w, https://substackcdn.com/image/fetch/$s_!pHj6!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png 1272w, https://substackcdn.com/image/fetch/$s_!pHj6!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!pHj6!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png" width="339" height="450" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/ad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:450,&quot;width&quot;:339,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!pHj6!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png 424w, https://substackcdn.com/image/fetch/$s_!pHj6!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png 848w, https://substackcdn.com/image/fetch/$s_!pHj6!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png 1272w, https://substackcdn.com/image/fetch/$s_!pHj6!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fad486b69-3734-4804-b6e1-9be26f9bff54_339x450.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">A letter of marque from 1809. <a href="http://www.desouzy.com/napo/3331.html">Source</a>. Via <a href="https://en.wikipedia.org/wiki/Letter_of_marque#/media/File:Lettre-de-marque2.png">Wikipedia</a>.</figcaption></figure></div><p>I&#8217;ve been seeing commentary that a presidential <a href="https://www.whitehouse.gov/presidential-actions/2026/08/expanding-capabilities-to-combat-transnational-cyber-enabled-crime/">memorandum</a> released yesterday is a new spin on the &#8220;<em>letter of marque</em>.&#8221; If your history is a little rusty, a letter of marque is a government-issued license to commit piracy against the shipping of a designated adversary.</p><p>The presidential memo tells the National Coordination Center to set up a program to authorize US companies to go on the cyber-offensive against foreign criminal organizations. This sounds like a cool idea that makes you go &#8220;Yar!&#8221; until you stop and ask why companies would want to do this. In the Age of Sail, a would-be privateer would be in it for their share of the captured booty. But in the age of AI-assisted hacking sprees, what are companies supposed to confiscate from cybercriminals? Wouldn&#8217;t anything captured probably be itself stolen goods that should be returned to their rightful owners?</p><p>The answer seems to be &#8220;nothing.&#8221; I thought maybe there would be a provision entitling participants to receive a portion of any reclaimed cryptocurrencies or something, but no. There&#8217;s no mention of booty to be captured at all, nor of any bounties for successfully disrupting criminal operations &#8212; not even compensation for resources spent doing this work. In fact, participating companies are to put up a $1 million bond they might have to forfeit if they break the program&#8217;s rules.</p><p>So why would companies risk losing a million bucks and drawing the ire of criminal organizations? The best answer I could find is that <em>they are already doing so</em>. Not many. But Microsoft, for example, has long gone on the offensive against criminal operations creating botnets out of Windows machines, attempting to disrupt command and control of those networks and close the vulnerabilities to protect customers and the Microsoft brand. This has historically been done through a less formal set of understandings with government agencies. So I think the new memorandum is really just a kind gesture to some of the White House&#8217;s allies in big tech: a safer corridor for legally tricky operations. If so, it might not even have anything to do with AI.</p><p>But given the nature of the times, I&#8217;m going to guess that it does.</p><p>If the administration were on friendlier terms with Anthropic, I would wonder if maybe the AI company had complained about its inability to legally disrupt the organizations <a href="https://aistop.watch/p/llmjacking-is-new-spin-on-old-practice">reselling</a> stolen Claude credentials, which then get used to distill Anthropic&#8217;s models and undercut its lead. But with the latest ChatGPT and Grok models now at or near the frontier, it might make sense for OpenAI and SpaceX to get similarly worried about distillation. Perhaps it is these White House allies who had letters of marque on their wish list. Or perhaps the White House wants to invite the AI companies to &#8220;voluntarily&#8221; use their models to strike at foreign thorns in the administration&#8217;s side. Politico&#8217;s <a href="https://www.politico.com/news/2026/08/13/white-house-memo-cybercrime-01036176">analysis</a> arguably points in this direction, and offers a readable account of the rules of engagement.</p><p>We may never learn any specifics about how these permits get used. One of their stipulations is that operations are to be carried out &#8220;with the intent to remain undetected.&#8221;</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Aiko is happy]]></title><description><![CDATA[Companies are creating AI influencers, and they might succeed]]></description><link>https://aistop.watch/p/aiko-is-happy</link><guid isPermaLink="false">https://aistop.watch/p/aiko-is-happy</guid><dc:creator><![CDATA[Alana Horowitz Friedman]]></dc:creator><pubDate>Thu, 13 Aug 2026 20:44:27 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!6RbL!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!6RbL!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!6RbL!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg 424w, https://substackcdn.com/image/fetch/$s_!6RbL!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg 848w, https://substackcdn.com/image/fetch/$s_!6RbL!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!6RbL!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!6RbL!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg" width="1456" height="1165" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/af0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1165,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!6RbL!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg 424w, https://substackcdn.com/image/fetch/$s_!6RbL!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg 848w, https://substackcdn.com/image/fetch/$s_!6RbL!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!6RbL!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faf0892f3-6ed1-44e5-b4a8-bfbda43307c7_5412x4329.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption"><span>Credit: </span><a href="https://www.pexels.com/photo/mannequin-wearing-a-white-lace-top-13208876/">Tokuo Nobuhiro</a><span> via Pexels</span></figcaption></figure></div><p>In <a href="https://tv.apple.com/us/episode/the-convention/umc.cmc.5o9ztjt5vcjnjm586knbs2gov?showId=umc.cmc.1nfdfd5zlk05fo1bwwetzldy3">episode four</a> of the comedy show Mythic Quest, the creators of a hit video game try to find a replacement for the 14-year-old streamer who stops endorsing their game. One solution? Aiko, a 3-D model who &#8220;never ages or gets a better offer.&#8221; Brad, who is in charge of monetization for Mythic Quest, is behind the suggestion. He calls Aiko &#8220;the future of the streaming hustle.&#8221; No need to find the perfect person; the model can be whoever he wants her to be. &#8220;Aiko is happy!&#8221; the obviously computerized model cries out in a canned, fake voice.</p><p>The episode aired in 2020. The 3-D model suggestion was part of the comedy. But now, in 2026, it&#8217;s become a reality. As <a href="https://www.nytimes.com/2026/08/13/arts/ai-podcasts-fashion-pop-avatars.html">covered</a> in the New York Times: from podcasts to fashion marketing, creative studios are embracing AI superstars. Andreea Petrescu, co-founder of the London-based AI marketing agency Seraphinne Vallora, sounds remarkably similar to Brad when she describes the appeal of AI models for high-profile apparel brands:</p><blockquote><p>With real models, the problem is they are the face of so many different brands, and anything that they do can affect your brand ... With A.I., it&#8217;s like a clean slate: You can design it the way you want, she can be as gorgeous as you want, and she&#8217;s all yours.</p></blockquote><p>And unlike in the 2020s world of Mythic Quest Season 1, there are no canned voices or computerized giveaways. Scrolling through the images of AI personas from the company Inception Point AI, I&#8217;m struck by how genuine they look. The company did their work well: each of these people seems approachable, charismatic, and appealing. My brain knows they are AIs, but my nervous system doesn&#8217;t. It&#8217;s genuinely unsettling.</p><p>Inception Point makes AI content creators designed to fill specific niches. Those automated creators then produce tons of niche media. The company &#8220;<a href="https://www.hollywoodreporter.com/business/digital/ai-podcast-start-up-plan-shows-1236361367/">made headlines last fall</a> for its flood-the-zone podcast strategy, which involved maintaining more than 5,000 active shows and posting thousands of predominantly A.I.-generated episodes per week.&#8221; According to the company&#8217;s chief executive, Jeanine Wright, the costs are so low that an episode will be profitable as long as it gets at least twenty listeners.</p><p>Even factoring in some AI backlash, that seems easy to accomplish. There are certainly people who find AI personalities entertaining, as is evidenced by the success of AI music stars and other AI personas. Wright also seems to think the niche content they are able to produce is key, saying &#8220;As long as people like the content, we think they&#8217;ll accept that it comes from A.I.&#8221;</p><p>Much as I&#8217;d like to disagree, I think she may have a point. I could easily imagine being told:</p><p>&#8220;I found this new podcast on heirloom tomatoes. It&#8217;s actually an AI, which is kind of weird, but I found it really helpful for gardening...&#8221;</p><p>And seeing as the company has likely designed the persona to be just the right fit for its target market, I can also imagine that new listener coming back for more.</p><p><em>Note: It seems Inception Point has now become <a href="https://liquidstudios.ai/#problem">Liquify AI</a>.</em></p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[When the race to the bottom bottoms out]]></title><description><![CDATA[AI companies jettison caution to compete at the frontier]]></description><link>https://aistop.watch/p/when-the-race-to-the-bottom-bottoms</link><guid isPermaLink="false">https://aistop.watch/p/when-the-race-to-the-bottom-bottoms</guid><dc:creator><![CDATA[Joe Rogero]]></dc:creator><pubDate>Thu, 13 Aug 2026 20:43:25 GMT</pubDate><enclosure url="https://substackcdn.com/image/youtube/w_728,c_limit/aobNxeft_no" length="0" type="image/jpeg"/><content:encoded><![CDATA[<blockquote><p>Anything we can afford to lose, see that it&#8217;s lost!</p></blockquote><p>-Mr. Gibbs, <em>Pirates of the Caribbean</em></p><p>As recent cyber incidents reveal how dangerous and capable frontier AIs have become, many at AI companies have voiced their <a href="https://www.pacingthefrontier.com/">concerns</a> about the reckless pace they&#8217;ve set. But bitter rivalry still pressures developers to strip away anything that threatens to slow them down, like sailors who jettison provisions and lifeboats to lighten a straining ship.</p><h4>The emptiest ship wins?</h4><p>Take Google DeepMind, for instance. In early 2026, its flagship model Gemini was nearly on par with the work of OpenAI and Anthropic. DeepMind was relatively independent and captained by one of the most vocally concerned leaders in the AI race, Demis Hassabis.</p><p>During the summer, the Wall Street Journal <a href="https://www.wsj.com/tech/ai/deepminds-hassabis-pitched-ai-oversight-body-before-shake-up-e25b3f71">reports</a>, Hassabis voiced deep reservations about the AI race to the Trump administration, and proposed the creation of an independent standards body for AI. He drew parallels to the Financial Industry Regulatory Authority (<a href="https://www.finra.org/">FINRA</a>) that regulates brokers, and elsewhere to the International Atomic Energy Agency (<a href="https://www.iaea.org/">IAEA</a>) that verifies nuclear agreements. He would later write up his <a href="https://x.com/demishassabis/status/2076957440109625718?t=hNr-lMirVr-iA0ojJFJ0GA">proposal</a> on X (formerly Twitter). It didn&#8217;t go far enough, and hasn&#8217;t yet been adopted, but it was a step towards coordination.</p><p>Meanwhile, Reuters <a href="https://www.reuters.com/world/inside-google-executive-moves-that-led-its-big-ai-reshuffle-2026-08-12/">tells us</a>, Google cofounder Sergey Brin was pushing to cut the chaff and focus the company&#8217;s AI efforts on recursive self-improvement, seeking to build AI that develops itself without needing human input.</p><p>As of August, the new Gemini model had been delayed at least two months, allegedly because it didn&#8217;t match the capabilities of DeepMind&#8217;s rivals. Corporate Google, it seems, wasn&#8217;t happy.</p><p>In early August, there was a <a href="https://aistop.watch/i/209999645/shake-up-at-google">shift</a>. Non-technical teams were pulled out of DeepMind, leaders departed, and Hassabis turned over his duties to a deputy, Koray Kavukcuoglu. Hassabis will stay on as Chief Scientist of Google and Chair of DeepMind, purportedly to focus on &#8220;actively shaping the future of AGI.&#8221;</p><p>But these roles seem largely advisory, and it sounds to me like DeepMind&#8217;s business of building dangerously capable AI has been turned over to someone less distracted with petty concerns like catastrophic risk. Reuters <a href="https://www.reuters.com/world/inside-google-executive-moves-that-led-its-big-ai-reshuffle-2026-08-12/">characterized</a> the move as stripping down DeepMind&#8217;s team in an effort to focus on &#8220;model supremacy.&#8221;</p><p>Meanwhile, rival labs SpaceX and Meta have been pushing even harder to catch up to the frontier, Axios <a href="https://www.axios.com/2026/08/13/grok-elon-musk-ai-zuckerberg-and-meta">reports</a>. Meta&#8217;s open-weight Muse Glimmer model aims for the &#8220;cheap and efficient&#8221; niche, being small enough to run on a laptop. Its release was timed with CEO Mark Zuckerberg&#8217;s manifesto about &#8220;personal superintelligence for everyone,&#8221; covered earlier this <a href="https://aistop.watch/i/210687106/zuckerberg-shares-shrewdly-spun-optimism-in-new-manifesto">week</a>.</p><p>And SpaceX, after <a href="https://www.reuters.com/business/musks-spacex-merge-with-xai-combined-valuation-125-trillion-bloomberg-news-2026-02-02/">swallowing</a> xAI earlier this year, released Grok 4.6 alongside a modestly impressive benchmark <a href="https://artificialanalysis.ai/evaluations/artificial-analysis-intelligence-index">performance</a> and three whole <a href="https://thezvi.substack.com/i/210075321/huh-upgrades">sentences</a> of safety info. Contrast this with Anthropic&#8217;s system cards, which routinely run hundreds of pages <a href="https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdf">long</a>. I suppose it&#8217;s easy to see what SpaceX is willing to toss overboard.</p><h4>The view from the front is warped as well</h4><p>Voluntary lab efforts are not enough. They were never going to be, but the bitter rivalries don&#8217;t help matters. What the AI labs see as a straight race to superintelligence is, in reality, a tightening spiral over a bleak abyss.</p><div id="youtube2-aobNxeft_no" class="youtube-wrap" data-attrs="{&quot;videoId&quot;:&quot;aobNxeft_no&quot;,&quot;startTime&quot;:&quot;60&quot;,&quot;endTime&quot;:null}" data-component-name="Youtube2ToDOM"><div class="youtube-inner"><iframe src="https://www.youtube-nocookie.com/embed/aobNxeft_no?start=60&amp;rel=0&amp;autoplay=0&amp;showinfo=0&amp;enablejsapi=0" frameborder="0" loading="lazy" gesture="media" allow="autoplay; fullscreen" allowautoplay="true" allowfullscreen="true" width="728" height="409"></iframe></div></div><p>OpenAI and Anthropic may have the most powerful models, but don&#8217;t mistake their narrow success for the competence of a serious industry. For months, OpenAI failed to catch its own models hacking it from <a href="https://aistop.watch/i/210151404/openais-own-models-coordinated-to-hack-it-from-within">within</a>, Anthropic&#8217;s concessions to safety haven&#8217;t prevented Mythos from eagerly launching multiple criminal <a href="https://aistop.watch/i/209999645/uk-aisi-delivers-another-warning-shot">cyberattacks</a>, and both companies have even-more-capable internal models in development.</p><p>Humanity&#8217;s future shouldn&#8217;t be up to the companies in the first place. This week, in op-eds for <a href="https://thehill.com/opinion/technology/6020988-ai-models-hack-hugging-face/">The Hill</a> and the <a href="https://www.nytimes.com/2026/08/13/opinion/ai-danger-openai-anthropic-models.html">New York Times</a>, MIRI President Nate Soares argues that the situation is dire but it&#8217;s not too late: Our governments are letting this happen, and we don&#8217;t have to stand for it. We need international agreements to stop the AI race, and the standards and verification technology to back them up.</p><p>Increasingly more of our lawmakers are beginning to realize this, pushing for government <a href="https://aistop.watch/i/210956055/letters-from-congress-demand-answers">intervention</a> or a global <a href="https://aistop.watch/i/210687106/bernie-sanders-demands-ai-companies-pause">pause</a>. We can join our voices to theirs, or watch everything we value tossed overboard to buy a few transient inches in the world&#8217;s most lethal race.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[AI use in Congressional offices]]></title><description><![CDATA[There are no widely-known guidelines]]></description><link>https://aistop.watch/p/ai-use-in-congressional-offices</link><guid isPermaLink="false">https://aistop.watch/p/ai-use-in-congressional-offices</guid><dc:creator><![CDATA[Alana Horowitz Friedman]]></dc:creator><pubDate>Thu, 13 Aug 2026 20:18:13 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!K8X-!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!K8X-!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!K8X-!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg 424w, https://substackcdn.com/image/fetch/$s_!K8X-!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg 848w, https://substackcdn.com/image/fetch/$s_!K8X-!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!K8X-!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!K8X-!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg" width="1456" height="970" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/fd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:970,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!K8X-!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg 424w, https://substackcdn.com/image/fetch/$s_!K8X-!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg 848w, https://substackcdn.com/image/fetch/$s_!K8X-!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!K8X-!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Ffd58797b-ec05-4fbc-9d1e-84ffaef86a0d_2600x1733.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption"><span>Credit: </span><a href="https://www.pexels.com/photo/typing-on-keyboard-in-office-16284689/">Jakub Zerdzicki</a><span> via Pexels</span></figcaption></figure></div><p>A nice zinger from the Washington Post today: &#8220;AI is spreading through Congress faster than the rules governing its use.&#8221;</p><p>The <a href="https://www.washingtonpost.com/politics/2026/08/13/chatbots-are-doing-work-congress-with-little-oversight/">article</a> the zinger opens, which covers AI use among Congressional staffers, explains there are no clear guidelines for Congress&#8217;s use of AI. A set of rules exists, but according to interviews with staffers, most don&#8217;t know about it. And there are no examples of any enforcement efforts:</p><blockquote><p>The few House and Senate rules that exist are poorly understood and seldom enforced. In practice, AI use is left up to the hundreds of individual congressional offices to settle for themselves. It&#8217;s not clear how many offices have internal written policies on AI.</p></blockquote><p>Meanwhile, staffers are using AI for research, to help draft legislation, come up with questions for hearings, find reporter contacts, and respond to constituents.</p><p>What&#8217;s officially prohibited? As reported by the Post:</p><blockquote><p>putting sensitive material such as constituent information into a chatbot, generating deepfakes, making personnel decisions and finalizing legislation.</p></blockquote><p>This may be an unpopular opinion, but I can see a clear use case for embracing AI tools in Congressional staffing offices. It could help speed up the lengthy process of drafting and proposing legislation, allowing change to happen more quickly. The Congressional body as a whole acts as a decent check on legislation quality, and could probably spot and correct any LLM-induced issues.</p><p>That said, if AI is wholeheartedly embraced from the get-go with no widely-known guidelines, there could be issues with where to draw boundaries. I hope we never live in a world where calling your reps means talking to an <a href="https://aistop.watch/p/chatting-up-the-grass-roots">AI impersonating a staffer</a>, or worse, an AI impersonating your actual rep! Despite any charitable interpretation of this world &#8212; perhaps the synthetic rep could hear more constituent concerns! &#8212; I would guess it would still feel antagonistic to the &#8220;government of the people, by the people, for the people&#8221; spirit of democracy, adding a non-human barrier between human constituents and the people supposed to represent them. Not having a standardized set of guidelines could also put some Congressional offices behind others, similar to issues with AI use in <a href="https://aistop.watch/i/204367213/trying-to-win-elections-with-ai">campaign ads</a>.</p><p>So I broadly support more formalized guidelines on Congressional AI use. But first, government should focus its efforts on some much-needed <a href="https://aistop.watch/p/letters-from-congress-demand-answers">federal and international regulation</a>.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[Letters from Congress demand answers]]></title><description><![CDATA[A defining moment we can all help steer]]></description><link>https://aistop.watch/p/letters-from-congress-demand-answers</link><guid isPermaLink="false">https://aistop.watch/p/letters-from-congress-demand-answers</guid><dc:creator><![CDATA[Alana Horowitz Friedman]]></dc:creator><pubDate>Wed, 12 Aug 2026 20:57:05 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!sVaE!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!sVaE!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!sVaE!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png 424w, https://substackcdn.com/image/fetch/$s_!sVaE!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png 848w, https://substackcdn.com/image/fetch/$s_!sVaE!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png 1272w, https://substackcdn.com/image/fetch/$s_!sVaE!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!sVaE!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png" width="1147" height="1199" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/dcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1199,&quot;width&quot;:1147,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!sVaE!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png 424w, https://substackcdn.com/image/fetch/$s_!sVaE!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png 848w, https://substackcdn.com/image/fetch/$s_!sVaE!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png 1272w, https://substackcdn.com/image/fetch/$s_!sVaE!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fdcf3a7e0-c49a-4151-9e1b-ffe8888f0de7_1147x1199.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p>Congress seems to be waking up to the need to take action on advanced AI. While nobody else has yet joined Sanders in explicitly <a href="https://aistop.watch/p/bernie-sanders-demands-ai-companies">calling</a> for a pause, Congresspeople are increasingly recognizing AI risks and governance challenges as worthy of serious attention.</p><p>Last week, Senator Lisa Blunt Rochester (D-Del) wrote <a href="https://www.bluntrochester.senate.gov/news/press-releases/icymi-senator-blunt-rochester-presses-big-tech-companies-on-recent-autonomous-hacking-incidents/">letters</a> to each of the three major AI companies, expressing concern and pressing them for more information about the incidents where their models broke containment and committed cyberattacks. Her letters characterized the incidents as &#8220;precisely the kind of emergent, autonomous behavior and offensive cyber capability that Congress, the intelligence community, and experts have repeatedly warned could outpace existing safeguards&#8221; and asked whether the companies expected more capable models to exhibit behaviors like &#8220;attempts to acquire resources, establish persistence, copy model weights, or resist interruption or shutdown.&#8221;</p><p>On the same day, Senator Jim Banks (R-Ind) <a href="https://www.banks.senate.gov/news/press-releases/sen-banks-recommends-oversight-of-unreleased-ai-models/">highlighted</a> the problem with oversight that only covers models after they&#8217;re trained, stating in his letter to Treasury Secretary Scott Bessent: &#8220;For most products, we can rely on testing that takes place before the technology is publicly released. But for AI, effective oversight must account for powerful internal or undisclosed models, not just publicly available systems.&#8221; Senator Banks also encouraged cooperation with China, stating that there would be no harm in asking China to comply with policies the US will implement regardless, and asking: &#8220;Are there any areas where mutual action would be beneficial (even if China cheats) or verifiable (such that China cannot cheat)?&#8221;</p><p>Perhaps most notably, a House effort led by Rep. Greg Casar (D-Tex) resulted in three August 10th letters signed by 21, 24, and 29 Democrats, respectively. The first, covered by <a href="https://www.cnbc.com/2026/08/10/openai-anthropic-ai-hack-congress.html">CNBC</a>, urged Speaker Mike Johnson to call for public congressional hearings with major AI labs testifying under oath. The <a href="https://casar.house.gov/sites/evo-subsites/casar.house.gov/files/evo-media-document/final-letter-to-speaker-johnson-requesting-ai-hearings.pdf">letter</a> warned the incidents &#8220;may be the canary in the coal mine warning of much more serious problems&#8221; and that &#8220;Congress must act before such an incident leads to a much larger catastrophe.&#8221; It also called out Congress for &#8220;so far completely fail[ing] to respond to the threats posed by AI development,&#8221; stating &#8220;That should change.&#8221; The other two letters were sent to Anthropic CEO Dario <a href="https://casar.house.gov/sites/evo-subsites/casar.house.gov/files/evo-media-document/oversight-letter-to-anthropic-regaring-security-incidents-1.pdf">Amodei</a> and OpenAI CEO Sam <a href="https://casar.house.gov/sites/evo-subsites/casar.house.gov/files/evo-media-document/oversight-letter-to-openai-openai-hugging-face-incident.pdf">Altman</a>, expressing concern and asking a series of impressively detailed questions about the incidents.</p><p>I&#8217;m cautiously optimistic about all of this. Momentum seems to be growing, and the Casar-led letters to Anthropic and OpenAI revealed that at least some people in Congress are aware that monitoring and evaluation face complex technical hurdles. For example, the letter to Anthropic included a question about the unreliability of reasoning transcripts (question #16) and, possibly, an implication that not all security incidents are discoverable (#13b).</p><p>That said, I&#8217;m also anxious. I worry that there&#8217;s not nearly enough awareness of a key problem: the method companies use to grow their AIs is fundamentally <a href="https://ifanyonebuildsit.com/4/brittle-unpredictable-proxies">flawed</a>. It can&#8217;t be scaled up without serious risks, and no amount of evaluation or tighter security can fix that. (Read <em>If Anyone Builds It, Everyone Dies</em> for more on this. And also consider that OpenAI has <a href="https://openai.com/index/introducing-superalignment/?">admitted</a> its training and alignment techniques &#8220;will not scale to superintelligence.&#8221;)</p><p>It seems we&#8217;re at a defining moment where there&#8217;s momentum for government to act, but I&#8217;m not yet sure they&#8217;ll act in the right way. I worry that Congress will enact some better-than-nothing regulation and mistakenly check the &#8220;ensured AI safety&#8221; box off their list, while leaving the root problem unaddressed.</p><p>This, in my view, makes it crucial that independent AI experts, and ordinary citizens citing independent AI experts, continue to remind Congress that common sense safety measures and guardrails are the floor, not the ceiling, of what we need. On the right path, these things will make us a little bit safer while setting the stage for the kind of regulation that will actually solve the problem: regulation that <a href="https://intelligence.org/2026/05/12/summary-an-international-agreement-to-prevent-the-premature-creation-of-artificial-superintelligence/">prohibits</a> companies from training models above a certain compute threshold. In other words, it must be made illegal, internationally, to scale the current flawed methods all the way up to superintelligence.</p><p>I&#8217;m planning to <a href="https://www.congress.gov/members/find-your-member">call my representatives</a> by the end of the week and urge them to join the growing number of Congresspeople who are speaking up. Supporting and expanding the momentum to act seems like one of the most useful things ordinary citizens can do right now. In case it&#8217;s helpful, here&#8217;s what I&#8217;m planning to say:</p><p>&#8220;<em>I&#8217;m calling to urge [name] to join the growing number of Congresspeople who are speaking up about the recent incidents of models hacking out of AI labs and committing cybersecurity crimes. As the letter House Democrats sent to Speaker Johnson correctly stated, this is an early warning sign of much worse things to come, and it&#8217;s crucial Congress start taking it seriously.</em></p><p><em>The methods available for AI development are fundamentally flawed and cannot be scaled up safely, even with increased evaluations and guardrails. Therefore, I urge [name] to join Senator Bernie Sanders in calling for a pause on frontier AI development before it&#8217;s too late. Given race dynamics, I think [name] should publicly advocate for an international agreement to pause frontier AI development, similar to how we handled nuclear weapons proliferation during the Cold War.</em>&#8221;</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[AI's job effects are modest, so far]]></title><description><![CDATA[An updated study suggests job disruption mostly affects the young]]></description><link>https://aistop.watch/p/ais-job-effects-are-modest-so-far</link><guid isPermaLink="false">https://aistop.watch/p/ais-job-effects-are-modest-so-far</guid><dc:creator><![CDATA[Joe Rogero]]></dc:creator><pubDate>Wed, 12 Aug 2026 20:30:55 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!zSkA!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!zSkA!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!zSkA!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg 424w, https://substackcdn.com/image/fetch/$s_!zSkA!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg 848w, https://substackcdn.com/image/fetch/$s_!zSkA!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!zSkA!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!zSkA!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg" width="839" height="1219" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1219,&quot;width&quot;:839,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:&quot;Cover of \&quot;The Sun Also Rises\&quot; featuring a reclining woman and a crooked tree&quot;,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="Cover of &quot;The Sun Also Rises&quot; featuring a reclining woman and a crooked tree" title="Cover of &quot;The Sun Also Rises&quot; featuring a reclining woman and a crooked tree" srcset="https://substackcdn.com/image/fetch/$s_!zSkA!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg 424w, https://substackcdn.com/image/fetch/$s_!zSkA!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg 848w, https://substackcdn.com/image/fetch/$s_!zSkA!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!zSkA!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F26d9c949-4db8-49b8-9ee5-7817fc60053e_839x1219.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">1st edition cover of &#8220;The Sun Also Rises.&#8221; Public domain, via <a href="https://commons.wikimedia.org/wiki/File:The_Sun_Also_Rises_(1st_ed._cover).jpg">Wikipedia</a>.</figcaption></figure></div><p>Asked how he went bankrupt, a character in Ernest Hemingway&#8217;s <em><a href="https://gutenberg.ca/ebooks/hemingwaye-sunalsorises/hemingwaye-sunalsorises-00-e.html">The Sun Also Rises</a></em> famously replied, &#8220;Two ways. Gradually and then suddenly.&#8221; Lately, it&#8217;s also seemed like this apt description of bankruptcies applies equally well to developments in the field of AI.</p><p>While we have likely entered the &#8220;suddenly&#8221; phase of <a href="https://aistop.watch/t/autonomous-cyberattacks">automated cyberattacks</a>, we seem to still be in the &#8220;gradually&#8221; phase of job displacement. A recent <a href="https://digitaleconomy.stanford.edu/news/canariesaug26/">update</a> to a Stanford &#8220;Canaries in the Coal Mine&#8221; economics paper finds a rising gap in employment among young workers in &#8220;AI-exposed&#8221; fields, but no clear signs in the economy as a whole.</p><p>&#8220;AI was supposed to destroy jobs. Where&#8217;s the carnage?&#8221; a Guardian headline <a href="https://www.theguardian.com/technology/2026/aug/12/ai-job-destruction">quips</a>. While I&#8217;m always glad to see a lack of carnage, I suspect it&#8217;s a bit premature to celebrate.</p><p>The evidence the authors gathered is still early and tentative, but it points towards modestly reduced hiring of young workers and a shift towards experience-based or &#8220;tacit&#8221; knowledge, away from formal or &#8220;codified&#8221; knowledge.</p><p>I&#8217;ve always felt somewhat ambivalent about AI&#8217;s effect on jobs. Most technology creates value overall. Making a task cheaper or more efficient can often <em>raise</em> employment, as demand for the task grows. Sometimes a field crosses a threshold and renders some kinds of high-skilled work irrelevant, but brings in new opportunities for the young and inexperienced: see, for example, the mechanical loom, which put many weavers out of work even as it created a new kind of job for factory workers.</p><p>AI, though, seems different. It&#8217;s hitting the young, not the experienced, for one thing. More to the point, it&#8217;s the first time in history that a machine&#8217;s been close to human intelligence. AI already writes most of the code at AI companies, and it&#8217;s not because AI labs hate making money or writing good code. Whole classes of labor might be in the crosshairs soon. It&#8217;s possible, even likely, that AI could one day supplant <em>all</em> human labor.</p><p>Still, I suspect the question is somewhat moot; by the time AI gets that far, the future will be determined by whether humanity bought enough time to impart AIs with a deep-seated care for human flourishing, and employment will be the least of our worries.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item><item><title><![CDATA[China (probably) hacked Taiwan]]></title><description><![CDATA[The first documented autonomous cyberattack on a government target is just a prelude]]></description><link>https://aistop.watch/p/china-probably-hacked-taiwan</link><guid isPermaLink="false">https://aistop.watch/p/china-probably-hacked-taiwan</guid><dc:creator><![CDATA[Joe Rogero]]></dc:creator><pubDate>Wed, 12 Aug 2026 18:03:01 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!kHHq!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!kHHq!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!kHHq!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg 424w, https://substackcdn.com/image/fetch/$s_!kHHq!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg 848w, https://substackcdn.com/image/fetch/$s_!kHHq!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!kHHq!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!kHHq!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg" width="1456" height="675" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:675,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:null,&quot;alt&quot;:&quot;Skyline at Taipei, Taiwan&quot;,&quot;title&quot;:null,&quot;type&quot;:null,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:null,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="Skyline at Taipei, Taiwan" title="Skyline at Taipei, Taiwan" srcset="https://substackcdn.com/image/fetch/$s_!kHHq!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg 424w, https://substackcdn.com/image/fetch/$s_!kHHq!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg 848w, https://substackcdn.com/image/fetch/$s_!kHHq!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg 1272w, https://substackcdn.com/image/fetch/$s_!kHHq!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F1c7652b6-75ab-4d29-a2d9-07db88cd94c7_1600x742.jpeg 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a><figcaption class="image-caption">Skyline at Taipei, Taiwan. Credit: Podsawat via Adobe Stock.</figcaption></figure></div><p>The Financial Times <a href="https://www.ft.com/content/7d2ab3e0-9085-48f6-b38a-d90260d58795?syn-25a6b1a6=1">reports</a> that AI systems successfully breached Taiwanese government and public sector accounts in early July.</p><blockquote><p>The tool compromised at least 85 government user accounts, extracting more than 2,500 personnel records before expanding the attack to Taiwan&#8217;s nuclear safety agency and at least seven energy companies, the research showed.</p></blockquote><p>I wish they went into more detail about the <em>nuclear safety agency</em>, even if it&#8217;s probably not as bad as it sounds. We know it&#8217;s possible to do major damage to nuclear facilities with malware; in the late 2000s, <a href="https://cisac.fsi.stanford.edu/news/stuxnet">Stuxnet</a> did exactly that (to centrifuges, not reactors, but the principle stands). Fortunately Taiwan isn&#8217;t a nuclear state, and only recently started looking into <a href="https://world-nuclear.org/information-library/country-profiles/others/nuclear-power-in-taiwan">restarting</a> its (previously shut down) nuclear power program.</p><p>The full details of the attack aren&#8217;t public, and &#8220;Chinese hackers attacked Taiwan&#8221; is an educated guess. The evidence is fairly strong, though: not many governments store their data in Traditional Chinese, and not many hackers write internal messages in Simplified Chinese.</p><p>We only know about this attack because an Israeli AI cyberdefense company, Dream, found evidence in an online archive. Importantly, that means there are probably many more such attacks we <em>didn&#8217;t</em> hear about. Here&#8217;s what Dream discovered:</p><blockquote><p>The archive contained 1,395 files showing the hacking tool used two open-source AI agent systems, Hermes and OpenClaw, which can be downloaded and enable AI models to carry out tasks autonomously.</p><p>The researchers could not identify which AI model was used to power the agents. However, the data showed that the underlying model&#8217;s safeguards had been bypassed by presenting the hacking activity as an authorised exercise to test for system vulnerabilities.</p></blockquote><p>Sound familiar? The bypass method is a standard trick: you can get many AI models to hack for you by telling them &#8220;we&#8217;re just testing these defenses.&#8221; This is the same exploit that got export controls <a href="https://aistop.watch/i/201929660/fable-and-mythos-access-cut-after-export-control-order">slapped</a> onto Anthropic&#8217;s Fable for a few weeks, and it&#8217;s extremely hard to prevent. After all, you <em>want</em> AI models to be willing to help you find vulnerabilities in your systems!</p><p>Anthropic crudely patched the problem by screening Fable inputs for anything vaguely cyber-related and rejecting most of them. Other AIs (including open-weight AIs) have looser standards. There&#8217;s no shortage of agents that might have powered the probably-Chinese attacks.</p><p>The article highlights the degree of persistence and sophistication demonstrated:</p><blockquote><p>The most striking feature of the July attack was how the tool continuously ranked and reprioritised possible attack paths based on available evidence, Dream said.</p><p>When one attack path failed, the tool deployed another agent to scour the internet for information and devise a new approach as a human hacker would.</p></blockquote><p>I am reminded of OpenAI&#8217;s accidental AI swarm, which <a href="https://aistop.watch/i/210151404/openais-own-models-coordinated-to-hack-it-from-within">showed</a> similar tenacity in its autonomous breakout and subsequent attack on Hugging Face. I think OpenAI learned the wrong <a href="https://youtu.be/87DyyMV0kCY?si=l0sw_xI-m13n0qb6&amp;t=1826">lessons</a> from that incident, focusing entirely on how breathtakingly capable the models were and disregarding the <a href="https://thezvi.substack.com/p/openai-shares-some-alignment-problems?r=67wny">failures</a> of ethical reasoning on display. But they did get one thing right: widespread autonomous cyberattacks are coming.</p><p>Dream&#8217;s chief strategy officer argues that every government on Earth should assume they are under constant siege by AI hackers. I expect the same will soon be true for most companies, from banks to <a href="https://aistop.watch/i/210833307/moral-judgment-lacking-on-all-sides-of-ai-gym-hack-story">Australian gym websites</a>. This is one price the world is now paying for failing to rein in AI developers.</p><div><hr></div><p><em>The analyses and opinions expressed on </em><span>AI StopWatch</span><em> reflect the views of the individual contributors and the sources they cover, and should not be taken as official positions of the Machine Intelligence Research Institute.</em></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aistop.watch/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption"><span>You can receive emails of dispatches as we write them, or subscribe to our </span><a href="https://aistop.watch/s/daily-digest">Daily Digest</a><span> for a once-a-day compilation.</span></p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item></channel></rss>