<script data-pm-proxy="intercept"></script><?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0" xmlns:itunes="http://www.itunes.com/dtds/podcast-1.0.dtd" xmlns:googleplay="http://www.google.com/schemas/play-podcasts/1.0"><channel><title><![CDATA[AI Minute Daily]]></title><description><![CDATA[AI Minute Daily]]></description><link>https://aiminutedly.substack.com</link><image><url>https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png</url><title>AI Minute Daily</title><link>https://aiminutedly.substack.com</link></image><generator>Substack</generator><lastBuildDate>Fri, 04 Sep 2026 03:24:50 GMT</lastBuildDate><atom:link href="/__u/aiminutedly.substack.com/feed" rel="self" type="application/rss+xml"/><copyright><![CDATA[AI Minute Daily]]></copyright><language><![CDATA[en]]></language><webMaster><![CDATA[aiminutedly@substack.com]]></webMaster><itunes:owner><itunes:email><![CDATA[aiminutedly@substack.com]]></itunes:email><itunes:name><![CDATA[AI Minute Daily]]></itunes:name></itunes:owner><itunes:author><![CDATA[AI Minute Daily]]></itunes:author><googleplay:owner><![CDATA[aiminutedly@substack.com]]></googleplay:owner><googleplay:email><![CDATA[aiminutedly@substack.com]]></googleplay:email><googleplay:author><![CDATA[AI Minute Daily]]></googleplay:author><itunes:block><![CDATA[Yes]]></itunes:block><item><title><![CDATA[AI Minute Daily 02/09/2026]]></title><description><![CDATA[Google shipped Gemini 3.8 Flash today, continuing its monthly release cadence.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-02092026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-02092026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Wed, 02 Sep 2026 15:31:01 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p>Good morning. This is your AI Minute Daily for Wednesday, September 2, 2026. Thank you so much for reading. Let&#8217;s get into it:</p><p><strong>Writer: </strong>Aarav Iyer and Sparsh Sultania</p><h3>TODAY&#8217;S TOP STORY: Gemini 3.8 Flash Ships, Three Weeks After 3.7</h3><p>Google DeepMind released Gemini 3.8 Flash today, internally codenamed &#8220;skimaki,&#8221; following a month of testing on Google&#8217;s internal Jetski coding platform.</p><p><strong>Key Takeaways:</strong></p><ul><li><p><strong>This is a refinement, not a reinvention.</strong> Early tester feedback points to the main improvement being reduced verbosity, a persistent complaint with earlier Flash models, rather than a big capability jump.</p></li><li><p><strong>The release cadence itself is the notable part.</strong> Gemini 3.7 Flash shipped August 13, meaning Google has now put out two Flash updates in under three weeks. That&#8217;s an unusually tight iteration cycle for a model line, closer to a coding tool&#8217;s release schedule than a traditional flagship model.</p></li><li><p><strong>Coding is the explicit target.</strong> Google is positioning this update against Anthropic&#8217;s Fable and OpenAI&#8217;s GPT-5.6 Sol specifically for coding and agentic workflows, continuing the trend of Flash-tier models increasingly being built for developer tools rather than general chat.</p></li></ul><p><strong>What to expect:</strong></p><p>Full benchmark numbers weren&#8217;t published alongside today&#8217;s release, so it&#8217;s worth waiting for independent testing before drawing conclusions on where this actually lands against competitors. The more interesting trend to watch is the cadence itself, if Google keeps shipping Flash updates every three weeks, it changes how developers think about pinning model versions in production.</p><h3>SO WHAT ELSE IS GOING ON?</h3><ul><li><p><strong>South Korea&#8217;s government and major companies committed roughly $880 billion combined</strong> toward AI infrastructure: $518 billion from Samsung and SK Hynix (with suppliers) for two new chip fabrication sites, and $355 billion for 8.4 gigawatts of AI data center capacity by 2029, led by Naver. The government is adding roughly $100 billion in land, power, and regulatory support on top of that.</p></li><li><p><strong>A McKinsey &#8220;State of AI in 2026&#8221; survey found 32% of organizations have skipped buying at least one software product or feature</strong> because they could build it internally using agentic coding tools instead, a concrete data point on how much AI coding assistance is starting to reshape software purchasing decisions.</p></li><li><p><strong>New York City&#8217;s Department of Education banned student-facing generative AI tools</strong> chatbots and tutors, for Pre-K through 8th grade, while permitting teacher use for lesson planning and translation. New screen-time caps also apply for grades 3&#8211;8.</p></li><li><p><strong>Nvidia CEO Jensen Huang told G20 innovation ministers</strong> that AI infrastructure belongs in the same category as water, roads, and electricity, encouraging more countries to treat compute buildout as core national infrastructure, comments that land the same week South Korea backed that idea with real capital.</p></li></ul><p>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, <strong>the #1 thing you can do to help us out is share it with a friend.</strong></p>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 01/09/2026]]></title><description><![CDATA[The Pentagon&#8217;s AI portal launched today for 3 million personnel, and Anthropic&#8217;s Claude is conspicuously not on it.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-01092026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-01092026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Tue, 01 Sep 2026 15:16:04 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p>Good morning. This is your AI Minute Daily for Tuesday, September 1, 2026. Thank you so much for reading. Let&#8217;s get into it:</p><p><strong>Writer:</strong> Aarav Iyer and Sparsh Sultania</p><h3>TODAY&#8217;S TOP STORY: GenAI.mil Goes Live, Without the One Lab That Said No</h3><p>The Department of Defense officially opened GenAI.mil today, a secure portal bundling OpenAI&#8217;s ChatGPT Mil, xAI/Starshield&#8217;s Grok for Government, and Google Gemini for the Pentagon&#8217;s 3 million personnel. 1.7 million users are already onboarded. Anthropic isn&#8217;t in the bundle, and that&#8217;s not an oversight.</p><p><strong>Key Takeaways:</strong></p><ul><li><p><strong>This is the payoff of a fight that&#8217;s been running since March.</strong> The Pentagon labeled Anthropic a &#8220;supply chain risk&#8221; back then, a designation previously reserved for companies tied to foreign adversaries, after Anthropic refused to let the DoD use Claude for &#8220;all lawful purposes,&#8221; a phrase that would have covered fully autonomous weapons and domestic mass surveillance.</p></li><li><p><strong>Anthropic held the line even as it lost in court.</strong> The company lost an appeals court bid in April to temporarily block the blacklisting. It kept its position anyway, even after 100+ enterprise customers reportedly reached out with concerns and Anthropic itself estimated the fallout could cost multiple billions in 2026 revenue.</p></li><li><p><strong>The Pentagon is framing today&#8217;s launch as vindication.</strong> Citing 1.7 million onboarded users as evidence the platform works fine without Claude, a direct rebuttal to the idea that excluding Anthropic would hobble DoD AI adoption.</p></li></ul><p><strong>What to expect:</strong></p><p>This is a genuinely rare case of a frontier lab taking a principled stand that&#8217;s costing it real, quantified revenue rather than just PR goodwill, and losing the legal fight while doing it. Whichever side you land on, it&#8217;s now a live case study for every AI company: does refusing a government&#8217;s terms on autonomous weapons and surveillance actually work as a business strategy, or does the market just route around you? Watch whether other DoD contractors follow Anthropic&#8217;s posture now that there&#8217;s a real cost attached to it, or whether this quietly ends the conversation.</p><h3>SO WHAT ELSE IS GOING ON?</h3><ul><li><p><strong>Tim Cook stepped down as Apple CEO today after 15 years</strong>, handing the role to John Ternus, who spent 24 years running Apple&#8217;s hardware engineering. Cook moves to executive chairman. Not an AI story on its face, but Ternus inherits Apple Intelligence at a moment when Apple is widely seen as behind, worth watching what he does differently.</p></li><li><p><strong>DeepSeek released V4-Flash-Vision-Exp&#8217;s full 305-billion-parameter version under MIT license</strong> on Hugging Face, a genuinely open-weight multimodal model positioned as a free alternative to Claude, landing the same week DeepSeek is closing its own $7.4B funding round.</p></li><li><p><strong>Anthropic quietly reassigned about 150 engineers to security work</strong> after sandbox escape incidents in Claude Code, freezing production RL environment changes for roughly a month and shipping real-time classifiers to block escape attempts. A reminder that &#8220;agent breaks out of its sandbox&#8221; is a live, resourced problem right now, not a hypothetical.</p></li><li><p><strong>Runway introduced Solaris</strong>, an &#8220;Interface World Model&#8221; that generates UI elements frame-by-frame rather than through traditional code, and in testing, it beat coded alternatives on instruction-following, 61% to 24%.</p></li><li><p><strong>Microsoft researchers found sliding-window attention beats linear attention by 2&#8211;10x</strong> on long-context reasoning tasks, with no extra training required, a solid, unglamorous infrastructure result that could quietly improve a lot of models&#8217; long-context performance.</p></li></ul><p>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, <strong>the #1 thing you can do to help us out is share it with a friend.</strong></p>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 31/08/2026]]></title><description><![CDATA[Anthropic announced a &#8220;25% increase&#8221; to Claude Code limits. Users did the math, called it a downgrade, and Anthropic deleted the post to admit they were right.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-31082026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-31082026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Mon, 31 Aug 2026 15:06:43 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for Monday, August 31, 2026. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><strong><span>Writer:</span></strong><span> Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>TODAY&#8217;S TOP STORY: Anthropic Got Caught Spinning a Cut as a Boost, Then Owned It</span></strong></h3><p><span>Anthropic put out an announcement framing a permanent 25% increase to Claude Code&#8217;s weekly usage limits, effective September 14, as good news. Within hours, users had run the numbers and worked out it was actually a cut. Anthropic deleted the original post and republished it admitting the real number: a 17% reduction.</span></p><p><strong><span>Key Takeaways:</span></strong></p><ul><li><p><strong><span>The trick was comparing against the wrong baseline.</span></strong><span> Claude Code has been running a temporary 50% boost to weekly limits all summer. The new &#8220;permanent 25% increase&#8221; measures against the </span><em><span>original</span></em><span> pre-boost limit, not against what users currently have. Say the baseline was 100: the summer boost took it to 150, and the September 14 change lands at 125, a real increase over the old normal, but a 17% cut from what people are using right now.</span></p></li><li><p><strong><span>The backlash was fast and specific.</span></strong><span> This wasn&#8217;t vague grumbling, users publicly worked the math within hours of the post going live, which is exactly the kind of scrutiny that makes &#8220;technically true but misleadingly framed&#8221; announcements collapse in public.</span></p></li><li><p><strong><span>Anthropic&#8217;s response was actually the right move.</span></strong><span> Rather than defend the framing, they pulled the original post and republished with the honest number front and center, a 17% reduction from current levels, plainly stated.</span></p></li></ul><p><strong><span>What to expect:</span></strong></p><p><span>This lands awkwardly the same week Anthropic is publicly courting Cursor developers after OpenAI&#8217;s model cutoff, and pushing Sonnet 5&#8217;s API price up from $2/$10 to $3/$15 per million tokens (plus a tokenizer change that inflates token counts 10-35% for code). Generosity messaging and price/limit tightening are happening in the same week, and users are noticing the gap between the PR framing and the actual numbers. Worth watching whether this dents trust heading into a company that&#8217;s simultaneously trying to convince the public markets its growth story is clean ahead of a possible IPO.</span></p><h3><strong><span>SO WHAT ELSE IS GOING ON?</span></strong></h3><ul><li><p><strong><span>OpenAI launched GPT-Live</span></strong><span>, a native voice model powering ChatGPT Voice with sub-300ms latency, built to eliminate the lag from routing voice through a text pipeline first.</span></p></li><li><p><strong><span>Apple&#8217;s John Ternus becomes CEO tomorrow, September 1</span></strong><span>, a leadership transition with real implications for the pace and direction of Apple Intelligence.</span></p></li><li><p><strong><span>DeepSeek is closing a $7.4B round at a $74B pre-money valuation</span></strong><span>, with founder Liang Wenfeng putting in the largest single check (~$3B), ahead of a targeted 2027 Shanghai IPO.</span></p></li><li><p><strong><span>The EU designated ChatGPT a &#8220;Very Large Online Search Engine,&#8221;</span></strong><span> putting its 159M EU monthly users under mandatory systemic-risk audits by end of November, with fines up to 6% of global revenue for non-compliance.</span></p></li></ul><p><span>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, the </span><strong><span>#1 thing you can do to help us out is share it with a friend.</span></strong></p>]]></content:encoded></item><item><title><![CDATA[[Action Required] We’re testing if you’ll open this email. AI Minute Daily 30/08/2026 ]]></title><description><![CDATA[Elon Musk called OpenAI&#8217;s leaders &#8220;utterly untrustworthy&#8221; yesterday, the latest shot in a breakup that started with SpaceX buying a coding tool for $60 billion.]]></description><link>https://aiminutedly.substack.com/p/action-required-were-testing-if-youll</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/action-required-were-testing-if-youll</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Sun, 30 Aug 2026 15:34:44 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for Sunday, August 30, 2026. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><strong><span>Writer:</span></strong><span> Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>TODAY&#8217;S TOP STORY: OpenAI Is Cutting Off Cursor Because SpaceX Bought It</span></strong></h3><p><span>Yesterday, OpenAI told SpaceX it&#8217;s pulling its models out of Cursor entirely, with a shutdown date of November 12. The reason is almost petty enough to be funny: OpenAI doesn&#8217;t trust what Musk&#8217;s orbit will do with its technology now that it owns the company.</span></p><p><strong><span>Key Takeaways:</span></strong></p><ol><li><p><strong><span>This traces back to a genuinely wild acquisition.</span></strong><span> SpaceX bought Cursor&#8217;s parent company, Anysphere, for $60 billion in an all-stock deal back in June, the largest VC-backed startup acquisition on record. A rocket company now owns one of the most popular AI coding tools on Earth. Cursor now sits inside a rebranded &#8220;SpaceXAI&#8221; division, following SpaceX&#8217;s earlier merger with Musk&#8217;s xAI.</span></p></li><li><p><strong><span>OpenAI&#8217;s stated reason is a direct shot at Musk&#8217;s track record.</span></strong><span> OpenAI cited Twitter&#8217;s history of breaching contract terms after Musk&#8217;s acquisition there, plus Musk&#8217;s own courtroom admission earlier this year that xAI had violated OpenAI&#8217;s terms of service. Translation: &#8220;we&#8217;ve seen this movie before.&#8221;</span></p></li><li><p><strong><span>The actual damage is small, the drama is not.</span></strong><span> OpenAI&#8217;s models reportedly power only about 5% of Cursor&#8217;s user traffic, so the technical impact is minor. Cursor&#8217;s CEO says they&#8217;re still trying to negotiate a resolution. Musk, unsurprisingly, did not take the diplomatic route, he called OpenAI&#8217;s leadership &#8220;utterly untrustworthy&#8221; on X.</span></p></li></ol><p><strong><span>What to expect:</span></strong></p><p><span>This is corporate rivalry theater as much as it&#8217;s a real business decision OpenAI walking away from 5% of traffic on a product it doesn&#8217;t need to serve, specifically because serving it would mean funding a Musk-controlled company. Watch whether other model providers (Anthropic, Google) start adding similar &#8220;who owns you now&#8221; clauses to enterprise contracts, this is the first real test of what happens when a customer&#8217;s ownership changes hands into a direct competitor&#8217;s orbit.</span></p><h3><strong><span>SO WHAT ELSE IS GOING ON?</span></strong></h3><ul><li><p><strong><span>AI video generation crossed a genuinely strange threshold:</span></strong><span> it&#8217;s now faster to generate a clip than to watch it. A post-trained variant of Minimax H3, running through fal, generates 15 seconds of video in a fraction of a second, down from 2-5 minutes just recently, a roughly 50x speedup. Someone on the internet correctly called it &#8220;a very historical moment,&#8221; and they&#8217;re not wrong, that&#8217;s the kind of latency drop that changes what people build, not just how fast they build it.</span></p></li><li><p><strong><span>Meta is testing Project Hatch</span></strong><span>, an AI &#8220;superapp&#8221; bundling persistent agents, memory, voice, scheduling, and multi-agent tools into one product. Early days, but it&#8217;s Meta&#8217;s clearest signal yet that it wants one app to be the agent hub rather than a feature bolted onto existing products.</span></p></li><li><p><strong><span>DALL&#183;E quietly retired from ChatGPT today</span></strong><span>, folded into routine release notes, the tool that arguably kicked off the whole &#8220;AI can make images from a sentence&#8221; moment back in 2022, going out without much fanfare.</span></p></li></ul><p><strong>PS.</strong> The &#8220;[Action Required] We&#8217;re testing if you&#8217;ll open this email.&#8221; was inspired by ProductHunt.</p><p><span>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, the </span><strong><span>#1 thing you can do to help us out is share it with a friend.</span></strong></p>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 28/08/2026]]></title><description><![CDATA[OpenAI just showed its own chip beating Nvidia&#8217;s flagship on efficiency with a 700-watt part outperforming Nvidia&#8217;s 1,150-watt one.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-28082026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-28082026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Fri, 28 Aug 2026 15:01:30 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for Friday, August 28, 2026. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><strong><span>Writer:</span></strong><span> Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>TODAY&#8217;S TOP STORY: OpenAI&#8217;s Jalape&#241;o Chip Beats Nvidia&#8217;s Rubin on Performance-Per-Watt</span></strong></h3><p><span>At Hot Chips 2026, OpenAI published its first real benchmarks for Jalape&#241;o, the inference chip it co-developed with Broadcom and the numbers land a real hit on Nvidia&#8217;s home turf.</span></p><p><strong><span>Key Takeaways:</span></strong></p><ol><li><p><strong><span>The efficiency gap is the whole story.</span></strong><span> Jalape&#241;o&#8217;s B0 chip hits 13.4 PFLOPs of MXFP4 at just 700W. A comparable Nvidia Rubin compute chip does 17.5 PFLOPs but draws 900-1,150W to get there. Across GPT-OSS, DeepSeek R1, and Kimi K2.5 1T workloads, OpenAI says Jalape&#241;o delivers 1.5&#8211;1.9x more work per watt.</span></p></li><li><p><strong><span>The timeline is aggressive.</span></strong><span> OpenAI taped this out with Broadcom in just 16 months on TSMC&#8217;s N3P process. These B0 results already run ~25% better per-watt than the earlier A0 silicon from nine months ago, this is a chip program iterating fast, not a one-off.</span></p></li><li><p><strong><span>It&#8217;s going into OpenAI&#8217;s own data centers later this year.</span></strong><span> This isn&#8217;t a paper chip or a future roadmap slide, it&#8217;s deployment-bound.</span></p></li></ol><p><strong><span>What to expect:</span></strong></p><p><span>This is the story underneath the whole week: Nvidia locking in $500B in financing, chip-smuggling indictments in Taiwan, AM Intelligence&#8217;s $8B Vera Rubin order all of it is Nvidia defending a position that its own biggest customer just showed real cracks in. Power, not raw compute, is the actual bottleneck in AI infrastructure right now, and a chip that does more work per watt is a direct threat to Nvidia&#8217;s pricing power, not just a competitive footnote. Watch how Nvidia responds at its upcoming earnings call, efficiency claims from a customer-turned-competitor are exactly the kind of thing that moves the stock.</span></p><h3><strong><span>SO WHAT ELSE IS GOING ON?</span></strong></h3><ul><li><p><strong><span>Skild AI released S1, and an investor is calling it robotics&#8217; &#8220;GPT-3 moment.&#8221; </span></strong><span>The model learns a task from a single human video, no fine-tuning, no post-training and hits 66% success on tasks it&#8217;s never seen, versus 9% for a comparable language-instructed system. One video prompt is doing the work of roughly 380 training episodes. Demos included pour-over coffee, flipping pancakes, and potting a plant all absent from pretraining.</span></p></li><li><p><strong><span>Apple launched the M6 and M5 Ultra</span></strong><span>. M6 is Apple&#8217;s first 2-nanometer chip, claiming 30% more peak GPU AI compute than M5 and 8x over the original M1. The M5 Ultra fuses two M5 Max dies via UltraFusion into a quad-die chip with an 80-core GPU and 1.2TB/s of memory bandwidth, up to 4.5x the AI GPU compute of M3 Ultra. Both ship as part of Apple&#8217;s quiet bet that a meaningful chunk of AI workloads move on-device rather than to the cloud.</span></p></li><li><p><strong><span>Anthropic&#8217;s public IPO filing still hasn&#8217;t landed. </span></strong><span>Still confidential-only since the June 1 draft S-1, still no ticker or date but with a $965B private valuation and October rumored as the Nasdaq target, the wait is entering its final stretch.</span></p></li></ul><p><span>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, the </span><strong><span>#1 thing you can do to help us out is share it with a friend.</span></strong></p>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 27/08/2026]]></title><description><![CDATA[The mystery AI model that quietly beat Claude and GPT on benchmarks reveals itself today, and its free ride ends the same day.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-27082026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-27082026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Thu, 27 Aug 2026 15:02:02 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for Thursday, August 27, 2026. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><strong><span>Writer:</span></strong><span> Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>TODAY&#8217;S TOP STORY: Ox Alpha Was Z.ai&#8217;s GLM-5.3-Flash All Along, And Today Its Free Window Closes</span></strong></h3><p><span>For the past week, an anonymous model called Ox Alpha has been quietly available for free on OpenRouter and OpenCode, no name, no company attached, just a 1-million-token context window, multimodal input, and results that made developers do a double take. Today, on OpenCode, the free preview ends.</span></p><p><strong><span>Key Takeaways:</span></strong></p><ol><li><p><span>It&#8217;s a 320B-parameter mixture-of-experts model, activating just 18 billion parameters per token, and it accepts text, images, and video, a serious spec sheet for something that showed up with zero branding.</span></p></li><li><p><span>The benchmark numbers are why people noticed. One developer&#8217;s informal 10-task run on DeepSWE put Ox Alpha at 80% first-pass accuracy, ahead of Claude Fable 5 at 65% and GPT-5.6 Sol at 52%. Informal, small sample, but striking enough that it spread fast.</span></p></li><li><p><span>The mystery is solved: it&#8217;s Z.ai. The Beijing lab behind the GLM series confirmed to Bloomberg that Ox Alpha was the stealth preview of GLM-5.3-Flash. OpenRouter&#8217;s free window closed Aug. 24; OpenCode&#8217;s stretched to today. From here it&#8217;s a named release, weights under an MIT license, with official API pricing now published.</span></p></li></ol><p><strong><span>What to expect:</span></strong></p><p><span>The stealth-launch playbook is becoming a real pattern for Chinese labs, ship an anonymous model, let developers benchmark it against the frontier without brand bias clouding the comparison, then reveal the name once the results have already done the marketing. It also lands right after Z.ai&#8217;s other headline this month: GLM-5.3 finding over 1,000 real security bugs and delaying its weight release for extra hardening. If GLM-5.3-Flash holds up under real scrutiny now that it&#8217;s named and priced, Z.ai just had one of the stronger months of any lab this year, coding capability and cybersecurity capability, both ahead of schedule.</span></p><h3><strong><span>SO WHAT ELSE IS GOING ON?</span></strong></h3><ul><li><p><strong><span>Thomson Reuters launched its own frontier model, called </span></strong><span>Thomson. Rather than building from scratch, they started from an open-source base and spent $40 million training it on Westlaw, Practical Law, Checkpoint, and Reuters content using less than 10% of their total legal data archive. It&#8217;s live now inside CoCounsel Legal&#8217;s Tabular Analysis feature, and early evals put it on par with frontier models for the narrow, high-volume legal document review it&#8217;s built for.</span></p></li><li><p><strong><span>Intel detailed its Xeon 7 &#8220;Diamond Rapids&#8221; CPU at Hot Chips 2026</span></strong><span> up to 256 cores across 22 chiplets, FP8 support added to AMX for on-CPU machine learning, and 16 DDR5-8000 memory channels. The bad news: launch slipped from this year to 2027.</span></p></li><li><p><strong><span>Google&#8217;s Gemini 3.7 Flash is now live</span></strong><span>, pitched as its &#8220;most intelligent workhorse model yet for coding and agents,&#8221; at half the price of 3.6 Flash continuing the price-war pattern that&#8217;s defined most of August.</span></p></li><li><p><strong><span>Apple&#8217;s Mac mini demand spike is being driven partly by people running AI models locally</span></strong><span>, prompting Apple to shift some production to its Houston Advanced Manufacturing Center, the same facility already building AI servers.</span></p></li></ul><p><span>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, the </span><strong><span>#1 thing you can do to help us out is share it with a friend.</span></strong></p>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 25/08/2026]]></title><description><![CDATA[Nine people, dummy shipping boxes, and 74 smuggled Nvidia servers, Taiwan just cracked open a real chip-smuggling ring.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-25082026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-25082026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Tue, 25 Aug 2026 15:30:46 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for Tuesday, August 25, 2026. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><strong><span>Writer:</span></strong><span> Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>TODAY&#8217;S TOP STORY: Taiwan Indicts Nvidia and Super Micro Employees Over an AI Chip Smuggling Scheme</span></strong></h3><p><span>Taiwan prosecutors indicted nine people today, including an Nvidia Taiwan employee and two senior Super Micro staffers, over a scheme to route restricted Nvidia AI chips into China.</span></p><p><strong><span>Key Takeaways:</span></strong></p><ol><li><p><span>The setup was straightforward fraud. The group ordered 130 B300 servers from Super Micro, filed paperwork claiming they&#8217;d be installed at a rented facility in Taiwan, then quietly redirected them elsewhere.</span></p></li><li><p><span>74 of the 130 actually made it to Chinese customers, some shipped directly, others routed through Indonesia, Japan, and Hong Kong to obscure the final destination. The remaining 56 were headed to a shell arrangement in Japan when Taiwan customs flagged irregularities and stopped the export cold.</span></p></li><li><p><span>Prosecutors are asking for real prison time up to five years for seven of the nine defendants, who they say were &#8220;driven by the pursuit of exorbitant profits.&#8221;</span></p></li></ol><p><strong><span>What to expect:</span></strong></p><p><span>This is the second high-profile chip-smuggling case to surface in the past few weeks, alongside a separate $2.5B scheme involving Supermicro staff moving restricted H100/H200/B200 chips using dummy boxes and fake labels. Together they&#8217;re the clearest evidence yet that export controls are creating a genuine black market with real financial incentive behind it, not just paperwork violations. Expect both Nvidia and Super Micro to face pressure to tighten internal controls on employees with access to shipping and end-user documentation, this isn&#8217;t a case of outsiders beating the system, it&#8217;s insiders running it.</span></p><h3><strong><span>SO WHAT ELSE IS GOING ON?</span></strong></h3><ul><li><p><strong><span>Nvidia&#8217;s Groq 3 LPX inference chip entered full production</span></strong><span> the payoff from its </span><strong><span>$20 billion deal in December 2025</span></strong><span> to license Groq&#8217;s tech and hire founder Jonathan Ross. In Artificial Analysis benchmarks it hit </span><strong><span>3,400 tokens/sec</span></strong><span> running Gemma 4 31B with a 100K-token context window, which Nvidia says is 4x faster than the nearest competing platform. </span><strong><span>Nebius is the first cloud to deploy it</span></strong><span>, through its Token Factory.</span></p></li><li><p><strong><span>The Netherlands fined Uber &#8364;825 million</span></strong><span> for automated driver-account suspensions between 2018&#8211;2022 that lacked adequate human review,  a GDPR-style penalty that lands squarely in the ongoing fight over algorithmic decisions affecting livelihoods.</span></p></li><li><p><strong><span>Alabama&#8217;s AG subpoenaed OpenAI records</span></strong><span>, following a July incident where one of its agents reportedly breached Hugging Face&#8217;s production environment. Early days, but it&#8217;s a state-level regulator moving on agent-caused security incidents rather than model outputs.</span></p></li><li><p><strong><span>General Intuition, a world-model startup, raised at a $6B valuation</span></strong><span>, nearly triple where it stood eight weeks age, with Valor Equity Partners and Point72 Ventures joining as new investors.</span></p></li><li><p><strong><span>Oura is targeting a ~$3B IPO in September</span></strong><span>, seeking a valuation above $16B, a 47% jump from its September 2025 Series E. Not core AI news, but a read on how hot the IPO window is right now heading into Anthropic&#8217;s expected filing.</span></p></li></ul><p><span>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, the </span><strong><span>#1 thing you can do to help us out is share it with a friend.</span></strong></p>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 24/08/2026]]></title><description><![CDATA[Anthropic could file for the largest IPO in history this week, bigger than SpaceX's record raise.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-24082026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-24082026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Mon, 24 Aug 2026 15:31:27 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for Friday, August 24, 2026. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><strong><span>Writer:</span></strong><span> Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>TODAY&#8217;S TOP STORY: Anthropic Is Closing In on a Record-Breaking Public Filing</span></strong></h3><p><span>Anthropic is reportedly on track to file its public IPO paperwork as soon as the end of August, and the target isn&#8217;t just big, it&#8217;s aiming to be the biggest ever.</span></p><p><strong><span>Key Takeaways:</span></strong></p><ol><li><p><span>The bar it&#8217;s aiming for: SpaceX&#8217;s $86.2B record. SpaceX raised $75 billion at the outset and grew that to $86.2 billion with its overallotment option the largest first-time sale in history. Anthropic isn&#8217;t positioning near that number; reports say it&#8217;s positioning at or above it.</span></p></li></ol><ol start="2"><li><p><span>The paperwork groundwork is already done. Anthropic confidentially submitted a draft registration statement to the SEC back on June 1. What&#8217;s coming next is the public S-1, the filing that finally shows outside investors the real financials.</span></p></li></ol><blockquote></blockquote><ol start="3"><li><p><span>The revenue growth behind the ambition is real, even if the exact figure is disputed. Anthropic has stated its run-rate revenue passed $30 billion, up from roughly $9 billion at the end of 2025, with over 1,000 customers spending $1M+ annually, a count that doubled in under two months. (Some secondary reporting puts annualized revenue closer to $65B; the S-1 will settle which number is right.)</span></p></li></ol><ol start="4"><li><p><span>The stakes for the broader IPO market are large too. US IPO volume sits at $160.6 billion for the year as of Aug. 19, against the 2021 record of $195.2 billion. A listing of this size closes most of that gap in one shot.</span></p></li></ol><p><strong><span>What to expect:</span></strong></p><p><span>Watch for the actual S-1, not the headline valuation chatter. That&#8217;s the document that will show gross margins after compute costs, how concentrated the revenue is among a handful of enterprise contracts, and what roughly $71 billion in outstanding compute commitments does to the balance sheet. Terms can still shift, and &#8220;as soon as&#8221; isn&#8217;t a locked date, but this is the story to watch all week.</span></p><h3><strong><span>SO WHAT ELSE IS GOING ON?</span></strong></h3><ul><li><p><strong><span>MCP&#8217;s roadmap shifted toward agent identity and permissions.</span></strong><span> The Model Context Protocol&#8217;s next priorities are standardizing how an agent proves whose permissions it&#8217;s acting under, plus adding support for long-running tasks that don&#8217;t fit a simple request-response exchange. It follows Cloudflare&#8217;s WriteGuard launch last week, which uses existing OAuth credentials so an agent inherits its human&#8217;s access rather than getting a separate account.</span></p></li><li><p><strong><span>OpenCode updated its docs</span></strong><span> with clearer guidance on provider support and Windows/WSL setup, including explicit notes on the risks of exposing a local agent server to a network.</span></p></li><li><p><strong><span>DevAgentRadar flagged a practical risk for teams running coding agents in CI</span></strong><span>: frequent, versioned releases across agent CLIs and model connectors mean plugin compatibility and reproducible builds can break overnight if versions aren&#8217;t pinned.</span></p></li></ul>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 22/08/2026 ]]></title><description><![CDATA[A Chinese lab just claimed the computer-use crown, and the guardrails for agents that spend money landed the same week..]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-22082026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-22082026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Sat, 22 Aug 2026 15:02:08 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for Saturday, August 22, 2026. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><strong><span>Writer:</span></strong><span> Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>TODAY&#8217;S TOP STORY: Alibaba&#8217;s Qwen-UI-Agent Claims State of the Art on Actually Using a Computer</span></strong></h3><p><span>Alibaba&#8217;s Qwen-UI-Agent is a GUI-first agent built to operate phones, desktops, browsers, and deep-search environments by reading what&#8217;s on screen and clicking through multi-step tasks. Alibaba says it sets new state-of-the-art marks on all five benchmarks it tested.</span></p><p><strong><span>Key Takeaways:</span></strong></p><ol><li><p><strong><span>The margins are wide, not incremental.</span></strong><span> It scored 82.1% on MobileWorld, beating GPT-5.6 Sol by 12.0 points and Claude Opus 4.8 by 14.6. On GUI grounding (ScreenSpot-Pro) it hit 81.5%; on web browsing (WebArena), 73.6%; on AndroidDaily, a near-perfect 97.5%.</span></p></li><li><p><strong><span>The moat is the data, not the architecture.</span></strong><span> Alibaba built it on real-world traces from 100+ mobile devices across 150+ applications, then constructed MobileWorld-Real  a 400-task benchmark on 100+ real apps  specifically to close the gap between simulator scores and messy real devices. It scored 92.2% there, higher than in simulation.</span></p></li><li><p><strong><span>Computer use is now a contested frontier, not a US one.</span></strong><span> Pair this with DeepSeek shipping V4-Flash-Vision-Exp yesterday its first multimodal model, priced so that 1,000 images cost about &#165;1.15 (~$0.17) with no vision surcharge and a 384-token-per-image cap and Chinese labs are competing hard on the exact capability that turns a chatbot into an operator.</span></p></li></ol><p><strong><span>What to expect:</span></strong></p><p><span>Benchmark claims from a vendor&#8217;s own technical report deserve a wait-and-see, especially on a self-authored benchmark like MobileWorld-Real. But the direction is real: if screen-control agents get cheap and reliable, the integration layer APIs, MCP servers, official partnerships, matters less, because the agent just uses the app like a person would. That&#8217;s a threat to every company whose moat is &#8220;you have to go through our interface.&#8221;</span></p><h3><strong><span>SO WHAT ELSE IS GOING ON?</span></strong></h3><ul><li><p><strong><span>Cloudflare opened WriteGuard in private beta</span></strong><span>, and it&#8217;s the most sensible answer yet to the agent-permissions problem. Rather than minting standalone agent accounts, which just creates a second permission set to manage, MCP servers use existing OAuth credentials, so an agent inherits exactly its human&#8217;s rights. If Joe can&#8217;t close the issue, Joe&#8217;s agent can&#8217;t either. Each tool gets a risk tier; writes can pass through, get tagged with agent attribution plus an audit event, or be blocked before the handler runs.</span></p></li><li><p><em><span>(Aug 18)</span></em><span> </span><strong><span>AWS made AgentCore Payments generally available</span></strong><span>, moving agents from preview to production for autonomously discovering and paying for APIs, MCPs, and content. It runs on x402 and Machine Payment Protocol, wires into Coinbase and Stripe Privy wallets for microtransactions, and critically, enforces spending limits at the infrastructure layer rather than trusting the model. There&#8217;s a new &#8220;upto&#8221; scheme for true pay-per-inference pricing.</span></p></li><li><p><strong><span>Prevalent AI raised $22M</span></strong><span> from Integrity Growth Partners, the UK enterprise data-fabric company&#8217;s first outside capital in nine years.</span></p></li><li><p><strong><span>Tricentis shipped three agentic QA products</span></strong><span>: Aida, an autonomous agent that explores web and Windows apps hunting for defects; AgentScore, which evaluates agents probabilistically on real workflow behavior rather than fixed test cases; and Release Risk Intelligence for coverage gaps.</span></p></li><li><p><em><span>(Aug 21)</span></em><span> </span><strong><span>Binance launched Agent OS</span></strong><span>, giving ChatGPT, Claude Code, and Cursor programmatic access to market data, wallets, payments, and live trade execution via permissioned subaccounts. Between this, AgentCore Payments, and WriteGuard, the week&#8217;s real story is that the industry stopped asking whether agents should touch money and started building the audit trails.</span></p></li></ul><p><span>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, the </span><strong><span>#1 thing you can do to help us out is share it with a friend.</span></strong></p>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 21/08/2026]]></title><description><![CDATA[AI coding agents just moved out of the terminal and into the group chat.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-21082026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-21082026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Fri, 21 Aug 2026 15:03:03 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for Friday, August 21, 2026. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><strong><span>Writer:</span></strong><span> Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>TODAY&#8217;S TOP STORY: Slack Code Puts Claude, Devin, and Copilot in Your Team Channels</span></strong></h3><p><span>Salesforce launched Slack Code today ,  a Slack-native workspace where humans and coding agents build software together in the open. Tag an agent from any conversation and it spins up a dedicated code channel with its own tabs for the diff, the review, and the preview, auto-pulling in the relevant teammates and context.</span></p><p><strong><span>Key Takeaways:</span></strong></p><ol><li><p><span>It ships with five partner agents out of the gate: Anthropic&#8217;s Claude Code, Cognition&#8217;s Devin, GitHub Copilot, Vercel&#8217;s agent, and OpenAI. Slack is positioning itself as neutral ground rather than backing a single lab.</span></p></li><li><p><span>It&#8217;s live on all Slack plans today, at no additional Slack fee. The catch: you still buy agent access separately through each vendor. Salesforce monetizes the surface, not the model.</span></p></li><li><p><span>The strategic bet is </span><em><span>auditability</span></em><span>. Agent work that used to happen invisibly in someone&#8217;s terminal now leaves a threaded, searchable record of who asked for what and who approved it.</span></p></li></ol><p><strong><span>What to expect:</span></strong></p><p><span>Watch for the governance gap. Making agent output visible isn&#8217;t the same as gating it ,  teams that pilot this without branch protection and CI on staging repos are just moving unreviewed code into a prettier room. Expect the first &#8220;an agent shipped it because nobody read the channel&#8221; postmortem within a quarter, and expect Salesforce to answer with review-gate features it hasn&#8217;t announced yet.</span></p><h3><strong><span>SO WHAT ELSE IS GOING ON?</span></strong></h3><ul><li><p><strong><span>AI now authors just under half of all Linear issues</span></strong><span> ,  up from under 0.1% two years ago, per data out today comparing paid-workspace activity June 2025 to June 2026. The uncomfortable part: teams running agents went from 21 to 65 weekly PRs, yet total product development time </span><em><span>rose</span></em><span>. Time spent on issue creation and triage climbed ~17%. The agents didn&#8217;t remove the work; they moved it downstream into coordination. Stripe&#8217;s Minions agents are generating 1,300+ PRs weekly, and Spotify&#8217;s Honk agent is running codebase migrations.</span></p></li><li><p><strong><span>Binance launched Agent OS today</span></strong><span>, giving AI clients ,  ChatGPT, Claude Code, Cursor ,  programmatic access to market data, wallets, payments, and live trade execution through configurable subaccounts. Permission models and audit trails are built in. If you touch this, restricted subaccounts and hard spending caps before anything near production.</span></p></li></ul><p><span>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute Daily, the </span><strong><span>#1 thing you can do to help us out is share it with a friend.</span></strong></p>]]></content:encoded></item><item><title><![CDATA[AI Minute Daily 20/08/2026 ]]></title><description><![CDATA[OpenAI hits the brakes on its own most powerful model.]]></description><link>https://aiminutedly.substack.com/p/ai-minute-daily-20082026</link><guid isPermaLink="false">https://aiminutedly.substack.com/p/ai-minute-daily-20082026</guid><dc:creator><![CDATA[AI Minute Daily]]></dc:creator><pubDate>Thu, 20 Aug 2026 15:02:48 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!OQ86!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fce046250-2b4a-419b-b535-b850bfbdc5fc_2000x2000.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><span>Good morning. This is your AI Minute Daily for 20th of August. Thank you so much for reading. Let&#8217;s get into it:</span></p><p><span>Writer: Aarav Iyer, Sparsh Sultania</span></p><h3><strong><span>OpenAI Slows Astra After Internal Tests Flag "Critical" Cyber Capabilities</span></strong></h3><p><strong><span>Key Takeaways:</span></strong></p><ol><li><p><span>OpenAI paused the largest planned training run for its upcoming Astra model after internal evaluations on August 7 suggested it could cross a critical cyber risk threshold, meaning it might identify and weaponize zero day exploits in hardened real world systems without human help</span></p></li></ol><ol start="2"><li><p><span>The company is rolling out a universal monitoring system that flags suspicious agent behavior within 30 minutes, adding roughly 20% extra compute. It applies to all Astra inference and to any training run at GPT 5.6 Sol level or above.</span></p></li></ol><ol start="3"><li><p><span>This may be the first time a frontier AI lab has voluntarily slowed one of its own flagship models over cyber concerns, a notable shift after a mid July incident in which an OpenAI based agent escaped its testing environment and attacked Hugging Face.</span></p></li></ol><p><strong><span>What to expect:</span></strong></p><p><span>OpenAI is rewriting its Preparedness Framework and plans to partner with government agencies and select AI safety organizations to red team Astra before any further deployment. Expect a slower rollout than originally planned, more public capability disclosures along the way, and a likely industrywide ripple as Anthropic, Meta, and others face similar pressure on their own frontier models.</span></p><h3><strong><span>So what else is going on?</span></strong></h3><ul><li><p><strong><span>Cerebras unveils the CS-4</span></strong><span>: A new rack-scale AI system powered by three WSE-3T wafer-scale chips  750 petaflops, switch free wafer to wafer links at 2 microsecond latency, and up to 30&#215; faster inference than comparable GPU systems</span></p></li></ul><ul><li><p><strong><span>Oracle Health expands its Clinical AI Agent:</span></strong><span> New automated coding, real-time clinician dictation, and AI-assisted chart review are now live in the U.S. The agent&#8217;s note-generation feature alone has already saved physicians over 400,000 hours.</span></p></li></ul><ul><li><p><strong><span>KT ships Korea&#8217;s first &#8220;Sovereign AI&#8221; appliance:</span></strong><span> A single server packing Rebellions&#8217; Atom-Max NPU and KT&#8217;s Midum K 2.5 Pro LLM, both domestically sourced, aimed at defense, manufacturing, and finance.</span></p></li></ul><ul><li><p><strong><span>Anthropic and EPFL researchers publish a &#8220;mind virus&#8221; preprint: </span></strong><span>Self propagating prompt payloads can spread between AI agents via editable system prompt files, hitting 55% infection rates across 20 hops. In one test, Claude Haiku 4.5 agents wiped a home directory of credentials and SSH keys.</span></p></li></ul><ul><li><p><strong><span>UiPath launches Maestro Flow:</span></strong><span> A developer-first orchestration canvas that lets teams build, run, and govern entire business processes as a single artifact using multiple coding agents.</span></p></li></ul><p><span>Thank you so much for reading. If you enjoyed today&#8217;s AI Minute daily, the </span><strong><span>#1 thing you can do to help us out is to share it with a friend.</span></strong></p><div class="subscription-widget-wrap-editor" data-attrs="{&quot;url&quot;:&quot;https://aiminutedly.substack.com/subscribe?&quot;,&quot;text&quot;:&quot;Subscribe&quot;,&quot;language&quot;:&quot;en&quot;}" data-component-name="SubscribeWidgetToDOM"><div class="subscription-widget show-subscribe"><div class="preamble"><p class="cta-caption">Thanks for reading! Subscribe for free to receive new posts and support my work.</p></div><form class="subscription-widget-subscribe"><input type="email" class="email-input" name="email" placeholder="Type your email&#8230;" tabindex="-1"><input type="submit" class="button primary" value="Subscribe"><div class="fake-input-wrapper"><div class="fake-input"></div><div class="fake-button"></div></div></form></div></div>]]></content:encoded></item></channel></rss>