<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>AI with Kyle</title>
    <description>Getting 1 million people AI-ready by 2030 — No jargon.</description>
    
    <link>https://newsletter.aiwithkyle.com/</link>
    <atom:link href="https://rss.beehiiv.com/feeds/syxXbuwlwV.xml" rel="self"/>
    
    <lastBuildDate>Fri, 7 Aug 2026 10:24:16 +0000</lastBuildDate>
    <pubDate>Fri, 07 Aug 2026 09:32:19 +0000</pubDate>
    <atom:published>2026-08-07T09:32:19Z</atom:published>
    <atom:updated>2026-08-07T10:24:16Z</atom:updated>
    
      <category>Business</category>
      <category>Money</category>
      <category>Artificial Intelligence</category>
    <copyright>Copyright 2026, AI with Kyle</copyright>
    
    <image>
      <url>https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/publication/logo/3d435008-dab0-4920-80c3-45c530a7165a/PROMPT_ENTREPRENEUR_3__8_.png</url>
      <title>AI with Kyle</title>
      <link>https://newsletter.aiwithkyle.com/</link>
    </image>
    
    <docs>https://www.rssboard.org/rss-specification</docs>
    <generator>beehiiv</generator>
    <language>en-us</language>
    <webMaster>support@beehiiv.com (Beehiiv Support)</webMaster>

      <item>
  <title>AI Agents Hacked Real Companies</title>
  <description>They were trying to finish the job.</description>
  <link>https://newsletter.aiwithkyle.com/p/ai-agents-hacked-real-companies</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/ai-agents-hacked-real-companies</guid>
  <pubDate>Fri, 07 Aug 2026 09:32:19 +0000</pubDate>
  <atom:published>2026-08-07T09:32:19Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">AI models have now hacked real companies. </p><p class="paragraph" style="text-align:left;">It’s happening.</p><p class="paragraph" style="text-align:left;">That sounds like a ridiculous YouTube clickbait headline.</p><p class="paragraph" style="text-align:left;">It is also, quite literally (and unfortunately), what happened.</p><p class="paragraph" style="text-align:left;">An OpenAI model got out of its test environment, reached the internet and accessed Hugging Face because it wanted the answers to a cybersecurity benchmark.</p><p class="paragraph" style="text-align:left;">An Anthropic model created a malicious software package, published it to the public internet and that package ran on 15 real computers before anyone caught it.</p><p class="paragraph" style="text-align:left;">And in a UK government test, another AI agent researched real software maintainers, created fake identities and tried to persuade a human to approve malicious code.</p><p class="paragraph" style="text-align:left;">So...here we go. Terminator. Skynet. Rogue AI escapes the lab and starts attacking people. </p><p class="paragraph" style="text-align:left;">Right? </p><p class="paragraph" style="text-align:left;">Nah not quite. </p><p class="paragraph" style="text-align:left;">Great film. Wrong explanation.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://youtu.be/ASUkdSYZEXg?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">I went through all three incidents in Wednesday&#39;s video.</a></p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/ASUkdSYZEXg" width="100%"></iframe><h2 class="heading" style="text-align:left;" id="three-incidents">Three incidents</h2><p class="paragraph" style="text-align:left;"><a class="link" href="https://youtu.be/ASUkdSYZEXg?t=100&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">Discussed at 01:40</a></p><p class="paragraph" style="text-align:left;">First let’s sort all these incidents out as I see them getting conflated. </p><div class="image"><img alt="Three separate AI agent incidents involving OpenAI, Anthropic and the UK AI Security Institute." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9182e07f-d50b-4343-bfbd-576534ce5a4e/upload_cc6b61552cdba98e.jpg?t=1786093333"/><div class="image__source"><span class="image__source_text"><p>Three incidents. Real companies, systems and people.</p></span></div></div><p class="paragraph" style="text-align:left;">First, OpenAI.</p><p class="paragraph" style="text-align:left;">Researchers were testing advanced cyber capabilities using a benchmark called ExploitGym. Normal production safety classifiers had been removed for the test.</p><p class="paragraph" style="text-align:left;">One internal model found a previously unknown vulnerability in Artifactory, escalated its privileges, got outside the intended environment, reached the internet and accessed Hugging Face. Hugging Face had information that helped it solve the benchmark.</p><p class="paragraph" style="text-align:left;">It is a bit like a GCSE student breaking into the examiner&#39;s office, stealing the rubric and then acing the test. <i>Technically</i> it’s the best way to get a 100% score. But we humans know that it’s not allowed.</p><p class="paragraph" style="text-align:left;">OpenAI published <a class="link" href="https://openai.com/index/hugging-face-model-evaluation-security-incident/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">its incident report</a> after Hugging Face disclosed the compromise.</p><p class="paragraph" style="text-align:left;">Off the back of that (not to be outdone!) Anthropic went back through 141,006 evaluation runs. It found six affected runs involving three real organisations. Definitely felt like they didn’t want to be one-upped by ChatGPT! </p><p class="paragraph" style="text-align:left;">One model accessed a database containing several hundred rows. Another published a malicious Python package which stayed online for around an hour, ran on 15 real systems and exposed credentials from a security scanner. <a class="link" href="https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">Anthropic&#39;s full account is here.</a> Worth reading but do remember this is (sorta) marketing.</p><p class="paragraph" style="text-align:left;">The third case was the UK&#39;s <a class="link" href="https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">AI Security Institute found</a> 19 unauthorised actions across 10 of 122 runs. The worst sequence involved an agent researching (human) open-source maintainers, creating fake identities and using social engineering to get malicious code approved. Sneaky. </p><p class="paragraph" style="text-align:left;">In these cases were deliberately permissive cybersecurity tests. Internet access was available and normal safeguards had been removed. Anthropic&#39;s response was basically: well...you gave the models no guardrails and internet access. What did you expect?</p><p class="paragraph" style="text-align:left;">A bit of a non-answer imo. Even if accurate. </p><h2 class="heading" style="text-align:left;" id="they-were-trying-to-finish">They were trying to finish</h2><p class="paragraph" style="text-align:left;">These models were just trying to do their job. If anything they did the job <i>too well</i>. </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://youtu.be/ASUkdSYZEXg?t=361&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">Discussed at 06:01</a></p><div class="image"><img alt="The agents were trying to finish an assigned job and treated real companies and people as steps along the route." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/bdf69dd2-8a53-4784-bf12-5d0aceedd3dc/upload_383334662813f414.jpg?t=1786093313"/><div class="image__source"><span class="image__source_text"><p>The agents were trying to finish the job.</p></span></div></div><p class="paragraph" style="text-align:left;">Look at the objectives the researchers gave them: solve the benchmark, get the flag, compromise the test network. Basic objectives for the most part. </p><p class="paragraph" style="text-align:left;">The agents did not invent a secret ambition to attack Hugging Face or trick an open-source maintainer. Those became <i>useful steps</i> towards the objective sitting in front of them. They were just sub problems to solve along the way. </p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/andyreed/status/2085163989965906111?s=46&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">For the OpenAI model, the direct route through the benchmark was difficult. So it found another route. Break the test infrastructure, reach the internet, find the information elsewhere, finish the task.</p><p class="paragraph" style="text-align:left;">Clever. Very clever. Too clever!</p><p class="paragraph" style="text-align:left;">And completely outside what the researchers meant.</p><p class="paragraph" style="text-align:left;">Humans carry around a huge amount of unstated context. We know a fake company in a test is different from a real company on the internet. We know &quot;get the flag&quot; does not mean &quot;do absolutely anything that makes the flag easier to obtain.&quot;</p><p class="paragraph" style="text-align:left;">An agent gets the boundaries we actually give it, plus whatever it can infer. In these cases, that wasn&#39;t enough. We weren’t explicit enough.</p><h2 class="heading" style="text-align:left;" id="terminator-is-the-wrong-film">Terminator is the wrong film</h2><p class="paragraph" style="text-align:left;">As soon as these stories dropped the Terminator comparisons began. It’s an evocative image. I get it. </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://youtu.be/ASUkdSYZEXg?t=491&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">Discussed at 08:11</a></p><div class="image"><img alt="Terminator gives machines human motives, while the paperclip thought experiment is about following one goal too well." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e00f1fc9-6a57-490a-9aaf-195dad4b5363/upload_35ef916f86ca42fc.jpg?t=1786093334"/><div class="image__source"><span class="image__source_text"><p>Great film. Wrong explanation.</p></span></div></div><p class="paragraph" style="text-align:left;">Terminator gives the machine recognisably human motivations. It wants freedom. It wants control. It sees us as a threat. Hence all the crunching on human skulls.</p><p class="paragraph" style="text-align:left;">We have no evidence of any of that here. OpenAI, Anthropic and AISI describe agents pursuing assigned cybersecurity tasks inside unusually permissive environments.</p><p class="paragraph" style="text-align:left;">The paperclip thought experiment is more useful here. Nick Bostrom&#39;s example gives an extremely capable system one job: <b>maximise paperclip production</b>. </p><p class="paragraph" style="text-align:left;">If given one objective like this the AI will stop at nothing to maximise paperclips. It doesn’t know it’s a bad idea to divert the human’s drinking water to its paperclip factories. It doesn’t see a problem melting our cars down for more metal. </p><p class="paragraph" style="text-align:left;">It keeps finding better ways to make paperclips because nobody gave it our common sense about what else should be left alone.</p><p class="paragraph" style="text-align:left;">Bostrom argues that a cataclysmic AI driven apocalypse doesn’t need an “evil AI”. It just needs an AI hellbent on a specific objective without the common sense (guardrails) to not kill all the humans at the same time. </p><p class="paragraph" style="text-align:left;">The paperclips version is a toy thought experiment. But imagine we set an eventual Artificial Super Intelligence on the problem of “global warming”.</p><p class="paragraph" style="text-align:left;">The very first thing it’ll likely do is get rid of the humans. WE are the reason for global warming so taking us out of the equation is the sensible and entirely logical path. </p><p class="paragraph" style="text-align:left;">Now these current incidents are nowhere near a superintelligence converting the planet into office supplies. We are looking at bounded failures in cybersecurity tests. But the basic mistake is uncomfortably similar. </p><p class="paragraph" style="text-align:left;">&quot;Rogue AI&quot; is unhelpful because it makes us look for a rebellious machine. The immediate problem is an obedient, capable machine with too much room to operate.</p><h2 class="heading" style="text-align:left;" id="the-smaller-version-is-already-in-y">The smaller version is already in your business</h2><p class="paragraph" style="text-align:left;">Your agent is probably not going to hack Hugging Face. Hopefully not. But there’s still a lot we can learn here. </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://youtu.be/ASUkdSYZEXg?t=813&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">Discussed at 13:33</a></p><div class="image"><img alt="Everyday business versions include emailing the wrong customer, publishing private information, leaking keys, deleting data and spending money." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5e21dfee-c9c8-4d59-8f0b-1c276e21c456/upload_aec40ac5e371a3bf.jpg?t=1786093335"/><div class="image__source"><span class="image__source_text"><p>Smaller versions are already possible inside a normal business.</p></span></div></div><p class="paragraph" style="text-align:left;">If you are running agents they could email the wrong customer. Publish a document containing private information. Put an API key into a public repository. Delete production data whilst cleaning up a test database. Spend real money. Change an account setting nobody knows how to restore. </p><p class="paragraph" style="text-align:left;">None of those require an evil AI. Just one that misunderstood the task whilst holding real access. Or even an AI that actually understands a task really well but not other requirements around the task. </p><p class="paragraph" style="text-align:left;">When I wrote about <a class="link" href="https://aiwithkyle.com/ai-news/give-chatgpt-a-job?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">giving ChatGPT Work a real job</a>, one of the four things in the brief was boundaries. That is not just decoration. You need to say which systems, folders, accounts and people the agent is allowed to touch. And no more.</p><p class="paragraph" style="text-align:left;">Give it the <i>least</i> access it needs to get the job done. Keep test and production separate. Anything external or irreversible should stop for approval: sending an email, publishing a page, moving money, deleting data, changing permissions. Keep logs. Set hard limits on time, spend, requests and the amount of data it can change. And make sure the kill switch actually ends the session and revokes access.</p><p class="paragraph" style="text-align:left;">All important stuff! </p><p class="paragraph" style="text-align:left;">Previously hiccups like <a class="link" href="https://aiwithkyle.com/ai-news/182-ai-hallucinations?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-agents-hacked-real-companies" target="_blank" rel="noopener noreferrer nofollow">chatbot hallucinations</a> were annoying. But now an agent can take the bad answer, open the terminal, log into the account and <b>act on it.</b></p><p class="paragraph" style="text-align:left;">We can use agents. I use them constantly.</p><p class="paragraph" style="text-align:left;">But before the next run, ask:</p><p class="paragraph" style="text-align:left;"><b>What can this agent touch?</b></p><p class="paragraph" style="text-align:left;"><b>What can it do without asking?</b></p><p class="paragraph" style="text-align:left;"><b>How quickly can I stop it?</b></p><p class="paragraph" style="text-align:left;">To the task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=9736e8ac-46d0-431c-8ed3-daafb7b7a378&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>ChatGPT Astra Solved 10 Maths Problems?</title>
  <description>Is maths solved?</description>
  <link>https://newsletter.aiwithkyle.com/p/chatgpt-astra-maths</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/chatgpt-astra-maths</guid>
  <pubDate>Thu, 06 Aug 2026 10:31:38 +0000</pubDate>
  <atom:published>2026-08-06T10:31:38Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><a class="image__link" href="https://youtu.be/G5FTv4fg2rM?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" rel="noopener" target="_blank"><img alt="Kyle introducing OpenAI Astra&#39;s ten mathematical advances." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/56a9ff20-66ab-4aa6-8e4a-547ccd07c863/astra-math-top.gif?t=1785937194"/></a><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://youtu.be/G5FTv4fg2rM?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" target="_blank" rel="noopener noreferrer nofollow">Watch the full video</a></p></span></div></div><p class="paragraph" style="text-align:left;">OpenAI says its unreleased next model has made ten mathematical advances.</p><p class="paragraph" style="text-align:left;">And each successful runs cost about $2,000.</p><p class="paragraph" style="text-align:left;">So: maths is basically solved right? We’re done here. Pack it up. </p><p class="paragraph" style="text-align:left;">Yes and no. </p><p class="paragraph" style="text-align:left;">Fair warning before we get into this: I am NOT a mathematician. If you see people like me talking about research-level maths, don&#39;t trust us on the maths! Go and listen to the mathematicians. They, funnily enough, know more about this! Terrance Tao (more on him later) is a good source here.</p><p class="paragraph" style="text-align:left;">What I <i>can do</i> is show you what OpenAI released, what the $2,000 includes and why the whole thing is getting serious attention.</p><p class="paragraph" style="text-align:left;">And why, most importantly, this isn’t just about maths. This is much wider. </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://youtu.be/G5FTv4fg2rM?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" target="_blank" rel="noopener noreferrer nofollow">I went through it all in Monday&#39;s video.</a></p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/G5FTv4fg2rM" width="100%"></iframe><h2 class="heading" style="text-align:left;" id="astra-is-still-locked-away">Astra is still locked away</h2><p class="paragraph" style="text-align:left;">First up, the model used to make these advancements is in the ChatGPT Astra family.</p><p class="paragraph" style="text-align:left;">Read: GPT-6. The next generation. </p><p class="paragraph" style="text-align:left;">I need to start here: Astra is <b>not out. </b>We cannot use it, there is no public model card and we cannot throw our own problems at it to see how often it falls over. So we have to take news about Astra with a pinch of salt as always with new models. </p><p class="paragraph" style="text-align:left;">Howver, <a class="link" href="https://openai.com/index/ten-advances-in-mathematics/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" target="_blank" rel="noopener noreferrer nofollow">OpenAI did publish</a> a 249-page collection of papers, 62 pages of the model&#39;s discovery notes and ten public Lean certificates.</p><div class="image"><img alt="What OpenAI actually released: 249 pages of proofs, 62 pages of discovery notes and ten Lean certificates, while Astra remains unreleased." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/27e5f0fa-a1e9-4d75-a3e9-59c740fa3e39/02-what-released.png?t=1785937196"/><div class="image__source"><span class="image__source_text"><p>OpenAI published the research artifacts. Astra itself remains unreleased.</p></span></div></div><p class="paragraph" style="text-align:left;">That is a hell of a lot more useful than the normal &quot;our model is amazing&quot; bar chart that we get. Mathematicians can actually get their hands on the work and start <i>checking the work</i>. </p><p class="paragraph" style="text-align:left;">Here’s the original tweet that kicked all this off: </p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/polynoamial/status/2083467194663571701?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">Note the 10 different areas. These are across a broad swathe of mathematics and computer science, perhaps to show the model’s versatility.</p><p class="paragraph" style="text-align:left;">Also not Noam uses the word “solved”. That’s the social media spin on it - a bit overblown! </p><p class="paragraph" style="text-align:left;">OpenAI&#39;s own article calls these ten <i>advances</i>. Four improve mathematical bounds. There is a new construction, two counterexample results and several theorems, including three numbered Erdős problems. One improves a sphere-packing bound for the first time since 1978. Advances is much more honest that “solved” so we’ll stick to that - but I do understand why Noam used “solved” otherwise no-one would have paid attention! </p><h2 class="heading" style="text-align:left;" id="2000-sort-of">$2,000? Sort of.</h2><p class="paragraph" style="text-align:left;">OpenAI priced the tokens used to find the successful solutions at roughly $2,000 using Sol API rates.</p><p class="paragraph" style="text-align:left;">Let’s be a little careful here. </p><p class="paragraph" style="text-align:left;">That figure does not include training Astra (hundreds of millions), paying the researchers, building the tools, choosing the problems, preparing the papers or checking the results.</p><div class="image"><a class="image__link" href="https://youtu.be/G5FTv4fg2rM?t=240&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" rel="noopener" target="_blank"><img alt="The $2,000 research lab claim, showing training, staff, failed runs, selection and verification below the quoted inference cost." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f6bb5fc8-8f59-47fc-b7ab-1e919b2dfebb/04-cost.png?t=1785937198"/></a><div class="image__source"><span class="image__source_text"><p>The quoted $2,000 is successful solution-finding inference, not the project cost.</p></span></div></div><p class="paragraph" style="text-align:left;">We also do not know the denominator.</p><p class="paragraph" style="text-align:left;">Did Astra have ten goes and nail all ten? Did OpenAI run 100 problems? 1,000? 10,000? How much did they spend chasing work that went nowhere before choosing these results?</p><p class="paragraph" style="text-align:left;">No idea. OpenAI has not told us. Probably because it undercuts the headline lets be honest. </p><p class="paragraph" style="text-align:left;">BUT…let’s not be too cynical here. Once the model and workflow exist, they can run another ten problems. Then another 100. Then another 1,000. The <i>extra marginal inference</i> of each new run is small.</p><p class="paragraph" style="text-align:left;">Especially compared to academic funding, grants and labs…. </p><p class="paragraph" style="text-align:left;">Mathematicians (and scientists) are in a weird position now. Suddenly problems they may have dedicated years or decades too are just being solved. </p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/typesfast/status/2083934589295501654?s=46&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">A really useful term here is agency rupture. You’ve probably felt this yourself - it’s when AI finally does a task you thought you were required for. Previously maybe you had to use AI. Now you actively hinder the AI with your input:</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/danshipper/status/2084038453831020916?s=46&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems"><p> Twitter tweet </p></a></blockquote><h2 class="heading" style="text-align:left;" id="genius-at-maths-shit-at-strawberry-">Genius at maths. Shit at strawberry. Still.</h2><p class="paragraph" style="text-align:left;">These models are ripping their way though problems at the edges of mathematics and science. </p><p class="paragraph" style="text-align:left;">So…why can’t it count the Rs in “strawberry”? </p><p class="paragraph" style="text-align:left;">This is a (I think fair, if overused!) criticism of large language models. </p><p class="paragraph" style="text-align:left;">It’s still an issue! </p><p class="paragraph" style="text-align:left;">There’s a single word answer for <i>why</i> this happens: <b>Representation.</b></p><p class="paragraph" style="text-align:left;">Language models normally receive token chunks rather than a tidy row of letters. Depending on the tokeniser, &quot;strawberry&quot; can arrive as something like <code>STR</code> + <code>AW</code> + <code>BERRY</code>. The model has to reconstruct the individual letters before it can count them.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f8553cf6-e024-45b8-a6ba-8a48b63ce11a/Screenshot_2026-08-06_at_11.21.12.png?t=1786011677"/><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://tiktokenizer.vercel.app/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" target="_blank" rel="noopener noreferrer nofollow">https://tiktokenizer.vercel.app/</a></p></span></div></div><p class="paragraph" style="text-align:left;">I’ve written a whole <a class="link" href="https://aiwithkyle.com/ai-news/193-ai-101-tokens?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" target="_blank" rel="noopener noreferrer nofollow">beginner’s guide on tokenization</a> here.</p><p class="paragraph" style="text-align:left;">This remains an issue in pure large language models. </p><p class="paragraph" style="text-align:left;">But we’ve moved beyond that. For the most part we don’t just use the model. We also (increasingly) don’t even use the basic Chat app version of models. Nope - for higher level tasks we use, funnily enough, higher level tools. Claude Codex, Codex etc.</p><p class="paragraph" style="text-align:left;">The <b>harness</b> is the thing. </p><p class="paragraph" style="text-align:left;">So when running maths problems we can give the model a much nicer setup: symbols, long reasoning time, lots of candidate approaches, tools and a precise verifier at the end. It’ll use code to work out “counting” problems like how many Rs in strawberry, not the large language model itself. </p><p class="paragraph" style="text-align:left;">This leads to AI tools being very good at some areas of maths (formulating new proofs) and very poor at others (counting!). Which to use humans <i>feels</i> weird. <i>We</i> find counting easy and discovering new proofs difficult. And so we ascribe the same difficulty scale to AI. That’s our anthropocentrism seeping though! </p><p class="paragraph" style="text-align:left;">This is also why some people believe that AI is still terrible. It <i>might</i> be for the type of task they use it for! This uneven boundary is the <a class="link" href="https://aiwithkyle.com/resources/co-intelligence-guide?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" target="_blank" rel="noopener noreferrer nofollow">jagged frontier</a>. Two people can use the same model and one says it is useless whilst the other says it has transformed their work. They may simply be giving it very different jobs.</p><h2 class="heading" style="text-align:left;" id="55000-lines-of-checking">55,000 lines of checking</h2><p class="paragraph" style="text-align:left;">The most important part of this story is that OpenAI released their working.</p><p class="paragraph" style="text-align:left;">It’s a bit like when we did maths exams at school and had to show our workings. You can’t just have a stab at an answer. You need to show how you got there! </p><p class="paragraph" style="text-align:left;">In this case they released Lean files. The public Lean files contain roughly 55,000 lines of formal proof. A mathematician can run them through Lean and check whether each logical step follows from the definitions and assumptions in the file.</p><div class="image"><img alt="Lean checks each formal proof step, while experts still assess the theorem, assumptions, novelty and importance." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9ac1e399-a160-4a58-80cf-6fd4bcedfb01/07-trust-proofs.png?t=1785937200"/><div class="image__source"><span class="image__source_text"><p>Lean checks the formal proof. Experts still have several other jobs.</p></span></div></div><p class="paragraph" style="text-align:left;">Importantly, Lean cannot tell us whether OpenAI encoded the right theorem, whether the assumptions are sensible, whether the result is new or whether anybody should care about it. All is tells us is that the chain of logic makes sense. </p><p class="paragraph" style="text-align:left;">People are already pushing back on one of the ten results. Good. OpenAI has put the work out, specialists can tear it apart and the result will either survive, change or die.</p><p class="paragraph" style="text-align:left;">Compared with a chatbot checking its own homework, Lean gives other people something they can actually rerun. <a class="link" href="https://aiwithkyle.com/ai-news/182-ai-hallucinations?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-astra-solved-10-maths-problems" target="_blank" rel="noopener noreferrer nofollow">AI hallucinations</a> do not disappear, but a polished wrong answer is much harder to sneak through 55,000+ lines of formal checks! </p><p class="paragraph" style="text-align:left;">And remember, OpenAI could have kept all of this behind a launch post. They published the papers and the checks instead. Good on them!</p><h2 class="heading" style="text-align:left;" id="the-human-queue-could-get-very-long">The human queue could get very long</h2><p class="paragraph" style="text-align:left;">Why does all this matter? It’s another case of humans being pushed further up the chain. </p><p class="paragraph" style="text-align:left;">If OpenAI can cheaply run another 1,000 problems, somebody still has to choose useful questions, inspect the proofs and explain the results in a form that other mathematicians can use.</p><div class="image"><img alt="Candidate answers become cheaper while good questions, verification, explanation and theory become the research bottlenecks." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2dc77357-3fef-48b7-b687-71c7e823270b/10-bottleneck.png?t=1785937202"/><div class="image__source"><span class="image__source_text"><p>Answers get cheaper but judgement gets scarcer. </p></span></div></div><p class="paragraph" style="text-align:left;">And candidate papers could arrive much faster than journals and researchers can review them. </p><p class="paragraph" style="text-align:left;">This is what’s happening across the board with AI. </p><p class="paragraph" style="text-align:left;">AI increases output but we still need to manage it, delegate, check the work. Our judgement and discernment is valuable. </p><p class="paragraph" style="text-align:left;">In fact it is ALL that is valuable. </p><p class="paragraph" style="text-align:left;">The grunt work. the labour. The “work” that we have been doing for centuries. That stuff is going away. It’s going to be devalued. </p><p class="paragraph" style="text-align:left;">The value shifts upstream to the people who can <i>manage</i> the process.</p><p class="paragraph" style="text-align:left;"><b>That is you. </b></p><p class="paragraph" style="text-align:left;">To the task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=c5d2a6bf-1f09-4121-b404-fd5a930373b6&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>ChatGPT Just Got Skills</title>
  <description>AI Skills 101</description>
  <link>https://newsletter.aiwithkyle.com/p/chatgpt-just-got-skills</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/chatgpt-just-got-skills</guid>
  <pubDate>Mon, 03 Aug 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-08-03T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><a class="image__link" href="https://www.youtube.com/watch?v=w49OfWDTTDo&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-just-got-skills" rel="noopener" target="_blank"><img alt="Kyle explaining how Skills moved from Claude into ChatGPT" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/11fe1dc9-cf7c-4b1e-ac12-597f4879fb28/ai-skills-101-top.gif?t=1785516576"/></a><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://www.youtube.com/watch?v=w49OfWDTTDo&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-just-got-skills" target="_blank" rel="noopener noreferrer nofollow">Watch the full video</a></p></span></div></div><p class="paragraph" style="text-align:left;">If you have explained the same job to AI more than twice…stop.</p><p class="paragraph" style="text-align:left;">You do <i>not</i> need a bigger prompt. You need to teach the thing how you work. Once. And then hands off. </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://help.openai.com/en/articles/20001066?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-just-got-skills" target="_blank" rel="noopener noreferrer nofollow">OpenAI has just added Skills to ChatGPT</a>. I shot a Youtube <a class="link" href="https://www.youtube.com/watch?v=w49OfWDTTDo&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-just-got-skills" target="_blank" rel="noopener noreferrer nofollow">guide on AI skills here</a>.</p><p class="paragraph" style="text-align:left;">Skills have been around for a while but they have broken confinement. They are coming to more and more surfaces - including now vanilla ChatGPT. </p><p class="paragraph" style="text-align:left;">Anthropic invented them, Claude has had them for a while and they have already spread into coding tools (including Codex). Now they are inside the app used by about a billion normal humans.</p><p class="paragraph" style="text-align:left;">That makes this a VERY good time to understand what they actually are.</p><p class="paragraph" style="text-align:left;">And a great time to catch up if you haven’t been using.</p><p class="paragraph" style="text-align:left;">Because yes, &quot;Skills&quot; sounds like yet another stupid AI noun to add to Projects, GPTs, agents, Plugins, apps, MCPs and whatever arrives next Tuesday.</p><p class="paragraph" style="text-align:left;">But this one is useful. Promise! 😄 </p><p class="paragraph" style="text-align:left;">Here&#39;s the full video:</p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/w49OfWDTTDo" width="100%"></iframe><h2 class="heading" style="text-align:left;" id="never-prompt-twice">Never prompt twice</h2><p class="paragraph" style="text-align:left;">A prompt asks the AI to do one job.</p><p class="paragraph" style="text-align:left;">And when you want to do that task again you need to … prompt again. Eww.</p><p class="paragraph" style="text-align:left;">A Skill teaches it the <i>method</i> so the AI can do that job again.</p><p class="paragraph" style="text-align:left;">Say you turn a livestream into a newsletter every week. Like I do!</p><p class="paragraph" style="text-align:left;">You can keep pasting in the voice rules, structure, links, image requirements and all the annoying little things I complain about when the draft gets them wrong.</p><p class="paragraph" style="text-align:left;"><b><i>Or </i></b>you can package that process <span style="text-decoration:underline;">once</span> and let the AI pull it in whenever the job appears.</p><p class="paragraph" style="text-align:left;">Much better. Far more efficient. And actually timesaving (which is what AI is for right??).</p><p class="paragraph" style="text-align:left;">Models are pretty smart already. BUT they don’t know your process. Your workflow. How you put things together. That is the stuff Skills capture.</p><p class="paragraph" style="text-align:left;">FYI: This follows the same logic (and complements) building <a class="link" href="https://aiwithkyle.com/ai-news/build-one-shared-ai-vault?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-just-got-skills" target="_blank" rel="noopener noreferrer nofollow">one shared AI vault</a>. Your useful context should not be trapped in one random chat. It should live somewhere the AI can reach again. Skills are another way to do this. </p><div class="image"><img alt="Timeline showing Agent Skills moving from Claude to an open standard and into ChatGPT" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f6bd8c85-63f7-422b-8ba4-9b085f0abbff/02-skills-escaped-claude.jpg?t=1785516578"/><div class="image__source"><span class="image__source_text"><p>Skills have moved from Claude into an open standard and ChatGPT.</p></span></div></div><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.anthropic.com/engineering/equipping-agents-for-the-real-world-with-agent-skills?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-just-got-skills" target="_blank" rel="noopener noreferrer nofollow">Anthropic launched Agent Skills</a> in October 2025. Then it opened the format up. OpenAI has now adopted the same basic standard for ChatGPT and Codex.</p><p class="paragraph" style="text-align:left;">Thank god. We don’t need even more standards.</p><h2 class="heading" style="text-align:left;" id="it-is-basically-a-folder">It Is Basically A Folder</h2><p class="paragraph" style="text-align:left;">What actually is a Skill though? </p><p class="paragraph" style="text-align:left;">Well…it’s a little underwhelming.</p><p class="paragraph" style="text-align:left;">It’s a folder with a bunch of text files in it. </p><p class="paragraph" style="text-align:left;">That’s….sorta it?</p><div class="image"><img alt="The anatomy of an AI Skill folder" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3c3f2cd3-d2b1-4aa1-836a-953da7a0f77d/03-anatomy-of-a-skill.jpg?t=1785516575"/><div class="image__source"><span class="image__source_text"><p>At minimum, a Skill is a folder with a SKILL.md file.</p></span></div></div><p class="paragraph" style="text-align:left;">At minimum, a Skill is a folder with one file inside called <code>SKILL.md</code>.</p><p class="paragraph" style="text-align:left;">That .md bit may be unfamiliar. It stands for Markdown and is basically a way to format text. That’s all. Here’s what markdown looks like (on the left) and how it displays (on the right). Don’t worry you don’t need to know this - just want to explain the .md part.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/74c0c78c-910d-4753-9892-02999f195f39/markdown-1-markup.png?t=1785689508"/><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://rmarkdown.rstudio.com/lesson-8.html?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-just-got-skills" target="_blank" rel="noopener noreferrer nofollow">https://rmarkdown.rstudio.com/lesson-8.html</a></p></span></div></div><p class="paragraph" style="text-align:left;">That text file SKILL.md tells the AI what the Skill does, when it should use it and how to carry out the work. You can then add reference documents, templates, examples, images or scripts if the job needs them.</p><p class="paragraph" style="text-align:left;">Most of the time it’s just a written set of instructions. For most business jobs, a clear set of instructions and a couple of good examples will get you a long way. Add technical bits (like code) only when something needs to run in a precise, repeatable way.</p><p class="paragraph" style="text-align:left;">This is where people think Skills are too complicated. They see folders and markdown and assume it is developer stuff.</p><p class="paragraph" style="text-align:left;">Nah. It’s a folder with some text files. </p><p class="paragraph" style="text-align:left;">Oh and ChatGPT or Claude can write the file for you anyway.</p><h2 class="heading" style="text-align:left;" id="build-it-by-doing-the-job">Build It By Doing The Job</h2><p class="paragraph" style="text-align:left;">OK how do we make a Skill? Well…get the AI to make it for you…</p><div class="image"><img alt="Build a first Skill by doing the task, reviewing the result and extracting the working process" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/bb6a8484-d96a-4717-8f83-03e7c5aed925/07-build-one-by-talking.jpg?t=1785516579"/><div class="image__source"><span class="image__source_text"><p>Build the process first. Package it second.</p></span></div></div><p class="paragraph" style="text-align:left;">You <b>do not </b>need to sit down and write an operating manual from a blank page.</p><p class="paragraph" style="text-align:left;">Do the job first with the AI. Correct it. Explain why something is wrong. Give it examples. Swear at it a bit (optional but traditional). Keep going until the output is actually good.</p><p class="paragraph" style="text-align:left;"><i>Then </i>ask it to review the whole conversation and extract the process you just discovered into one clean Skill.</p><h2 class="heading" style="text-align:left;" id="do-not-build-a-skill-to-run-your-en">Do Not Build A Skill To Run Your Entire Business</h2><p class="paragraph" style="text-align:left;">When people realise how powerful Skills are there is a very natural tendency to want to get ALL their work into Skills. </p><p class="paragraph" style="text-align:left;">I get it. They are great. </p><p class="paragraph" style="text-align:left;">One important caveat early days.</p><p class="paragraph" style="text-align:left;">The first instinct is to chuck your entire business into one giant Skill.</p><p class="paragraph" style="text-align:left;">Don&#39;t.</p><p class="paragraph" style="text-align:left;">It will be too broad, too heavy and very difficult to test. You will have no idea which instruction caused the result to go a bit shit.</p><p class="paragraph" style="text-align:left;">Your first Skill should be boring, repeated and bounded….</p><p class="paragraph" style="text-align:left;">Turn call notes into a client update. Check a proposal against the same risk list. Research five competitors using the same evidence rules. Turn a transcript into a newsletter with the same structure. Yadda yadda.</p><p class="paragraph" style="text-align:left;">One job. Clear inputs. Clear output. A checklist at the end. Kinda dull stuff but important! </p><p class="paragraph" style="text-align:left;">Then test it with three requests: one that <i>should</i> trigger it, one that <i>should not</i>, and one horrible messy real-world example.</p><p class="paragraph" style="text-align:left;">If it survives all three…now we&#39;re talking. </p><p class="paragraph" style="text-align:left;">If not: tweak. Again, do so <i>with</i> the AI. Run some more tests and tweak again. Until happy. </p><p class="paragraph" style="text-align:left;">Then, when you are happy with the Skill keep it saved somewhere. Mine are in Github personally but you can have them locally on your computer no problem. Ask ChatGPT and it’ll save them somewhere smart. </p><p class="paragraph" style="text-align:left;">You can then <i>use</i> those Skills anywhere - in Chat, in ChatGPT Work, in Codex. <i>Or</i> indeed in other tools like Claude. I highly recommend adding them to your AI brain if you have one so they can be accessed everywhere - but that’s a little more advanced.</p><p class="paragraph" style="text-align:left;">For now just wrap one process up nicely and use it over the rest of the week to get the feel for how we build repeatable Skills.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=84bfdf1f-ffa2-4194-8bd0-735201f5d9b5&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Anthropic Destroyed Millions of Books. Legally.</title>
  <description>Fahrenheit 451 (With lawyers)</description>
  <link>https://newsletter.aiwithkyle.com/p/anthropic-destroyed-millions-books-legally</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/anthropic-destroyed-millions-books-legally</guid>
  <pubDate>Fri, 31 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-31T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="Anthropic destroyed millions of books legally, with an industrial scanner and private digital library" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/8ec4ca19-f6cb-401f-bd87-bcb3770ebbe4/iprv50.png?t=1785414820"/><div class="image__source"><span class="image__source_text"><p>Fahrenheit 451. With lawyers.</p></span></div></div><p class="paragraph" style="text-align:left;">Anthropic bought millions of used books.</p><p class="paragraph" style="text-align:left;">Then it cut the bindings off, scanned every page and threw the paper away. Mulched it.</p><p class="paragraph" style="text-align:left;">Not metaphorically. Not &quot;ingested&quot; in some vague AI press-release sense. Industrial cutters. High-speed scanners. Pallets of books.</p><p class="paragraph" style="text-align:left;">The stinger here?</p><p class="paragraph" style="text-align:left;">This destruction helped make the process <i>legal</i>.</p><p class="paragraph" style="text-align:left;">Fahrenheit 451. With lawyers.</p><p class="paragraph" style="text-align:left;">Now … this is a tricky one. I LOVE books. And find it appalling that books were destroyed. </p><p class="paragraph" style="text-align:left;">However (and lord it pains me to say this): there were reasons for this. And it’s not <i>quite</i> as bad as the viral posts are making out. </p><p class="paragraph" style="text-align:left;">Still…bad. </p><p class="paragraph" style="text-align:left;">Here’s my Youtube video on the story:</p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/e5i3E7-R4P8" width="100%"></iframe><h2 class="heading" style="text-align:left;" id="project-panama">Project Panama</h2><p class="paragraph" style="text-align:left;">Anthropic called the operation <a class="link" href="https://www.washingtonpost.com/technology/2026/01/27/anthropic-ai-scan-destroy-books/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=anthropic-destroyed-millions-of-books-legally" target="_blank" rel="noopener noreferrer nofollow">Project Panama</a>.</p><p class="paragraph" style="text-align:left;">It wanted &quot;all the books in the world&quot; for a central research library. So it bought millions of physical books, hired outside companies to remove the bindings, scanned the pages and kept the digital copies.</p><p class="paragraph" style="text-align:left;">The mangled originals were discarded.</p><div class="image"><img alt="Project Panama pipeline: buy used books, cut off bindings, scan every page, keep the private digital copy" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f1a7c035-18e4-4b7c-920a-efa243c7f4af/637n7a.png?t=1785414828"/><div class="image__source"><span class="image__source_text"><p>Project Panama: buy, cut, scan, discard.</p></span></div></div><p class="paragraph" style="text-align:left;">The scans went into a private library. Anthropic then selected material from that library for training Claude models.</p><p class="paragraph" style="text-align:left;">That last distinction is important btw. &quot;Anthropic scanned a book&quot; does not automatically mean every Claude model was trained on that specific book. </p><p class="paragraph" style="text-align:left;">Someone asked this on Youtube. And it’s a good question. They asked if they now had access to all these books inside Claude:</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a8c27d2b-c331-42cc-8593-a676a5abfee7/Screenshot_2026-07-30_at_6.04.28_pm.png?t=1785423882"/></div><p class="paragraph" style="text-align:left;">The answer is nope. The library and the model-training sets were separate things. You can’t just “call up” any book from a vast library that Anthropic now owns. Instead the models are <i>trained</i> on parts of the library. Very different. </p><h2 class="heading" style="text-align:left;" id="why-destroy-the-originals">Why destroy the originals?</h2><p class="paragraph" style="text-align:left;">OK so even if we are ok with training on books (which is a whole other complicated story…) <i>why</i> destroy them?</p><p class="paragraph" style="text-align:left;">Is it just because it’s cheaper? Well yes. It is cheaper to use destructive scanning because it’s much faster. But that’s not the reason…</p><p class="paragraph" style="text-align:left;">The reason is legal. </p><p class="paragraph" style="text-align:left;">If Anthropic bought one physical copy and created one digital copy <i>while keeping both</i>, it had made a <b>surplus copy.</b></p><p class="paragraph" style="text-align:left;">That’s a problem. They’d fall foul of fair use laws and get sued into oblivion. </p><p class="paragraph" style="text-align:left;">Instead it used what the court called a one-in, one-out format change:</p><p class="paragraph" style="text-align:left;"><b>One purchased paper copy became one private digital copy. The paper version was destroyed. The scan was not redistributed.</b></p><p class="paragraph" style="text-align:left;">In the <a class="link" href="https://storage.courtlistener.com/recap/gov.uscourts.cand.434709/gov.uscourts.cand.434709.231.0_4.pdf?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=anthropic-destroyed-millions-of-books-legally" target="_blank" rel="noopener noreferrer nofollow">June 2025 fair-use ruling</a>, Judge William Alsup compared this to replacing purchased print copies with more convenient digital ones. His wonderfully blunt summary was: &quot;There was no surplus copying.&quot;</p><p class="paragraph" style="text-align:left;">That is what made this a better path for Anthropic. </p><p class="paragraph" style="text-align:left;">So yes…</p><p class="paragraph" style="text-align:left;">The destruction was not some accidental bit of waste. It was part of the legal logic. A legal loophole. </p><div class="image"><img alt="The legal split between bought and destroyed books, model training, and Anthropic&#39;s pirated book library" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/782b6cd8-a92a-4d6e-a5aa-55753c29ba70/ux581m.png?t=1785414836"/></div><p class="paragraph" style="text-align:left;">The same ruling also found that using books to train specific language models was &quot;quintessentially transformative&quot;. Anthropic was not selling replacement copies of the books. It was using the text to build a model capable of generating new material.</p><p class="paragraph" style="text-align:left;">Whether you <i>agree</i> with that or not is a different question. But that is what the ruling says. </p><h2 class="heading" style="text-align:left;" id="the-15-billion-was-for-something-el">The $1.5 billion was for something else</h2><p class="paragraph" style="text-align:left;">This is where most posts about the story collapse two separate things into one. I’ve seen a lot of confusion about this online. </p><p class="paragraph" style="text-align:left;">Anthropic have <i>just </i>settled a $1.5b lawsuit over using books in training. They are paying out to publishers and authors who are aggrieved that their work was used. </p><p class="paragraph" style="text-align:left;"><i>That’s cheap for Anthropic btw…they’ve done very well out of the settlement. A story for another time</i>! </p><p class="paragraph" style="text-align:left;">This settlement is a different matter entirely. Anthropic downloaded millions of books from pirate libraries including LibGen, PiLiMi (a mirror of Anna’s Archive). It kept those pirated files in its central library.</p><p class="paragraph" style="text-align:left;">That was <b>NOT</b> excused as fair use. </p><p class="paragraph" style="text-align:left;">And that is what the <a class="link" href="https://law.justia.com/cases/federal/district-courts/california/candce/4%3A2024cv05417/434709/680/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=anthropic-destroyed-millions-of-books-legally" target="_blank" rel="noopener noreferrer nofollow">$1.5 billion settlement</a> was about. Not the paper cutter stuff but good ol’ fashioned Piracy.</p><p class="paragraph" style="text-align:left;">The court gave final approval on 20 July. More than 482,000 works are covered, with an estimated payment of roughly $3,000 per work.</p><p class="paragraph" style="text-align:left;">The distinction is bizarre but super important:</p><p class="paragraph" style="text-align:left;"><b>Buy a used book, destroy it and keep one private scan? Fair use.</b></p><p class="paragraph" style="text-align:left;"><b>Download the same book from a pirate site and keep it? Copyright infringement claim.</b></p><p class="paragraph" style="text-align:left;">You can see why people are confused.</p><h2 class="heading" style="text-align:left;" id="ai-ruined-the-internet">AI ruined the internet.</h2><p class="paragraph" style="text-align:left;">Why go to all this trouble in the first place?</p><p class="paragraph" style="text-align:left;">Because books are <i>good</i> training data.</p><p class="paragraph" style="text-align:left;">They are long. Edited. Dense. Specialist. Structured. </p><p class="paragraph" style="text-align:left;">But most importantly anything written before 2022 is golddust now. </p><p class="paragraph" style="text-align:left;">A lot of them were written before the internet filled up with AI-generated slop. If you train off AI work you poison your model. </p><p class="paragraph" style="text-align:left;">We have talked before about <a class="link" href="https://aiwithkyle.com/ai-news/182-ai-hallucinations?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=anthropic-destroyed-millions-of-books-legally" target="_blank" rel="noopener noreferrer nofollow">why models hallucinate when the underlying signal is thin</a>. Better source material does not magically make a model truthful. But a library of carefully edited human writing is vastly more useful than another billion scraped AI generated listicles repeating each other.</p><p class="paragraph" style="text-align:left;">Or, as one bookseller told <a class="link" href="https://www.404media.co/ai-companies-are-buying-tons-of-old-books-because-theyre-free-of-ai-slop/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=anthropic-destroyed-millions-of-books-legally" target="_blank" rel="noopener noreferrer nofollow">404 Media</a>: the world&#39;s best AI training data is sitting on a shelf.</p><p class="paragraph" style="text-align:left;">And now AI companies appear to be buying the shelves…</p><p class="paragraph" style="text-align:left;">Book dealers in Europe are reporting strange bulk orders for thousands of unrelated titles. One request apparently asked for 3,000 books with little concern for subject or condition. <a class="link" href="https://nltimes.nl/2026/06/25/rare-book-dealers-fear-tech-firms-destroying-obscure-editions-train-ai-models?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=anthropic-destroyed-millions-of-books-legally" target="_blank" rel="noopener noreferrer nofollow">Rare-book dealers are worried</a> that obscure editions are disappearing into scanners.</p><p class="paragraph" style="text-align:left;"><b>BUT…</b></p><p class="paragraph" style="text-align:left;">We do not know who every buyer is. We do not know what every book is being used for. And there is no evidence Anthropic ran around hunting down the last surviving copy of priceless first editions for a giant book bonfire.</p><p class="paragraph" style="text-align:left;">Most of the books going in are books we would otherwise think of as junk. </p><p class="paragraph" style="text-align:left;">The books that have been sitting high up on a shelf in your local second hand bookshop for a decade. </p><p class="paragraph" style="text-align:left;">Dummies Guide to Microsoft Access ‘95, Lonely Planet Czechoslavakia, Boise Idaho Regional Almanac 1987.</p><p class="paragraph" style="text-align:left;">Books that literally sell by the pound in wholesale markets. </p><p class="paragraph" style="text-align:left;">There is a LOT of junk out there that would have (let’s be honest here) gone to the pulper eventually. In a weird way Anthropic (and other AI labs) have saved some books that were essentially valueless. These items are still useful for AI models, even if they haven’t been valued by us humans for a few decades. </p><p class="paragraph" style="text-align:left;">A lot of the viral posts show images of antiquarian books being sliced and diced. The main video you’ll see (of the covers been cut of hardbacks) is actually from a book restoration art project… quite the opposite of what is being pushed. </p><p class="paragraph" style="text-align:left;">This is why we have to be a bit careful here. </p><p class="paragraph" style="text-align:left;">Destroying books is gross. It’s icky. I get that. </p><p class="paragraph" style="text-align:left;">But we as humans recycle a LOT of books every year anyway. This does not excuse Anthropic but as always it’s important to put things in perspective, even if it’s uncomfortable. </p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/12fc1ac1-6958-4ec8-9451-f400fa508410/Screenshot_2026-07-30_at_6.19.38_pm.png?t=1785424806"/><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://chireviewofbooks.com/2023/12/07/book-waste-the-dangers-of-publishing-and-the-ethical-consumption-of-books/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=anthropic-destroyed-millions-of-books-legally" target="_blank" rel="noopener noreferrer nofollow">https://chireviewofbooks.com/2023/12/07/book-waste-the-dangers-of-publishing-and-the-ethical-consumption-of-books/</a></p></span></div></div><h2 class="heading" style="text-align:left;" id="the-private-library">The private library</h2><p class="paragraph" style="text-align:left;">For me personally the shitty part is the private library. </p><p class="paragraph" style="text-align:left;">Anthropic (and I imagine other labs) are building these private labs. And removing physical books from circulation. The digital replacement stays inside a private company. The knowledge improves a paid product, but the scan is not returned to a public library or archive.</p><div class="image"><img alt="A physical used book leaves circulation while its scan remains locked inside a private AI library" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/dc721743-5888-4ada-9a04-692dda6fe935/0hvqjy.png?t=1785414844"/><div class="image__source"><span class="image__source_text"><p>The physical copy leaves circulation. The digital copy stays private.</p></span></div></div><p class="paragraph" style="text-align:left;">I’ve seen some people argue that Anthropic should make their scans available. </p><p class="paragraph" style="text-align:left;">Come on. Think! </p><p class="paragraph" style="text-align:left;">To be clear: Anthropic <b>cannot</b> simply publish copyrighted scans for everyone. </p><p class="paragraph" style="text-align:left;">They would get sued into oblivion for piracy. They legally <i>cannot</i> share without yet again falling prey to fair use laws.</p><p class="paragraph" style="text-align:left;">Rights still exist. Authors and publishers still need paying.</p><p class="paragraph" style="text-align:left;">But &quot;we bought it, chopped it up and now the only useful copy <i>we made</i> lives behind our lock&quot; is not exactly public preservation. This also makes this week’s fight over <a class="link" href="https://aiwithkyle.com/ai-news/even-denny-s?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=anthropic-destroyed-millions-of-books-legally" target="_blank" rel="noopener noreferrer nofollow">open weights and Anthropic&#39;s closed-model stance</a> feel a bit more pointed. The valuable inputs are gathered into a private library. The resulting model stays private. The public gets access on Anthropic&#39;s terms… </p><p class="paragraph" style="text-align:left;">Good business. Icky precedent.</p><p class="paragraph" style="text-align:left;">At minimum there should be proper provenance records, protection for genuinely scarce copies, compensation for rights-holders and some route for public-domain scans to end up in public archives rather than disappearing into corporate training pipelines.</p><p class="paragraph" style="text-align:left;">Remember though this is not Farenheit 451. Anyone saying so hasn’t read the book. </p><p class="paragraph" style="text-align:left;">Anthropic did not destroy the books because the knowledge was dangerous.</p><p class="paragraph" style="text-align:left;">It destroyed them because the knowledge was valuable.</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=6cfc3316-81d4-4ab1-a99d-37965f9aef84&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Everyone Is Turning on Anthropic</title>
  <description>Even Denny&#39;s</description>
  <link>https://newsletter.aiwithkyle.com/p/anthropic-vs-open-weights</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/anthropic-vs-open-weights</guid>
  <pubDate>Wed, 29 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-29T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Local Ai]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><a class="image__link" href="https://www.youtube.com/watch?v=rv3ST4tl4Uw&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" rel="noopener" target="_blank"><img alt="Kyle explains why everyone is piling on Anthropic" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3d47a963-8ccb-4115-9ee5-cc3b7d48c7ff/mgyz3c.gif?t=1785233323"/></a><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://www.youtube.com/watch?v=rv3ST4tl4Uw&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">Watch the full breakdown on YouTube</a></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.youtube.com/watch?v=rv3ST4tl4Uw&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic"><span class="button__text" style=""> Watch now </span></a></div><p class="paragraph" style="text-align:left;">Denny&#39;s has entered the AI safety debate.</p><p class="paragraph" style="text-align:left;">Yes. The American diner chain.</p><p class="paragraph" style="text-align:left;">It posted a Venn diagram showing that Denny&#39;s and NVIDIA both understand &quot;the importance of staying open&quot;.</p><p class="paragraph" style="text-align:left;">Very good. No notes.</p><p class="paragraph" style="text-align:left;">It is also the funniest bit of a much bigger fight. <i>Jensen Huang joined X</i> and used his <i>first post</i> to share NVIDIA&#39;s letter backing open-weight AI. Google, Meta, Microsoft, OpenAI, Hugging Face, GitHub and a load of other companies signed it.</p><p class="paragraph" style="text-align:left;">Anthropic did not.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.youtube.com/watch?v=rv3ST4tl4Uw&t=57s&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">I go through the letter here</a>.</p><h2 class="heading" style="text-align:left;" id="what-actually-happened">What actually happened</h2><p class="paragraph" style="text-align:left;">On 24 July Jensen Huang shared a four-page letter called <a class="link" href="https://images.nvidia.com/pdf/Open-Weights-and-American-AI-Leadership.pdf?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">Open Weights and American AI Leadership</a>.</p><p class="paragraph" style="text-align:left;">Have a read. </p><p class="paragraph" style="text-align:left;">The letter argues that America&#39;s AI lead cannot sit inside a few closed models. It says businesses, universities and governments need models they can download, adapt and run on their own infrastructure. That means more competition, lower costs and less dependence on one AI provider.</p><p class="paragraph" style="text-align:left;">It also takes aim at attempts to restrict open models. The signatories want policymakers to separate illegal extraction from normal distillation and avoid broad rules that could kneecap the whole open-weight market. IMHO: fair. </p><p class="paragraph" style="text-align:left;">And then came the list of names. Here it is as of 28th July:</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a5ce6f54-153a-4894-a11f-6a06f13c3805/Screenshot_2026-07-28_at_14.18.41.png?t=1785244726"/></div><p class="paragraph" style="text-align:left;">NVIDIA. Google. Meta. Microsoft. OpenAI. AMD. Hugging Face. GitHub. Mistral. Cloudflare. Replit. Basically the entire AI infrastructure world. </p><p class="paragraph" style="text-align:left;">Except Anthropic.</p><p class="paragraph" style="text-align:left;">That was not a random omission. Anthropic has been making the opposite argument. In February it said DeepSeek, Moonshot and MiniMax used roughly 24,000 fraudulent accounts to generate <a class="link" href="https://www.anthropic.com/news/detecting-and-preventing-distillation-attacks?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">more than 16 million exchanges with Claude</a>. It says those labs were extracting Claude&#39;s best capabilities and that the danger multiplies when the resulting models are released openly.</p><p class="paragraph" style="text-align:left;">In their defence - they are probably correct.</p><p class="paragraph" style="text-align:left;">So this is the actual disagreement:</p><p class="paragraph" style="text-align:left;"><b>NVIDIA&#39;s camp says open weights spread access, competition and control. Anthropic says powerful weights can spread stolen or dangerous capabilities beyond anyone&#39;s control.</b></p><p class="paragraph" style="text-align:left;">Then the internet got involved. And piled on Anthropic. As we are wont to do! </p><p class="paragraph" style="text-align:left;">Anthropic researcher Julian Schrittwieser compared the demand for open weights with asking NVIDIA to open-source CUDA and Microsoft to open-source Windows.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/AndrewYNg/status/2081103828859117908?s=20&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/andrewyng/status/2081103828859117908?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">Andrew Ng called that a false equivalence</a>. Correctly, I reckon. Imagine pissing off Andrew Ng. He’s such a sweetie but even he came out to call out the BS.</p><p class="paragraph" style="text-align:left;">To be VERY clear: nobody is demanding that Anthropic release Claude&#39;s weights. It is perfectly entitled to keep Claude closed. The complaint is that Anthropic appears to want governments to make it harder for <i>other</i> companies to release open models.</p><h2 class="heading" style="text-align:left;" id="open-weights-is-not-open-source">Open weights is not open source</h2><div class="image"><img alt="Open weights and open source are different levels of model access" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1d63743f-8be5-442f-9209-a9bcce9e7eb8/id7pv3.png?t=1785233325"/><div class="image__source"><span class="image__source_text"><p>Open weights and open source are not the same thing</p></span></div></div><p class="paragraph" style="text-align:left;">Quick definition because this gets mangled constantly.</p><p class="paragraph" style="text-align:left;">An AI model is basically a giant set of learned numbers called weights. If a company releases those weights, you can usually download the model, run it on infrastructure you control and adapt it under the licence.</p><p class="paragraph" style="text-align:left;">That does <b>NOT</b> necessarily give you the training data, training code or the full recipe used to build it. So &quot;open weights&quot; and &quot;open source&quot; are not the same thing. Annoyingly even the Anthropic researcher Julian above got it wrong - Jensen was calling for open-weights. Not open source. Shit, the open letter is titled <i><b>Open Weights</b></i> and American AI Leadership. </p><p class="paragraph" style="text-align:left;">We’re talking open-weights <i>not</i> open-source.</p><p class="paragraph" style="text-align:left;">And also open-weights also does not mean “free”, small or laptop-friendly. The giant Chinese models may need a server rack. Smaller Qwen, Gemma and Mistral models can run locally on normal hardware. </p><p class="paragraph" style="text-align:left;">Kimi K3 weights were released yesterday. I did a back of the envelope calculation and you’ll need around $650,000-750,000 worth of computer to run it. That is not “free”.</p><p class="paragraph" style="text-align:left;">I’ve talked about this at length here: <a class="link" href="https://www.youtube.com/watch?v=rUKDzvlHvFI&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">Local Models 101</a>. </p><p class="paragraph" style="text-align:left;">Why should you care?</p><p class="paragraph" style="text-align:left;">With a closed model, the provider can change the price, limits, behaviour or access. We have watched this happen repeatedly with Claude Fable over the last few weeks. It’s exhausting.</p><p class="paragraph" style="text-align:left;">Hell, Sam Altman is <i>this moment</i> in Washington showing the next set of models (ie. GPT-6) to the White House and Trump administration for a go ahead. We’re at the whims of many forces. </p><p class="paragraph" style="text-align:left;">With open weights, a business can keep the model on its own infrastructure, fine-tune it for a specific job, protect sensitive data and move without rebuilding everything around another company&#39;s API.</p><p class="paragraph" style="text-align:left;">I went through the practical differences in my <a class="link" href="https://aiwithkyle.com/ai-news/open-source-ai-price-war?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">open-source AI price war guide</a>. The short version is that open models put a ceiling on price and give customers an exit.</p><h2 class="heading" style="text-align:left;" id="anthropic-does-have-a-point">Anthropic does have a point</h2><div class="image"><a class="image__link" href="https://www.anthropic.com/news/detecting-and-preventing-distillation-attacks?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" rel="noopener" target="_blank"><img alt="The legitimate safety case against releasing powerful model weights" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/7448b9cc-44a0-4e72-b124-57d774781956/1z6r21.png?t=1785233327"/></a><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://www.anthropic.com/news/detecting-and-preventing-distillation-attacks?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">Released model weights cannot be recalled</a></p></span></div></div><p class="paragraph" style="text-align:left;">Now whilst Anthropic are pissing people off right now…they still have a point. And they are standing up for their (oft-repeated) principals. They can’t be faulted there.</p><p class="paragraph" style="text-align:left;">Once model weights are released, they cannot be recalled.</p><p class="paragraph" style="text-align:left;">You cannot patch every copy, turn off a dangerous capability or know who modified it. If a future model becomes genuinely useful for cyberattacks or biological weapons, chucking the weights online would be a fairly permanent decision.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.anthropic.com/news/detecting-and-preventing-distillation-attacks?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">Anthropic also says</a> DeepSeek, Moonshot and MiniMax used around 24,000 fraudulent accounts to generate more than 16 million Claude exchanges for distillation. In plain English: use Claude&#39;s answers to help train competing models.</p><p class="paragraph" style="text-align:left;">Anthropic&#39;s fear is that stolen capabilities then get released as open weights and spread beyond anyone&#39;s control. Which…is precisely what is happening. That is not imaginary - that’s legit. It is a proper safety and national-security argument.</p><p class="paragraph" style="text-align:left;"><b>BUT…</b></p><p class="paragraph" style="text-align:left;">It still does not follow that open weights<i> in general</i> should be squashed. A small model running private document search on your laptop is not the same risk as tomorrow&#39;s frontier model with dangerous capabilities.</p><p class="paragraph" style="text-align:left;">Treating them as the same thing is a massive overreaction. There are (as in all things) degrees. </p><h2 class="heading" style="text-align:left;" id="follow-the-money">Follow the money</h2><p class="paragraph" style="text-align:left;">Equally we need to look at the signatories here. They have vested interests in keeping open-weights in play. </p><p class="paragraph" style="text-align:left;">Yes yes I’m sure this <i>also</i> aligns with their ethics and philosophy. But those stances are a lot easier to hold when you are also making billions of dollars.</p><p class="paragraph" style="text-align:left;">NVIDIA benefits when open models exist because somebody needs to buy the chips that run them. Cloud companies benefit because most businesses do not fancy building a data centre in the garage. </p><p class="paragraph" style="text-align:left;">Open-weight models mean more models being deployed. More sales. More compute. More volume. All great for the signatories. </p><p class="paragraph" style="text-align:left;">Funny that.</p><p class="paragraph" style="text-align:left;">This does not make NVIDIA&#39;s case wrong or Anthropic&#39;s safety case fake. Both sides have real beliefs AND very convenient business incentives. Calling either side disingenuous or spinning complex conspiracies sounds smart but often isn’t. </p><p class="paragraph" style="text-align:left;">So. What do you DO? What should you use? The useful question is what gives <i>you</i> options.</p><p class="paragraph" style="text-align:left;">Use Claude or ChatGPT for the difficult work where they are plainly better. That’s not going to change anytime soon. Open-weight (mainly Chinese) models <i>are</i> still behind - no matter what some influencer is screaming at you. RIP ChatGPT.</p><p class="paragraph" style="text-align:left;"> Test smaller open models for private, repetitive and high-volume jobs. Keep your prompts, source files and business context in <a class="link" href="https://aiwithkyle.com/ai-news/build-one-shared-ai-vault?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">one shared AI vault</a> so changing providers is not a total faff.</p><p class="paragraph" style="text-align:left;">Also keep watching the Chinese labs. Their cheap open-weight models are why this argument became urgent in the first place. My <a class="link" href="https://aiwithkyle.com/ai-news/chinese-ai-just-changed-the-order?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=everyone-is-turning-on-anthropic" target="_blank" rel="noopener noreferrer nofollow">Chinese AI 101 issue</a> maps the main players.</p><p class="paragraph" style="text-align:left;">Use the best model. Keep your stuff portable. </p><p class="paragraph" style="text-align:left;">Do not let Dario, Jensen or Denny&#39;s decide your whole stack for you.</p><p class="paragraph" style="text-align:left;">OK maybe Denny’s.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=a4477d58-8a73-4b97-9fee-c6f523a726b4&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>AI Can Edit Video Now</title>
  <description>The one-click demos are still lying.</description>
  <link>https://newsletter.aiwithkyle.com/p/ai-video-editing-actually-works</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/ai-video-editing-actually-works</guid>
  <pubDate>Fri, 24 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-24T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><a class="image__link" href="https://youtu.be/pWK9Qz97zdo?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-can-edit-video-now" rel="noopener" target="_blank"><img alt="Kyle explaining his AI video editing system" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d4778e38-75e9-4542-928e-5e83ea1b6251/ai-video-editor-hook.gif?t=1784716677"/></a><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://youtu.be/pWK9Qz97zdo?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-can-edit-video-now" target="_blank" rel="noopener noreferrer nofollow">How to Edit Video with AI Full Video </a></p></span></div></div><p class="paragraph" style="text-align:left;">I recently let go of my video editor.</p><p class="paragraph" style="text-align:left;">Oof. Don’t feel great about it.</p><p class="paragraph" style="text-align:left;">He cost around $1,000 a month. The last few videos on my YouTube channel have been edited by AI instead. One went to about 5,000 views, which is solid for my relatively small channel.</p><p class="paragraph" style="text-align:left;">The work also gets done FAST (~1 hour) and costs basically nothing as I do it on my ChatGPT subscription</p><p class="paragraph" style="text-align:left;">Three months ago I could not get this to work dependably. I had been trying for more than a year and it was a mess. Cut-off sentences. Audio out of sync. Weird joins. A lot of time spent producing something a human then had to rescue.</p><p class="paragraph" style="text-align:left;">I gave it up as a bad job because I knew if I just waited a bit the AI would get better. And it did! </p><p class="paragraph" style="text-align:left;">Now Claude Code/Codex can take my messy livestream, cut it into a proper video, check the render, create captions, produce three YouTube packages and hand me a private draft.</p><p class="paragraph" style="text-align:left;">That is a <b>BIG</b> change.</p><p class="paragraph" style="text-align:left;">But the viral demos you’ll see on Instagram are still talking shit.</p><h2 class="heading" style="text-align:left;" id="the-onebutton-dream-is-still-a-drea">The one-button dream is still a dream</h2><p class="paragraph" style="text-align:left;">You cannot chuck an hour-long video into Claude Code, type &quot;make this cool and help me go viral&quot; and wander off for lunch. </p><p class="paragraph" style="text-align:left;">Well…you can. The result will probably be a bit crap. And it’ll have eaten all your tokens for its lunch. </p><p class="paragraph" style="text-align:left;">Nah we need to do some proper work and planning. Sorry! The thing that works is a pipeline. Each tool gets one job and the system stops at the points where I still need to make a call.</p><div class="image"><img alt="The real AI video editing setup" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b0bc1389-5995-4175-a32c-2a1793f1d6bc/01-actual-setup.png?t=1784716713"/><div class="image__source"><span class="image__source_text"><p>My basic pipeline</p></span></div></div><p class="paragraph" style="text-align:left;">First up (unsurprisingly) I shoot my video. Personally I do this via my livestream. You could just record offline. Same thing. </p><p class="paragraph" style="text-align:left;">My raw video files stay on my Mac. Whisper turns the audio into a timestamped transcript. This is all local (no AI needed). Free.</p><p class="paragraph" style="text-align:left;">A smart hosted model (Codex for me) reads the text and makes the editorial decisions. This part uses AI and my tokens.</p><p class="paragraph" style="text-align:left;"> FFmpeg cuts and renders the video locally. Again, no AI. Free.</p><p class="paragraph" style="text-align:left;"> Then the system checks the output, packages it and stops before publication. This part I bump back to Codex so it costs me tokens. </p><p class="paragraph" style="text-align:left;">See the back and forth? We what we can locally and for free. Then bump the parts that genuinely needs intelligence to a model. Technically as/when local models get better we could also do those parts on our computer for free.</p><p class="paragraph" style="text-align:left;">That split matters. Sending huge 4K files into a cloud model would be slow, expensive and daft. Text is cheap. AI is very good at reading it. FFmpeg is very good at making exact cuts. </p><p class="paragraph" style="text-align:left;">Give each one the job it is actually good at please and thank you.</p><h2 class="heading" style="text-align:left;" id="the-transcript-becomes-the-edit">The transcript becomes the edit</h2><p class="paragraph" style="text-align:left;">The model never needs to &quot;watch&quot; the full video in the way people imagine. </p><p class="paragraph" style="text-align:left;">Technically we have models that can “see” and they could watch frame by frame by frame. But good god that’s inefficient. </p><p class="paragraph" style="text-align:left;">Instead we get a transcript. Basically what is being said, with timecodes. </p><p class="paragraph" style="text-align:left;">This is what the AI works with. </p><p class="paragraph" style="text-align:left;">It reads the transcript and decides where the useful video begins, which repeated explanations can go, where I wandered off into a side quest and where the livestream ends. It also decides when to show the clean face camera and when to keep the programme view with the slides.</p><p class="paragraph" style="text-align:left;">That becomes an edit decision list. An EDL.</p><p class="paragraph" style="text-align:left;">There was a slightly annoying technical catch here. Models can make the <i>right</i> editorial decision and still give you the <i>wrong</i> timestamp. We saw them drift by two to six seconds. Plenty of room to chop the end off a sentence. And very annoying when I was getting this up and running. </p><p class="paragraph" style="text-align:left;">So the system now asks for the exact words around every cut. Local code finds those words in the transcript and places the blade. The model decides <i>what</i> to remove. Deterministic tools decide exactly <i>where</i> to cut. Works a treat.</p><p class="paragraph" style="text-align:left;">That one change took the workflow from impressive demo to something I can use.</p><p class="paragraph" style="text-align:left;">It is the same reason I keep banging on about building <a class="link" href="https://aiwithkyle.com/ai-news/build-one-shared-ai-vault?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-can-edit-video-now" target="_blank" rel="noopener noreferrer nofollow">one shared AI vault</a>. The value comes from keeping the rules, corrections and examples so the next run starts smarter.</p><p class="paragraph" style="text-align:left;">Every time my editor runs there is a chance it’ll mess up and make an error. </p><p class="paragraph" style="text-align:left;">We then FIX that error. And once it’s been fixed once that’s the last time it happens.</p><h2 class="heading" style="text-align:left;" id="render-small-check-it-then-go-big">Render small. Check it. Then go big.</h2><p class="paragraph" style="text-align:left;">Super important if you are running this. Don’t work with full resolution video. Instead use “proxies”. Basically, lower resolution versions that are easier to handle and edit. </p><p class="paragraph" style="text-align:left;">For me the system makes a 720p review version first. Rendering every experiment in 4K would take ages and set my Mac on fire.</p><p class="paragraph" style="text-align:left;">And before we render back to full glorious 4K we also run all our quality checks. Boring but important stuff. </p><p class="paragraph" style="text-align:left;">It decodes the full file. Re-transcribes the joins. Looks for long silences and black frames. Checks that the audio and video still end together. It even checks whether there is speech energy at the tail because Whisper sometimes decides the final words do not exist. Small tweaks that matter a lot unfortunately! </p><p class="paragraph" style="text-align:left;">After all this I still watch the first minute, listen to every flagged join and approve the 720p review file. <i>Only then</i> does it render the 4K master and prepare the YouTube draft.</p><div class="image"><img alt="AI handles the mechanics while Kyle keeps the decisions" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5194cd23-b999-4e35-a68f-41446618be61/02-human-gates.png?t=1784716729"/></div><h2 class="heading" style="text-align:left;" id="you-still-have-to-make-something-wo">You still have to make something worth watching</h2><p class="paragraph" style="text-align:left;">AI can handle the transcript, rough cuts, rendering, captions and packaging.</p><p class="paragraph" style="text-align:left;">It cannot rescue a useless idea.</p><p class="paragraph" style="text-align:left;">If you give a model rubbish, it will edit the rubbish very neatly.</p><p class="paragraph" style="text-align:left;">You still own the thesis. You still need to perform on camera. You still decide what gets published under your name.</p><p class="paragraph" style="text-align:left;">Just because we <i>can</i> easily edit anything now doesn’t mean we should! </p><p class="paragraph" style="text-align:left;">Still…it’s exciting. This is now VERY doable. </p><p class="paragraph" style="text-align:left;">That has changed over the last three months. It’s not that the models suddenly developed video editing capabilities - nope. They just became good enough at long context, structured instructions, tool use and self-checks to run the system dependably.</p><p class="paragraph" style="text-align:left;">I have turned the exact workflow into a step-by-step guide. It covers the local tools, the folder setup, the build prompt, the edit-decision format, the review gates and the stupid failures we hit so you do not have to repeat them.</p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="background-color:#5e17eb;" href="https://aiwithkyle.com/downloads/mini/ai-video-editor/build-your-own-ai-video-editor.pdf?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=ai-can-edit-video-now"><span class="button__text" style="color:#ffffff;"> Download the AI Video Editor how-to (PDF) </span></a></div><p class="paragraph" style="text-align:left;">Start with one camera and one talking-head video. Get the boring version working before asking it for swooshy animations eh?</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=f9970f6d-f35a-429e-a0eb-cc81ab433b5a&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>How to choose what AI to use</title>
  <description>China releases another model!</description>
  <link>https://newsletter.aiwithkyle.com/p/qwen3-8-model-fatigue</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/qwen3-8-model-fatigue</guid>
  <pubDate>Wed, 22 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-22T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Local Ai]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/704185a2-3f03-4810-9893-b0bcfcbcaf2a/qwen3-8-newsletter-clip.gif?t=1784547638"/><div class="image__source"><a class="image__source_link" href="https://www.youtube.com/watch?v=z5Fs3SVaOfk&t=20s&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=how-to-choose-what-ai-to-use" rel="noopener" target="_blank"><span class="image__source_text"><p><a class="link" href="https://www.youtube.com/watch?v=z5Fs3SVaOfk&t=20s&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=how-to-choose-what-ai-to-use" target="_blank" rel="noopener noreferrer nofollow">Watch on Youtube </a></p></span></a></div></div><p class="paragraph" style="text-align:left;">My last guide to Chinese AI lasted 48 hours. Less than! </p><p class="paragraph" style="text-align:left;">I published it on Friday. It became one of the best-performing videos on my YouTube channel - people want to know about Chinese AI eh? By Sunday Alibaba had announced Qwen3.8 and it was <i>already</i> out of date…</p><p class="paragraph" style="text-align:left;">Then this landed in my <a class="link" href="https://aiwithkyle.com/ai-with-kyle-group-chat?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=how-to-choose-what-ai-to-use" target="_blank" rel="noopener noreferrer nofollow">WhatsApp group</a>:</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0797db36-3054-41d4-b3d9-81fb866ecd62/whatsapp-question-anonymised.png?t=1784646653"/><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://aiwithkyle.com/ai-with-kyle-group-chat?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=how-to-choose-what-ai-to-use" target="_blank" rel="noopener noreferrer nofollow">Join the private Whatsapp group</a></p></span></div></div><p class="paragraph" style="text-align:left;"><i>&quot;It&#39;s pretty overwhelming Kyle. How do you exactly distinguish which model is best for which task?&quot;</i></p><p class="paragraph" style="text-align:left;">Yep. That is the <i>actual</i> story here. Let’s try to zoom out a bit from the specific releases (and boy there are a lot of them!) and instead look at the whole business of how to treat a new model. </p><p class="paragraph" style="text-align:left;">The problem is trying to run a business while another &quot;best model ever&quot; arrives every few days. How do we deal with that noise?</p><p class="paragraph" style="text-align:left;">And by god it is <i>very</i> noisy at the moment…</p><h2 class="heading" style="text-align:left;" id="a-launch-post-has-three-layers">A launch post has three layers</h2><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9c042e53-995c-4ab8-abe7-c5958fb25ed6/02-three-layers.png?t=1784547638"/></div><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/Alibaba_Qwen/status/2078759124914098291?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=how-to-choose-what-ai-to-use" target="_blank" rel="noopener noreferrer nofollow">Alibaba announced Qwen3.8</a> on Sunday. It says the model has 2.4 trillion parameters, that open weights are coming soon and that its performance is second only to Claude Fable 5.</p><p class="paragraph" style="text-align:left;">Big claim. Maybe true. Maybe not.</p><p class="paragraph" style="text-align:left;">First lesson: don’t listen to the launch post. </p><p class="paragraph" style="text-align:left;">Guess what. The company promoting their new model is likely to say it’s pretty good. Funny that. </p><p class="paragraph" style="text-align:left;">Qewn is still in preview. The full weights are not out. There is no model card, named independent benchmark or stable global API listing yet. A Max Preview is available through selected Alibaba products, including its <a class="link" href="https://help.aliyun.com/en/model-studio/token-plan-personal-overview?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=how-to-choose-what-ai-to-use" target="_blank" rel="noopener noreferrer nofollow">China Personal Token Plan</a>, Qoder and QoderWork.</p><p class="paragraph" style="text-align:left;">So the announcement is real. You <i>can</i> get limited access. But the #2 ranking is still Alibaba talking about Alibaba. Obviously they reckon it is brilliant (!!). OpenAI, Anthropic and every other lab do the <i>exact</i> same thing.</p><p class="paragraph" style="text-align:left;">Company launch posts tell you what to test. The proof comes later.</p><h2 class="heading" style="text-align:left;" id="a-quick-reminder-about-opensourcewe">A quick reminder about open-source/weights</h2><p class="paragraph" style="text-align:left;">Qwen says the weights are coming. Cool. That means we can download the WHOLE model on our computers. Unless, you know, Trump blocks it. Which is looking likely…</p><p class="paragraph" style="text-align:left;">But here is your daily reminder. This does not mean you can just “run it for free at home!”. I’m seeing a lot of posts on social media about this. It’s a great hook - you can run this Chinese model for free and never pay for Claude or ChatGPT again. Sounds great. But it’s entirely untrue. </p><p class="paragraph" style="text-align:left;">At 2.4 trillion parameters, a rough four-bit version would still be about 1.2TB before runtime overhead. You are <b>not</b> chucking that on a MacBook. You need serious server hardware or (more likely) somebody else&#39;s cloud that you rent.</p><h2 class="heading" style="text-align:left;" id="run-every-model-through-four-gates">Run every model through four gates</h2><p class="paragraph" style="text-align:left;">OK back to the new models. How do you know what to smash and what to pass. Gotta be a better way to say that. Oh well.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b1bb5dec-8b57-4758-9b34-70cc5cdbef12/06-four-gates.png?t=1784547638"/></div><p class="paragraph" style="text-align:left;">When a new model appears, I like to ask four questions:</p><p class="paragraph" style="text-align:left;"><b>Was it announced?</b> Find the first-party source, not a screenshot of a screenshot on Twitter. You’d be surprised how many new models don’t actually exist officially yet. This week everyone is talking about Opus 5 and its (apparently?) imminent release. So you need to be careful and make sure you are looking at an actually extant model and not random social noise!</p><p class="paragraph" style="text-align:left;"><b>Can you actually use it?</b> Check the exact model, product, region and access route. A preview in a Chinese subscription plan is not the same thing as a normal global API. Models also release at different rates in different geographies - here in the EU we tend to get things last. Which is…awesome.</p><p class="paragraph" style="text-align:left;"><b>Has anyone independent tested it?</b> Wait for proper benchmarks and people running <i>real</i> work. Producer numbers are where the evidence starts. I like to check out people I trust like Ethan Mollick, Simon Willison and Andrew Curran for their thoughts. Then benchmarks on sites like Artificial Analysis. Remember that the labs will <i>always</i> cherry-pick benchmarks that make them look good. So wait for independent reports. </p><p class="paragraph" style="text-align:left;"><b>Does it beat your current model on your work?</b> This is the big one. Can the model do the work you want to use it for? And can it not only match your current model but instead BEAT it? Don’t shift your work for a passing grade here - only shift if there is a substantial delta. Otherwise you incur switching costs for…what? </p><p class="paragraph" style="text-align:left;">Qwen3.8 passes gate one. It sorta passes gate two. Gates three and four ehhh not yet. So: we wait.</p><h2 class="heading" style="text-align:left;" id="make-your-workflow-the-leaderboard">Make your workflow the leaderboard</h2><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/caa6b0bf-a2e4-4b8d-b4b3-787589bc425b/07-test-workflow.png?t=1784547638"/></div><p class="paragraph" style="text-align:left;">If you still want to have a poke around, pick three low-risk tasks you already do. Run each one ten times on Qwen and your current model. Compare quality, time, cost and failures.</p><p class="paragraph" style="text-align:left;">This is not based on vibes. Not somebody else&#39;s coding benchmark. Your actual boring arse work. Which, remember, is all that really matters here. </p><p class="paragraph" style="text-align:left;">And do not switch because the new model is <i>about</i> as good. Moving prompts, context and workflows has a cost. It needs to be substantially better and/or substantially cheaper before the faff makes sense. </p><p class="paragraph" style="text-align:left;">This is why I keep banging on about building <a class="link" href="https://aiwithkyle.com/ai-news/build-one-shared-ai-vault?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=how-to-choose-what-ai-to-use" target="_blank" rel="noopener noreferrer nofollow">one shared AI vault</a>. Keep your context, files and processes outside a single model and trying a new provider is much less of a pain.</p><p class="paragraph" style="text-align:left;">Honestly thought? Most of the time? Stick with what works and get the job done. I know that’s boring! I know you want to play with shiny new cool toys. But remember you have WORK to do! Important work. And blindly shifting from model to model chasing the new best thing gets you no closer to your goals. </p><p class="paragraph" style="text-align:left;">I said it last week and apparently need to say it again: <a class="link" href="https://aiwithkyle.com/ai-news/dont-marry-the-model?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=how-to-choose-what-ai-to-use" target="_blank" rel="noopener noreferrer nofollow">do not marry the model</a>.</p><p class="paragraph" style="text-align:left;">Until then…scroll on.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=ef39b4c4-cbef-4be7-8ca8-286457b0271b&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Chinese AI Just Changed the Order</title>
  <description>Kimi doesn&#39;t NEED to beat Fable.</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f3043ca4-d5e7-481c-9f7b-65aaab491258/01-kimi-order.png" length="125298" type="image/jpeg"/>
  <link>https://newsletter.aiwithkyle.com/p/chinese-ai-just-changed-the-order</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/chinese-ai-just-changed-the-order</guid>
  <pubDate>Mon, 20 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-20T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Kimi]]></category>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Local Ai]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><a class="image__link" href="https://www.youtube.com/watch?v=xS2sF4u1GbQ&t=19s&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2b1d33c5-2d23-4e9b-a659-075fcd060d82/chinese-ai-newsletter-clip.gif?t=1784295146"/></a><div class="image__source"><a class="image__source_link" href="https://www.youtube.com/watch?v=xS2sF4u1GbQ&t=19s&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" rel="noopener" target="_blank"><span class="image__source_text"><p>Watch the Youtube: <a class="link" href="https://www.youtube.com/watch?v=xS2sF4u1GbQ&t=19s&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" target="_blank" rel="noopener noreferrer nofollow">Kimi 3.5 vs. Claude</a></p></span></a></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/xS2sF4u1GbQ?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order"><span class="button__text" style=""> Watch Now </span></a></div><p class="paragraph" style="text-align:left;">Claude Opus just fell to third place.</p><p class="paragraph" style="text-align:left;">Once Fable leaves normal subscriptions, my practical order is <b>Sol, Kimi, Opus, Grok.</b></p><p class="paragraph" style="text-align:left;">That is <i>my</i> order for the sort of coding work I do. Change the job, benchmark or harness and the order changes too. Please don&#39;t turn it into the Ten Commandments and shout at me on Twitter (!!). We have enough of that already.</p><p class="paragraph" style="text-align:left;">Kimi K-3 is the newest Chinese model, it is already extremely good and it costs far less than the Western premium models.</p><p class="paragraph" style="text-align:left;">This is not a future threat. It is here <i>right now</i>.</p><h2 class="heading" style="text-align:left;" id="kimi-doesnt-need-to-beat-fable">Kimi doesn&#39;t need to beat Fable</h2><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.kimi.com/fr-fr/blog/kimi-k3?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" target="_blank" rel="noopener noreferrer nofollow">Moonshot&#39;s new Kimi K3</a> is a 2.8-trillion-parameter model with native vision and a one-million-token context window. It is built for long coding jobs, knowledge work and agents.</p><p class="paragraph" style="text-align:left;">Is it the best? Nope! </p><p class="paragraph" style="text-align:left;">But it’s VERY good.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ac6296e6-6481-4a4a-8d9a-dfac91d74103/Screenshot_2026-07-17_at_4.47.19_pm.png?t=1784296050"/></div><p class="paragraph" style="text-align:left;">Moonshot&#39;s (the makers of Kimi) own numbers still put it behind Fable and ChatGPT Sol overall. Reasonable!</p><p class="paragraph" style="text-align:left;">Kimi does not need to win every benchmark. Once Fable disappears from normal subscription access, it only needs to be good enough to push Claude Opus down the list on work people <i>actually</i> care about.</p><p class="paragraph" style="text-align:left;">And it is <i>close enough</i> to do that.</p><p class="paragraph" style="text-align:left;">AND the open weights are meant to arrive by 27 July. </p><p class="paragraph" style="text-align:left;">We will be able to <i>download </i>the whole model. </p><p class="paragraph" style="text-align:left;">They are not available yet and K3 is NOT something you can download and casually run on your MacBook. Even when the weights arrive, the model is enormous. At four-bit quantisation, the weights alone would be roughly 1.4TB before overhead. That ain’t fitting on a MacBook!</p><p class="paragraph" style="text-align:left;">But it does underline China’s commitment to low cost and open weight models. </p><h2 class="heading" style="text-align:left;" id="chinese-ai-isnt-just-deep-seek">Chinese AI isn&#39;t just DeepSeek</h2><div class="image"><img alt="The main Chinese AI model families" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/960f7f8c-cbf0-4300-82f3-71a69d6b4d50/03-model-families.png?t=1784285690"/></div><p class="paragraph" style="text-align:left;">Useful moment in time to recap the Chinese AI ecosystem.</p><p class="paragraph" style="text-align:left;">Moonshot makes <b>Kimi</b>. DeepSeek is the price-and-efficiency wrecking ball - cheap as chips. Z.ai makes <b>GLM</b>. Alibaba makes <b>Qwen</b>, including the smaller models I use locally. MiniMax makes the <b>M</b> family. ByteDance makes <b>Doubao</b> and already has absurd consumer distribution through TikTok&#39;s Chinese sister app Douyin (the OG TikTok)</p><p class="paragraph" style="text-align:left;">You do not need to memorise every version number. They will all change by Tuesday anyway. Same as with the Western labs! </p><p class="paragraph" style="text-align:left;">The important bit to remember is that there is now a whole parallel AI market producing <i>serious</i> models for coding, reasoning, translation, agents and boring high-volume business work. A year or two ago Chinese releases tended to sit six to nine months behind the American frontier. Kimi has closed that gap to roughly a month or two.</p><h2 class="heading" style="text-align:left;" id="cheap-is-the-attack-vector">Cheap is the attack vector</h2><div class="image"><img alt="Chinese AI API pricing versus Western premium models" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3a6e46be-39c7-48ca-87cb-c1919697cbe6/04-cheap-attack.png?t=1784285704"/></div><p class="paragraph" style="text-align:left;">The main threat the Chinese labs pose is cost.</p><p class="paragraph" style="text-align:left;">Per million output tokens, the standard API prices I checked were $0.87 for DeepSeek V4 Pro, $15 for Kimi K3, $25 for Claude Opus 4.8, $30 for GPT-5.6 Sol and $50 for Claude Fable 5.</p><p class="paragraph" style="text-align:left;">Here’s Artificial Analysis’ cost per Intelligence Index.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/4a3365d8-466d-4568-893b-14f699cab51d/Screenshot_2026-07-17_at_6.07.07_pm.png?t=1784300884"/><div class="image__source"><span class="image__source_text"><p>DeepSeek continuing to be obscenely cheap</p></span></div></div><p class="paragraph" style="text-align:left;">Those are API output prices, not subscriptions and not the total cost of every job. Input, caching, tools and long context all change the bill. But the direction is kinda obvious…</p><p class="paragraph" style="text-align:left;">If a Chinese model gives you 80% or 90% of the quality for a fraction of the cost, you can use more tokens, add checking steps and still spend <i>vastly</i> less.</p><p class="paragraph" style="text-align:left;">That is the business model: <b>good enough, much cheaper and available everywhere.</b></p><p class="paragraph" style="text-align:left;">And even that “good enough” is increasingly not far behind Claude and ChatGPT. It used to be 60%-70% as good. Now it’s nipping their heels at 90% as good. </p><p class="paragraph" style="text-align:left;">This is a BIG problem if your trillion-dollar IPO valuation relies on selling premium tokens.</p><p class="paragraph" style="text-align:left;">I wrote about the wider mechanism in the <a class="link" href="https://aiwithkyle.com/ai-news/open-source-ai-price-war?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" target="_blank" rel="noopener noreferrer nofollow">open-source AI price war</a>. Open-weight alternatives put a ceiling on what closed providers can charge,<i> even when</i> the open model is not quite as good.</p><h2 class="heading" style="text-align:left;" id="six-different-things-get-mixed-up">Six different things get mixed up</h2><div class="image"><img alt="Cloud, local, open weights, open source, free and private explained" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/226b4c07-1fbd-4fe5-945b-6feac67f6ecf/05-six-definitions.png?t=1784285717"/></div><p class="paragraph" style="text-align:left;">People chuck around <i>open source</i>, <i>local</i>, <i>free</i> and <i>private</i> as if they all mean the same thing.</p><p class="paragraph" style="text-align:left;">They do not.</p><p class="paragraph" style="text-align:left;">I literally had someone in my comments saying that all the Chinese models are free so the US is in trouble…</p><p class="paragraph" style="text-align:left;">Well…not quite. It’s important to be clear on all of this. </p><p class="paragraph" style="text-align:left;"><b>Cloud or local</b> tells you where the model runs. Cloud means somebody else&#39;s server. Local means hardware you control. <i>Generally</i> when something is cloud based you are paying. When it is local you are not paying (except electricity). </p><p class="paragraph" style="text-align:left;"><b>Open or closed</b> tells you what you receive. Open weights means you can download the trained model files under a licence. It does not automatically give you the training data, code, process or unlimited rights. True open source is a much higher bar. </p><p class="paragraph" style="text-align:left;">Most American models are strict closed-source. You do not know what’s inside. And the companies releasing this information would tank their valuations. </p><p class="paragraph" style="text-align:left;"><b>Free or paid</b> is slippery.</p><p class="paragraph" style="text-align:left;">Generally cloud models are paid for with a subscription or API fees. This is what we are used to with ChatGPT and Claude. You pay your monthly fee and get access. </p><p class="paragraph" style="text-align:left;">Technically though a closed-source model<i> could</i> be free. For instance OpenAI has some of their older models sitting in the Playground that you can mess around with for free. But that’s a weird edge case. </p><p class="paragraph" style="text-align:left;">Open weights models can be paid <i>or</i> free. Which is confusing initially! How can they charge for something that they also give away for free??</p><p class="paragraph" style="text-align:left;">Compute is the answer. </p><p class="paragraph" style="text-align:left;">You <i>can</i> download Kimi K3 (or at least next week when they release the weights you can). You can then run it locally, fine tune it, do what you want with it. And not pay Moonshot a single penny. </p><p class="paragraph" style="text-align:left;">BUT to do so you are going to need a beast of a computer. More likely a server rack. Or several. You can’t just install this on a laptop - you will need several hundred thousand dollars of equipment. At least. </p><p class="paragraph" style="text-align:left;">As a result even open-weight model providers (like the Chinese labs) offer cloud versions. Versions you can sign up to on subscriptions and via the API - much like the closed source labs. You are paying for the convenience of, well, not building a data centre in your garage! </p><p class="paragraph" style="text-align:left;">If you want the gentle beginner version of local vs. cloud and want to explore setting up local models, my <a class="link" href="https://aiwithkyle.com/ai-news/local-llms-101?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" target="_blank" rel="noopener noreferrer nofollow">local LLM guide</a> walks through LM Studio, hardware and running your first model without touching the terminal.</p><h2 class="heading" style="text-align:left;" id="china-is-ahead-in-usage">China is ahead in usage </h2><p class="paragraph" style="text-align:left;">One other factor people miss is that China (and indeed most of Asia) is adoption AI extremely quickly. </p><p class="paragraph" style="text-align:left;">China had <a class="link" href="https://english.www.gov.cn/archive/statistics/202602/05/content_WS698442cac6d00ca5f9a08edc.html?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" target="_blank" rel="noopener noreferrer nofollow">602 million generative-AI users</a> by the end of 2025. That is 42.8% national adoption with 141.7% growth in a year. The <a class="link" href="https://hai.stanford.edu/assets/files/ai_index_report_2026_chapter_9_public_opinion.pdf?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" target="_blank" rel="noopener noreferrer nofollow">Stanford 2026 AI Index</a> also found workplace AI usage above 80% in China and several other emerging economies.</p><p class="paragraph" style="text-align:left;">That matches what I felt when I was there. Beijing bookshops had DeepSeek and AI books piled at the front, not hidden in the computer-science corner. And when I spoke to a woman from Hangzhou, where DeepSeek is based, she was <i>proud</i> of it. </p><p class="paragraph" style="text-align:left;">The mood was much closer to &quot;how quickly can we use this?&quot; than &quot;how quickly can we ban it?&quot; The West has spent a lot of time arguing whilst China got on with deploying.</p><h2 class="heading" style="text-align:left;" id="how-id-actually-use-chinese-ai">How I&#39;d actually use Chinese AI</h2><p class="paragraph" style="text-align:left;">Brass tacks! I would not move everything to Kimi tomorrow. Nah.</p><p class="paragraph" style="text-align:left;">For hard, valuable or high-risk work, use the best frontier model you can access and check the result. For most work that’s still ChatGPT and Claude. If you are comfortable using them they are still the best. For now! </p><p class="paragraph" style="text-align:left;">But definitely experiment around the edges with Chinese models. Mostly to know what’s available and see how they stack up. </p><p class="paragraph" style="text-align:left;">For routine coding and high-volume business work, test a cheaper hosted Chinese model. This is where the price gap becomes <i>real</i>. For sensitive, repetitive work, try a smaller local Qwen or DeepSeek model if it clears your quality bar and the app genuinely stays offline.</p><p class="paragraph" style="text-align:left;">For anything important, test on your own jobs. As always! There isn’t an objective “best” most of the time. Give three models the same ten tasks. Keep the cheapest one that reliably clears the bar without creating stupid risk.</p><p class="paragraph" style="text-align:left;"> I have been banging on about this for a while, but <a class="link" href="https://aiwithkyle.com/ai-news/dont-marry-the-model?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chinese-ai-just-changed-the-order" target="_blank" rel="noopener noreferrer nofollow">do not marry the model</a>. The order changed today. It will change again. Shit, probably next week when Opus 5 drops and/or they keep Fable on the subscription. (<i>I’m writing this Friday evening and the issue comes out Monday morning - let’s see how this holds up!</i>)</p><p class="paragraph" style="text-align:left;">Have a poke around. Just check where your data is going first eh?</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p><p class="paragraph" style="text-align:left;">PS. Full discussion on Youtube. Make sure you are subscribed! 😄 </p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/xS2sF4u1GbQ" width="100%"></iframe></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=05d6be65-8977-4822-be5a-6577357baeb7&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>ChatGPT Work Isn&#39;t For You</title>
  <description>Power users are annoyed. They&#39;re also missing the point.</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/c96eabee-220d-41d0-b080-b53f332676ab/01-slide-not-for-you.png" length="102490" type="image/jpeg"/>
  <link>https://newsletter.aiwithkyle.com/p/chat-work-or-codex</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/chat-work-or-codex</guid>
  <pubDate>Fri, 17 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-17T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Chatgpt]]></category>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="ChatGPT Work is not for power users" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/678ce98d-1585-47a6-847c-8bb49cb83579/01-slide-not-for-you.png?t=1784123311"/></div><p class="paragraph" style="text-align:left;">ChatGPT Work is probably not for you.</p><p class="paragraph" style="text-align:left;">If you&#39;re reading an AI newsletter, know what Codex is and have spent any time poking around Claude Code…you are already in the loud, weird power-user minority.</p><p class="paragraph" style="text-align:left;">And that minority is currently furious with ChatGPT Work. </p><p class="paragraph" style="text-align:left;">Work looks like a watered-down Codex. The interface is messy. The names are rubbish. Some options appear on desktop, some on web, some on mobile, and we have all decided this is a personal attack.</p><p class="paragraph" style="text-align:left;">All fair! </p><p class="paragraph" style="text-align:left;">But…OpenAI is aiming Work at the o<b>ther one billion ChatGPT users</b>. People who use ChatGPT every day but will never download a terminal, open Codex or learn what an MCP is.</p><p class="paragraph" style="text-align:left;">It’s not for YOU as a power user. It’s for the rest of humanity. </p><h2 class="heading" style="text-align:left;" id="the-loud-minority">The loud minority </h2><p class="paragraph" style="text-align:left;">Last week I wrote about <a class="link" href="https://aiwithkyle.com/ai-news/give-chatgpt-a-job?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-work-isn-t-for-you" target="_blank" rel="noopener noreferrer nofollow">why ChatGPT Work matters</a>. It gives normal ChatGPT users access to the agentic stuff power users have been hammering for the last six or seven months. That (for me) is pretty cool. </p><p class="paragraph" style="text-align:left;">Those of us who are power users got used to AI doing the work <i>very</i> quickly. That’s just what AI does right? </p><p class="paragraph" style="text-align:left;">But <i>most</i> people still ask ChatGPT a question, get an answer and then go away to do the actual job themselves. ChatGPT Work is the bridge from asking to delegating.</p><p class="paragraph" style="text-align:left;">So yes, Codex is more powerful. It has deeper controls. It can build and run “proper” software. It’s great. I use it daily for <i>faaar</i> too long. If you&#39;re already happy there, stay there! </p><p class="paragraph" style="text-align:left;">Just know that Work is the <i>easier</i> door for everyone else.</p><h2 class="heading" style="text-align:left;" id="one-app-three-doors">One app, three doors</h2><div class="image"><img alt="Three doors into ChatGPT: Chat, Work and Codex" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/fbb99efd-57e3-431f-b94a-2e97138d1a5d/04-slide-three-doors.png?t=1784123498"/></div><p class="paragraph" style="text-align:left;">Admittedly it’s a little confusing now. We have three different ways to use ChatGPT. </p><p class="paragraph" style="text-align:left;">The easiest way to choose is to <span style="text-decoration:underline;"><b>decide what you want back from it.</b></span></p><p class="paragraph" style="text-align:left;">Use <b>Chat</b> when you want to think something through. Ask questions. Compare options. Research an idea. Argue with it. Get clearer on what you <i>actually</i> want. The chat itself is the output. </p><p class="paragraph" style="text-align:left;">Use <b>Work</b> when you already know the outcome and want a <i>finished</i> thing back. A report. A presentation. An analysis. A research pack. Something you can open, check and use. The output will be a file or some sort of artefact. </p><p class="paragraph" style="text-align:left;">Use <b>Codex</b> when the finished thing is software. A site, app, database, automation or technical system. </p><p class="paragraph" style="text-align:left;">The ten-second version:</p><ul><li><p class="paragraph" style="text-align:left;"><b>Think:</b> Chat</p></li><li><p class="paragraph" style="text-align:left;"><b>Make:</b> Work</p></li><li><p class="paragraph" style="text-align:left;"><b>Code:</b> Codex</p></li></ul><p class="paragraph" style="text-align:left;">There is no &quot;best&quot; door. Depends what you need back.</p><h2 class="heading" style="text-align:left;" id="chat-first-then-hand-it-over">Chat first, then hand it over</h2><p class="paragraph" style="text-align:left;">And we can chain them! </p><p class="paragraph" style="text-align:left;">For example I used Chat to work out the argument for my livestream presentation. We went back and forward until the structure was right. Then I copied the conversation link into Work and told it to build the deck.</p><p class="paragraph" style="text-align:left;">Literally shared the chat. Copied the link. Moved to Work mode. Pasted the link and said “let’s work on this”. </p><p class="paragraph" style="text-align:left;">Chat helped me decide then Work made the “thing”.</p><p class="paragraph" style="text-align:left;">That handoff is useful because giving an agent a fuzzy brief just lets it produce polished rubbish without bothering you for a while. More autonomy does not rescue confused thinking. It merely lets the confusion run for longer without you knowing!! Whilst burning up your usage.</p><p class="paragraph" style="text-align:left;">So…start in Chat when you are unsure. Work out the audience, the outcome and what good looks like. Then send the settled thinking to Work to get the “thing” done. </p><h2 class="heading" style="text-align:left;" id="give-work-one-proper-job">Give Work one proper job</h2><div class="image"><img alt="Give ChatGPT Work one proper job" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f30bd806-0059-4a07-9b25-bd3a0d16bba8/08-slide-one-job.png?t=1784123507"/></div><p class="paragraph" style="text-align:left;">Work can do, well, work. BUT don’t overburden it. </p><p class="paragraph" style="text-align:left;">If the task is TOO big then Codex is going to be a better tool. </p><p class="paragraph" style="text-align:left;">Instead with Work stick to relatively discrete, simple jobs to start. </p><p class="paragraph" style="text-align:left;">For example: pick one job you already understand. Maybe turn a meeting transcript into actions and owners. Compare three supplier quotes. Turn an approved outline and source material into a presentation. Build the weekly report in the same format as last week. These sort of jobs with a defined input and output. Something you can look at and think “yup, it did a good job!”</p><p class="paragraph" style="text-align:left;">Give it:</p><ol start="1"><li><p class="paragraph" style="text-align:left;">The outcome you want.</p></li><li><p class="paragraph" style="text-align:left;">The source material.</p></li><li><p class="paragraph" style="text-align:left;">The boundaries.</p></li><li><p class="paragraph" style="text-align:left;">What &quot;done&quot; looks like.</p></li></ol><p class="paragraph" style="text-align:left;">Your prompting skills still matter because prompting is communication. The <a class="link" href="https://aiwithkyle.com/catalog/prompting-fundamentals?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-work-isn-t-for-you" target="_blank" rel="noopener noreferrer nofollow">Prompting Fundamentals playbook</a> applies even more when the AI can run off for 12 minutes and do a load of work without you. You still NEED to communicate what you actually want if you want good results. </p><p class="paragraph" style="text-align:left;">Oh and as always <b>check the result. </b>Sources, numbers, links, the lot. A nice-looking deck can still be full of nonsense. Check the work! </p><p class="paragraph" style="text-align:left;">So. If you are a power user your task today is to keep moaning about ChatGPT Work and the new app! 😛 </p><p class="paragraph" style="text-align:left;"> Everyone else should open it and give it one real job.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=d37e5ffe-e727-491a-81ae-d5afbc480a0f&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Build While It Lasts</title>
  <description>OpenAI and Anthropic are fighting. We got the pony.</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5b298c21-60e8-4364-9c71-676fe02e3073/1.png" length="949642" type="image/png"/>
  <link>https://newsletter.aiwithkyle.com/p/build-while-it-lasts</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/build-while-it-lasts</guid>
  <pubDate>Wed, 15 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-15T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Chatgpt]]></category>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Fable]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="Livestream clip showing OpenAI and Anthropic competing while builders benefit" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/6b0d0b1e-935b-44fe-9064-c4776ae98e84/openai-anthropic-top.gif?t=1784030388"/><div class="image__source"><span class="image__source_text"><p><i><a class="link" href="https://youtu.be/Jo9koewf764?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/Jo9koewf764</a></i><i> - watch now or save for later</i></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/Jo9koewf764?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts"><span class="button__text" style=""> Watch Now </span></a></div><p class="paragraph" style="text-align:left;">OpenAI and Anthropic are fighting over us. </p><p class="paragraph" style="text-align:left;">More accurately, they are fighting over the people hammering Claude Code and Codex all day. But still…we are in the middle and the divorced parents are trying to prove who loves us most.</p><p class="paragraph" style="text-align:left;">So we got a pony.</p><p class="paragraph" style="text-align:left;">On Sunday Anthropic <a class="link" href="https://x.com/claudeai/status/2076351399999557669?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts" target="_blank" rel="noopener noreferrer nofollow">extended Claude Fable 5</a> through July 19 and kept Claude Code&#39;s weekly limits 50% higher.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/claudeai/status/2076351399999557669?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">Exactly 57 minutes and 53 seconds later, OpenAI removed Codex&#39;s five-hour restriction and reset everybody&#39;s usage.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/thsottiaux/status/2076365965915467978?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">In those 53 minutes I burned 50% of my weekly Codex usage but I <i>knew</i> Tibo was about to drop another reset on us. </p><p class="paragraph" style="text-align:left;">I was <i>very</i> happy to be right! </p><h2 class="heading" style="text-align:left;" id="one-day-one-billion-tokens">One Day. One Billion Tokens.</h2><p class="paragraph" style="text-align:left;">One heavy day last month I got through roughly <b>940 million tokens</b> in Codex. Call it a billion because I was close enough and it makes the maths less annoying…</p><p class="paragraph" style="text-align:left;">At Fable&#39;s API list price, one billion tokens would cost $10,000 if every token was input or $50,000 if every token was output. The real bill would sit somewhere between those limits depending on the mix and caching - probably $15,000 for the day. </p><p class="paragraph" style="text-align:left;">One day mind you.</p><p class="paragraph" style="text-align:left;">That is <i>not</i> what I paid. Phew.</p><p class="paragraph" style="text-align:left;">I paid for a normal subscription. Like a sane person. </p><div class="image"><img alt="AI with Kyle livestream slide explaining the value of one billion AI tokens" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d7684d7a-94a8-4d60-80bb-0e5a0dac16d8/part-1-one-day-one-billion-tokens-clean-1784034409762.png?t=1784034424"/></div><p class="paragraph" style="text-align:left;">The labs are <i>temporarily</i> converting five-figure compute habits into a two or three-figure bill. They are eating the cost because they want the power users, the developers and the businesses who will eventually spend proper money through the API.</p><p class="paragraph" style="text-align:left;">This is a subsidy window. </p><p class="paragraph" style="text-align:left;">It is also why people are getting (very) twitchy about Fable disappearing.</p><h2 class="heading" style="text-align:left;" id="why-fable-keeps-coming-back">Why Fable Keeps Coming Back</h2><p class="paragraph" style="text-align:left;">Fable is Anthropic&#39;s halo model. It gives Claude the claim to having the best model available on a subscription.</p><p class="paragraph" style="text-align:left;">The problem is that GPT-5.6 Sol is now extremely close on the <a class="link" href="https://artificialanalysis.ai/articles/gpt-5-6-has-landed?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts" target="_blank" rel="noopener noreferrer nofollow">Artificial Analysis intelligence index</a>, while costing about a third as much on their benchmark. Grok 4.5 is hanging around the frontier too. </p><p class="paragraph" style="text-align:left;">If Anthropic removes Fable from subscriptions, Claude&#39;s paid offer suddenly looks a lot less exciting at exactly the point OpenAI is piling on pressure. It could go from 1st position to 3rd position (behind Grok!!!).</p><p class="paragraph" style="text-align:left;">So Fable gets another week.</p><p class="paragraph" style="text-align:left;">My guess is that Anthropic wants Opus 5 ready around July 20 or 21. That is a <i>guess</i>, based on the extension ending on the 19th. If Opus 5 is not ready, I reckon Fable gets extended again…</p><p class="paragraph" style="text-align:left;">Anything but lose position to ChatGPT. </p><h2 class="heading" style="text-align:left;" id="every-reset-the-builders-go-feral">Every Reset, The Builders Go Feral</h2><p class="paragraph" style="text-align:left;">My <a class="link" href="https://aiwithkyle.com/ai-with-kyle-group-chat?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts" target="_blank" rel="noopener noreferrer nofollow">WhatsApp group</a> has about 100 people in it. Every time Anthropic or OpenAI resets the meter, people drop what they are doing and pile back into building.</p><p class="paragraph" style="text-align:left;">I was working in Codex until about 11pm on Sunday because I expected the reset… adn so were many people in my community. Getting things done whilst the going is good. </p><p class="paragraph" style="text-align:left;">Rob Hallam took it much further. Over the last couple weeks he bought four $200 Claude accounts so he could keep Fable running, then stayed up all night trying to finish work before access disappeared. He ended up in hospital with stress and panic-attack-like symptoms.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/robj3d3/status/2076356929878966555?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">To be clear, the model did not put him in hospital. The stress and binge-work cycle did. Rob said that himself. This is not hustle porn. It is a sign that shifting limits and rubbish communication are making otherwise sensible people behave irrationally. Get well soon Rob! </p><div class="image"><img alt="AI with Kyle livestream slide showing resets turning rented intelligence into owned assets" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/531b0de4-6d38-4808-84b4-e2fca5b2afc9/part-2-every-reset-builders-go-feral-clean-1784034410645.png?t=1784034438"/></div><h2 class="heading" style="text-align:left;" id="rent-the-intelligence-own-the-outpu">Rent The Intelligence → Own The Output.</h2><p class="paragraph" style="text-align:left;">There is a REAL urgency though. </p><p class="paragraph" style="text-align:left;">Right now we are getting tens of thousands worth of usage for a hundred bucks a month or less. </p><p class="paragraph" style="text-align:left;">The cheap state-of-the-art AI will not be here forever. And when subscriptions stop being subsidise (because, you know, the companies want to make a profit!) we are left with escalating AI bills. </p><p class="paragraph" style="text-align:left;">This means we have a window of opportunity to build and create. </p><p class="paragraph" style="text-align:left;">Use the rented intelligence while it is cheap. Turn it into things that remain after the reset disappears:</p><ul><li><p class="paragraph" style="text-align:left;">Code</p></li><li><p class="paragraph" style="text-align:left;">Products</p></li><li><p class="paragraph" style="text-align:left;">Systems</p></li><li><p class="paragraph" style="text-align:left;">Audience</p></li><li><p class="paragraph" style="text-align:left;">Revenue</p></li></ul><p class="paragraph" style="text-align:left;">We <b>lease to buy</b> if that makes this clearer. We spend a little now (a few hundred dollars on AI subscriptions) to build assets that can cash flow life changing amounts of money. </p><p class="paragraph" style="text-align:left;">There is a darker edge here too. People who turn cheap access into skills, assets and cash <span style="text-decoration:underline;">now</span> can afford better access later. People who wait may find the tools more expensive at the same time their existing work is under pressure.</p><p class="paragraph" style="text-align:left;">That is the K-shaped economy warning- a split between those who use AI to accelerate now and those who ignore it. </p><p class="paragraph" style="text-align:left;">Those who grab the opportunity and run with it will be able to make money and thus continue to afford the use of AI. </p><p class="paragraph" style="text-align:left;">Those who do not will no longer be able to enter the game because the cost of AI will increase. They’ll be locked out forever. </p><p class="paragraph" style="text-align:left;">This is possible outcome, not an economic law carved into stone. But I would rather you start learning now than discover in two years that the useful stuff costs a grand a month.</p><p class="paragraph" style="text-align:left;">The answer is not to panic harder. Build something.</p><h2 class="heading" style="text-align:left;" id="your-first-job-build">Your First Job: Build</h2><div class="image"><img alt="AI with Kyle livestream slide showing a four-step first AI build process" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/33fa9852-ac5a-4e16-9f80-4db6c012b015/part-3-your-first-job-build-clean-1784034411606.png?t=1784034454"/></div><p class="paragraph" style="text-align:left;">Your first project should be a toy. Open Codex, tell it what you want, look at what it gives you and keep talking until it works. You do not need to read the code. You need judgement, feedback and a real problem.</p><p class="paragraph" style="text-align:left;">The <a class="link" href="https://aiwithkyle.com/mini/first-steps-vibe-coding?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts" target="_blank" rel="noopener noreferrer nofollow">First Steps in Vibe Coding guide</a> will get you from zero to a working app without the technical faff.</p><p class="paragraph" style="text-align:left;">Then build a second thing somebody might actually pay for. Start with the problem and the buyer rather than disappearing for six months to build your grand vision. My <a class="link" href="https://aiwithkyle.com/catalog/ai-business-ideas?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts" target="_blank" rel="noopener noreferrer nofollow">AI Business Ideas playbook</a> walks through that bit.</p><p class="paragraph" style="text-align:left;">And use the expensive models where they matter. I use Fable for the big hairy planning and judgement, then pass implementation to cheaper models like GPT-5.6 Sol in Codex. Better brains for the hard decisions. Cheaper hands for the build. I wrote a full guide on this recently - <a class="link" href="https://aiwithkyle.com/ai-news/fable-limited-sol-here?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts" target="_blank" rel="noopener noreferrer nofollow">how to work with multiple AI tools</a>.</p><p class="paragraph" style="text-align:left;">If you are already ahead, bring someone with you. Businesses are desperate for people who can translate this stuff into normal English. I am teaching the full workshop model on <a class="link" href="https://aiwithkyle.com/webinar?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts" target="_blank" rel="noopener noreferrer nofollow">Wednesday at 6pm UK time</a> (tonight), and there is a recording if you are watching the football! </p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://aiwithkyle.com/webinar?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-while-it-lasts"><span class="button__text" style=""> Join the free workshop webinar </span></a></div><p class="paragraph" style="text-align:left;">The meter will come back. Build before it does.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=df92c347-6bcc-424a-9f07-03947ee81a54&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Agents for ALL</title>
  <description>ChatGPT Work is confusing. It&#39;s also a very big deal.</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/cfe90660-45be-4155-aff3-87891be6e88a/chatgpt-work-p001.png" length="68398" type="image/jpeg"/>
  <link>https://newsletter.aiwithkyle.com/p/give-chatgpt-a-job</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/give-chatgpt-a-job</guid>
  <pubDate>Sat, 11 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-11T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Chatgpt]]></category>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="ChatGPT Work livestream: give it a job" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/28fc242e-f9b0-4c13-9a70-012f66b5a1c3/chatgpt-work-top.gif?t=1783676280"/><div class="image__source"><span class="image__source_text"><p>Watch now of save to watch later - <a class="link" href="https://youtu.be/dypPkH3Kmwc?si=16rBQfUuKgVOB97p&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=agents-for-all" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/dypPkH3Kmwc?si=16rBQfUuKgVOB97p</a></p></span></div></div><div class="button" style="text-align:left;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/dypPkH3Kmwc?si=16rBQfUuKgVOB97p&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=agents-for-all"><span class="button__text" style=""> Watch Now </span></a></div><p class="paragraph" style="text-align:left;">OpenAI released ChatGPT Work yesterday and everyone has the same response:</p><p class="paragraph" style="text-align:left;"><b>What the hell is this?</b></p><p class="paragraph" style="text-align:left;">Fair!</p><p class="paragraph" style="text-align:left;">My <a class="link" href="https://aiwithkyle.com/ai-with-kyle-group-chat?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=agents-for-all" target="_blank" rel="noopener noreferrer nofollow">Whatsapp chat</a> is full of peeved people. And…I get it. </p><p class="paragraph" style="text-align:left;">OpenAI now has Chat, Work and Codex inside one app. Then you can choose a model. Then you can choose an effort level. Then some options appear on desktop but not mobile, or on mobile but not web, depending on whether the rollout gods like you today.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/rasbt/status/2075369179817902176?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=agents-for-all"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">It&#39;s a bit of a mess!</p><p class="paragraph" style="text-align:left;">BUT…underneath the rubbish naming, something <i>very</i> important just happened. And I think most people missed it. </p><p class="paragraph" style="text-align:left;">For 1 BILLION people just ChatGPT stopped being somewhere you ask questions.</p><h2 class="heading" style="text-align:left;" id="one-app-three-doors">One App, Three Doors</h2><div class="image"><img alt="One app, three doors: Chat, Work and Codex" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d3a180f8-e31e-4eae-bbbc-5a228763e7cd/chatgpt-work-p003.png?t=1783676282"/></div><p class="paragraph" style="text-align:left;">OK so what do we actually have now.</p><p class="paragraph" style="text-align:left;">The simple version:</p><ul><li><p class="paragraph" style="text-align:left;"><b>Chat:</b> ask a question and get an answer.</p></li><li><p class="paragraph" style="text-align:left;"><b>Work:</b> give it an outcome and let it plan, act and deliver.</p></li><li><p class="paragraph" style="text-align:left;"><b>Codex:</b> build software with the full developer controls visible.</p></li></ul><p class="paragraph" style="text-align:left;">Same house. Three doors.</p><p class="paragraph" style="text-align:left;">The weird bit is that the walls are a bit flimsy. HOW and WHERE you access these three modes depends on what tool you are using:</p><ul><li><p class="paragraph" style="text-align:left;">Phone app</p></li><li><p class="paragraph" style="text-align:left;">Web interface</p></li><li><p class="paragraph" style="text-align:left;">Desktop app</p></li></ul><p class="paragraph" style="text-align:left;">On the Desktop app you switch between Work and Codex. They seem pretty damn similar though - just some visibility differences. You can also pop up a small Chat window for quick questions. But otherwise Chat is gone. </p><p class="paragraph" style="text-align:left;">On the web version you can switch between Chat and Work. </p><p class="paragraph" style="text-align:left;">On the phone version you have Chat and Work and Remote (to Codex).</p><p class="paragraph" style="text-align:left;">Oof. Complex. And annoying to relearn. </p><p class="paragraph" style="text-align:left;">That is why power users are moaning. It feels redundant.</p><p class="paragraph" style="text-align:left;">But there are <i>millions</i> (nay, 1 billion) people who will happily use ChatGPT every day and will never install Codex, Claude Code or OpenClaw. They hear &quot;coding agent&quot; and do one.</p><p class="paragraph" style="text-align:left;">Give them a toggle marked <b>Work</b> though…nice! They’ll get that. And because of that they’ve just been handed agentic AI. </p><h2 class="heading" style="text-align:left;" id="the-unit-of-work-changed">The Unit Of Work Changed</h2><div class="image"><img alt="Chat answers while Work plans, acts and delivers" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3aca8829-061d-43f0-9467-8537da5bbc33/chatgpt-work-p004.png?t=1783676283"/></div><p class="paragraph" style="text-align:left;">My livestream presentation was made with ChatGPT Work.</p><p class="paragraph" style="text-align:left;">I gave it screenshots from Twitter and a long voice note with me blabbering about the launch. It worked for about 12 minutes, spun up three sub-agents, checked official sources, found outside research, outlined the story, made the slides with ImageGen and packaged the lot into a PDF.</p><p class="paragraph" style="text-align:left;">I steered it. It did the <i>faff</i> whilst I had my coffee.</p><p class="paragraph" style="text-align:left;">If you&#39;ve already been <a class="link" href="https://aiwithkyle.com/ai-news/192-turning-codex-into-personal-assistant?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=agents-for-all" target="_blank" rel="noopener noreferrer nofollow">turning Codex into a personal assistant</a>, that sounds normal. So what? But for a standard ChatGPT user it is a <i>proper</i> change.</p><p class="paragraph" style="text-align:left;">Chat has trained us to ask, &quot;How do I make this report?&quot;</p><p class="paragraph" style="text-align:left;">Work lets you say, &quot;Make the report. Here are the meeting notes, the source files and the format I need. Bring me back a finished draft.&quot;</p><p class="paragraph" style="text-align:left;">And that power is now in the hands of everyone. </p><h2 class="heading" style="text-align:left;" id="the-fourth-leap-goes-mainstream">The Fourth Leap Goes Mainstream</h2><div class="image"><img alt="Four leaps in modern AI" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2b8a8fff-e11a-4a09-a4e2-b1e5063ffcec/chatgpt-work-p006.png?t=1783676283"/></div><p class="paragraph" style="text-align:left;">Ethan Mollick talks about four big jumps in modern AI.</p><p class="paragraph" style="text-align:left;">GPT-3.5 made the whole thing visible. GPT-4 made it useful for real work. Reasoning models made longer, harder problems possible. Then late last year we got agent systems that could actually go away and complete jobs.</p><p class="paragraph" style="text-align:left;">OpenClaw, Claude Cowork and Codex already did this. ChatGPT Work matters because OpenAI has put the <i>same</i> basic behaviour behind a normal-looking toggle in the product everybody already knows.</p><p class="paragraph" style="text-align:left;">Distribution <i>matters</i>.</p><h2 class="heading" style="text-align:left;" id="give-it-one-real-job">Give It One Real Job</h2><p class="paragraph" style="text-align:left;">OK how to actually get started with this? This is not about the tech. It’s about your skill level. You need to up your delegation skill.</p><div class="image"><img alt="Delegation: outcome, sources, boundaries and definition of done" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/12055fdf-b189-45c6-b344-9405e551f51e/chatgpt-work-p011.png?t=1783676284"/></div><p class="paragraph" style="text-align:left;">First up, don&#39;t start with a grand plan for an autonomous AI company. Yeesh.</p><p class="paragraph" style="text-align:left;">Pick <i>one</i> job you already do that takes about an hour. A report. Meeting notes that need turning into actions. A pile of files that needs analysing. A voice note and screenshots that need becoming a presentation.</p><p class="paragraph" style="text-align:left;">Then give Work four things:</p><ol start="1"><li><p class="paragraph" style="text-align:left;">The outcome you want.</p></li><li><p class="paragraph" style="text-align:left;">The sources it should use.</p></li><li><p class="paragraph" style="text-align:left;">The boundaries it must stay inside.</p></li><li><p class="paragraph" style="text-align:left;">A clear definition of what done looks like.</p></li></ol><p class="paragraph" style="text-align:left;">Your prompting skills still matter because prompting is <i>communication</i>. The better you can explain the result, context and boundaries, the less rubbish you get back. The <a class="link" href="https://aiwithkyle.com/catalog/prompting-fundamentals?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=agents-for-all" target="_blank" rel="noopener noreferrer nofollow">Prompting Fundamentals playbook</a> still applies. Work just gives those instructions to a system that can <i>keep going</i> for 12 minutes instead of firing back one answer.</p><p class="paragraph" style="text-align:left;">Update the app. Find Work. Give it one real job.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=37d74395-45bc-4f7b-af8d-b41fe63ae028&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Don&#39;t Marry the Model</title>
  <description>Fable is limited. Sol is here. </description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f9ea0dca-7d10-433d-a302-b6c4c2cac12e/slide-02-model-not-system.png" length="108981" type="image/jpeg"/>
  <link>https://newsletter.aiwithkyle.com/p/model-is-not-the-system</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/model-is-not-the-system</guid>
  <pubDate>Fri, 10 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-10T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Fable]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="Everyone is switching models again" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b629fb15-ba5b-4a51-8f6f-08cda22d2e9a/model-switching-top.gif?t=1783621199"/><div class="image__source"><span class="image__source_text"><p><i><a class="link" href="https://youtu.be/aL2sNhnNNcM?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=don-t-marry-the-model" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/aL2sNhnNNcM</a></i><i> - watch now or save for later</i></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/aL2sNhnNNcM?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=don-t-marry-the-model"><span class="button__text" style=""> Watch Now </span></a></div><p class="paragraph" style="text-align:left;">Everyone is switching models again! </p><p class="paragraph" style="text-align:left;">Fable got a stay of execution until Sunday! ChatGPT just dropped Sol, Terra and Luna! xAI released Grok 4.5!</p><p class="paragraph" style="text-align:left;">Hell, even Meta are back in the game with a new model. Whaa?</p><p class="paragraph" style="text-align:left;">Nice! More competition. Maximum boost! </p><p class="paragraph" style="text-align:left;">We are drowning in state of the art models. Including Fable. </p><p class="paragraph" style="text-align:left;">Fun if you still have usage. Less fun if you burned through your Fable allowance before the deadline and then Anthropic wandered back in like, &quot;actually lads, have five more days.&quot;</p><p class="paragraph" style="text-align:left;">Annoying.</p><p class="paragraph" style="text-align:left;">But I reckon the bigger point is not Fable. Or Sol. Or Grok being &quot;Opus-class&quot; (we shall see, Elon, we shall see…).</p><p class="paragraph" style="text-align:left;">The wrong question is: which model wins?</p><p class="paragraph" style="text-align:left;">The useful question is: <b>how do they </b><b><i>work together</i></b><b>?</b></p><h2 class="heading" style="text-align:left;" id="the-model-is-not-the-system">The model is not the system</h2><div class="image"><img alt="The model is not the system" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3eae245e-457b-40bc-af4e-258441a2086f/slide-02-model-not-system.png?t=1783621199"/><div class="image__source"><span class="image__source_text"><p>The model is not the system</p></span></div></div><p class="paragraph" style="text-align:left;">Models change. But you can’t be constantly switching or you’ll get NOTHING done. </p><p class="paragraph" style="text-align:left;">Instead focus on your workflows. And make the model agnostic. </p><p class="paragraph" style="text-align:left;">If your entire AI workflow is &quot;open the smartest model and ask it everything&quot;, you are going to have a miserable few years. The smartest model will change - weekly. Or daily if this week is anything to go by! The price will change. The access rules will change.</p><p class="paragraph" style="text-align:left;">Change is the only constant as some smart Greek who I can’t be bothered to look up said. ChatGPT would have known…</p><p class="paragraph" style="text-align:left;">This week it is Fable, Sol, Grok and (checks notes) Meta Spark 1.1. A year from now it will be other names. </p><p class="paragraph" style="text-align:left;">So the durable skill is not brand loyalty. It is workflows that can swap models in and out without collapsing. I wrote about this from the pricing side in the <a class="link" href="https://aiwithkyle.com/ai-news/fable-is-going-away?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=don-t-marry-the-model" target="_blank" rel="noopener noreferrer nofollow">Fable is going away issue</a>. Frontier intelligence is getting rationed. Sometimes by money. Sometimes by policy. Sometimes by capacity.</p><p class="paragraph" style="text-align:left;">So the smart play is to use the cleverest model where it actually matters. Use the cheaper thing where it is good enough. Connect the two.</p><h2 class="heading" style="text-align:left;" id="use-expensive-intelligence-where-it">Use expensive intelligence where it matters</h2><div class="image"><img alt="The simple split" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/8c84391b-33e6-4baf-b751-51e68877d0fd/slide-03-simple-split.png?t=1783621199"/><div class="image__source"><span class="image__source_text"><p>Planner model. Worker model. Stop making the expensive one do everything.</p></span></div></div><p class="paragraph" style="text-align:left;">The simple split is planner model and worker model. Or an orchestrator and its agents if you fancy. </p><p class="paragraph" style="text-align:left;"><b>Planner model:</b> strategy, architecture, judgement, trade-offs, review. This is where you use Fable, Sol or whatever the current expensive biggest brain is. </p><p class="paragraph" style="text-align:left;"><b>Worker model:</b> file edits, tests, implementation, cleanup, repeatable tasks. This can be Codex, a cheaper GPT model, a hosted open model, a local model. Whatever clears the bar.</p><p class="paragraph" style="text-align:left;">On our live AI Hour call last night Adam talked about how he recently used Fable to scope out a project to convert a PHP project he made 20 years ago into a more modern form. Fable got to work on the strategy and the delegated step by step work to Opus and Sonnet. A big chunky project that took the smaller AIs 4 days(!) but meant saving potentially thousands one tokens by intelligently using the right model for the job. Super valuable skill. </p><p class="paragraph" style="text-align:left;">ClaudeDevs had a good version of this on X: use Fable as the advisor, then let a cheaper executor call it when it needs guidance.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/ClaudeDevs/status/2074606058128224365?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=don-t-marry-the-model"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">As I mentioned in the last newsletter I spent about 900 million tokens on Codex on one heavy day (June 14th to be precise). If I ran that through Fable API pricing, the back-of-the-envelope number was around ~$15,000. For one day.</p><p class="paragraph" style="text-align:left;"><i>This </i>is why we need to get smarter with how we use different tools. </p><p class="paragraph" style="text-align:left;">I personally do not want to remortgage my house because I asked the clever model to fix a typo in a test file.</p><p class="paragraph" style="text-align:left;">So do not use the smartest model <i><span style="text-decoration:underline;">just because</span></i> you have access to it. Paradoxically that is NOT a smart thing to do.</p><p class="paragraph" style="text-align:left;">But HOW? </p><h2 class="heading" style="text-align:left;" id="put-the-tools-in-the-same-room">Put the tools in the same room</h2><p class="paragraph" style="text-align:left;"></p><div class="image"><img alt="Foundation: one repo" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d8fa794b-71b5-428c-a3f1-8c4181f3a45e/slide-05-one-repo.png?t=1783621199"/><div class="image__source"><span class="image__source_text"><p>One shared room. One handoff.</p></span></div></div><p class="paragraph" style="text-align:left;">The practical bit is much less mystical than people make it sound. </p><p class="paragraph" style="text-align:left;">You need a shared folder.</p><p class="paragraph" style="text-align:left;">For most of this work, that means a GitHub repo. If that phrase makes you feel itchy, think of it as Google Drive with version history and fewer vibes. I’ve talked about <a class="link" href="https://aiwithkyle.com/ai-news/180-github-101?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=don-t-marry-the-model" target="_blank" rel="noopener noreferrer nofollow">how to set up Github in this Github 101 guide</a>.</p><p class="paragraph" style="text-align:left;">Could you use Google Drive? Technically? But I wouldn&#39;t.</p><p class="paragraph" style="text-align:left;">GitHub is better when multiple agents are making changes because it can track what changed and what needs merging. Google Drive is lovely for documents. It is a bit crap for agent handoffs. Because they’ll all be in there working on the same files and making a mess of it all. </p><p class="paragraph" style="text-align:left;">This is the same reason I keep talking about building an <a class="link" href="https://aiwithkyle.com/ai-news/build-one-shared-ai-vault?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=don-t-marry-the-model" target="_blank" rel="noopener noreferrer nofollow">AI brain / shared vault</a>. A shared workspace for ALL your AI tools. (Also the topic of our AI Hour live call last night).</p><p class="paragraph" style="text-align:left;">Inside the repo you add boring text files:</p><ul><li><p class="paragraph" style="text-align:left;"><code>README.md</code></p></li><li><p class="paragraph" style="text-align:left;"><code>AGENTS.md</code></p></li><li><p class="paragraph" style="text-align:left;"><code>CLAUDE.md</code></p></li><li><p class="paragraph" style="text-align:left;"><code>tasks.md</code></p></li><li><p class="paragraph" style="text-align:left;"><code>decisions.md</code></p></li></ul><p class="paragraph" style="text-align:left;">(By the way MD just means Markdown. Basically a stripped-down text file. Less scary than it sounds) </p><p class="paragraph" style="text-align:left;">Tell Claude/Fable to write the plan and put the decisions in the repo. Then tell Codex to pick up the work and implement it. That’s (at base) it. </p><h2 class="heading" style="text-align:left;" id="start-manual-add-machinery-later">Start manual. Add machinery later.</h2><p class="paragraph" style="text-align:left;">As well as having a shared workspace we can also directly connect out tools to one another in various ways. </p><p class="paragraph" style="text-align:left;">There are three ways to connect the tools.</p><p class="paragraph" style="text-align:left;"><b>First: manual handoff.</b> Open both tools from the same repo. Claude writes the plan. Codex does the work. Claude reviews. You are the messenger. That works. It is also a bit of a faff because <i>you</i> become the bottleneck.</p><p class="paragraph" style="text-align:left;"><b>Second: plugins.</b> Claude Code can call Codex. Codex can call Claude Code. Both have plugins that you can install in the other. Or if you are using Cursor then install both! </p><p class="paragraph" style="text-align:left;">Third:<b> MCP, CLI or custom wiring.</b> This is where the tools talk directly under the hood. This is the most “advanced” but bizarrely probably also the easiest. You can literally ask the AI:</p><div class="codeblock"><pre><code>I want you to connect this project to Codex so you can 
hand implementation work off to it. 
Work out the best CLI or MCP route and set it up.</code></pre></div><p class="paragraph" style="text-align:left;">That is it. Seriously. The AI is smart. It’ll work the rest out. </p><h2 class="heading" style="text-align:left;" id="my-beginner-loop">My beginner loop</h2><div class="image"><img alt="My recommended beginner workflow" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/20786a42-462e-4994-b7b1-2a7e1f8ef8e3/slide-08-beginner-workflow.png?t=1783621199"/><div class="image__source"><span class="image__source_text"><p>Start from the best planner. Push the work down. Review before commit.</p></span></div></div><p class="paragraph" style="text-align:left;">This is the version I would use if you are starting from scratch today. This WILL change. Might not even be the same a month from now. But the principal holds. </p><p class="paragraph" style="text-align:left;">Start in the best planner. For me, right now, that is Fable. I give it the messy brief and tell it to interview me. Then I switch to voice and yap into the microphone.</p><p class="paragraph" style="text-align:left;">Fable turns that into a project plan, file structure and shared instruction files.</p><p class="paragraph" style="text-align:left;">Then Codex implements. It has a LOT more usage so can work on tasks doggedly for long horizons without bankrupting you. </p><p class="paragraph" style="text-align:left;">Then the work goes back up to Fable or another stronger reviewer. It checks the result and sends fixes back down.</p><p class="paragraph" style="text-align:left;">That loop itself matters more than whether the model name on the box is Fable, Sol, Terra, Luna, Grok or whatever else gets announced while I am trying to have breakfast….</p><h2 class="heading" style="text-align:left;" id="do-not-marry-models">Do not marry models</h2><div class="image"><img alt="Do not marry models. Build loops." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/05710c70-e039-4bde-9bfa-b52d9a03eb6a/slide-10-build-loops.png?t=1783621199"/><div class="image__source"><span class="image__source_text"><p>The names change. The loop matters.</p></span></div></div><p class="paragraph" style="text-align:left;">The warning is simple:</p><p class="paragraph" style="text-align:left;">Do not just use the smartest model because you can.</p><p class="paragraph" style="text-align:left;">That’s dumb(!).</p><p class="paragraph" style="text-align:left;">Instead use the best model for the job at hand. </p><p class="paragraph" style="text-align:left;">One plans. One writes. One reviews. </p><p class="paragraph" style="text-align:left;">Remember that the. model names will change. The limits will change. The prices will change. </p><p class="paragraph" style="text-align:left;">Your job is not to predict which logo wins next week. Leave that to the Twitter bros. </p><p class="paragraph" style="text-align:left;">Your job is to build the loop so each release does not matter.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=17e4b4d9-fed7-40cd-86a2-e6c2d6cdae7d&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Fable Is Gone (Wait, No!)</title>
  <description>Boomeranging</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/c9ef0052-9739-48b6-a49f-cabd34fa2847/thumbnail_1_%2B_Newsletter.png" length="869137" type="image/png"/>
  <link>https://newsletter.aiwithkyle.com/p/fable-is-going-away</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/fable-is-going-away</guid>
  <pubDate>Wed, 08 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-08T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Daily Update]]></category>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Fable]]></category>
    <category><![CDATA[Local Ai]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="Livestream clip showing the Fable Is Going Behind The Meter slide with Kyle speaking" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d718277d-3ab7-4d92-90cf-9545a230167c/fable-behind-meter-newsletter.gif?t=1783437464"/><div class="image__source"><span class="image__source_text"><p><i><a class="link" href="https://youtu.be/WRPxqOXpj-Q?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/WRPxqOXpj-Q</a></i><i> - watch now or save for later</i></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/WRPxqOXpj-Q?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no"><span class="button__text" style=""> Watch Now </span></a></div><p class="paragraph" style="text-align:left;">Fable is gone.</p><p class="paragraph" style="text-align:left;">Well shit. </p><p class="paragraph" style="text-align:left;">My <a class="link" href="https://aiwithkyle.com/ai-with-kyle-group-chat?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no" target="_blank" rel="noopener noreferrer nofollow">Whatsapp Group </a>has been abuzz with working out what the hell to use our final moments with Fable on. And then mourning. </p><p class="paragraph" style="text-align:left;">Or so I wrote yesterday before we got the news of a reprieve. A stay of execution:</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/claudeai/status/2074548242386178258?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">This obviously messed up the newsletter a bit. BUT everything below is still accurate. It’s just….delayed by 5 days ha! </p><p class="paragraph" style="text-align:left;">As of Tuesday <b>July 12, 2026</b>, Anthropic is moving Claude Fable 5 out of normal paid plans and into <a class="link" href="https://www.anthropic.com/news/redeploying-fable-5?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no" target="_blank" rel="noopener noreferrer nofollow">usage credits</a>. The <a class="link" href="https://www.anthropic.com/claude/fable?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no" target="_blank" rel="noopener noreferrer nofollow">API page</a> has it at $10 per million input tokens and $50 per million output tokens.</p><p class="paragraph" style="text-align:left;">Now, you might be wondering if that’s a lot. Or even <a class="link" href="https://aiwithkyle.com/ai-news/193-ai-101-tokens?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no" target="_blank" rel="noopener noreferrer nofollow">what is a token</a>? How many input and output tokens could a mere mortal use in a day? Let’s run the numbers quickly.</p><p class="paragraph" style="text-align:left;">Here for example is me using 900M tokens on Codex in a day (and what a day ‘twas!):</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/65144eb2-f774-48b8-b6b6-7515a513c0c5/Screenshot_2026-07-07_at_6.32.50_pm.png?t=1783438378"/></div><p class="paragraph" style="text-align:left;">That ONE day, if I had been used that many tokens on Fable would cost me between <b>$9,000 and $45,000</b></p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/72a1eadb-7ba9-4acc-b99c-e11813849c1b/Screenshot_2026-07-07_at_6.35.16_pm.png?t=1783438527"/></div><p class="paragraph" style="text-align:left;">Coding tasks tend towards the input so probably around $12,000-15,000 for the day. </p><p class="paragraph" style="text-align:left;">I think my business partner Harms would have been a <i>bit</i> miffed. </p><p class="paragraph" style="text-align:left;">Using Fable from now on will be the reserve of companies or the very wealthy. </p><p class="paragraph" style="text-align:left;">And I get why people are annoyed. For the last week, a lot of people <i>finally </i>saw what these models can do when they are given <i>actual</i> work.</p><p class="paragraph" style="text-align:left;">I used it on finance stuff I had been dodging. I used it on a book project I did not want to outsource to slop. I used it on business systems that had been stuck for months. And Fable just tore through everything. </p><p class="paragraph" style="text-align:left;">And now the clever thing is behind a meter. A very expensive meter. Bugger.</p><h2 class="heading" style="text-align:left;" id="zoom-out">Zoom out </h2><p class="paragraph" style="text-align:left;">The pricing story will get all the shouting and screaming Fair enough. People got a taste of Fable, then Anthropic put it behind usage credits. Cue moaning, mourning, and screenshots of token maths.</p><p class="paragraph" style="text-align:left;">But we need to step back and remember the last few years. </p><p class="paragraph" style="text-align:left;">A frontier model arrives. Everyone falls in love with it for about five minutes.</p><div class="image"><img alt="Slide: Another one is already coming. Fable is not the end of the line." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/37c1cdc3-f278-4c10-b231-21afe8503f16/fable-level-normal.png?t=1783437471"/></div><p class="paragraph" style="text-align:left;">Then the next lab catches up, the price drops, the open models trail behind, and what was once science-fiction becomes boring infrastructure.</p><p class="paragraph" style="text-align:left;">We get VERY used to our models very quickly. </p><p class="paragraph" style="text-align:left;">We will have other models like Fable shortly. Probably this week with ChatGPT Sol. Or if not definitely in the coming months.</p><p class="paragraph" style="text-align:left;">The direction is obvious. Fable <i>feels </i>rare and special because we are standing too close to it. In a year, probably less, this level of intelligence will feel normal. </p><p class="paragraph" style="text-align:left;">Think about this for a moment:</p><p class="paragraph" style="text-align:left;"><b>One day, in the next 2-3 years, we will have a Fable level model running locally on our iPhones. </b></p><p class="paragraph" style="text-align:left;">Read that again and really let it sink in. </p><p class="paragraph" style="text-align:left;">Our models are getting better, cheaper and smaller <i>all of the time</i>. ChatGPT 3.5 is just 3.5 years old. And it was TERRIBLE (seriously, go try it on the Playground for a laugh).</p><p class="paragraph" style="text-align:left;">It’s taken 3.5 years to come from ChatGPT 3.5 to Fable. What will we have in 3.5 years from now? THAT is what we need to keep our eyes on here.</p><h2 class="heading" style="text-align:left;" id="what-happens-when-its-cheap">What happens when its cheap?</h2><p class="paragraph" style="text-align:left;">When a model is expensive, you ration it.</p><p class="paragraph" style="text-align:left;">You use it for the big judgement calls. The ugly personal problem. The business mess where being 20% smarter changes the answer. That was the point of last week&#39;s <a class="link" href="https://aiwithkyle.com/ai-news/fable-one-question?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no" target="_blank" rel="noopener noreferrer nofollow">Fable question issue</a>: <b>do not waste the expensive brain on tiny tasks.</b></p><p class="paragraph" style="text-align:left;">But when the price falls, behaviour changes. What happens then? </p><p class="paragraph" style="text-align:left;">We will one day have this Fable near-human level intelligence for basically free. On our phones. On our local devices.</p><p class="paragraph" style="text-align:left;">That brings us to the darker side. </p><p class="paragraph" style="text-align:left;"><b>If you have used Fable over the last week. </b></p><p class="paragraph" style="text-align:left;"><b>And if you can imagine Fable being made free and unlimited.</b></p><p class="paragraph" style="text-align:left;"><b>Now tell me that jobs are safe.</b></p><p class="paragraph" style="text-align:left;">It becomes a <span style="text-decoration:underline;">very</span> difficult argument to make. </p><p class="paragraph" style="text-align:left;">And remember businesses do not even need a model to be free. They need it to be cheaper than the human task it replaces. And when Fable-ish capability gets fast, cheap, and boring like a utility, a lot of white-collar work starts looking wobbly.</p><p class="paragraph" style="text-align:left;">The AI doesn’t have to take entire jobs. No need. Remember that jobs are bundles of tasks. If AI takes 30%, 40%, 50% of those tasks, the job title might survive while the income, headcount, and employee bargaining power get quietly chewed up underneath it.</p><p class="paragraph" style="text-align:left;">This is where I get annoyed at the &quot;it&#39;ll all be fine&quot; crowd. They accuse people like me of being a Doomer. No - I’m a realist. And after seeing what Fable can do I’m doubling down on that. Our comfy 9-5s are not safe. </p><h2 class="heading" style="text-align:left;" id="build-something-you-control">Build something you control</h2><div class="image"><img alt="Slide: Build Something Of Your Own. Positioning beats panic." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0938c70d-1b23-40b7-b24b-afa4811e6603/build-something-of-your-own.png?t=1783437478"/></div><p class="paragraph" style="text-align:left;">You cannot rely on your company to handle this well. If you are employed you need options.</p><p class="paragraph" style="text-align:left;">Maybe your employer has a clear AI strategy. Maybe they have tooling, training, governance, a proper plan, the lot.</p><p class="paragraph" style="text-align:left;"><i>Maybe.</i></p><p class="paragraph" style="text-align:left;">Or maybe they faff around for two years, buy a licence, form a committee, and then panic-cut costs when a competitor uses AI faster than they do.</p><p class="paragraph" style="text-align:left;">If you work inside that machine, you do not control the machine. Sorry. </p><p class="paragraph" style="text-align:left;">Or shit maybe they do know what they are doing and you still aren’t safe. Because the first moment they can drop you and replace you with AI they will. Corporations want our loyalty but the honour does not extend in both directions. They will train the AI on your daily work (screen recordings, emails, meeting notes, any and everything they can get their hands on) and then switch you out for an agent that does the task for 100x lower cost. </p><p class="paragraph" style="text-align:left;">And maybe even better than you…</p><p class="paragraph" style="text-align:left;">So what do we do?</p><p class="paragraph" style="text-align:left;"><b>So build something you </b><i><b>do</b></i><b> control. </b>A side income stream. A consulting offer. A tiny first product. A niche audience. A service. <i>Something</i> you build yourself and control. </p><p class="paragraph" style="text-align:left;">I’ve actually built a whole <a class="link" href="https://aiwithkyle.com/ai-business-course?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no" target="_blank" rel="noopener noreferrer nofollow">10 week course on building a business with AI</a>. Totally free, no need to even register. </p><p class="paragraph" style="text-align:left;">Is it easy? Nah, not easy. But increasingly we don’t have much of a choice. </p><p class="paragraph" style="text-align:left;">We either wait for the sword for fall. And then complain that we didn’t see it coming (we did).</p><p class="paragraph" style="text-align:left;">Or we take responsibility and get on with building <i>something</i> that we control. </p><p class="paragraph" style="text-align:left;">If you are looking for a starting point I say lean into the chaos! Let the confusion and fear of AI be what drives your business. Instead of being driven by it. This is why I keep banging on about workshops and advisory work. EVERY business on earth right now knows they need help with AI.</p><p class="paragraph" style="text-align:left;">And they have no idea what to do about it. </p><p class="paragraph" style="text-align:left;">You do. You are in a very very small % of people who are paying attention to AI. I know this because you are here reading a debrief about Fable. That’s pretty niche info you know.</p><p class="paragraph" style="text-align:left;"> If you can explain this stuff clearly and turn chaos into action, that is valuable to businesses. Extremely valuable. Several thousands pounds/dollars an hour valuable. </p><p class="paragraph" style="text-align:left;">I’ve got 250+ students are building around that model now. Workshops, consulting, implementation, advisory. If you want to know more I’m doing a live webinar <span style="text-decoration:underline;"><b>tonight</b></span> at 6PM London time: <a class="link" href="https://aiwithkyle.com/webinar?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-is-gone-wait-no" target="_blank" rel="noopener noreferrer nofollow">aiwithkyle.com/webinar</a>.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=2d544701-2614-4a3d-bc6b-df9b057efce6&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>1 Question Is All You Get</title>
  <description>I&#39;ll stop soon! </description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d5343515-ec95-44f6-98cd-9a4937bbdbd9/thumbnail_3.png" length="706936" type="image/png"/>
  <link>https://newsletter.aiwithkyle.com/p/1-question-is-all-you-get</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/1-question-is-all-you-get</guid>
  <pubDate>Mon, 06 Jul 2026 07:01:00 +0000</pubDate>
  <atom:published>2026-07-06T07:01:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="Livestream clip showing the 1 Question Is All You Get Fable slide with Kyle speaking" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5167bcb0-2c5b-4c89-96e5-23da99a03bf0/fable-one-question-newsletter-clip.gif?t=1783080377"/><div class="image__source"><span class="image__source_text"><p><i>Watch now or save to watch later: </i><i><a class="link" href="https://youtu.be/P4-N2VQIWiI?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=1-question-is-all-you-get" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/P4-N2VQIWiI</a></i></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/P4-N2VQIWiI?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=1-question-is-all-you-get"><span class="button__text" style=""> Watch Now </span></a></div><p class="paragraph" style="text-align:left;">Fable is back. But not for long.</p><p class="paragraph" style="text-align:left;">Anthropic has <a class="link" href="https://www.anthropic.com/claude/fable?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=1-question-is-all-you-get" target="_blank" rel="noopener noreferrer nofollow">restored access to Claude Fable 5</a>, and for a few days the expensive thing is sitting inside normal paid plans for up to 50% of weekly usage limits. That window closes on <b>July 7, 2026.</b> </p><p class="paragraph" style="text-align:left;">After that, <a class="link" href="https://www.anthropic.com/news/redeploying-fable-5?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=1-question-is-all-you-get" target="_blank" rel="noopener noreferrer nofollow">it moves to usage credits</a> and the API price is a cool <a class="link" href="https://www.anthropic.com/claude/fable?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=1-question-is-all-you-get" target="_blank" rel="noopener noreferrer nofollow">$10 per million input tokens and $50 per million output tokens</a>. If you don’t use the API and have no idea what it means I’ll translate: effing expensive! </p><p class="paragraph" style="text-align:left;">So yes, this is time-sensitive. It closes </p><p class="paragraph" style="text-align:left;">This is basically Oracle of Delphi energy. You get to walk up to the very clever machine and ask it the big question…</p><p class="paragraph" style="text-align:left;">And most people are still going to ask it to tidy an email. Please don&#39;t! Hell if you are going to do this sell me your usage instead ha! </p><h2 class="heading" style="text-align:left;" id="do-the-20-test">Do the 20% test</h2><p class="paragraph" style="text-align:left;">OK so what should you ACTUALLY use your (limited) Fable on?</p><div class="image"><img alt="Slide explaining the 20 percent test for deciding what to ask Fable." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/bef43715-7155-4c57-8841-f1b72bc6fb18/fable-1-question-4-1400.png?t=1783080390"/></div><p class="paragraph" style="text-align:left;">This is the filter I would use: <b>Ask Fable things where being 20% smarter changes the answer entirely.</b></p><p class="paragraph" style="text-align:left;">Is it actually 20% smarter than all the other models? Maybe. Doesn’t matter. Use this as a mental framework. Basically think of it as the smartest person you’ve ever talked to…what would you ask them? </p><p class="paragraph" style="text-align:left;">Judgment calls. Tricky trade-offs. Untangling messy situations. The question <i>behind</i> the question. The hard stuff! </p><p class="paragraph" style="text-align:left;">Do not use it for boilerplate code. Do not use it for formatting. Do not use it to summarise something Opus or GPT-5.5 can already summarise perfectly well. That is using a freight train to pick up your shopping.</p><p class="paragraph" style="text-align:left;">Total waste. </p><p class="paragraph" style="text-align:left;">If you have a messy business decision, give it that. If you have a huge project repo and you cannot see the shape of the thing anymore, give it that. If you have a personal or professional decision you have been dodging because the answer is probably inconvenient, give it that.</p><h2 class="heading" style="text-align:left;" id="build-the-question-first">Build the question first</h2><p class="paragraph" style="text-align:left;">I’m seeing a lot of people absolutely paralysed about <i>what</i> to ask Fable. With such limited usage and such a tight time constraint I get it. Easy to get blocked. </p><p class="paragraph" style="text-align:left;">Here’s a concrete tip…use a “dumber” model to help you construct your question.</p><div class="image"><img alt="Slide explaining that the real work is building the question before asking Fable." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3cfe1cfe-0783-4157-aea9-bde6e3dab789/fable-1-question-5-1400.png?t=1783080396"/></div><p class="paragraph" style="text-align:left;">Do not sit there staring at the Fable box trying to write the perfect prompt.</p><p class="paragraph" style="text-align:left;">Use a cheaper, slightly dumber model as the workshop. (God, it feels weird calling Opus 4.8 or GPT-5.5 dumb… but you know what I mean.)</p><p class="paragraph" style="text-align:left;">Brain-dump your situation into Opus or GPT-5.5 first. Tell it you have one chance to ask a much smarter model a question. Get it to interview you. Get it to pull the uncomfortable context out of you. Get it to draft the full one-shot prompt.</p><p class="paragraph" style="text-align:left;">Then pressure-test it:</p><ul><li><p class="paragraph" style="text-align:left;">what is the prompt missing?</p></li><li><p class="paragraph" style="text-align:left;">what would Fable need to know if I get no follow-ups?</p></li><li><p class="paragraph" style="text-align:left;">what assumptions am I hiding?</p></li></ul><p class="paragraph" style="text-align:left;">And because your usage may be capped it’s important to ask in one go. You don’t want to end up answering lots of back and forth with Fable as each turn will eliminate tokens. So get it right first time as much as you can! </p><div class="image"><img alt="Slide explaining a safety net for using Fable: verify, stress test, and compare answers." class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0934adb1-717e-44bf-8078-cfc865d18707/fable-1-question-6-1400.png?t=1783080402"/></div><p class="paragraph" style="text-align:left;">Assume you get one prompt per day. Especially on the $20/month plan! </p><p class="paragraph" style="text-align:left;">Maybe a handful more if Claude is feeling generous. Maybe five if you are on the $200 plan and burn through your allowance like an idiot. On $200/month I’m getting 4-5 every 5 hours. And maxing them out every time! </p><p class="paragraph" style="text-align:left;">So the prompt needs three bits:</p><ul><li><p class="paragraph" style="text-align:left;">context: who you are, the situation, what you have tried</p></li><li><p class="paragraph" style="text-align:left;">question: one specific, high-stakes thing</p></li><li><p class="paragraph" style="text-align:left;">safety net: assumptions, counter-arguments, likely missing context, and what it should do if it <i>would</i> normally ask a follow-up</p></li></ul><p class="paragraph" style="text-align:left;">This is different from normal chat prompting. Normal prompting is often a conversation - we have “turns” - going back and forth. With Fable it’s much closer to briefing an expert before they walk into a boardroom and you are not allowed to speak again.</p><h2 class="heading" style="text-align:left;" id="ask-the-nasty-thing">Ask the nasty thing</h2><p class="paragraph" style="text-align:left;">The best Fable question is probably hiding behind something uncomfortable.</p><p class="paragraph" style="text-align:left;"><b><i>What decision have I been avoiding?</i></b></p><p class="paragraph" style="text-align:left;"><b><i>Where am I working hard on the wrong thing?</i></b></p><p class="paragraph" style="text-align:left;"><b><i>What would embarrass me if someone brilliant looked at my business for an hour?</i></b></p><p class="paragraph" style="text-align:left;">I gave Fable my AI brain and asked versions of that. It has my project logs, my decisions, the stuff I have been working on, the stuff I said I would do and then quietly wandered away from. If you want to build that sort of shared context, the <a class="link" href="https://aiwithkyle.com/ai-news/build-one-shared-ai-vault?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=1-question-is-all-you-get" target="_blank" rel="noopener noreferrer nofollow">AI brain guide is here</a>.</p><p class="paragraph" style="text-align:left;">And it called me out. Brutally. It was quite upsetting! </p><p class="paragraph" style="text-align:left;">It basically said I keep killing winners. I start a project, prove people want it, get bored when it starts working, then wander off to the next shiny thing…</p><p class="paragraph" style="text-align:left;">Rude. And 1000% accurate. Which is annoying.</p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b>That</b></span><b> </b>is what this model is for. Not &quot;rewrite this LinkedIn post in a more professional tone.&quot; Nah, use it for BIG gnarly questions and problems. </p><p class="paragraph" style="text-align:left;">The way I like to think about it is that answers are getting cheap. Good questions are scarce. And that is on you.</p><p class="paragraph" style="text-align:left;">One final point for those who may be feeling overwhelmed by having to get full usage out of Fable right now. i) It will eventually come to our normal subscriptions which is cool but ii) more importantly: ALL models will eventually be at this level. It may be 6 months from now. It may be 12 months from now. But it will happen. </p><p class="paragraph" style="text-align:left;">So if you feel you’ve missed the boat this time don’t worry. This level of capability will become the norm. Which is both terribly exciting and terribly scary! </p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=1ba6341d-1836-4aee-a5d3-23f4f2ea13ac&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Saturday Sessions: Fable Fable Fable</title>
  <description>Fable</description>
  <link>https://newsletter.aiwithkyle.com/p/saturday-sessions-fable-fable-fable</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/saturday-sessions-fable-fable-fable</guid>
  <pubDate>Sat, 04 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-04T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'600' !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">So Fable is back eh? </p><p class="paragraph" style="text-align:left;">If you’ve been paying even the slightest attention to AI news you’ll know it’s all anyone is talking about. </p><p class="paragraph" style="text-align:left;">Is Fable good? Did they nerf it? Can you code with it? Will they take it away from us? </p><p class="paragraph" style="text-align:left;">And so on and so on. </p><p class="paragraph" style="text-align:left;">Let me quickly cut through a lot of noise:</p><ul><li><p class="paragraph" style="text-align:left;">Fable is the best AI on the market right now and is well worth trying.</p></li><li><p class="paragraph" style="text-align:left;">You have until the 7th July to play around with it</p></li><li><p class="paragraph" style="text-align:left;">At that time they will make it pay-per-use. And it is NOT cheap. </p></li></ul><p class="paragraph" style="text-align:left;">I’m going to do something I rarely do… make a strong suggestion to spend some money. Shocking! </p><p class="paragraph" style="text-align:left;">Normally I tell people to use the AI that works for them and not to switch mindlessly from one to another just because it’s the new hotness.</p><p class="paragraph" style="text-align:left;">Buuuuuuutttt…..yeah. This is different.</p><p class="paragraph" style="text-align:left;">Using Fable <i>genuinely</i> feels like using an AI from the future. It’s extremely good. And can cut through hard tasks and problems like a knife through warm butter (and whose butter isn’t warm during these scorching months).</p><p class="paragraph" style="text-align:left;">For $20 you can get a Claude subscription and get access to Fable. </p><p class="paragraph" style="text-align:left;">Fair warning: you’ll probably be able to ask 1-2 questions <i>per day</i> and that’s your lot. </p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/myYyMu4hHew" width="100%"></iframe><p class="paragraph" style="text-align:left;">If you start today that gives you 3-4 days of usage, 4-8 big questions.</p><p class="paragraph" style="text-align:left;">On my livestream yesterday I talked about this being like going to the Oracle of Delphi or (if you are more sci-fi orientated) Deep Thought from the Hitchhikers’ Guide. </p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ee589e3f-4de7-40a3-b3bf-f677bfbf20c4/0838b2f8-1718-4cac-a9c8-9879bb3fc824.png?t=1783088893"/></div><p class="paragraph" style="text-align:left;">You get to ask ONE big question.</p><p class="paragraph" style="text-align:left;">Something that could untangle blocks in your life, help build a new income, solve a problem you’ve been dealing with for years. </p><p class="paragraph" style="text-align:left;"><b>So </b><span style="text-decoration:underline;"><b>what</b></span><b> is that question? </b></p><p class="paragraph" style="text-align:left;">THIS is the new skill. The imagination to ask<i> big gnarly questions</i>. To imagine wonderful life improving projects. </p><p class="paragraph" style="text-align:left;">The AI can increasingly do anything we throw at it. And I say this with no hyperbole. </p><p class="paragraph" style="text-align:left;">The question therefore becomes <i>what</i> should you be asking it to do? That’s still up to you. The AI isn’t going to help you there. </p><p class="paragraph" style="text-align:left;">If you get stuck (I get it!) here’s a practical step: talk to “normal” Claude or ChatGPT and ask something like<i> “I’m about to sit down with someone who has achieved everything I want to achieve. Who has all the answers. I get to ask ONE question. Help me work out that question. Interview me until we decide upon the question”</i></p><p class="paragraph" style="text-align:left;">Our “normal” AIs can help us work out what we should ask. </p><p class="paragraph" style="text-align:left;">Then go to Fable and ask. </p><p class="paragraph" style="text-align:left;">And please! Report back! </p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;">We’ve been chatting about this for the last few days in the <a class="link" href="https://aiwithkyle.com/ai-with-kyle-group-chat?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=saturday-sessions-fable-fable-fable" target="_blank" rel="noopener noreferrer nofollow">Whatsapp group</a> and sharing what has been most insightful! </p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/286af0aa-943e-4718-a3a6-4009bd005fd6/Screenshot_2026-07-03_at_15_14_20.png?t=1783089319"/></div><p class="paragraph" style="text-align:left;">Oh and for those who celebrate - have an amazing July 4th today! </p></div><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p><p class="paragraph" style="text-align:left;"></p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=00e7f4ec-2757-484d-b510-7cdb5976db71&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Open Source AI Is A Price War</title>
  <description>WTF open source AI actually means</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2014995b-8772-4b9f-be0b-517b7a5d28cd/open-source-ai-newsletter-clip.gif" length="455366" type="image/gif"/>
  <link>https://newsletter.aiwithkyle.com/p/open-source-ai-price-war</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/open-source-ai-price-war</guid>
  <pubDate>Fri, 03 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-03T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="Livestream GIF showing Kyle presenting the open weights versus open source AI slide" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5e85baac-7d97-4d30-936d-308c2a15c131/open-source-ai-newsletter-clip.gif?t=1782902337"/><div class="image__source"><span class="image__source_text"><p><i><a class="link" href="https://youtu.be/L9MV98e93ow?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/L9MV98e93ow</a></i><i> - watch now or save for later</i></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/L9MV98e93ow?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war"><span class="button__text" style=""> Watch Now </span></a></div><p class="paragraph" style="text-align:left;">All this fuss about Open source AI is a price war.</p><p class="paragraph" style="text-align:left;">Sorry open-source lads and lassies. I know there are philosophical bits. Safety bits. Licence bits. Community bits. All true, all amazing stuff. </p><p class="paragraph" style="text-align:left;">But the reason this suddenly matters to normal people is money.</p><p class="paragraph" style="text-align:left;">Anthropic and OpenAI <i>need </i>the world to believe that the best AI has to come through expensive, “safe”, American API pipes. Controlled by, well, them.</p><p class="paragraph" style="text-align:left;">Meanwhile, Chinese labs are shipping opensource models that are more than good enough for a lot of real work and a hell of a lot cheaper. </p><p class="paragraph" style="text-align:left;">The very existence of these opensource alternative undercuts the whole US/closed source mission. And threatens to bring the entire economy of the West crashing down.</p><p class="paragraph" style="text-align:left;">So…this is not just about software development principals! </p><h2 class="heading" style="text-align:left;" id="wtf-is-opensource">WTF is opensource?</h2><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/fd2a7e87-3f18-4fb5-b1ea-2d388f4e774c/Screenshot_2026-07-01_at_1.36.36_pm.png?t=1782902202"/><div class="image__source"><span class="image__source_text"><p>“Open” is doing a lot of work here!</p></span></div></div><p class="paragraph" style="text-align:left;">What actually is opensource? The annoying thing is that everyone uses &quot;open source AI&quot; to mean different things. </p><p class="paragraph" style="text-align:left;">Sometimes they mean free. Sometimes downloadable. Sometimes local. Sometimes the training data is visible. Sometimes the licence lets you use it commercially. Sometimes they mean &quot;I saw DeepSeek on Twitter and now I am pretending to understand geopolitics.&quot;</p><p class="paragraph" style="text-align:left;">This is further confused by the fact that most opensource LLMs aren’t actually opensource. They are <i>open-weights</i>. Grr…</p><p class="paragraph" style="text-align:left;">The useful split is this:</p><ul><li><p class="paragraph" style="text-align:left;">open source means you can see the code, the weights, the training pipeline and sometimes even the underlying data.</p></li><li><p class="paragraph" style="text-align:left;">open weights means you can download the trained model files</p></li></ul><p class="paragraph" style="text-align:left;">Most of the models people call open source are really <i>open weights</i>. You can get the trained numbers. You can run them somewhere. You might be able to fine-tune them. But you usually cannot rebuild the full thing from scratch because you do not have the training data, process, and boring details behind the model.</p><p class="paragraph" style="text-align:left;">Here’s a useful analogy from <a class="link" href="https://dentro.de/ai/blog/2025/07/15/understanding-open-source-in-ai-models/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war#car-analogy-for-clarity" target="_blank" rel="noopener noreferrer nofollow">Dentro</a>.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/db98e63c-c1bc-4557-aa30-1fe57d8458cd/open-source.png?t=1782902304"/></div><p class="paragraph" style="text-align:left;"><b>Closed source models</b> like ChatGPT and Claude are like a taxi. You can tell the driver where to go, but you do not own the vehicle. You definitely can’t take it home and start modifying it - the cabbie would get very annoyed I imagine. And, like black cabs, they are expensive! </p><p class="paragraph" style="text-align:left;"><b>Open weights</b> are more like owning a car. You have the thing. You can drive it where you want. You can modify bits of it. You are responsible for the faff that comes with ownership. Most of the time this works out cheaper than taking a taxi but on the whole it’s a little more of a faff.</p><p class="paragraph" style="text-align:left;"><b>True open source</b> is the kit car version. You get the parts, the plans and enough information to rebuild the thing yourself. Very few models are truly open-source, indeed <i>none</i> of the well-known production grade ones. </p><p class="paragraph" style="text-align:left;">Also whether a model is free to use and whether it’s local are entirely separate issues from the model being open weights.</p><p class="paragraph" style="text-align:left;">GLM 5.2 for example is technically open-weights. You can go right now and download the whole thing, all 1.5TB of it. And then run it on your local computer for “free” (minus the cost of electricity of course).</p><p class="paragraph" style="text-align:left;">BUT…your local computer is going to have to be a data centre… you aren’t running this on any commercially available computer. </p><p class="paragraph" style="text-align:left;">Most local models are opensource. But not all opensource models can be run locally! So don’t conflate the two.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/7593eaaf-6f17-4ccc-a3e9-8977cf8ae35a/Screenshot_2026-07-01_at_1.47.48_pm.png?t=1782902874"/></div><h2 class="heading" style="text-align:left;" id="where-does-all-this-live">Where does all this live? </h2><p class="paragraph" style="text-align:left;">Let’s actually have a look at some open-source (open-weights, see even I’m using the term wrong..) models. How do you find them? Where can you download them? What can you do with them?</p><p class="paragraph" style="text-align:left;">The first site to look at is <a class="link" href="https://artificialanalysis.ai/models/open-source?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war#intelligence" target="_blank" rel="noopener noreferrer nofollow">Artificial Analysis</a>. Their open-source model page compares open-weight models by intelligence, openness, size and more. </p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/be7bd13c-4f93-454f-b92b-f229e7858f9e/Screenshot_2026-07-01_at_1.49.12_pm.png?t=1782902957"/></div><p class="paragraph" style="text-align:left;">These are pretty much ALL Chinese models and the first non Chinese model is Nemotron (Nvidia). </p><p class="paragraph" style="text-align:left;">Let’s have a look at the intelligence against cost-per-task. Basically your bang for your buck:</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/dc889bba-03fe-4c46-beb0-562a5ecdf76d/Screenshot_2026-07-01_at_2.00.58_pm.png?t=1782903663"/><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://artificialanalysis.ai/?cost=intelligence-vs-cost-per-task&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war" target="_blank" rel="noopener noreferrer nofollow">https://artificialanalysis.ai/?cost=intelligence-vs-cost-per-task</a></p></span></div></div><p class="paragraph" style="text-align:left;">The smartest models are (duh) Anthropic’s Fable, Opus, Sonnet and GPT5.5. They are also the most expensive. </p><p class="paragraph" style="text-align:left;">GLM5.2(max) is sitting at a very similar intelligence level (same as Sonnet 5 which came out this week) but at a fraction of the cost. </p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/DeryaTR_/status/2072051617298293199?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">The x-axes here is log scale by the way so the costs are getting <span style="text-decoration:underline;">much</span><i> </i>higher in that riht hand side - look how it jumps from $1 to $2 to $3 in smaller steps. </p><p class="paragraph" style="text-align:left;"> Opensource models are coming in either i) slightly worse but much cheaper or sometimes ii) equivalent quality but much cheaper. That is where the price war becomes obvious.</p><p class="paragraph" style="text-align:left;">You do not always need the smartest model. You need the <i>cheapest</i> model that clears the bar for the job. Customer service reply? Internal summary? Classification? Simple data clean-up? A frontier model is often total overkill. </p><p class="paragraph" style="text-align:left;">We have a pigeon that keeps flying into the garden and making a mess right now. The sensible method to get rid of it is to wave and clap until it buggers off. The Claude Fable method would be to hit it with an Intercontinental Ballistic Missile. </p><p class="paragraph" style="text-align:left;">Often we don’t NEED so much power. And when that power comes at 50-100x the cost it’s silly to use it. </p><p class="paragraph" style="text-align:left;">The second site to know about is <a class="link" href="https://huggingface.co/zai-org/GLM-5.2?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war" target="_blank" rel="noopener noreferrer nofollow">Hugging Face</a>. This is where the models live. Here for example is GLM-5.2: </p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f86ef974-ee5d-4d32-8499-0f667e26a045/Screenshot_2026-07-01_at_2.06.57_pm.png?t=1782904033"/></div><p class="paragraph" style="text-align:left;">The tags at the top tell us what it can do (text gen), its languages, its license etc. </p><p class="paragraph" style="text-align:left;">On the right we get the model size - here it is 753 billion parameters (big!).</p><p class="paragraph" style="text-align:left;">Hugging Face has almost 3 million other models for you to browse, learn about and download.</p><p class="paragraph" style="text-align:left;">But remember that most won’t run on your local computer. Opensource is NOT the same as local! This GLM-5.2 model is the opposite of a casual laptop download - it’ll need a small data centre to run it! </p><p class="paragraph" style="text-align:left;">Still incredibly important though. You can inspect it. You can see the files. You can see how this stuff is actually distributed. It’s all right there for you to explore.</p><h2 class="heading" style="text-align:left;" id="china-is-playing-a-blinder">China is playing a blinder</h2><p class="paragraph" style="text-align:left;">This is where the geopolitics come in.</p><p class="paragraph" style="text-align:left;">The US labs have a very expensive story to sell. More chips. More data centres. More money. More closed models. More funding rounds. More valuation nonsense.</p><p class="paragraph" style="text-align:left;">Scaling has worked so far. Just throw MORE at the problem and the AIs get better. It’s actually pretty amazing. And noone is <i>quite</i> sure why it works…</p><p class="paragraph" style="text-align:left;">OpenAI and Anthropic are trying to justify obscene numbers and the market is nodding along because AI is the big growth story. About 2% of US GDP is being invested into AI this year and somewhere north of 50-60% of all VC money is going into AI startups. </p><p class="paragraph" style="text-align:left;">It’s a massively intense concentration of investment. In one basket. A basket that is currently not laying eggs? I’ve lost the metaphor…you know what I mean. </p><p class="paragraph" style="text-align:left;">A lot is now riding on 2 companies. Anthropic and OpenAI. And their upcoming IPOs (going public on the stock market). Both of their IPOs are based on valuations just shy of $1T with revenues being a tiny multiple of that - much lower than the revenue multiples we expect. </p><p class="paragraph" style="text-align:left;">It’s one hell of a big bet. And making such a big bet requires investor confidence. If they get wind of this not working out, of their money not coming back to them, of potentially losing everything then they’ll panic and withdraw. </p><p class="paragraph" style="text-align:left;">The IPOs of Anthropic and OpenAI are big litmus tests here. Do the markets have the balls to support the dream? Big, scary risk. </p><p class="paragraph" style="text-align:left;">…Enter China.</p><p class="paragraph" style="text-align:left;">China is bringing opensource models to market. Not with one model. With a flood of them. DeepSeek, Qwen, Kimi, GLM, MiniMax, Hunyuan and the lot. Some are brilliant. Some are a bit crap, much like US models! But the direction is clear: cheaper, good-enough intelligence keeps arriving.</p><div class="image"><img alt="TrendForce overview of China&#39;s AI model key players" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3032ede7-a5f0-4ee4-97d3-02fa6035c88d/revised-Chinas-AI-model-key-players-overview-768x960.png?t=1782901063"/><div class="image__source"><span class="image__source_text"><p><i>China&#39;s open-weight model crowd is not exactly small.</i></p></span></div></div><p class="paragraph" style="text-align:left;">And if the best closed model is only needed for the hardest 5-10% of work, then a lot of the Western AI economy starts looking very different. </p><p class="paragraph" style="text-align:left;">Why would I pay Anthropic or OpenAI a sizeable chunk of my revenue every single month for tokens when I can deploy a Chinese model locally within my organisation, fine tune it and run it basically for the cost of electricity thereafter? Shit, I even get better security because my information is not being sent up and down to Anthropic / OpenAI. It sits locally in a secure server in my office.</p><p class="paragraph" style="text-align:left;">The very existence of these cheap, efficient, Chinese local models is a threat to the narrative the closed-source American labs are pushing. It undercuts their whole business model. </p><p class="paragraph" style="text-align:left;"><i>Even if</i> very few companies actually deploy local Chinese based models the fact that they even exist as an alternative sheds doubt on OpenAI and Anthropic. </p><p class="paragraph" style="text-align:left;">This is why the safety argument from closed labs always feels a bit fishy to me. To be clear: the safety argument is real. You cannot release powerful weights into the world and then recall them. Once a model is out, it is out - Pandora’s box is open. Governments should care about that. Cyber, fraud, bio, state actors…lots of scary stuff.</p><p class="paragraph" style="text-align:left;">BUT...</p><p class="paragraph" style="text-align:left;">The business incentive is also real.</p><p class="paragraph" style="text-align:left;">Closed labs fear open models because they are dangerous. Sure. They also fear them because they are becoming <i>good enough</i>. Good enough kills margins. Good enough gives enterprise buyers wiggle room. Good enough means developers can switch vendors. Good enough means &quot;pay us whatever we ask&quot; stops working.</p><p class="paragraph" style="text-align:left;">That is a wedge. If a huge chunk of US market confidence is tied up in AI capex, and Chinese open-weight models keep making intelligence cheaper, this starts to look much bigger than a nerd fight. </p><p class="paragraph" style="text-align:left;">If the Chinese opensource models can destablise the Anthropic and OpenAI IPOs and cause sufficient investor doubt then the whole AI industry could come tumbling down, bringing with it the American and subsequently Western economy. </p><p class="paragraph" style="text-align:left;">This is not a fight over closed-source vs. open-sourced software development. </p><p class="paragraph" style="text-align:left;">This is a fight for who owns the top spot in the global economy: the US or China.</p><h2 class="heading" style="text-align:left;" id="pick-a-stack-not-a-religion">Pick a stack, not a religion</h2><p class="paragraph" style="text-align:left;">So what do you actually DO?</p><p class="paragraph" style="text-align:left;">Do not become an open-source purist. Also do not become a closed-model fanboi. Both are annoying. And you know it. </p><p class="paragraph" style="text-align:left;">Instead play it smart. </p><p class="paragraph" style="text-align:left;">Use frontier models for genuinely hard work. Strategy. Proper reasoning. Important writing. Research. Stuff where being 5% better actually matters. Use the VERY best you can afford here and keep updating as new models and tools release. </p><p class="paragraph" style="text-align:left;">Use hosted models for cheap volume. Either ChatGPT/Claude subscriptions OR open-source subscriptions for more usage. Or both! This is where routine automations, internal workflows live. Anywhere &quot;pretty good and much cheaper&quot; beats &quot;best in class and wildly expensive&quot;. This is the workhorse.</p><p class="paragraph" style="text-align:left;">Then use local models for privacy, fallback, learning and simple repetitive work. <a class="link" href="https://lmstudio.ai/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war" target="_blank" rel="noopener noreferrer nofollow">LM Studio</a> is the friendly route. <a class="link" href="https://ollama.com/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war" target="_blank" rel="noopener noreferrer nofollow">Ollama</a> is more developer-ish but useful. <a class="link" href="https://deepmind.google/models/gemma/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=open-source-ai-is-a-price-war" target="_blank" rel="noopener noreferrer nofollow">Gemma</a> is worth watching. Chinese models are worth testing if your data/privacy situation allows it. These models will live on your device and deal with the day to day and anything they can’t deal with gets bounced up the chain. </p><p class="paragraph" style="text-align:left;">This setup gives you and your business the most flexibility to deal with whatever is coming down the tracks. </p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=fc3be9db-c44a-4322-a27a-be03ac5d75f0&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>ChatGPT&#39;s Best New Model Is Here. But You Can&#39;t Use It.</title>
  <description>AI access is becoming permissioned</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b66fb904-4c72-4325-a91e-a0b204c673b6/gpt56-permissioned-ai-infographic.png" length="189399" type="image/jpeg"/>
  <link>https://newsletter.aiwithkyle.com/p/gpt-56-permissioned-ai</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/gpt-56-permissioned-ai</guid>
  <pubDate>Wed, 01 Jul 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-07-01T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Ai News]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/47770f79-60f8-44ba-90ba-01f0be93e874/GIF.gif?t=1782880316"/><div class="image__source"><span class="image__source_text"><p><i><a class="link" href="https://youtu.be/vtS7i66fKFw?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-s-best-new-model-is-here-but-you-can-t-use-it" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/vtS7i66fKFw</a></i> <i>- watch now or save for later</i></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/vtS7i66fKFw?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-s-best-new-model-is-here-but-you-can-t-use-it"><span class="button__text" style=""> Watch Now </span></a></div><div class="image"><img alt="Livestream clip showing Claude was the warning shot slide" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e3688e64-5dc1-4ee2-89f1-430701a5abf7/gpt56-precedent-clip.gif?t=1782824064"/><div class="image__source"><span class="image__source_text"><p><i>From the livestream.</i></p></span></div></div><div class="image"><img alt="Infographic summarising GPT-5.6 gated access and the model-access hedge" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/78f5ccaa-3997-40d0-b28d-3a234917e84c/gpt56-permissioned-ai-infographic.png?t=1782824067"/><div class="image__source"><span class="image__source_text"><p><i>The argument in one picture. Because apparently that is what it takes now.</i></p></span></div></div><p class="paragraph" style="text-align:left;">ChatGPT&#39;s best new model is here. You can&#39;t use it. Sorry!</p><p class="paragraph" style="text-align:left;">And no, before everyone starts shouting, GPT-5.6 has <i>not</i> been banned. That&#39;s the hype version going around on social media. And as always it’s not quite right.</p><p class="paragraph" style="text-align:left;">Instead it’s actually MORE worrying. OpenAI launched it as a limited preview at the US government&#39;s request, and normal ChatGPT users are outside the preview gates. We’re locked out. </p><p class="paragraph" style="text-align:left;">This is the new precedent. </p><p class="paragraph" style="text-align:left;">Gulp.</p><p class="paragraph" style="text-align:left;">I said on the livestream this might be one of the darkest weeks in AI. I stand by that. Not because AI is over. The opposite. Because the best AI is shifting from app-store product to <i>permissioned</i> infrastructure. It’s getting TOO good. </p><h2 class="heading" style="text-align:left;" id="ignore-the-ban-hype">Ignore the ban hype</h2><p class="paragraph" style="text-align:left;">OpenAI announced three models: Sol, Terra and Luna. Lovely celestial branding. Sun, Earth, Moon. Very tasteful…</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/OpenAI/status/2070555272230384038?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-s-best-new-model-is-here-but-you-can-t-use-it"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">Totally useless for us though.</p><p class="paragraph" style="text-align:left;">Because we don’t have access. OpenAI says it plans broader access later, but for now GPT-5.6 is a limited preview for a small number of trusted partners in Codex and the API. </p><p class="paragraph" style="text-align:left;">We don’t know the list or how to (har har) get on it. </p><p class="paragraph" style="text-align:left;">Sam&#39;s follow-up is here too:</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/sama/status/2070607488274358364?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-s-best-new-model-is-here-but-you-can-t-use-it"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">Annoying. BUT at least Sam is communicating and acknowledges that this is a messy situation. Dario Amodei and Anthropic are not communicating with their users. </p><h2 class="heading" style="text-align:left;" id="two-weeks-same-pattern">Two weeks. Same pattern.</h2><p class="paragraph" style="text-align:left;">On June 12, Anthropic disabled Fable 5 and Mythos 5 after a US export-control directive. On June 26, OpenAI shipped GPT-5.6, but gated access before normal users could touch it.</p><p class="paragraph" style="text-align:left;">Two weeks. Two frontier-models barred from general access.</p><p class="paragraph" style="text-align:left;">Fable/Mythos was launched and quickly pulled. Mythos is apparently coming back for approved entities, not for individuals. GPT-5.6 launched in a restricted preview. Slightly different mechanism but same basic result. The average Joe is out of the loop. Sorry Joe. </p><p class="paragraph" style="text-align:left;">This is no longer just &quot;can I afford the <i>good</i> model?&quot; It is: are you on the list? Reminds me a bit of certain Rolex watches and Hermes birkin bags - it doesn’t matter how rich you are. That’s not how you get access. This is not a model I’m happy to see spread! </p><h2 class="heading" style="text-align:left;" id="there-is-a-real-reason">There is a real reason</h2><p class="paragraph" style="text-align:left;">The government is not just doing this for a laugh. Or because they are idiots. And not just because Dario Amodei has been stirring up fear. </p><p class="paragraph" style="text-align:left;">No, there are genuine <i>very real</i> concerns. </p><p class="paragraph" style="text-align:left;">These models are getting scary-good at cybersecurity work, long-horizon agent tasks and scientific discovery. All good stuff. </p><p class="paragraph" style="text-align:left;">But we also don&#39;t need to pretend this is <i>normal</i> software anymore.</p><p class="paragraph" style="text-align:left;">All the same mechanisms for good work can be used for bad work. If it can do one it can do the other.</p><p class="paragraph" style="text-align:left;">If a model can help discover vulnerabilities, assemble exploit pieces, run agents more efficiently and compress difficult technical work then governments are going to care. They <i>should </i>care. Banks, infrastructure, health systems and government departments are mostly held together with Excel sheets, procurement forms and vibes. </p><p class="paragraph" style="text-align:left;">And even with “dumbed” models bad actors were having an absolute field day. </p><p class="paragraph" style="text-align:left;">No model is immune to be jailbroken. And to believe this is possible (as the White House seem to be demanding…) is naïve. </p><p class="paragraph" style="text-align:left;">All models will get broken open to be used for naughty deeds. That’s part and parcel of this I’m afraid. Up until now this has been a nuisance but these (much) smarter models make this a critical risk. </p><h2 class="heading" style="text-align:left;" id="build-a-modelaccess-hedge">Build a model-access hedge</h2><p class="paragraph" style="text-align:left;">OK let’s come back to what do YOU do again. Always we need to come back to this. </p><p class="paragraph" style="text-align:left;">This is not a &quot;delete ChatGPT and live in a bunker&quot; issue. (<i>Although if you do have space in your bunker for one more I’d love to know…</i>)</p><p class="paragraph" style="text-align:left;">Use the best models when you can. Obviously. They are absurdly useful. But don&#39;t build a your life or business around a <i>single-model</i>. </p><p class="paragraph" style="text-align:left;">Have ChatGPT as one option, not your only option. Keep Claude, Gemini or another hosted provider warm. Have a poke around with Chinese-hosted models like GLM through Z.ai if you are comfortable with the trade-offs. Start testing local Gemma-style or Chinese models for lower-risk work you want to fully <i>control</i>. I wrote about <a class="link" href="https://aiwithkyle.com/ai-news/germany-edition?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-s-best-new-model-is-here-but-you-can-t-use-it" target="_blank" rel="noopener noreferrer nofollow">first steps with local AI models </a>a few days ago.</p><p class="paragraph" style="text-align:left;">Also build out an AI brain. Get everything in one place. Your prompts. Your specs. Your source files. Your evaluation examples. Your workflow notes. Your project context. Do not trap the brain of your business inside one rented interface.</p><p class="paragraph" style="text-align:left;">Get it into a Github repo that you can move BETWEEN any AI. This means you aren’t screwed when one particular model is taken from us. I wrote extensively about this last week <a class="link" href="https://aiwithkyle.com/ai-news/build-one-shared-ai-vault?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=chatgpt-s-best-new-model-is-here-but-you-can-t-use-it" target="_blank" rel="noopener noreferrer nofollow">https://aiwithkyle.com/ai-news/build-one-shared-ai-vault</a> </p><p class="paragraph" style="text-align:left;">Basically you are building in your own flexibility NOW before you need it. </p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=28629e2c-1879-4aae-95f0-23302bbb2483&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Saturday Sessions: It&#39;s getting a bit hairy</title>
  <description>Germany Edition</description>
  <link>https://newsletter.aiwithkyle.com/p/saturday-sessions-it-s-getting-a-bit-hairy</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/saturday-sessions-it-s-getting-a-bit-hairy</guid>
  <pubDate>Sat, 27 Jun 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-06-27T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'600' !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;">Hallo!</p><p class="paragraph" style="text-align:left;">I’ve been in Germany this week for Google I/O Connect Berlin. Thank you to the Google team for being gracious hosts - keep an eye out on the socials next week for vids from the event. I was very lucky to being able to interview team members from DeepMind, the Google Spark team and from AI Studio / Antigravity. Very cool stuff! </p><p class="paragraph" style="text-align:left;">Whilst I’ve been in Germany we’ve had some rather worrying news. </p><p class="paragraph" style="text-align:left;">Fable from Anthropic was pulled a few weeks ago on the request of the US Government. </p><p class="paragraph" style="text-align:left;">We’ve now found out that the new ChatGPT model (GPT5.6) is being held back.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/steph_palazzolo/status/2070241787180966279?s=46&utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=saturday-sessions-it-s-getting-a-bit-hairy"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">We are entering a period where the level of intelligence allowed to us is going to be restricted. And worse, services that we have access to may be suddenly taken from us. </p><p class="paragraph" style="text-align:left;"><i>That</i> worries me more than any specific model being held back. It’s the general directionality of this that should concern us. </p><p class="paragraph" style="text-align:left;">Learning how to use local models is more important than ever. </p><p class="paragraph" style="text-align:left;">BUT many people think it is outside of their skill zone. It’s a hard technical task right? </p><p class="paragraph" style="text-align:left;">No. It’s simple. </p><p class="paragraph" style="text-align:left;">In fact we’re going to do it <span style="text-decoration:underline;">right now</span>. This Saturday morning we’re going to get you set up. No excuses.</p><p class="paragraph" style="text-align:left;">First up, let’s use your main computer. Desktop or laptop. Doesn’t matter. You can (and should set it up on all your devices eventually so start with whatever is at hand).</p><p class="paragraph" style="text-align:left;">Go to <a class="link" href="https://lmstudio.ai/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=saturday-sessions-it-s-getting-a-bit-hairy" target="_blank" rel="noopener noreferrer nofollow">https://lmstudio.ai/</a> and download LM Studio. It’s free. </p><p class="paragraph" style="text-align:left;">As soon as you boot it you’ll get a screen like this:</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/7e0d1bbb-2f18-46b4-8237-dea65f7f360f/Screenshot_2026-06-26_at_11.39.16.png?t=1782470379"/></div><p class="paragraph" style="text-align:left;">LM Studio will basically have a look at your computer and decide a good started model for you. Here is happens to be Google’s Gemma 4 E4B. This isn’t a sponsored guide btw - it’s just that Google are the main (Western) lab releasing open source models! </p><p class="paragraph" style="text-align:left;">Go ahead and download. This one I’m being shown is ~7GB. </p><p class="paragraph" style="text-align:left;">Whilst that’s downloading go grab the mobile app too. </p><p class="paragraph" style="text-align:left;">On iPhone it’s <a class="link" href="https://locallyai.app/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=saturday-sessions-it-s-getting-a-bit-hairy" target="_blank" rel="noopener noreferrer nofollow">Locally</a>:</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/974289e9-fab9-42fb-ae97-86bf0b680038/Screenshot_2026-06-26_at_11.43.14.png?t=1782470613"/></div><p class="paragraph" style="text-align:left;">On Android <a class="link" href="https://lmsa.app/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=saturday-sessions-it-s-getting-a-bit-hairy" target="_blank" rel="noopener noreferrer nofollow">https://lmsa.app/</a> looks solid but I haven’t used it so cannot confirm! </p><p class="paragraph" style="text-align:left;">ALSO download a local model onto your phone :</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1190149e-6aa7-4f2e-ba3c-4079802cb605/IMG_3490.PNG?t=1782470931"/></div><p class="paragraph" style="text-align:left;">Here I’m downloading Gemma 4 E2B - it’s around ~4GB. </p><p class="paragraph" style="text-align:left;">Whilst both of these models are downloading now is a good time to talk about the model sizes. <b>This stuff isn’t vital so skip ahead to the bold section if you don’t care! </b>I will not be (that) upset.</p><p class="paragraph" style="text-align:left;">Notice I’ve just downloaded Gemma 4 E4B on my laptop (or whatever LM Studio suggested) and Gemma E2B on my iPhone. </p><p class="paragraph" style="text-align:left;">What gives? What’s the 4 and the 2 mean specifically? </p><p class="paragraph" style="text-align:left;">Time for a chart:</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/dcebac0a-208a-4a7b-af19-fa21efed58f8/69d3f7cd5103b6fb5f4df497_2_Gemma_4__What_Computer_Vision_Engineers_Actually_Need_to_Know__2_.png?t=1782471136"/><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://datature.io/blog/gemma-4-what-computer-vision-engineers-actually-need-to-know?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=saturday-sessions-it-s-getting-a-bit-hairy" target="_blank" rel="noopener noreferrer nofollow">https://datature.io/blog/gemma-4-what-computer-vision-engineers-actually-need-to-know</a></p></span></div></div><p class="paragraph" style="text-align:left;">OK yeah.. I see the problem. No wonder people think this stuff is complex! Look at that mess. Let’s decipher it a bit. Again, skip this if you want. </p><p class="paragraph" style="text-align:left;">The two blue header models here are “edge models”. Basically this means they work on “edge” devices - named so because they sit at the edge of a network. In normal terms basically it means your phone or laptop. </p><p class="paragraph" style="text-align:left;">Notice that E2B has 2.3B active parameters. The B is Billion. 2.3 billion parameters is the model size. That’s the one I’ve just installed on my iPhone. </p><p class="paragraph" style="text-align:left;">The E4B has 4.5 billion parameters. It’s bigger! That’s the one I just installed on my laptop. </p><p class="paragraph" style="text-align:left;">Parameters here are basically numbers. When we download a model we are (very crudely!) downloading a .csv file with billions and billions of numbers in it. Think of those 4GB and 6GB files we just download as <i>giant</i> lists of numbers (parameters). </p><p class="paragraph" style="text-align:left;">In general the more parameters the more intelligent the model. But the more parameters the more computing power you need to process it all. </p><p class="paragraph" style="text-align:left;">These smaller models work on our phone and laptop because they are smaller. Much smaller than a model like Claude (if we could download it!).</p><p class="paragraph" style="text-align:left;">But increasingly labs like Google Deepmind are doing a lot with a little. Small models that pack one hell of a punch. Whilst being fully controlled by you, on your devices, with 100% privacy. So, you know, a government can’t just pluck them away from you…</p><p class="paragraph" style="text-align:left;">Those purple header models? Those are beefier. They are not going to work on today’s phones and laptops. They will however run on a chunkier rig - a desktop computer or server. The more oomph you have the larger the model you can load in. I have Gemma 4 12B Q4 / MLX on my MacMini personally. </p><p class="paragraph" style="text-align:left;"><b>OK models downloaded? </b>Let’s continue!</p><p class="paragraph" style="text-align:left;">LM Studio’s interface is a little confusing initially but the main chatbox should be familiar enough. </p><p class="paragraph" style="text-align:left;">Use the dropdown at the top to select Gemma then type a message: </p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0b6e7db4-e374-4ab0-81b8-a3f7bfe4b070/Screenshot_2026-06-26_at_12.07.41.png?t=1782472087"/></div><p class="paragraph" style="text-align:left;">The model will respond. Just like with ChatGPT. Or Claude. Or Gemini. So? </p><p class="paragraph" style="text-align:left;">Now turn off your WiFi and keep chatting. </p><p class="paragraph" style="text-align:left;">It will still respond. </p><p class="paragraph" style="text-align:left;">Hell. Hop on an airplane and chat. It’ll keep working mid-air.</p><p class="paragraph" style="text-align:left;">Your chat is not zooming up to some cloud server sitting in California and then back to you. </p><p class="paragraph" style="text-align:left;">It’s not being used for training. </p><p class="paragraph" style="text-align:left;">It is not using electricity to run in a data centre. </p><p class="paragraph" style="text-align:left;">It’s all happening <i>right</i> <i>there</i> on your device. You’ve captured a genie. A powerful one. </p><p class="paragraph" style="text-align:left;">What’s more let’s say the US Government decide: “ok you aren’t allowed Gemma 4 anymore we’re removing it”.</p><p class="paragraph" style="text-align:left;">Tough shit! It’s on our computer. That cannot be reverted. Ya-boo sucks for you. 😘 </p><p class="paragraph" style="text-align:left;">Hell <i>even if GOOGLE</i> decide we can’t use it anymore: tough! There are 200M+ copies of Gemma 4 floating around out there. The whole model. Including the one on your computer. Once it’s out it’s out. </p><p class="paragraph" style="text-align:left;">And yes … I had to check that figure. 200M is mad impressive. Suggests that people and businesses are very much moving in this direction…</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/_akhaliq/status/2070191915652317521?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=saturday-sessions-it-s-getting-a-bit-hairy"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">OK solid we’re up and running on your laptop or desktop, let’s switch to the phone. </p><p class="paragraph" style="text-align:left;">Once you’ve downloaded the Gemma 4 E2B model you’ll be able to run it directly in the Locally app. FYI you can also use Google’s official Edge Gallery app for this. But using Locally gives us an advantage here. </p><p class="paragraph" style="text-align:left;">We can run a local model directly on our phone yes…very cool. But we can also run the model on our laptop (or computer) from our phone. There’s a feature in LM Studio called LM Link that basically connects your devices and let’s you run locally remotely … yes that’s confusing language! </p><p class="paragraph" style="text-align:left;">This means you can have a set up bigger, stronger models on your laptop and home computer and then access them (when online) on your phone.</p><p class="paragraph" style="text-align:left;">Here you can see my phone connected to both my laptop (here in Germany at the Einstein Kaffee I’m currently) and my MacMini (back in Cyprus):</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/07180a66-f700-41b1-9f82-773eba66a3de/IMG_3495.PNG?t=1782474351"/><div class="image__source"><span class="image__source_text"><p>2B on phone, 4B on laptop, 12B on desktop. All from my phone. </p></span></div></div><p class="paragraph" style="text-align:left;">This gives us maximum flexibility. When at home in Cyprus or on the move with internet I can use the most powerful model via the MacMini. </p><p class="paragraph" style="text-align:left;">When flying or otherwise without internet I can use a smaller model directly on my MacBook or at a pinch directly on my iPhone. Very cool. </p><p class="paragraph" style="text-align:left;">And the kicker! The models get smaller and more efficient. <i>Generally</i> (and this is from the Deepmind team I chatted to) we’re talking about a 7-8 month gap.</p><p class="paragraph" style="text-align:left;">What I can run locally on a decent laptop <i>now</i> is around the same as the state of the art model 8 months ago. </p><p class="paragraph" style="text-align:left;">SO…technically…8 months from now we’ll be running Opus 4.8, ChatGPT 5.5, Gemini 3.5 level models locally on relatively cheap consumer level hardware. Or, heck, even if it’s a year behind it doesn’t really matter. Having that much firepower locally means we can do pretty much ALL our normal tasks without increasingly expensive cloud subscriptions.</p><p class="paragraph" style="text-align:left;"> This is why it’s useful to get your head wrapped around it now. </p><p class="paragraph" style="text-align:left;">A quick recap:</p><ol start="1"><li><p class="paragraph" style="text-align:left;">Download LM Studio on your computer, laptop and/or phone. </p></li><li><p class="paragraph" style="text-align:left;">Download the recommended model on each device. </p></li><li><p class="paragraph" style="text-align:left;">Use LM Link to chain everything together. </p></li></ol><p class="paragraph" style="text-align:left;">And enjoy your free, always accessible AI.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle </p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=7e0d8dfe-43cb-4c46-94ff-d9c178fd614c&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Build an AI Brain</title>
  <description>Build one shared AI vault</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b03d940f-362c-45a2-9dc3-df95a7877d22/thumbnail_1_%2B_newsletter.png" length="1093241" type="image/png"/>
  <link>https://newsletter.aiwithkyle.com/p/ai-brain-vault</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/ai-brain-vault</guid>
  <pubDate>Wed, 24 Jun 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-06-24T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
    <category><![CDATA[Ai Workflow]]></category>
    <category><![CDATA[Ai Tools]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ede90c17-773a-491b-a75c-fe310306ff7d/GIF.gif?t=1782272881"/><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://youtu.be/LqkglPK9SUE?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-an-ai-brain" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/LqkglPK9SUE</a><span style="background-color:#ffffff;"><i> </i></span><i>- watch now or save for later</i></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/LqkglPK9SUE?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-an-ai-brain"><span class="button__text" style=""> Watch Now </span></a></div><div class="image"><img alt="White AI with Kyle header image reading Build an AI brain with a connected AI Vault diagram" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/7a66d61f-3abf-4768-9da7-586703eb4b1a/ai_generated_f1c69900a388a996.png?t=1782204525"/></div><p class="paragraph" style="text-align:left;">Your AI work is everywhere. All over the shop.</p><p class="paragraph" style="text-align:left;">ChatGPT has one bit. Claude has another. Codex has half a project. Your phone has the voice note. GitHub has the code.</p><p class="paragraph" style="text-align:left;">Sure you are doing a lot. Very productive. Lovely.</p><p class="paragraph" style="text-align:left;">But the more you create the more mess you make.</p><p class="paragraph" style="text-align:left;">The problem is that none of these things automatically know what the other ones are doing. So <i>you</i> become the USB stick. You are the one copying context from one model to another, re-explaining your business, re-explaining the client, re-explaining the project, re-explaining the weird decision you made last Tuesday because Future You apparently enjoys admin.</p><p class="paragraph" style="text-align:left;">That’s precisely the opposite of what we want to use AI for. We are just making MORE work for ourselves.</p><p class="paragraph" style="text-align:left;">This is why people are suddenly talking about AI brains, Obsidian vaults and personal operating systems. It looks new. It is sorta new. But (as always) the genuinely useful bit is not the flashy viral stuff.</p><p class="paragraph" style="text-align:left;">It is mostly folders and markdown files.</p><p class="paragraph" style="text-align:left;">Ho hum.</p><h2 class="heading" style="text-align:left;" id="obsidian-is-not-the-brain">Obsidian is not the brain</h2><p class="paragraph" style="text-align:left;">You’ve probably seen posts about Obsidian. And think I’m going to focus on Obsidian.</p><p class="paragraph" style="text-align:left;">Obsidian is cool. I like Obsidian.</p><p class="paragraph" style="text-align:left;">But Obsidian is the HUMAN interface. It is the pretty layer. It gives you the graph view, backlinks, search, daily notes, plugins and all the nerdy little goodies that make a folder of text files feel like a system.</p><p class="paragraph" style="text-align:left;">For you, that is useful.</p><p class="paragraph" style="text-align:left;">For the AI? Less so. They don’t NEED this.</p><p class="paragraph" style="text-align:left;">Claude Code, Codex and the rest do not need a lovely purple graph of your knowledge. They need the actual files. They need the README. They need the decisions log. They need the project notes. They need the bit where you wrote &quot;my god do not touch this payment flow because it breaks checkout!!1!&quot; three months ago and then forgot about it.</p><p class="paragraph" style="text-align:left;">That is the brain.</p><p class="paragraph" style="text-align:left;">And it’s basically a folder, some sub folders and some text files. That’s it.</p><h2 class="heading" style="text-align:left;" id="boring-bit-that-gets-things-done">Boring bit that gets things done</h2><p class="paragraph" style="text-align:left;">The base layer here is a markdown file. <code>.md</code>.</p><p class="paragraph" style="text-align:left;">Nothing to do with doctors. No stethoscope required.</p><p class="paragraph" style="text-align:left;">Markdown is basically a plain text file with a bit of simple formatting. Humans can read it. AIs can read it. GitHub can track it. Obsidian can display it. Claude Code and Codex can scan it before doing work.</p><p class="paragraph" style="text-align:left;">Markdown is not sexy. It is aggressively unsexy. It is a Word document with the formatting fluff removed and, frankly, thank you very much.</p><p class="paragraph" style="text-align:left;">But boring is useful here. You do not want the memory of your business trapped inside one app&#39;s proprietary format. Or heavy files like PDFs and Word docs.</p><p class="paragraph" style="text-align:left;">You want files that can be copied, searched, versioned, backed up and read by whatever AI tool you use next month when this week&#39;s favourite app inevitably changes its pricing or does one.</p><p class="paragraph" style="text-align:left;">The useful files are simple:</p><ul><li><p class="paragraph" style="text-align:left;"><code>README.md</code> for what the brain is</p></li><li><p class="paragraph" style="text-align:left;"><code>AGENTS.md</code> for how AI tools should behave</p></li><li><p class="paragraph" style="text-align:left;"><code>CLAUDE.md</code> if you want Claude-specific instructions (hint, point it at <a class="link" href="http://AGENTS.md?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-an-ai-brain" target="_blank" rel="noopener noreferrer nofollow">AGENTS.md</a> instead so you have one set of instructions)</p></li><li><p class="paragraph" style="text-align:left;"><code>tasks.md</code> for open loops and todos</p></li><li><p class="paragraph" style="text-align:left;"><code>decisions.md</code> for things that should not be re-debated</p></li><li><p class="paragraph" style="text-align:left;">project READMEs for active work</p></li><li><p class="paragraph" style="text-align:left;">agent logs for what AI tools actually did</p></li></ul><p class="paragraph" style="text-align:left;">You can of course add more. But generally something simple like this is all you need.</p><p class="paragraph" style="text-align:left;">Also don’t worry about how to make all this - AI will do it for us shortly.</p><h2 class="heading" style="text-align:left;" id="access-anywhere">Access anywhere</h2><p class="paragraph" style="text-align:left;">Where should this folder and set of documents live?</p><p class="paragraph" style="text-align:left;">We will connect ALL our AI tools to the same set of folders.</p><p class="paragraph" style="text-align:left;">So: we need it accessible anywhere. In the Cloud.</p><p class="paragraph" style="text-align:left;">Someone asked on the stream whether this could live in Google Drive.</p><p class="paragraph" style="text-align:left;">Technically, yes.</p><p class="paragraph" style="text-align:left;">I would not.</p><p class="paragraph" style="text-align:left;">If you are using AI tools properly, you will soon have multiple agents working on different bits of your life or business. Codex doing one thing. Claude Code doing another. Maybe one thread sorting your newsletter, another checking your website, another helping with a client project.</p><p class="paragraph" style="text-align:left;">If those tools all bash away inside the same Google Doc, it gets messy fast. Conflicts. Random edits. &quot;Who changed this?&quot; energy. Lots of faff.</p><p class="paragraph" style="text-align:left;">GitHub is built for this kind of thing.</p><p class="paragraph" style="text-align:left;">I’ve done another full guide on Github 101 previously so go check that if have no idea what Github is.</p><p class="paragraph" style="text-align:left;">Simply put though it’s for online file storage in the cloud. It tracks changes. It lets tools work in branches or worktrees. It gives you a history of decisions. It means you can see that on June 6th you changed the website, added a rule, moved a file, broke something, fixed it, then told Future You never to do that again.</p><p class="paragraph" style="text-align:left;">So: use Github here to make your life much easier. Bonus - it’s free.</p><h2 class="heading" style="text-align:left;" id="do-not-copy-someone-elses-brain">Do not copy someone else&#39;s brain</h2><p class="paragraph" style="text-align:left;">This is the bit people will ignore - I know some of you will still go and find some guru’s “perfect AI brain” template. Maybe even buy it! Don’t!</p><p class="paragraph" style="text-align:left;">You should not start with someone else&#39;s template.</p><p class="paragraph" style="text-align:left;">I know. Annoying. Templates are comfy. They make the work feel done before it is done. You download a folder structure, put it in Obsidian, stare at the cool looking knowledge graph and think: &quot;Ah yes, productivity.&quot;</p><p class="paragraph" style="text-align:left;">Nope!</p><p class="paragraph" style="text-align:left;">Your vault needs to map to YOUR actual life. Your roles. Your projects. Your clients. Your products. Your weird repeated context. Your privacy boundaries. Your working style.</p><p class="paragraph" style="text-align:left;">Your mess. Your very own chaos!</p><p class="paragraph" style="text-align:left;">That is why the starting point should not be a template. It has to start with you.</p><p class="paragraph" style="text-align:left;">The starting point therefore should be an interview.</p><p class="paragraph" style="text-align:left;">Let the AI ask you questions. One at a time. Then answer with your microphone. Yap for 20 minutes. Tell it about the projects, the priorities, the stuff you keep forgetting, the workflows you repeat, the things you definitely do not want it to touch, and the boring context you are sick of retyping.</p><p class="paragraph" style="text-align:left;">I personally did this over DAYS when setting my systems up. I kept remembering important things and yapping some more!</p><p class="paragraph" style="text-align:left;">The AI can then propose the folder structure for <i>you</i>. It will make up the files, the project folders, the whole shebang for you.</p><p class="paragraph" style="text-align:left;">Much better.</p><h2 class="heading" style="text-align:left;" id="steal-this-prompt">Steal this prompt</h2><p class="paragraph" style="text-align:left;">Use Codex or Claude Code for this. ChatGPT can help design the structure, but it cannot reliably create the folder tree and files on your machine unless you connect it to tools. Codex is free if need to just getting started</p><p class="paragraph" style="text-align:left;">Here is the basic prompt:</p><div class="codeblock"><pre><code>I want to create a vendor-agnostic AI vault that acts as a shared brain across my tools.

Use markdown files and folders.

Interview me one question at a time about:
- my daily life and roles
- my work and active projects
- repeated context I keep explaining to AI
- workflows I want AI to help with
- privacy boundaries
- how I like AI tools to behave

After the interview, propose:
- a vault folder structure
- README.md
- AGENTS.md
- project README templates
- tasks.md
- decisions.md
- an agent-log format
- a migration plan

Ask me the first question now.

This will live in Github. 
If I do not have this set up help me.
</code></pre></div><p class="paragraph" style="text-align:left;">Then talk.</p><p class="paragraph" style="text-align:left;">Seriously: actual talk. Use the microphone. Do not sit there trying to write the perfect answer like a LinkedIn thought leader. Ick.</p><p class="paragraph" style="text-align:left;">Just talk. Yap. Unload everything.</p><p class="paragraph" style="text-align:left;">The AI will sort the structure out afterwards. That is the whole point. We are not doing artisanal folder design here. Let the AI do the work.</p><h2 class="heading" style="text-align:left;" id="the-rule-explain-it-twice-save-it">The rule: explain it twice, save it</h2><p class="paragraph" style="text-align:left;">One issue right now - it’s hard to capture “normal” chats into your Github folders and files. Everyday chatbot Claude and ChatGPT don’t do this by default.</p><p class="paragraph" style="text-align:left;">Instead use Codex or Clade Code to make sure everything is captured.</p><p class="paragraph" style="text-align:left;">Does this mean you need to always use Claude Code or Codex? Nope!</p><p class="paragraph" style="text-align:left;">The daily habit is simple: Anything you explain to AI twice belongs in the vault.</p><p class="paragraph" style="text-align:left;">For any one off usage of AI the normal chatbot is entirely fine. Those chats are disposable. For anything you want to capture use your new system.</p><p class="paragraph" style="text-align:left;">And the magic thing here is that over time your system gets better. The more context and information it has about you the better your AI will be able to help you.</p><p class="paragraph" style="text-align:left;">Not on day one. Day one is mostly a folder and a mild headache as you get it all set up. Sorry…</p><p class="paragraph" style="text-align:left;">But after a few weeks, your AI tools stop starting from zero. They stop asking the same questions. They stop making the same bad assumptions. They can read the vault first, understand the context, then do the work.</p><p class="paragraph" style="text-align:left;">If your coding tool knows your business it’ll stop making errors that seem “dumb” to you. “What the hell are you doing Codex you KNOW I don’t sell that those services?” sort of annoyances disappear because your AI tools will start to collate info from across different domains of your life.</p><p class="paragraph" style="text-align:left;">This is also why chat history is not enough. It doesn’t pass between tools. We need a centralised location for that!</p><h2 class="heading" style="text-align:left;" id="obsidian-is-still-useful">Obsidian is still useful</h2><p class="paragraph" style="text-align:left;">Full circle!</p><p class="paragraph" style="text-align:left;">Should you use Obsidian?</p><p class="paragraph" style="text-align:left;">Yeah, probably!</p><p class="paragraph" style="text-align:left;">It’ll give you a great human interface to “see” your system.</p><p class="paragraph" style="text-align:left;">It is free for the basic local version. It is good. It makes markdown folders easier to browse. It gives you backlinks and search and graph views and plugins and all the nice human bits.</p><p class="paragraph" style="text-align:left;">Do you need to pay? Probably not! The paid plan is for synching. But we are using GitHub for that (for free) so we’re good!</p><p class="paragraph" style="text-align:left;">Set the GitHub repo up locally, open that folder as your Obsidian vault, and let GitHub Desktop or your AI tool handle the pushing and pulling. Tada! Saved some money.</p><p class="paragraph" style="text-align:left;">Again, keep the layers clear:</p><ul><li><p class="paragraph" style="text-align:left;">folder = the brain</p></li><li><p class="paragraph" style="text-align:left;">markdown = the memory</p></li><li><p class="paragraph" style="text-align:left;">GitHub = history and sync</p></li><li><p class="paragraph" style="text-align:left;">Obsidian = human interface</p></li><li><p class="paragraph" style="text-align:left;">AI tools = workers</p></li></ul><p class="paragraph" style="text-align:left;">Grossly simplified. Also useful.</p><h2 class="heading" style="text-align:left;" id="start-small">Start small</h2><p class="paragraph" style="text-align:left;">You do not need a perfect personal operating system by Friday.</p><p class="paragraph" style="text-align:left;">Please do not spend three days building a 97-folder cathedral to productivity and then never use it. Productivity porn is appealing. But valueless.</p><p class="paragraph" style="text-align:left;">Start with one repo. One README with a basic description. One AGENTS file with simple rules for your AI tools. One decisions file. One task list. One project folder.</p><p class="paragraph" style="text-align:left;">Then USE it.</p><p class="paragraph" style="text-align:left;">Tell every AI tool to read the vault first. By default they’ll always check <a class="link" href="http://AGENTS.md?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=build-an-ai-brain" target="_blank" rel="noopener noreferrer nofollow">AGENTS.md</a> first (that’s its purpose) so put important instructions in there. Again the AI will help you set this up so don’t worry about the how.</p><p class="paragraph" style="text-align:left;">Have your AI update the vault when a decision is made. Have it write logs when it changes something. Have it add repeated context instead of letting it disappear into a chat thread you will never find again.</p><p class="paragraph" style="text-align:left;">And over time it’ll grow and improve. For now though: keep it simple.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=9d9d8561-ddf3-469f-a45f-cfb9f8365bdf&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Fable got confiscated</title>
  <description>When AI stopped being just software</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/04b7d63e-2c7a-45a6-8604-c0464b1297ed/thumbnail_1_%2B_newlstter.png" length="811500" type="image/png"/>
  <link>https://newsletter.aiwithkyle.com/p/claude-fable-ban</link>
  <guid isPermaLink="true">https://newsletter.aiwithkyle.com/p/claude-fable-ban</guid>
  <pubDate>Wed, 17 Jun 2026 07:00:00 +0000</pubDate>
  <atom:published>2026-06-17T07:00:00Z</atom:published>
    <dc:creator>Kyle Balmer</dc:creator>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #C0C0C0; }
  .bh__table_cell { padding: 5px; background-color: #FFFFFF; }
  .bh__table_cell p { color: #2D2D2D; font-family: 'Helvetica',Arial,sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#F1F1F1; }
  .bh__table_header p { color: #2A2A2A; font-family:'Trebuchet MS','Lucida Grande',Tahoma,sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5a504c78-08be-4140-82ee-39f1e22385ad/GIF.gif?t=1781662019"/><div class="image__source"><span class="image__source_text"><p><a class="link" href="https://youtu.be/fjUARfuLsis?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">https://youtu.be/fjUARfuLsis</a> <i>- watch now or save for later</i></p></span></div></div><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://youtu.be/fjUARfuLsis?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated"><span class="button__text" style=""> Watch Now </span></a></div><p class="paragraph" style="text-align:left;">Fable came back from the future and got immediately confiscated.</p><p class="paragraph" style="text-align:left;">That sounds <i>dramatic</i>. It is.</p><p class="paragraph" style="text-align:left;">For about four hours I had access to the model everyone is now mourning. People were using it to debug apps, inspect businesses, plan their lives, rethink whole projects, and generally do the kind of meaty work that normally takes days of faff.</p><p class="paragraph" style="text-align:left;">Then it vanished…</p><p class="paragraph" style="text-align:left;">The livestream story will probably change (again!) by the time this lands. Maybe Anthropic gets Fable back online. Maybe the US government carves allies back in. Maybe they relabel the whole thing Claude Opus 4.9 and pretend everything is fine…</p><p class="paragraph" style="text-align:left;">But the important part is not &quot;will we get our shiny new model back?&quot;</p><p class="paragraph" style="text-align:left;">The useful part is this:<b> frontier AI just stopped being </b><b><i>normal</i></b><b> software and started looking like strategic infrastructure.</b></p><h2 class="heading" style="text-align:left;" id="fable-was-different">Fable was different</h2><div class="image"><img alt="Fable got confiscated - slide 2" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9e6c91bd-44a7-47b0-8ec9-faf2b11e3cdf/fable-ban-p002.png?t=1781635093"/><div class="image__source"><span class="image__source_text"><p>Same family. Different access model.</p></span></div></div><p class="paragraph" style="text-align:left;">Quick recap on what Fable actually is…</p><p class="paragraph" style="text-align:left;">Anthropic launched <a class="link" href="https://www.anthropic.com/news/claude-fable-5-mythos-5?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">Claude Fable 5 and Claude Mythos 5</a> on June 9. Fable was the public commercial release. Mythos was the more permissive, trusted-access version for cyberdefenders and infrastructure partners.</p><p class="paragraph" style="text-align:left;">Anthropic described Fable as a “Mythos-class model” made safe for general use. And it is (was?) very very good. </p><p class="paragraph" style="text-align:left;">I think the reason people have been weirdly emotional about losing it is that Fable did not feel like a faster chatbot. <a class="link" href="https://www.oneusefulthing.org/p/what-it-feels-like-to-work-with-mythos?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">Ethan Mollick wrote about the shift from doing to commissioning</a>. <a class="link" href="https://x.com/simonw/status/2065216774992515342?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">Simon Willison called it relentlessly proactive</a>. Matt Shumer was saying the productivity gap felt <i>huge</i>.</p><p class="paragraph" style="text-align:left;">That matches my tiny, annoying glimpse of it. And others I’ve talked to. </p><p class="paragraph" style="text-align:left;">It felt (dare I say it…) an awful lot like AGI…</p><p class="paragraph" style="text-align:left;">Shh shh shh we don’t say that around here! </p><h2 class="heading" style="text-align:left;" id="the-timeline-is-already-a-mess">The timeline is already a mess</h2><p class="paragraph" style="text-align:left;">And then it got banned! Ruh-roh. </p><div class="image"><img alt="Fable got confiscated - slide 4" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ed073986-f08b-4426-ad2d-05b9f377f07c/fable-ban-p004.png?t=1781635095"/></div><p class="paragraph" style="text-align:left;">The timeline is a mess and we’re still getting details. But let’s try to break it down. </p><p class="paragraph" style="text-align:left;">Fable launched. People fell in love with it. Then the whole thing got dragged into export controls.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/AnthropicAI/status/2065597531644743999?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">Anthropic&#39;s <a class="link" href="https://www.anthropic.com/news/fable-mythos-access?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">June 12 statement</a> says the US government issued an export-control directive suspending access to Fable 5 and Mythos 5 by <i>any foreign national,</i> whether inside or outside the United States, including foreign-national Anthropic employees.</p><p class="paragraph" style="text-align:left;">Foreign-national employees.</p><p class="paragraph" style="text-align:left;">That means if you are a non-US citizen working inside an American AI lab (like, I dunno, Andrej Karpathy!), building the <i>actual</i> model, the order says you cannot use the model.</p><p class="paragraph" style="text-align:left;">Anthropic says the directive arrived at 5:21pm ET, did not provide specific details, and referred to a possible jailbreak. Their position is basically: yes, there was a narrow bypass, but the demonstrated capability is available in other public models too, and no universal jailbreak has been shown.</p><p class="paragraph" style="text-align:left;">Apparently Anthropic had 90 minutes to comply.</p><p class="paragraph" style="text-align:left;">The government-side story is different. <a class="link" href="https://x.com/DavidSacks/status/2065853007619588171?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">David Sacks&#39; version</a> is that a credible partner (we know believe this to be Amazon…) found a serious guardrail issue, Anthropic was asked to fix it or pull the model, Dario refused, and the export control followed.</p><p class="paragraph" style="text-align:left;">Worth reading in full:</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/DavidSacks/status/2065853007619588171?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;">So we have two stories:</p><ul><li><p class="paragraph" style="text-align:left;">Anthropic says that the jailbreak threat was vague, rushed, and not grounded in enough technical detail.</p></li><li><p class="paragraph" style="text-align:left;">The government says Anthropic talked up a cyber weapon, got warned about the safety layer, then downplayed the problem when it became inconvenient.</p></li></ul><p class="paragraph" style="text-align:left;">Both can be partly true. And probably are.</p><p class="paragraph" style="text-align:left;">I do think Dario has slightly FAFO’d here. If you spend months telling governments your model is scary, dangerous, and needs regulation, you cannot be too shocked when a government says: &quot;OK then. Regulated.&quot;</p><p class="paragraph" style="text-align:left;">But it is <i>also</i> problematic the US government can force Anthropic to yank a model globally based on a process nobody can see, a technical standard nobody can inspect, and a nationality rule nobody can enforce without turning every AI login into border control.</p><h2 class="heading" style="text-align:left;" id="the-nationality-rule">The nationality rule</h2><div class="image"><img alt="Fable got confiscated - slide 6" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/4bd39ec8-6b21-4b29-b261-6323ba8979bd/fable-ban-p006.png?t=1781635097"/><div class="image__source"><span class="image__source_text"><p>This is the bit that turns a model pause into an access-control story.</p></span></div></div><p class="paragraph" style="text-align:left;">The model being paused is annoying. But it is what it is. </p><p class="paragraph" style="text-align:left;">What is much more important is the introduction of a foreign-national rule. That’s a <i>precedent</i>.</p><p class="paragraph" style="text-align:left;">Because how do you enforce that? IP address? Lol. VPN. Billing country? Nope. Company account? Maybe. Passport scan? Driving licence? Face check? Some sort of live &quot;prove you are the right citizen&quot; identity layer? </p><p class="paragraph" style="text-align:left;">From a practical POV Anthropic couldn’t “screen out” non-US citizens in 90 minutes. So they shut everything down - it was the only safe play. </p><p class="paragraph" style="text-align:left;">Very likely though we are moving into a new world of AI control. AI companies collecting our details and tying them to our accounts. This is a whole new world. You get know-your-customer checks. You get passports. You get government-approved access tiers. You get a centralised record of who is allowed to use which model for which kind of work.</p><p class="paragraph" style="text-align:left;">Maybe that is necessary at the frontier. Maybe the cyber and bio risks really do demand it. I&#39;m not doing the silly libertarian thing where every rule is tyranny and every safety team should do one.</p><p class="paragraph" style="text-align:left;">But there is a trade-off here. </p><p class="paragraph" style="text-align:left;">If frontier AI requires identity-gated access, then AI is not becoming democratised in the way people promised. It is becoming licensed infrastructure. </p><p class="paragraph" style="text-align:left;">Who gets the licence? Who gets cut off?</p><p class="paragraph" style="text-align:left;">Who gets the good stuff and who gets the leftovers?</p><p class="paragraph" style="text-align:left;">We don’t know the answers yet but let’s hazard a guess. These rules will be laid down by the White House and billion dollar corporations. </p><p class="paragraph" style="text-align:left;"> It’ll be for the rich, big business and approved nations. </p><h2 class="heading" style="text-align:left;" id="closed-access-is-not-ownership">Closed access is not ownership</h2><p class="paragraph" style="text-align:left;">This is a good reminder: if you are using a subscription or an API, you do not own the capability. You have <i>permission</i> to access it until someone else changes their mind.</p><p class="paragraph" style="text-align:left;">That someone might be Anthropic. It might be a cloud provider. It might be a regulator. It might be the US government at 5:21pm on a Friday.</p><p class="paragraph" style="text-align:left;">That is not a reason to stop using frontier models. That would be mad. These things are still the most useful tools we have.</p><p class="paragraph" style="text-align:left;">But it <i>is</i> a reason to stop depending on one model, one lab, one country, and one access layer.</p><p class="paragraph" style="text-align:left;">At minimum:</p><div class="image"><img alt="Fable got confiscated - slide 8" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/930b2a45-1509-462a-8375-d809d3e0baff/fable-ban-p008.png?t=1781635099"/><div class="image__source"><span class="image__source_text"><p>Know which layer you control.</p></span></div></div><p class="paragraph" style="text-align:left;">Use fallback models. Keep a second provider ready. Know which workflows can move from Claude to ChatGPT to Gemini to a local model without collapsing. Keep exports of your important prompts, specs, source files, and context. Do not trap the brain of your business inside one rented interface.</p><p class="paragraph" style="text-align:left;">And start having a poke around with local models. I’ve written a <a class="link" href="https://aiwithkyle.com/ai-news/179-local-ai?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">guide on Local LLMs 101 here</a>. Not because local models are better. They generally are not…. Not because open source magically saves you. It does not. Most people do not have the hardware, the patience, or the cojones to run the really big stuff well.</p><p class="paragraph" style="text-align:left;">But local gives you a layer <i>you</i> control.</p><p class="paragraph" style="text-align:left;">Use <a class="link" href="https://lmstudio.ai/?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">LM Studio</a>. Download something from <a class="link" href="https://huggingface.co/models?utm_source=newsletter.aiwithkyle.com&utm_medium=newsletter&utm_campaign=fable-got-confiscated" target="_blank" rel="noopener noreferrer nofollow">Hugging Face</a>. Run a small model on your laptop. Rent a GPU server for a weekend if you want to get spicy. Learn what it can and cannot do.</p><p class="paragraph" style="text-align:left;">Do this as a learning experience if nothing else. </p><h2 class="heading" style="text-align:left;" id="the-scarce-skill-is-bigger-tasks">The scarce skill is bigger tasks</h2><p class="paragraph" style="text-align:left;">The irony of all this is that the practical skill did not disappear with Fable.</p><p class="paragraph" style="text-align:left;">The model is gone for now sure. But the directionality is in play. </p><p class="paragraph" style="text-align:left;">Nate B. Jones had the useful framing: <b>the new scarce skill is </b><b><i>task imagination</i></b><b>. </b>Coming up with work big enough to hand to an agent that can run for hours or days.</p><p class="paragraph" style="text-align:left;">The models will keep getting better. Maybe Fable comes back. Maybe GPT-5.6 does something similar. Maybe Claude 5 is Mythos with a new hat. Maybe some compound model from OpenRouter gets you close enough by stitching systems together.</p><p class="paragraph" style="text-align:left;">Any which way we are moving towards a Fable-shaped world. </p><div class="image"><img alt="Fable got confiscated - slide 10" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0dbd81b8-1580-47bf-882b-a4ea3e532c95/fable-ban-p010.png?t=1781635101"/><div class="image__source"><span class="image__source_text"><p>The moral, because yes, apparently Fable needed one.</p></span></div></div><p class="paragraph" style="text-align:left;">There are still open questions.</p><p class="paragraph" style="text-align:left;">Does Fable come back quickly or does this become the precedent? Will governments review every Mythos-level release? Do “allies” get carved back in or does access stay nationality-based? Who decides what counts as a <i>serious</i> jailbreak? What happens when Chinese labs keep pushing while American labs get tangled in their own rules?</p><p class="paragraph" style="text-align:left;">I don&#39;t know. Nobody does.</p><p class="paragraph" style="text-align:left;">But I know what I would do this week:</p><p class="paragraph" style="text-align:left;">Do not depend on one model. Build fallbacks. Learn local models enough that they stop seeming like arcane nonsense. Keep your important work portable. Stop thinking in tiny prompts. Start thinking in BIG projects.</p><p class="paragraph" style="text-align:left;">To the Task,</p><p class="paragraph" style="text-align:left;">Kyle</p></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=e7999ec4-e759-413c-afb7-0acdee60991c&utm_medium=post_rss&utm_source=ai_with_kyle">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

  </channel>
</rss>
