<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>unwind ai</title>
    <description>Open-source Ecosystem for High-Leverage AI Builders</description>
    
    <link>https://www.theunwindai.com/</link>
    <atom:link href="https://rss.beehiiv.com/feeds/fMHDv0Uk41.xml" rel="self"/>
    
    <lastBuildDate>Thu, 13 Aug 2026 04:51:20 +0000</lastBuildDate>
    <pubDate>Tue, 23 Jun 2026 12:30:00 +0000</pubDate>
    <atom:published>2026-06-23T12:30:00Z</atom:published>
    <atom:updated>2026-08-13T04:51:20Z</atom:updated>
    
      <category>Machine Learning</category>
      <category>Artificial Intelligence</category>
      <category>Technology</category>
    <copyright>Copyright 2026, unwind ai</copyright>
    
    <image>
      <url>https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/publication/logo/84ac330c-894c-4f61-ac66-e747ce8b32eb/logo.png</url>
      <title>unwind ai</title>
      <link>https://www.theunwindai.com/</link>
    </image>
    
    <docs>https://www.rssboard.org/rss-specification</docs>
    <generator>beehiiv</generator>
    <language>en-us</language>
    <webMaster>support@beehiiv.com (Beehiiv Support)</webMaster>

      <item>
  <title>Opus 4.8-level model now runs locally for FREE</title>
  <description>+ OpenRouter Fusion, GLM-5.2 locally, Loop Engineering</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/506a4fa0-c094-4d73-bfaf-db42127ddb46/Opus_4.8-level_model_now_runs_locally_for_FREE.jpg" length="186942" type="image/jpeg"/>
  <link>https://www.theunwindai.com/p/opus-4-8-level-model-now-runs-locally-for-free</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/opus-4-8-level-model-now-runs-locally-for-free</guid>
  <pubDate>Tue, 23 Jun 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-06-23T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Google Cloud turns scattered knowledge into agent-readable files</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Vibe is here: one agent for work and code</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Run GLM 5.2 locally</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Telegram bots can now talk to other bots</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Open-source alternative to Loom, Granola, and Wisprflow</b></p></li></ol><p class="paragraph" style="text-align:start;">& a lot more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/generative-ui-is-the-new-frontend?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">Generative UI Is the New Frontend</a></b></p><p class="paragraph" style="text-align:left;">The frontend used to be a fixed thing. Designers drew it. Engineers built it. Users got what shipped.</p><p class="paragraph" style="text-align:left;">That&#39;s over.</p><p class="paragraph" style="text-align:left;">The interfaces shipping in 2026 are drawn partly by the agent itself, in real time, from what the user actually asked for. Ask for a table, get a table. Not a paragraph describing one.</p><p class="paragraph" style="text-align:left;">Generative UI is the layer that lets agents stop describing and start showing. </p><p class="paragraph" style="text-align:left;">This guide walks you through three patterns that have emerged on how to build it, and the differences between them matter more than most teams realize.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/generative-ui-is-the-new-frontend?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/6c3faaf3-c6cf-4f53-806b-8a114703b66a/Generative_UI_Is_the_New_Frontend.png?t=1780532864"/><div class="embed__content"><p class="embed__title"> Generative UI Is the New Frontend </p><p class="embed__description"> How AI agents stop describing and start showing </p></div></a></div><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://cloud.google.com/blog/products/data-analytics/how-the-open-knowledge-format-can-improve-data-sharing?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>Google Cloud Turns Company Knowledge Into Agent-Readable Files</b></a><b> </b>🧠📁</h3><div class="image"><a class="image__link" href="https://cloud.google.com/blog/products/data-analytics/how-the-open-knowledge-format-can-improve-data-sharing?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/695ed768-c228-4a03-a336-9a11c34a28f5/image.png?t=1782192445"/></a></div><p class="paragraph" style="text-align:left;">Every team building internal agents eventually hits the same wall: the model is smart, but the context is scattered everywhere.</p><p class="paragraph" style="text-align:left;">Part of it lives in data catalogs. Part of it lives in wikis. Part of it lives in code comments, dashboards, tribal knowledge, and that one senior engineer&#39;s brain.</p><p class="paragraph" style="text-align:left;"><b>Google Cloud just introduced Open Knowledge Format (OKF)</b> to make that context portable. It is a vendor-neutral spec that turns enterprise knowledge into Markdown files with YAML frontmatter, so agents can read it, search it, version it, and move it between tools without another custom integration.</p><p class="paragraph" style="text-align:left;">The nice part is how boring the format is. Just Markdown. Just files. Just a small set of structured fields like type, title, description, resource, tags, and timestamp.</p><p class="paragraph" style="text-align:left;">That is exactly why it could work. Agents do not need another complex metadata platform; they need context they can actually open, inspect, and use.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Markdown-first</b>: OKF represents context as human-readable Markdown, so engineers can review it in normal editors and agents can index it without special tooling.</p></li><li><p class="paragraph" style="text-align:left;"><b>Structured enough for agents</b>: YAML frontmatter adds queryable fields like type, title, resource, tags, and timestamp without turning the whole thing into a heavy schema project.</p></li><li><p class="paragraph" style="text-align:left;"><b>Portable by default</b>: OKF bundles can live in Git, ship as files, mount on a filesystem, or move across tools without locking context inside one vendor&#39;s catalog.</p></li><li><p class="paragraph" style="text-align:left;"><b>Reference implementations included</b>: Google shipped examples including BigQuery enrichment, a static HTML visualizer, sample bundles, and Knowledge Catalog ingestion support.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://mistr.al/vibe-unwindai-nl?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>Vibe is here: one agent for work and code</b></a></h3><div class="image"><a class="image__link" href="https://mistr.al/vibe-unwindai-nl?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b7f36dbd-9c8e-45e7-8b4a-b6fba210569c/image.jpeg?t=1782192515"/></a></div><p class="paragraph" style="text-align:left;">Meet <a class="link" href="https://mistr.al/vibe-unwindai-nl?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>Vibe by Mistral</b></a>, one agent and one licence across work and code. Vibe takes on long-running, multi-step work: catching up across your inbox and calendar, running deep research, drafting deliverables, and taking coding work from request to merged change, across the web app, your editor, and your terminal.</p><p class="paragraph" style="text-align:left;"><b>Key highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Work Mode for complex, multi-stage tasks</b>: Maps out a plan, gets your sign-off, then works across your connectors to carry it through. Every tool call and reasoning step is visible and expandable as it runs. </p></li><li><p class="paragraph" style="text-align:left;"><b>Code Mode for remote coding sessions</b>: Connect to GitHub, start sessions, and see them through to a pull request. Sessions run in an isolated sandbox, persist while your machine is off, and can run in parallel.</p></li><li><p class="paragraph" style="text-align:left;"><b>VS Code extension</b>: Vibe now works across your whole project inside VS Code. Reads, edits, and runs commands in a side panel. Open files attach automatically, @ mentions pull in context from anywhere in your repo.</p></li><li><p class="paragraph" style="text-align:left;"><b>CLI updates</b>: Skills become / commands. Permissions are session-scoped. /teleport moves a live session between your terminal and the cloud, history and approvals intact.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://openrouter.ai/blog/announcements/fusion-beats-frontier/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>OpenRouter Fusion Makes Model Panels a One-Call Primitive </b></a>🧪🤝</h3><div class="image"><a class="image__link" href="https://openrouter.ai/blog/announcements/fusion-beats-frontier/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3580c67f-125f-485f-9850-4859e9a0e5bb/image.png?t=1782192336"/></a></div><p class="paragraph" style="text-align:left;">The most annoying part of using multiple models is that you usually have to become the router yourself.</p><p class="paragraph" style="text-align:left;">You ask one model, compare it with another, maybe try a third, then manually decide which answer is right. </p><p class="paragraph" style="text-align:left;"><b>OpenRouter&#39;s new Fusion API</b> turns that pattern into a single call. You send a prompt to Fusion, it dispatches the task to a panel of models in parallel, gives them web search and web fetch, then uses a judge model to compare the answers before producing the final response.</p><p class="paragraph" style="text-align:left;">The results are worth paying attention to: Fable 5 + GPT-5.5 fused together scored 69.0% on DRACO, beating every individual model in OpenRouter&#39;s test, including Fable 5 alone at 65.3% and GPT-5.5 alone at 60.0%. That matters because Fable 5 is Anthropic&#39;s strongest model, and Fusion still found extra lift by pairing it with another frontier model instead of treating one model as the ceiling.</p><p class="paragraph" style="text-align:left;">OpenRouter also tested a budget panel with Gemini 3 Flash, Kimi K2.6, and DeepSeek V4 Pro. That panel scored 64.7%, beating GPT-5.5 and Claude Opus 4.8 individually, coming within about one point of Fable 5 alone, and doing it at roughly half the cost.</p><p class="paragraph" style="text-align:left;">The real story is not just the benchmark. It is that model diversity is becoming a product primitive. Instead of picking one model and hoping it is the right one, builders can start treating models like a small research team.</p><p class="paragraph" style="text-align:left;">You can use it through the normal OpenRouter API. Just call openrouter/fusion directly or configure the Fusion plugin with your own analysis models and judge model.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/addyosmani/status/2064127981161959567?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>Everyone is talking about loop engineering, but Addy&#39;s version makes it usable</b></a><br>Addy Osmani&#39;s piece is useful because it turns the phrase into an actual operating model. The shift is from &quot;I prompt the agent&quot; to &quot;I design the loop that finds work, hands it to agents, checks the result, records state, and decides the next step.&quot;</p><p class="paragraph" style="text-align:left;">That is a better frame for where coding agents are going. The prompt is no longer the main artifact. The loop is. If you are building agent workflows, this gives you a cleaner way to think about retries, memory, evaluation, escalation, and token cost before you wire everything together.</p><p class="paragraph" style="text-align:left;"> </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://sakana.ai/fugu/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>Sakana Fugu explores orchestration as the model</b></a><br>Sakana&#39;s Fugu is interesting because it is less about launching another standalone model and more about coordinating multiple models into a stronger system. The bet is that intelligence can come from routing, combining, challenging, and arbitrating models, not only from scaling one model in isolation.</p><p class="paragraph" style="text-align:left;">That makes it rhyme with the Fusion story, but from a research direction rather than an API product. The useful takeaway for builders is simple: the next frontier may be systems that know which model to use, when to ask for disagreement, and how to merge partial answers without making the user manage the whole process.</p><p class="paragraph" style="text-align:left;"> </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://telegram.org/blog/ai-bot-revolution-11-new-features?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>Telegram bots can now talk to other bots</b></a><br>Telegram&#39;s latest bot update lets bots respond to other bots, not just humans. That sounds small, but it changes what Telegram can be used for: not just a chat UI, but a lightweight coordination layer for agent workflows.</p><p class="paragraph" style="text-align:left;">You could mention one bot, that bot could hand work to another bot, and the whole exchange stays visible in a normal chat thread.</p><p class="paragraph" style="text-align:left;"> </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://unsloth.ai/docs/models/glm-5.2?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>Run GLM-5.2 locally with Unsloth</b></a><br>Unsloth just published a guide for running GLM-5.2 with Dynamic GGUFs, including 1-bit and 2-bit quant options, llama.cpp instructions, and Unsloth Studio support. The 2-bit build is still huge at around 239GB, but that is dramatically smaller than the full 1.51TB model.</p><p class="paragraph" style="text-align:left;">This is not casual laptop territory, but it is meaningful for local-agent builders with serious memory available. A 744B-parameter open model with 40B active parameters and a 1M context window is already being squeezed into setups that advanced users can actually experiment with.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://birdclaw.sh/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">Birdclaw</a></b>: A local-first Twitter workspace that imports your X archive, syncs timeline/bookmarks/mentions, and stores everything in SQLite. The useful part is that your X memory becomes searchable and agent-readable: you can full-text search old likes and bookmarks, triage mentions with AI ranking, generate local digests, and keep a Git-friendly backup instead of losing everything inside the platform UI.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.agent-native.com/templates/clips?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">Agent-Native Clips</a></b>: An open-source Loom + Granola + Wisprflow-style app for screen recordings, meeting notes, and dictation. Every clip gets transcripts, summaries, timestamped frames, and searchable history, so an agent can understand what happened in a video or meeting without needing raw audio/video ingestion.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://docs.stripe.com/directory?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow"><b>Stripe Directory</b></a>: A public-preview Stripe CLI directory for discovering businesses and services on the Stripe network. Developers and agents can search providers by keyword, then get structured results for Stripe Apps, <a class="link" href="https://Projects.dev?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">Projects.dev</a> providers, machine-payment endpoints, and business profiles instead of manually hunting across docs and marketplaces.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> (113k+ </b>🌟 <b>) </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=opus-4-8-level-model-now-runs-locally-for-free"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=a7bd1c91-2f6a-4dcb-bced-c715cceda390&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Claude Code now spins up 100s of parallel agents on one task</title>
  <description>+ Apple drops a native AI framework for on-deivce AI agents</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5ab91e0a-f42b-49b2-9cbd-13cd544f9fe8/Claude_Code_now_spins_up_100s_of_parallel_agents_on_a_task.png" length="2323869" type="image/png"/>
  <link>https://www.theunwindai.com/p/claude-code-now-spins-up-100s-of-parallel-agents-on-one-task</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/claude-code-now-spins-up-100s-of-parallel-agents-on-one-task</guid>
  <pubDate>Tue, 09 Jun 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-06-09T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Apple Core AI Framework for on-device agents</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Printing Press: Print agent-native CLIs from a single prompt</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Claude Code Dynamic Workflows with massive parallelism</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Your job is to write agent loops now</b></p></li><li><p class="paragraph" style="text-align:left;"><b>ChatGPT now &quot;dreams&quot; to build better memory</b></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/generative-ui-is-the-new-frontend?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow">Generative UI Is the New Frontend</a></b></p><p class="paragraph" style="text-align:left;">The frontend used to be a fixed thing. Designers drew it. Engineers built it. Users got what shipped.</p><p class="paragraph" style="text-align:left;">That&#39;s over.</p><p class="paragraph" style="text-align:left;">The interfaces shipping in 2026 are drawn partly by the agent itself, in real time, from what the user actually asked for. Ask for a table, get a table. Not a paragraph describing one.</p><p class="paragraph" style="text-align:left;">Generative UI is the layer that lets agents stop describing and start showing. </p><p class="paragraph" style="text-align:left;">This guide walks you through three patterns that have emerged on how to build it, and the differences between them matter more than most teams realize.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/generative-ui-is-the-new-frontend?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/6c3faaf3-c6cf-4f53-806b-8a114703b66a/Generative_UI_Is_the_New_Frontend.png?t=1780532864"/><div class="embed__content"><p class="embed__title"> Generative UI Is the New Frontend </p><p class="embed__description"> How AI agents stop describing and start showing </p></div></a></div><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://developer.apple.com/documentation/coreai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>Apple Just Gave Developers Their Own On-Device AI Framework</b></a></h3><div class="image"><a class="image__link" href="https://developer.apple.com/documentation/coreai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a3bdfe49-348d-44ee-92d5-4d60fc268372/Screenshot_2026-06-08_at_11.04.15_PM.png?t=1780985059"/></a></div><p class="paragraph" style="text-align:left;">Forget calling external APIs. Apple&#39;s new Core AI framework, announced at WWDC 2026, gives developers Swift-native access to Apple&#39;s on-device foundation models with tool calling, structured generation, and full Apple Intelligence integration.</p><p class="paragraph" style="text-align:left;">This is the first time Apple has opened up its on-device models as a developer-facing framework. You write Swift, define tools, and the model runs locally on the device with zero cloud round-trips. Privacy-first by default, no API keys, no usage-based pricing, no latency from network calls. For anyone building iOS or macOS apps, this changes how you think about adding intelligence to your product.</p><p class="paragraph" style="text-align:left;">The bigger picture: Apple also revealed that Apple Intelligence is now co-developed with Google using Gemini models under the hood, running both on-device and through Private Cloud Compute. A system orchestrator automatically coordinates AI features across apps.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Tool calling in Swift</b>: Define custom tools that the on-device model can invoke, enabling agentic workflows entirely on the user&#39;s device without a server.</p></li><li><p class="paragraph" style="text-align:left;"><b>Structured generation</b>: Get typed, schema-conforming outputs from the model, not just raw text. Build reliable features without post-processing hacks.</p></li><li><p class="paragraph" style="text-align:left;"><b>Gemini under the hood</b>: Apple Intelligence now runs on foundation models co-developed with Google, giving the platform multimodal capabilities, including image generation, visual Q&A, and speech generation.</p></li><li><p class="paragraph" style="text-align:left;"><b>No cloud dependency</b>: Models run locally. Your users&#39; data stays on their devices. No API costs, rate limits, or cold starts.</p></li><li><p class="paragraph" style="text-align:left;"><b>Available now</b>: Core AI ships with iOS 27, macOS 27, and the latest Xcode. Documentation is live at <a class="link" href="https://developer.apple.com/documentation/coreai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow">developer.apple.com/documentation/coreai</a>.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://printingpress.dev/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>Print an Agent-Native CLI for Any API from a Single Prompt</b></a></h3><div class="image"><a class="image__link" href="https://printingpress.dev/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/0852d831-0bca-4caf-88ce-0db19cb32ce2/printing_press.png?t=1780984943"/></a></div><p class="paragraph" style="text-align:left;">What if every app, API, and website your agent needs came as a purpose-built CLI with a local SQLite mirror, compound commands, and token-efficient output?</p><p class="paragraph" style="text-align:left;">That&#39;s Printing Press by Matt Van Horn and Trevin Chow. Point it at an API spec, a website URL, or even a service with no public API, and it generates a Go CLI, a Claude Code skill, an OpenClaw skill, and an MCP server. All from one prompt.</p><p class="paragraph" style="text-align:left;">Super interesting concept: a local SQLite mirror beats a remote API call. Compound commands beat ten round trips. An agent-native CLI beats raw HTTP. When you &quot;print&quot; an ESPN CLI, you don&#39;t get a thin wrapper around REST endpoints. You get live scores, series state, leading scorers, and injury news in one call, all queried from a local database that syncs incrementally. Same goes for Linear, Slack, Notion, or any of the 237+ CLIs in their Public Library.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>No API needed</b>: For services without a public API, Printing Press launches a browser, captures traffic, reverse-engineers the endpoints, and generates the spec automatically. If you can click through it, the press can build a CLI.</p></li><li><p class="paragraph" style="text-align:left;"><b>Local-first data layer</b>: High-gravity resources get domain-specific SQLite tables with FTS5 full-text search and incremental sync. Queries run in milliseconds offline. Your agent never waits for a 429.</p></li><li><p class="paragraph" style="text-align:left;"><b>237+ community CLIs</b>: The Public Library ships pre-built CLIs across 19 categories, from flight search to restaurant reservations to eBay auctions. Install with one command or let your agent browse and pick what it needs.</p></li><li><p class="paragraph" style="text-align:left;"><b>Token-efficient by default</b>: --compact mode cuts 60-80% of tokens. Auto-JSON when piped. Typed exit codes for agent self-correction. The CLI is built for agents first, humans second.</p></li><li><p class="paragraph" style="text-align:left;"><b>Try it now</b>: Install via Go, add the Claude Code or OpenClaw skills, and run /printing-press &lt;app&gt; inside your agent. Check it out at <a class="link" href="https://printingpress.dev?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow">printingpress.dev</a>.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://claude.com/blog/introducing-dynamic-workflows-in-claude-code?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>Claude Code Can Now Orchestrate 100s of Parallel Agents on a Single Task</b></a></h3><div class="image"><a class="image__link" href="https://claude.com/blog/introducing-dynamic-workflows-in-claude-code?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/fac16e46-09a2-43cb-82d9-df69753554b3/image.png?t=1780985221"/></a></div><p class="paragraph" style="text-align:left;">Some problems are too big for one agent in one pass. A bug hunt across an entire service. A migration that touches hundreds of files. A plan you want stress-tested from every angle before committing.</p><p class="paragraph" style="text-align:left;">Anthropic&#39;s Dynamic Workflows for Claude Code changes the math entirely. Claude writes a custom JavaScript orchestration script on the fly, fans work out across 100s of parallel subagents, has independent agents try to break each other&#39;s results, and keeps iterating until answers converge. The coordination lives in code, not context, so the plan stays on track no matter how big the task gets.</p><p class="paragraph" style="text-align:left;">The proof of concept is wild: Jarred Sumner used Dynamic Workflows to port Bun from Zig to Rust. 750,000 lines of Rust, 99.8% test suite passing, eleven days from first commit to merge. Hundreds of agents worked in parallel with two reviewers on each file.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Claude writes the orchestrator</b>: No pre-built templates. Claude generates a bespoke JS script tailored to your specific task, then runs it. Every workflow is custom.</p></li><li><p class="paragraph" style="text-align:left;"><b>Independent verification built in</b>: Agents tackle the problem from different angles, other agents try to refute what they found, and the run iterates until results converge. This is how it catches things a single pass misses.</p></li><li><p class="paragraph" style="text-align:left;"><b>Resumable and saveable</b>: Progress is checkpointed. Interrupted jobs pick up where they left off. Save a workflow as a reusable /command for future sessions with structured input parameters.</p></li><li><p class="paragraph" style="text-align:left;"><b>Token warning</b>: These workflows consume meaningfully more tokens than a typical session. Start with a scoped task to get a feel for usage before throwing it at your whole codebase.</p></li><li><p class="paragraph" style="text-align:left;"><b>Available now</b>: Research preview on Max, Team, and Enterprise plans, plus the Claude API, Amazon Bedrock, Vertex AI, and Microsoft Foundry. Requires Claude Code v2.1.154+. Turn on ultracode effort level or just ask Claude to &quot;create a workflow.&quot;</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/mvanhorn/article/2063865685558903149?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>Don’t Prompt Agent, Design Loops that prompt your agent</b></a><b> (WTF!)</b><br>Peter Steinberger posted six words on Saturday that hit 6.3 million views: &quot;You should be designing loops that prompt your agents.&quot; Boris Cherny, the creator of Claude Code, said the same thing a few days earlier: &quot;I don&#39;t prompt Claude anymore. I have loops running. They&#39;re the ones prompting Claude.&quot; If all of this chatter left you wondering what the hell a loop even is, Matt Van Horn wrote the definitive explainer. He traces the concept all the way back to the 2022 ReAct paper, through Geoffrey Huntley&#39;s ralph loop, to today&#39;s multi-agent orchestration loops that run on cron and survive restarts. Worth the read.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.theverge.com/tech/944245/apple-wwdc-2026-ai-siri-gemini?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>Apple Intelligence, take two</b></a><br>Remember when Apple announced Apple Intelligence with ChatGPT as the backup brain at WWDC 2024? The &quot;smart Siri&quot; with personal context, on-screen awareness, etc? Most of it never shipped, and apparently, “it wasn&#39;t good enough&quot; and &quot;didn&#39;t converge quality-wise.&quot; Two years later, they&#39;re trying again, this time with Google. The new Apple Intelligence, announced at WWDC yesterday, is co-developed with Google using Gemini as the foundation, not just a fallback. On-device and Private Cloud Compute, multimodal everything, and a conversational Siri that&#39;s getting its own standalone app. </p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="http://openai.com/index/chatgpt-memory-dreaming/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>ChatGPT now &quot;dreams&quot; to remember you better</b></a><br>OpenAI shipped Dreaming V3, and the name is apt. ChatGPT now runs a background memory synthesis process when you&#39;re not chatting, analyzing your conversation history and building a unified &quot;Memory Summary&quot; that stays current over time. The old &quot;saved memories&quot; approach (manually saying &quot;remember this&quot;) is gone. Now it automatically captures context from natural conversation and updates temporal facts, so it knows your Singapore trip is in the past, not upcoming. Available on Plus and Pro now, rolling out to free users soon after a 5x compute reduction made it feasible at scale. The direction is clear: persistent, stateful AI assistants where memory is infrastructure, not a feature you toggle on.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="http://research.google/blog/unlocking-dependable-responses-with-gemini-enterprise-agent-platforms-agentic-rag/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>Google Research tackles RAG&#39;s biggest problem</b></a><br>Standard RAG retrieves once and hopes for the best. Google Research&#39;s new &quot;Agentic RAG&quot; for their Gemini Enterprise Agent Platform retrieves, checks if it got enough, and goes back for more. The key innovation is a &quot;Sufficient Context Agent&quot; that inspects retrieved snippets, evaluates a draft response, identifies exactly what&#39;s missing, and sends targeted follow-up searches. It&#39;s a multi-agent pipeline: orchestrator, planner, query rewriter, search fanout, and synthesis. The result is a 34% accuracy improvement over standard RAG on factuality benchmarks with negligible latency overhead. The pattern is the real takeaway here, even if you&#39;re not on Google Cloud.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/Panniantong/Agent-Reach?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>Agent-Reach</b></a>: Deploy one AI agent across Telegram, Discord, Slack, WhatsApp, Web, and CLI simultaneously from a single codebase. It normalizes messages into a consistent format per platform and auto-adapts responses. Ships with ready-made adapters for LangChain, CrewAI, and AutoGen.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/luongnv89/claude-howto?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>claude-howto</b></a>: A visual, example-driven guide to Claude Code covering everything from basic setup to advanced workflows like multi-agent orchestration and custom skills. Super useful whether you&#39;re just starting with Claude Code or trying to level up. </p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/mvanhorn/last30days-skill?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>last30days-skill</b></a>: An agent skill that synthesizes research across X, Reddit, HN, YouTube, TikTok, and GitHub from the last 30 days on any topic you give it. Matt Van Horn used it to write that viral loops article, running it against the word everyone was fighting about. </p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/mvanhorn/agentcookie?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow"><b>agentcookie</b></a>: Continuously syncs your browser cookies, bearer tokens, and API keys from your daily-driver Mac to a second Mac where your agents run, encrypted over Tailscale. Your agents wake up authenticated to every service you use, zero per-site login ceremony, no cloud middleman. </p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> (113k+ </b>🌟 <b>) </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=claude-code-now-spins-up-100s-of-parallel-agents-on-one-task"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=4fe39004-4fe4-4eb2-90b9-f63062e2ff13&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Generative UI Is the New Frontend </title>
  <description>How AI agents stop describing and start showing</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/6c3faaf3-c6cf-4f53-806b-8a114703b66a/Generative_UI_Is_the_New_Frontend.png" length="1328478" type="image/png"/>
  <link>https://www.theunwindai.com/p/generative-ui-is-the-new-frontend</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/generative-ui-is-the-new-frontend</guid>
  <pubDate>Thu, 04 Jun 2026 00:28:39 +0000</pubDate>
  <atom:published>2026-06-04T00:28:39Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <category><![CDATA[Ai Blogs]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">The frontend used to be a fixed thing. Designers drew it. Engineers built it. Users got what shipped.</p><p class="paragraph" style="text-align:left;">That&#39;s over. </p><p class="paragraph" style="text-align:left;">The interfaces shipping in 2026 are drawn partly by the agent itself, in real time, from what the user actually asked for. Ask for a table, get a table. Not a paragraph describing one.</p><p class="paragraph" style="text-align:left;">Generative UI is the layer that lets agents stop describing and start showing. Three patterns have emerged for how to build it, and the differences between them matter more than most teams realize.</p><p class="paragraph" style="text-align:left;">But there isn&#39;t one way to build this. There are three. And most teams pick one without knowing they chose.</p><h2 class="heading" style="text-align:left;"><b>The protocol stack</b></h2><p class="paragraph" style="text-align:left;">Three protocols. Each does one job.</p><p class="paragraph" style="text-align:left;"><b>MCP</b> connects agents to tools. <b>A2A</b> connects agents to each other. <b>AG-UI</b> connects agents to users.</p><p class="paragraph" style="text-align:left;"><b>AG-UI </b>is the streaming layer that carries everything you&#39;ll see below: tool calls, A2UI schemas, MCP App events, state deltas. Runs over SSE. State flows both ways on the same stream. User edits, agent sees. Agent mutates, user sees.</p><p class="paragraph" style="text-align:left;"><b>A2UI</b> is Google&#39;s spec for agents emitting UI as schema. It rides on AG-UI. CopilotKit ships it in production.</p><p class="paragraph" style="text-align:left;">You don&#39;t write a parser for any of this. CopilotKit is an AG-UI client and decodes the stream for you.</p><h2 class="heading" style="text-align:left;"><b>The three patterns most teams confuse</b></h2><p class="paragraph" style="text-align:left;">Ask ten developers what Generative UI is. You get ten answers. Most of them are describing whichever pattern their current framework ships.</p><p class="paragraph" style="text-align:left;">There are just three. The spectrum runs from more control to more flexibility.</p><ul><li><p class="paragraph" style="text-align:left;"><b>Controlled:</b> You pre-build the components. The agent picks which to render.</p></li><li><p class="paragraph" style="text-align:left;"><b>Declarative:</b> The agent emits a schema. Your app maps it to components.</p></li><li><p class="paragraph" style="text-align:left;"><b>Open-ended:</b> The agent writes raw HTML. Your app renders it in a sandbox.</p></li></ul><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e36bdb41-766b-4733-b369-a32f63087161/image.png?t=1780531905"/></div><p class="paragraph" style="text-align:left;">Every Gen UI framework in 2026 lives somewhere on this line. The differences are architectural, not cosmetic. Each pattern breaks your app in a different way at scale.</p><p class="paragraph" style="text-align:left;">I tried different stacks. Most cover one pattern well. Landed on CopilotKit because it supports all three on the same runtime, riding AG-UI. That&#39;s the stack everything below runs on.</p><h2 class="heading" style="text-align:left;"><b>Pattern 1: Controlled, frontend owns the UI</b></h2><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/03ae9963-4371-45ed-94ae-1fb876e8e803/image.png?t=1780531917"/></div><p class="paragraph" style="text-align:left;">This is where most teams start. It&#39;s also where most teams get stuck.</p><p class="paragraph" style="text-align:left;">You pre-build a React component. You bind it to a tool name. The agent picks that tool and the component renders inline in chat with the agent&#39;s args as props.</p><p class="paragraph" style="text-align:left;">One frontend hook. Zero agent code. That&#39;s it.</p><div class="codeblock"><pre><code>&quot;use client&quot;;
import &#123; z &#125; from &quot;zod&quot;;
import &#123; useComponent &#125; from &quot;@copilotkit/react-core/v2&quot;;

const expenseChartSchema = z.object(&#123;
  title: z.string(),
  data: z.array(z.object(&#123; label: z.string(), value: z.number() &#125;)),
&#125;);

function ExpenseChart(&#123; title, data &#125;: z.infer&lt;typeof expenseChartSchema&gt;) &#123;
  return (
    &lt;section className=&quot;rounded-xl border p-4&quot;&gt;
      &lt;h3 className=&quot;text-sm font-medium&quot;&gt;&#123;title&#125;&lt;/h3&gt;
      &lt;ul className=&quot;mt-2 grid gap-1&quot;&gt;
        &#123;data.map((d) =&gt; (
          &lt;li key=&#123;d.label&#125; className=&quot;flex justify-between text-sm&quot;&gt;
            &lt;span&gt;&#123;d.label&#125;&lt;/span&gt;
            &lt;span&gt;$&#123;d.value&#125;&lt;/span&gt;
          &lt;/li&gt;
        ))&#125;
      &lt;/ul&gt;
    &lt;/section&gt;
  );
&#125;

export function ExpensesCopilot() &#123;
  useComponent(&#123;
    name: &quot;showExpenseChart&quot;,
    description: &quot;Render a breakdown of expenses by category.&quot;,
    parameters: expenseChartSchema,
    render: ExpenseChart,
  &#125;);

  return null;
&#125;</code></pre></div><p class="paragraph" style="text-align:left;">The hook registers the tool with CopilotKit&#39;s runtime. The runtime advertises it to the agent over AG-UI. When the agent calls it, the args stream in and your component renders inline. No Python tool to write, no schema to wire, no API route to add.</p><p class="paragraph" style="text-align:left;">Your design system stays in charge.</p><p class="paragraph" style="text-align:left;">That expense chart isn&#39;t a mockup. The <b><span style="text-decoration:underline;"><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents/ai-financial-coach-agent?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #0f1419">AI Financial Coach Agent</a></span></b> renders cards just like it for real budgets, savings plans, and debt payoff.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a472bcaf-c810-4942-bdba-2e62ae865ef5/ezgif-745fa91cbc1343a4.gif?t=1780532184"/></div><p class="paragraph" style="text-align:left;">Want the bare hook first? It&#39;s &#39;<i>use-generative-ui-examples.tsx&#39;</i> in the <span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents/generative-ui-starter-project?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #0f1419">Generative UI Starter Project</a></b></span>.</p><p class="paragraph" style="text-align:left;"><b>The token tax</b></p><p class="paragraph" style="text-align:left;">Every component you register sits in the agent&#39;s context window before the user has said anything. A typical tool description with its JSON schema runs around 400 tokens. 25 components are 10,000 tokens on every turn. You pay that tax per request.</p><p class="paragraph" style="text-align:left;">The agent picks the wrong component too. Too many look similar. Pie chart and donut chart both &quot;show proportions.&quot; It guesses.</p><p class="paragraph" style="text-align:left;"><b>When to add agent-side state</b></p><p class="paragraph" style="text-align:left;">Shared state is the one case where writing a Python tool is worth it. The agent writes to session state. Other parts of the UI subscribe and re-render with no second LLM call. Pin a metric, the dashboard updates. Add a row, the table redraws.</p><div class="codeblock"><pre><code>from google.adk.agents import LlmAgent
from google.adk.tools import ToolContext

def pin_metric(tool_context: ToolContext, label: str, value: float) -&gt; dict:
    &quot;&quot;&quot;Pin a metric to the user&#39;s dashboard.&quot;&quot;&quot;
    pinned = tool_context.state.get(&quot;pinnedMetrics&quot;, [])
    tool_context.state[&quot;pinnedMetrics&quot;] = pinned + [&#123;&quot;label&quot;: label, &quot;value&quot;: value&#125;]
    return &#123;&quot;status&quot;: &quot;pinned&quot;&#125;

agent = LlmAgent(name=&quot;dashboard_agent&quot;, model=&quot;gemini-3.5-flash&quot;, tools=[pin_metric])</code></pre></div><p class="paragraph" style="text-align:left;">The frontend reads pinned metrics through CopilotKit&#39;s shared-state hook. The chat component still renders inline because the same tool name is wired with the frontend hook. </p><p class="paragraph" style="text-align:left;">Pin a metric in chat. The panel redraws with no second model call. That&#39;s the <span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents/ai-dashboard-canvas-agent?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #0f1419">AI Dashboard Canvas Agent</a></b></span>.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/50e2bf8b-cc05-4c92-969b-31d36decce7c/image.png?t=1780532080"/></div><p class="paragraph" style="text-align:left;">The <b><span style="text-decoration:underline;"><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents/ai-deep-research-agent?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #0f1419">AI Deep Research Agent</a></span></b> takes it further. The plan, every search, each file write, all of it streams in as live cards. For everything else, the frontend hook is the whole story.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/c963401e-9442-4d0c-8bb8-4d2095ae93b7/ezgif-73e872c02cea0875.gif?t=1780532486"/></div><p class="paragraph" style="text-align:left;"><b>When to ship Controlled:</b> Ten or fewer high-value flows. Design precision matters. You know the exact UIs you need.</p><p class="paragraph" style="text-align:left;"><b>When not to: </b>Your codebase grows linearly with use cases. 25 components means 25 tool definitions sitting in every agent turn.</p><p class="paragraph" style="text-align:left;"><b>What breaks:</b> Agent picks the wrong component. Two tool descriptions overlap semantically. Past 15 tools, two of them probably read like &quot;displays data.&quot; Fix: rewrite descriptions to name the user intent, not the visual. &quot;Use when the user asks to compare proportions of a whole&quot; beats &quot;renders a pie chart.&quot;</p><h2 class="heading" style="text-align:left;"><b>Pattern 2: Declarative (A2UI), agent emits schema</b></h2><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f4e7cdf6-e242-4e1f-9d61-71838d4583e4/image.png?t=1780532206"/></div><p class="paragraph" style="text-align:left;">This is the pattern most production agent apps end up needing.</p><p class="paragraph" style="text-align:left;">The agent emits a JSON schema describing the UI. Your app has a catalog of components that maps schema nodes to React (or Svelte, Flutter, anything). One tool. Many UIs.</p><p class="paragraph" style="text-align:left;">A2UI is the standard spec. CopilotKit ships the runtime. ADK runs the agent. AG-UI is the wire.</p><p class="paragraph" style="text-align:left;">The agent tool returns three operations in order: create a surface, push the component tree, push the data.</p><div class="codeblock"><pre><code>def search_flights(flights: list[Flight]) -&gt; dict[str, Any]:
    &quot;&quot;&quot;Search flights and display them as rich cards.&quot;&quot;&quot;
    return &#123;
        &quot;a2ui_operations&quot;: [
            &#123;&quot;type&quot;: &quot;create_surface&quot;, &quot;surfaceId&quot;: SURFACE_ID, &quot;catalogId&quot;: CATALOG_ID&#125;,
            &#123;&quot;type&quot;: &quot;update_components&quot;, &quot;surfaceId&quot;: SURFACE_ID, &quot;components&quot;: FLIGHT_SCHEMA&#125;,
            &#123;&quot;type&quot;: &quot;update_data_model&quot;, &quot;surfaceId&quot;: SURFACE_ID, &quot;data&quot;: &#123;&quot;flights&quot;: flights&#125;&#125;,
        ]
    &#125;</code></pre></div><p class="paragraph" style="text-align:left;">The component tree above lives in flights.json. You wrote it. The agent only fills in the data. That&#39;s a fixed schema.</p><p class="paragraph" style="text-align:left;">Dynamic schema flips it: a secondary LLM writes the component tree per turn from conversation context. Same a2ui_operations container at the end. The Google ADK showcase ships both.</p><p class="paragraph" style="text-align:left;"><b>The catalog is the contract</b></p><p class="paragraph" style="text-align:left;">Definitions list the components the agent is allowed to emit, with Zod schemas for the props. Renderers fill in React. Typos become build errors instead of blank screens.</p><div class="codeblock"><pre><code>const renderers: CatalogRenderers&lt;TravelDefinitions&gt; = &#123;
  FlightCard: (&#123; props &#125;) =&gt; (
    &lt;article className=&quot;rounded-xl border p-4&quot;&gt;
      &lt;header className=&quot;flex justify-between&quot;&gt;
        &lt;span&gt;&#123;(props as any).airline&#125;&lt;/span&gt;
        &lt;span&gt;&#123;(props as any).price&#125;&lt;/span&gt;
      &lt;/header&gt;
      &lt;div className=&quot;text-sm text-muted-foreground&quot;&gt;
        &#123;(props as any).origin&#125; → &#123;(props as any).destination&#125; · &#123;(props as any).departureTime&#125;
      &lt;/div&gt;
    &lt;/article&gt;
  ),
&#125;;

export const travelCatalog = createCatalog(travelDefinitions, renderers, &#123;
  catalogId: &quot;copilotkit://travel-catalog&quot;,
  includeBasicCatalog: true,
&#125;);</code></pre></div><p class="paragraph" style="text-align:left;">Both halves live in the <span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents/generative-ui-starter-project?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #0f1419">Generative UI Starter Project</a></b></span>, wired and matched. search_flights in <i>&#39;a2ui_fixed_</i><i><a class="link" href="https://schema.py?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow">schema.py</a></i><i>&#39;</i>, the FlightCard catalog in <i>&#39;renderers.tsx</i>&#39;. Ask for flights. Watch the cards stream into chat.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/02ac044e-0776-44a2-a5f6-24d8fc2567d4/ezgif-7659c7a92e59344a.gif?t=1780532462"/></div><p class="paragraph" style="text-align:left;">Buttons and other interactive components carry an action in the schema. The basic catalog wires it to onClick. Click fires an event back to the agent over AG-UI. The agent decides what to render next. Zero click handlers.</p><p class="paragraph" style="text-align:left;"><b>The token math</b></p><p class="paragraph" style="text-align:left;">50 card types or 500, the agent sees one function. Tokens per turn stay flat as your component library grows.</p><p class="paragraph" style="text-align:left;">Extensible to any rendering framework because it&#39;s just JSON. Any agent that already speaks AG-UI can drive A2UI on day zero. You don&#39;t touch agent code to wire this up.</p><p class="paragraph" style="text-align:left;"><b>Trade-off: </b>The LLM owns the layout. Output varies run to run within your catalog. If you&#39;re shipping legal disclosures, marketing surfaces, or anything where exact pixel placement matters, this is not your bucket.</p><p class="paragraph" style="text-align:left;">Declarative is the pattern built for the long tail. Dashboards, results, forms, cards, widgets.</p><p class="paragraph" style="text-align:left;"><b>When to ship Declarative:</b> You have more use cases than time to pre-build. You care about token economics past the prototype stage.</p><p class="paragraph" style="text-align:left;"><b>What breaks:</b> Built a custom FlightCard. Every flight renders as the basic catalog&#39;s generic card. No error in the console. The CATALOG_ID on the agent and catalogId in createCatalog on the frontend don&#39;t match. Frontend doesn&#39;t recognize the catalog the agent is targeting, falls back to basic. Match the strings exactly on both sides.</p><h2 class="heading" style="text-align:left;"><b>Pattern 3: Open-ended, no catalog, no rules</b></h2><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/23782a72-b59b-4fa7-9418-4c7b74e49791/image.png?t=1780532583"/></div><p class="paragraph" style="text-align:left;">The third pattern is the opposite extreme. No catalog. No schema. Just a blank canvas.</p><p class="paragraph" style="text-align:left;">Two sub-patterns live in this bucket.</p><p class="paragraph" style="text-align:left;"><b>MCP Apps</b></p><p class="paragraph" style="text-align:left;">An MCP server exposes UI surfaces that the agent drives. Excalidraw is the example that stuck with me. The agent gets full control of the canvas. Draws diagrams from your context. Owns every pixel on the board.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/bfa32d2f-d7f5-4248-88fa-bef5c2cf1f7f/image.png?t=1780532597"/></div><p class="paragraph" style="text-align:left;">Implementing the client protocol from scratch is painful, so CopilotKit ships an <i>MCPAppsMiddleware</i>. Attach it to your agent and point it at any MCP Apps server.</p><div class="codeblock"><pre><code>const agent = new BuiltInAgent(&#123;
  model: &quot;openai/gpt-5.5&quot;,
  prompt: &quot;You are a helpful assistant.&quot;,
&#125;).use(
  new MCPAppsMiddleware(&#123;
    mcpServers: [&#123; type: &quot;http&quot;, url: &quot;https://mcp.excalidraw.com/mcp&quot;, serverId: &quot;my-server&quot; &#125;],
  &#125;),
);</code></pre></div><p class="paragraph" style="text-align:left;">Spin up the <span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents/mcp-apps-generative-ui-showcase?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #0f1419">MCP Apps Showcase</a></b></span><span style="text-decoration:underline;"><b> </b></span>and you&#39;re booking flights and reserving hotels inside the chat window. Same middleware, real MCP servers. Or go further.</p><p class="paragraph" style="text-align:left;">The <span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents/ai-mcp-app-builder?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #0f1419">AI MCP App Builder</a></b></span><span style="text-decoration:underline;"><b> </b></span>lets the agent write a brand-new app into an E2B sandbox, then renders it live.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2f1d5956-4540-4697-a36e-fbcf2bafd0f9/WEGGWcRPy3RhPaAV.jpg?t=1780531876"/></div><p class="paragraph" style="text-align:left;"><b>Sandboxed HTML</b></p><p class="paragraph" style="text-align:left;">The agent writes raw HTML. Your app renders it inside a sandboxed iframe so it can&#39;t hijack the session.</p><p class="paragraph" style="text-align:left;">The runtime registers an HTML rendering tool and ships it to the agent over AG-UI. The agent calls it with whatever markup it wants. There is no HTML tool to define on the agent side. The runtime injects it.</p><p class="paragraph" style="text-align:left;">Agent-side instruction is doing real work:</p><div class="codeblock"><pre><code>canvas_agent = LlmAgent(
    name=&quot;canvas_agent&quot;,
    model=&quot;gemini-3.5-flash&quot;,
    instruction=(
        &quot;You are a visualization assistant. When the user asks to see, &quot;
        &quot;draw, or visualize anything, generate an interactive HTML UI. &quot;
        &quot;Use Tailwind classes only. No external fonts. Stick to neutral &quot;
        &quot;colors unless the user names one.&quot;
    ),
)</code></pre></div><p class="paragraph" style="text-align:left;">Without those style rules, the model defaults to whatever aesthetic was loudest in its training data that week. With them, you get something close to your brand most of the time. Not always.</p><p class="paragraph" style="text-align:left;"><b>The brand inconsistency problem</b></p><p class="paragraph" style="text-align:left;">I tried shipping Open-ended as the primary UI for an agent. Pulled it in a week.</p><p class="paragraph" style="text-align:left;">&quot;Neo-brutalist&quot; on Tuesday. &quot;iOS 4 clone&quot; on Wednesday. Style rules in the prompt nudge the agent toward your brand. They don&#39;t guarantee it. The brand kept changing. The product felt unserious.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/bb2c6e2f-ebce-4d7d-b3e8-e8555f97e51d/image.png?t=1780532628"/></div><p class="paragraph" style="text-align:left;">Open-ended isn&#39;t useless. It&#39;s misapplied.</p><p class="paragraph" style="text-align:left;">Right call for one thing: throwaway interactions where the user doesn&#39;t care what the interface looks like and will never see it again. &quot;Show me how electrons work.&quot; &quot;Give me a weird bar chart of my last 10 queries.&quot; &quot;Visualize this API response.&quot; The kind of thing you see in Google AI overviews.</p><p class="paragraph" style="text-align:left;"><b>When to ship Open-ended:</b> One-shot queries. Disposable visualizations. Sandboxed experiments. Never as the primary surface.</p><p class="paragraph" style="text-align:left;"><b>What breaks:</b> The iframe renders. Buttons don&#39;t click. Forms don&#39;t submit. Sandbox flags are too tight, or too loose in a way the browser refuses. Set the iframe sandbox to allow scripts and allow forms. Nothing else. Never allow-same-origin.</p><h2 class="heading" style="text-align:left;"><b>How to pick</b></h2><p class="paragraph" style="text-align:left;">Run the decision tree before you write code.</p><p class="paragraph" style="text-align:left;">Designer has pixel-perfect mockups for this flow? Controlled.</p><p class="paragraph" style="text-align:left;">Dozens of card types or widgets to ship? Declarative.</p><p class="paragraph" style="text-align:left;">One-shot, throwaway visualization the user will never see twice? Open-ended.</p><p class="paragraph" style="text-align:left;">Can&#39;t decide? Default to Declarative. Upgrade to Controlled for the top 3 flows. Never Open-ended as the default.</p><p class="paragraph" style="text-align:left;">If you&#39;re already shipping and not sure where you landed, count the render tools. Past 15, you&#39;re in Controlled and the wall is close. Start wiring A2UI this week.</p><h2 class="heading" style="text-align:left;"><b>Three patterns. Three bets.</b></h2><p class="paragraph" style="text-align:left;">Controlled bets on you. Pre-built components, pixel-perfect. Expensive past 25 of them.</p><p class="paragraph" style="text-align:left;">Declarative bets on the schema. The schema is the contract. The agent fills it in. Scales flat.</p><p class="paragraph" style="text-align:left;">Open-ended bets on the model. No catalog, no schema, raw HTML. Good for throwaway. Brittle for anything that ships twice.</p><p class="paragraph" style="text-align:left;">The mistake isn&#39;t picking the wrong pattern. It&#39;s not knowing you picked one.</p><p class="paragraph" style="text-align:left;">Most teams default to Controlled because the framework defaults to Controlled. They hit the wall at 25 components and reach for Open-ended because it looks compelling in demos. Neither was a decision. Both were drift.</p><p class="paragraph" style="text-align:left;">Pick on purpose. Match the pattern to the problem. Controlled for the flows that need to be exact. Declarative for the long tail. Open-ended for the disposable.</p><p class="paragraph" style="text-align:left;">🚨<b> </b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow"><b>Open Source Generative UI Agent Templates</b></a></p><p class="paragraph" style="text-align:left;">The reference for all three lives in the new <span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/generative_ui_agents?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #0f1419">Generative UI Agents</a></b></span><span style="text-decoration:underline;"><b> </b></span>section of awesome-llm-apps. Clone what you need. Rip out what you don&#39;t.</p><hr class="content_break"><p class="paragraph" style="text-align:left;">I&#39;ll be publishing more about shipping agents in production, AG-UI, and the patterns that scale. </p><p class="paragraph" style="text-align:left;"><b>Follow me </b><b><a class="link" href="https://twitter.com/Saboo_Shubham_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow">@Saboo_Shubham_</a></b><b> to stay tuned.</b></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">We share in-depth blogs and tutorials like this 2-3 times a week, to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;"><b>Don’t forget to share this tutorial on your social channels and tag Unwind AI (</b><span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span><b>) to support us!</b></p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=generative-ui-is-the-new-frontend"><span class="button__text" style=""> Subscribe now for FREE - Get instant access to more LLM, RAG & AI Agent tutorials </span></a></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=d832bac9-409e-44b7-81fe-e065a8760e9e&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>OpenAI Codex Can Now Ship Live Shareable Websites</title>
  <description>+ Hermes Agent goes native on your desktop</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/fa7fcb33-1bd9-42bd-8bc0-e405ff8dc08d/OpenAI_Codex_Can_Now_Ship_Live_Shareable_Websites.png" length="2599513" type="image/png"/>
  <link>https://www.theunwindai.com/p/openai-codex-can-now-ship-live-shareable-websites</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/openai-codex-can-now-ship-live-shareable-websites</guid>
  <pubDate>Wed, 03 Jun 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-06-03T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>OpenAI Codex Can Now Ship Live Websites</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Hermes Agent now has a desktop app</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Open-source Alternative to Exa Websets</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Microsoft launches 7 in-house MAI models at Build</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Open-source code review tool from the OpenClaw team</b></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/the-ultimate-guide-to-goal?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">The Ultimate Guide to /goal</a></b></p><p class="paragraph" style="text-align:left;">HTTP is a primitive. JSON is a primitive. /goal is becoming one for coding agents.</p><p class="paragraph" style="text-align:left;">A few weeks ago, OpenAI&#39;s Codex CLI added /goal as a way to give the coding worker a job with a defined done state. Claude Code added it this week.</p><p class="paragraph" style="text-align:left;">Hermes Agent, the orchestrator I run on a Mac Mini to coordinate work between coding workers, has had /goal built in for a while.</p><p class="paragraph" style="text-align:left;">This guide walks through what /goal actually is, the three roles in a multi-agent setup, a real end-to-end run, the verification rule, and how to run goals in parallel without workers stepping on each other. </p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/the-ultimate-guide-to-goal?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Read The Ultimate Guide to /goal</a></b></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://openai.com/index/codex-for-every-role-tool-workflow/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow"><b>OpenAI Codex Can Now Ship Live Websites</b></a></h3><div class="image"><a class="image__link" href="https://openai.com/index/codex-for-every-role-tool-workflow/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/28d8c157-b2a8-4bd0-acb5-d11ec1d06d9b/Notion__1_.gif?t=1780468720"/></a></div><p class="paragraph" style="text-align:left;">Your Codex session doesn&#39;t have to end with a file sitting on your machine anymore.</p><p class="paragraph" style="text-align:left;"><b>Sites</b> lets Codex publish its work as a hosted, interactive website with a shareable URL. Dashboards, scenario planners, project trackers, launch hubs, whatever you&#39;re building, it goes live with a link you hand to your team. </p><p class="paragraph" style="text-align:left;">You can even ask Codex to keep the site up to date as things change. Not static pages either. These are collaborative canvases your whole workspace can explore and contribute to.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Annotations for precise refinement</b>: Point to the exact part of a site, doc, spreadsheet, or slide you want changed. Codex updates just that piece without starting over. Think inline editing, but AI-powered.</p></li><li><p class="paragraph" style="text-align:left;"><b>Rolling out now</b>: Sites are in preview for Business and Enterprise teams, expanding broadly soon.</p></li></ol><p class="paragraph" style="text-align:left;">And it&#39;s not just Sites. OpenAI is going really heavy on making Codex the tool for non-technical knowledge work.</p><p class="paragraph" style="text-align:left;">They shipped six <b><a class="link" href="https://github.com/openai/role-based-plugins?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">open-source, role-specific plugins</a></b> that each bundle apps, skills, and workflows for a specific job function. Data Analytics connects Snowflake, Databricks, Hex, and Tableau. Creative Production hooks into Figma, Canva, and Picsart. Sales brings in Salesforce, HubSpot, and Clay. There&#39;s also Product Design, Equity Investing, and Investment Banking. </p><p class="paragraph" style="text-align:left;">62 apps and 110 skills across all six. </p><p class="paragraph" style="text-align:left;"><b>Plugins work out of the box, but you own them</b>. Every plugin can be adapted to your team&#39;s workflows. Build and share custom ones too. Corporate Finance, Private Equity, Marketing Strategy, and Legal plugins are coming next.</p><p class="paragraph" style="text-align:left;">Non-developers already make up 20% of Codex&#39;s 5 million weekly users and are growing 3x faster than developers. OpenAI is clearly building for that curve.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://hermes-agent.nousresearch.com/desktop?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Hermes Desktop is Here</a></b></h3><div class="image"><a class="image__link" href="https://hermes-agent.nousresearch.com/desktop?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a8b06dd4-818f-4c32-99c0-c3e3868d71f1/ezgif-2d3bf08c0d6a8647.gif?t=1780469136"/></a></div><p class="paragraph" style="text-align:left;"><b>Hermes Agent</b> just got a native desktop app on macOS and Windows.</p><p class="paragraph" style="text-align:left;">First demoed during Jensen Huang&#39;s GTC keynote, <b>Hermes Desktop</b> is now in public preview. Same agent, same memory, same skills, same everything, just no terminal required. Download the .dmg or .exe, and you&#39;re running.</p><p class="paragraph" style="text-align:left;">The community has been using Hermes as a single interface for all their workflows, and has been actively growing the ecosystem around it. A native app lowers the floor for everyone who wants in but doesn&#39;t live in a terminal. If you&#39;re already running Hermes via CLI or messaging platforms, nothing changes. Desktop is just another surface, same agent underneath.</p><p class="paragraph" style="text-align:left;">Download at <a class="link" href="https://hermes-agent.nousresearch.com/desktop?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">hermes-agent.nousresearch.com/desktop</a>. macOS 12+, Windows 10/11, or install via terminal on Linux.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/tinyfish-io/bigset?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow"><b>Open-source Alternative to Exa Websets</b></a></h3><div class="image"><a class="image__link" href="https://github.com/tinyfish-io/bigset?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b6ab27d5-85af-470b-aa58-c83aff379db6/image.png?t=1780469235"/></a></div><p class="paragraph" style="text-align:left;">Describe the dataset you want in one sentence. AI agents go build it for you.</p><p class="paragraph" style="text-align:left;"><b>BigSet</b> is a new open-source tool from <a class="link" href="https://accounts.tinyfish.ai/api-keys?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow"><b>TinyFish</b></a> that turns a natural language prompt into a structured, verified dataset pulled from the live web. </p><p class="paragraph" style="text-align:left;">Say &quot;YC companies currently hiring engineers, with their funding stage, location, and number of open roles&quot; and BigSet infers the schema, fans out AI agents to research in parallel, deduplicates, and returns a clean table with citations that you can export as CSV or XLSX.</p><p class="paragraph" style="text-align:left;">The real fun bit is you can set a refresh cadence (30 minutes to weekly) so the dataset stays fresh always. </p><p class="paragraph" style="text-align:left;">Self-hosted via Docker in one command. </p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Schema inference from English</b>: You describe what you want, BigSet figures out column names, types, and primary keys. No manual schema design.</p></li><li><p class="paragraph" style="text-align:left;"><b>Parallel agent research</b>: Multiple AI agents fan out across the web simultaneously, verify data against real sources, and deduplicate before returning results.</p></li><li><p class="paragraph" style="text-align:left;"><b>Auto-refresh schedules</b>: Set it and forget it. Datasets update on a cadence you choose, from every 30 minutes to weekly.</p></li><li><p class="paragraph" style="text-align:left;"><b>Full stack, open source</b>: Next.js 16 frontend, Fastify backend, Mastra workflows for agent orchestration, powered by TinyFish&#39;s Search and Fetch APIs under the hood. AGPL-3.0 licensed.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://microsoft.ai/news/introducingmai-code-1-flash/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Microsoft finally built its own frontier models</a></b><b> </b><br>Microsoft dropped the new <b>MAI family of models</b> at Build, across text, image, voice, and speech. <b>MAI-Code-1-Flash</b> is the one to watch for devs, optimized specifically for fast, efficient coding tasks. <b>MAI-Thinking-1</b> handles heavier reasoning and SWE work. Even the Image model is debuting at No. 3 on <a class="link" href="https://Arena.ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Arena.ai</a>. Both are cheaper alternatives to OpenAI and Anthropic models. The word on the street is that Microsoft built these because relying on Anthropic&#39;s Claude was forcing them to raise GitHub Copilot prices and cap developer usage.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://factory.ai/news/factory-router?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Your coding agent doesn&#39;t need the most expensive model for every task</a></b><b> </b>Factory just shipped Factory Router, and the idea is overdue: stop burning frontier-model tokens on tasks that a smaller model handles just as well. Router automatically picks the right model for each coding session and escalates to a more capable one only if the first choice struggles. On their benchmarks, it hits 99% of Claude Opus 4.7&#39;s pass rate at 20% lower cost. If the first model can&#39;t crack it, Router bumps to a heavier one automatically. Available now in Factory CLI and Desktop in private research preview.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/perplexity_ai/status/2061861293569765847?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Perplexity Computer splits work between local and cloud</a></b><b> </b><br>Perplexity announced hybrid agentic inference for Perplexity Computer: the system can now split tasks between a local model running on your machine and frontier models in the cloud. Private data stays on-device, token efficiency goes up, and you stop sending everything through an API. Coming soon, but architecturally, this is the direction everyone&#39;s heading.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/jpschroeder/status/2061484426387677268?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Use Cursor&#39;s Composer 2.5 in any agent harness</a></b><b> </b><br>Someone built an open-source macOS app that exposes Cursor&#39;s Composer 2.5 as an API. That means you can now use Cursor&#39;s model routing in Codex, OpenCode, Cline, or whatever harness you prefer. If you&#39;ve been locked into Cursor&#39;s editor just for the model quality, this unbundles it.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://clawpatch.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">ClawPatch</a></b>: Open-source code review from the OpenClaw team that thinks in &quot;feature slices&quot; instead of files. It maps your codebase into semantic units (routes, commands, packages), sends bounded context to an AI for review, and then runs an explicit fix loop. Every finding gets a severity, confidence score, and audit trail. Works with Codex as the default AI provider.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/nesquena/hermes-webui?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Hermes WebUI</a></b>: A 12.7K-star open-source web interface for Hermes Agent by Nathan Esquenazi (CodePath co-founder). Full CLI parity in a three-panel browser layout: sessions on the left, chat in the center, workspace file browser on the right. No build step, no framework, just Python and vanilla JS. MIT-licensed.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/leodev/status/2061417039949099205?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Email SDK</a></b>: Unified TypeScript SDK that lets you send email through any provider: Resend, Postmark, SendGrid, Mailgun, Brevo, or raw SMTP. One clean API, swap providers by changing a config line. Built-in formatting, error handling, and type safety so you stop writing provider-specific glue code.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> (111k+ </b>🌟 <b>) </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=openai-codex-can-now-ship-live-shareable-websites"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=8cf1ddf1-7902-466a-8b4a-e67bf87950a3&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Every Software Just Became Agent-Native</title>
  <description>+ Self-Evolving Agent Skills by Microsoft</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a34b1984-af14-4808-a05e-43abfe5ff084/Every_Software_Just_Became_Agent-Native.png" length="2102905" type="image/png"/>
  <link>https://www.theunwindai.com/p/every-software-just-became-agent-native</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/every-software-just-became-agent-native</guid>
  <pubDate>Thu, 28 May 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-05-28T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>CLI-Anything: Every Software Just Became Agent-Native</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Shared AI agents for teams in Slack</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Microsoft SkillOpt: Train the Skill, not the model</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Inference price war is here (and it’s starting from China)</b></p></li><li><p class="paragraph" style="text-align:left;"><b>LangChain gives agents a code layer between tool calls</b></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/the-ultimate-guide-to-goal?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">The Ultimate Guide to /goal</a></b></p><p class="paragraph" style="text-align:left;">HTTP is a primitive. JSON is a primitive. /goal is becoming one for coding agents.</p><p class="paragraph" style="text-align:left;">A few weeks ago, OpenAI&#39;s Codex CLI added /goal as a way to give the coding worker a job with a defined done state. Claude Code added it this week.</p><p class="paragraph" style="text-align:left;">Hermes Agent, the orchestrator I run on a Mac Mini to coordinate work between coding workers, has had /goal built in for a while.</p><p class="paragraph" style="text-align:left;">This guide walks through what /goal actually is, the three roles in a multi-agent setup, a real end-to-end run, the verification rule, and how to run goals in parallel without workers stepping on each other. </p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/the-ultimate-guide-to-goal?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Read The Ultimate Guide to /goal</a></b></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://github.com/HKUDS/CLI-Anything?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Every Software Just Became Agent-Native</a></b></h3><div class="image"><a class="image__link" href="https://github.com/HKUDS/CLI-Anything?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a6ebc981-aafe-4d33-8682-6ca32be8635e/Screenshot_2026-05-27_at_11.40.28_PM.png?t=1779950434"/></a></div><p class="paragraph" style="text-align:left;">Your agent can write code, search the web, and manage files. But ask it to edit a Blender scene, export a MuseScore sheet, or automate Rekordbox, and it hits a wall. The software doesn&#39;t speak agent.</p><p class="paragraph" style="text-align:left;">CLI-Anything from the HKUDS lab fixes this by generating full CLI harnesses for any software, turning GUI-only apps into agent-controllable tools. One command analyzes the target app&#39;s source code, architects a CLI, implements it with tests, and publishes it to PATH. The project ships with a growing registry of 50+ ready-made CLIs covering GIMP, Blender, LibreOffice, OBS, Obsidian, Kdenlive, QGIS, and more.</p><p class="paragraph" style="text-align:left;">The idea is simple but the implications are huge: if CLI is the universal interface both humans and LLMs already speak, then wrapping every piece of software in a CLI makes the entire software ecosystem agent-accessible overnight.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>CLI-Hub package manager</b>: pip install cli-anything-hub, then browse, search, and install any harness with cli-hub install &lt;name&gt;. Supports pip, npm, brew, and system tools.</p></li><li><p class="paragraph" style="text-align:left;"><b>7-phase generation pipeline</b>: Point it at a repo or app and it runs through analyze, design, implement, plan tests, write tests, document, and publish, fully automated by your coding agent.</p></li><li><p class="paragraph" style="text-align:left;"><b>Works with every major agent</b>: Claude Code plugin, Pi extension, OpenCode commands, Codex, and OpenClaw skill. Each gets a native integration path.</p></li><li><p class="paragraph" style="text-align:left;"><b>Skills baked in</b>: Every generated CLI ships with a <a class="link" href="https://SKILL.md?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">SKILL.md</a> so agents can discover and use it autonomously, no manual wiring needed.</p></li><li><p class="paragraph" style="text-align:left;"><b>Try it now</b>: Install from PyPI or clone the repo. The CLI-Hub web registry is live at <a class="link" href="https://clianything.cc?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">clianything.cc</a>.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://github.com/paradigmxyz/centaur?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Shared AI agents for teams in Slack</a></b></h3><div class="image"><a class="image__link" href="https://github.com/paradigmxyz/centaur?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f890b5a7-aafb-4b23-80e9-4380b934c979/Screenshot_2026-05-27_at_11.41.43_PM.png?t=1779950507"/></a></div><p class="paragraph" style="text-align:left;">We love our personal Hermes and OpenClaws. But how many of us have been running it for our professional work, with our teams, cross-functionally? </p><p class="paragraph" style="text-align:left;">It’s a completely different set of problems: surviving laptop closures, handling real credentials securely, running for hours or days, and being reachable where the team actually works. </p><p class="paragraph" style="text-align:left;">Centaur is the self-hosted runtime Paradigm and Tempo have been running internally since January, now open-sourced under Apache 2.0. It&#39;s a Slack-native multiplayer agent system where: </p><ul><li><p class="paragraph" style="text-align:left;">every thread gets its own isolated Kubernetes sandbox, </p></li><li><p class="paragraph" style="text-align:left;">tools you add are instantly available to every conversation, and </p></li><li><p class="paragraph" style="text-align:left;">a credential firewall injects secrets in-flight so agents can never exfiltrate raw keys.</p></li></ul><p class="paragraph" style="text-align:left;">Tools are plain Python drop-ins that hot-reload across your org, workflows checkpoint to Postgres and resume exactly where they left off after a crash, and every night the system reviews its own performance and ships fixes to its own skills (super interesting!). </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/paradigmxyz/centaur?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Clone from GitHub</a> or visit <a class="link" href="https://centaur.run?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">centaur.run</a> to get started.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/microsoft/SkillOpt?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow"><b>Microsoft SkillOpt: Train the Skill, Not the Model</b></a></h3><div class="image"><a class="image__link" href="https://github.com/microsoft/SkillOpt?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/c3a18eb4-c5d4-4d07-9687-31730403ff68/Screenshot_2026-05-27_at_11.43.09_PM.png?t=1779950594"/></a></div><p class="paragraph" style="text-align:left;">What if you could train agent skills the same way you train neural networks, with learning rates, mini-batches, epochs, and momentum, but entirely in text space?</p><p class="paragraph" style="text-align:left;">SkillOpt from Microsoft Research does exactly that. Instead of fine-tuning model weights, it treats SKILL.md as a trainable external parameter. The frozen target model executes tasks, records scored trajectories, and a separate optimizer model proposes structured edits to the skill. Edits are accepted only when held-out validation performance improves. </p><p class="paragraph" style="text-align:left;">The whole thing mirrors a training loop: rollouts are forward passes, reflection is a backward pass, and a textual edit budget acts as a learning rate to prevent destructive rewrites.</p><p class="paragraph" style="text-align:left;">Evaluated across 6 benchmarks and 7 models, including real agent execution loops with Codex and Claude Code, SkillOpt achieves best or tied-best results in all 52 settings tested.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Improvement with Claude Code</b>: On GPT-5.5 target model running through Claude Code, SkillOpt skills boosted performance by an average of 18.6 points across benchmarks, with Spreadsheet tasks jumping +58.3%.</p></li><li><p class="paragraph" style="text-align:left;"><b>Cross-model and cross-harness transfer</b>: A skill trained with Codex transfers directly into Claude Code and gains +31.8% on SpreadsheetBench. Trained on GPT-5.4, transfers to GPT-5.4-nano and still gains +15.2%.</p></li><li><p class="paragraph" style="text-align:left;"><b>Self-optimizer mode works</b>: Even when the target model is its own optimizer, the constrained, validated update loop still discovers useful edits.</p></li><li><p class="paragraph" style="text-align:left;"><b>Exports a single file</b>: The whole optimization produces one best_skill.md file. The target model at deployment never sees the optimizer memory, rejected edits, or training state.</p></li><li><p class="paragraph" style="text-align:left;"><b>Open-source</b>: The whole thing is open-sourced under MIT license. Go and try it out!</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/kimmonismus/status/2059578380329394292?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">The inference price war is here</a></b><br>DeepSeek just made its 75% price cut on V4-Pro permanent. Xiaomi&#39;s MiMo slashed V2.5 pricing by up to 99%, effective today. But this isn&#39;t a loss-leader race to the bottom. V4-Pro&#39;s hybrid attention architecture compresses its KV cache at 1M tokens to 10% of V3.2&#39;s, with single-token inference FLOPs at 27% of previous. V4-Pro now sits at $0.87 per million output tokens. A year ago, sub-dollar output pricing meant you were using a small distilled model with real capability tradeoffs. These are frontier-class reasoners.</p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/ElevenLabs/status/2059312414198235642?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">ElevenLabs Launches Music v2</a></b><br>ElevenLabs just shipped Music v2 with better vocals, instrumentation, and arrangement across every genre, plus improved multilingual support and capabilities that weren&#39;t possible before. If you&#39;ve used their v1 for music generation, this is a huge upgrade.</p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://cohere.com/blog/command-a-plus?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Cohere Drops Command A+: 218B MoE Under Apache 2.0</a></b><br>Cohere just open-sourced Command A+, a 218B parameter MoE model with only 25B active per token. It unifies all previous Command A variants (reasoning, vision, translation) into a single model, supports 48 languages, and runs on as little as two H100s at W4A4 quantization. On τ²-Bench Telecom, it jumped from 37% to 85% over Command A Reasoning. Apache 2.0, available on Hugging Face in BF16, FP8, and W4A4.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.microsoft.com/en-us/research/articles/fara1-5-computer-use-agent/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow"><b>Microsoft Open-Sources Fara1.5 Browser Agents</b></a><br>Microsoft Research just dropped Fara1.5, a family of three open computer use agent models (4B, 9B, 27B) built on Qwen3.5 for browser automation. The 27B variant hits 72% on Online-Mind2Web, outperforming OpenAI Operator, Gemini 2.5 Computer Use, and Yutori Navigator n1. Even the 9B model at 63.4% beats every proprietary competitor. Available on the Microsoft Foundry now.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://developers.openai.com/api/docs/guides/secure-mcp-tunnels?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow"><b>OpenAI Launches Secure MCP Tunnels</b></a><br>Your private MCP servers can now stay inside your network while ChatGPT, Codex, and the Responses API connect through outbound-only HTTPS. No inbound ports, no public endpoints, no VPN. If you&#39;ve been holding off on connecting internal tools to OpenAI products because of network security concerns, this removes that blocker.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.langchain.com/blog/give-your-agents-an-interpreter?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow"><b>LangChain Gives Agents a Code Layer Between Tool Calls</b></a><br>Your agent calls a tool, reads the result, reasons, calls the next tool, reads, reasons, repeat. Every step is a model round trip. LangChain&#39;s Deep Agents now ships with interpreters — small QuickJS runtimes where the agent writes code that coordinates multiple tool calls, keeps intermediate state in the runtime, and returns only what matters. Early testing showed up to 35% fewer tokens on some tasks. Available in both Python and TypeScript.</p></div><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/colbymchenry/codegraph?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Codegraph</a></b>: Pre-indexed code knowledge graph for Claude Code, Codex, Cursor, and more. Agents query symbol relationships and call graphs instead of scanning files, averaging 35% cheaper and 70% fewer tool calls. 100% local, MIT licensed.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/perplexityai/bumblebee?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow"><b>Bumblebee</b></a>: Perplexity&#39;s open-source supply chain scanner for developer machines. A single Go binary that checks lockfiles, package metadata, extension manifests, and MCP configs against exposure catalogs. Apache 2.0.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://apps.apple.com/us/app/sieve-secret-scanner/id6767409365?mt=12&utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Sieve</a></b>: macOS app that scans your Claude Code, Cursor, Copilot, Windsurf, and Codex chat history for accidentally leaked API keys, tokens, and passwords. Ships with an MCP server so Claude can check for exposed secrets itself. $9.99.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> (111k+ </b>🌟 <b>) </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=every-software-just-became-agent-native"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=f50074a4-f6da-4920-ae86-db783084ba74&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Stop giving agents the whole computer</title>
  <description>+ GitHub Spec Kit, Qwen 3.7 Max</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a28ee767-15d5-4e90-b8a1-7c708330a9ce/Stop_giving_agents_the_whole_computer.png" length="2774943" type="image/png"/>
  <link>https://www.theunwindai.com/p/stop-giving-agents-the-whole-computer</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/stop-giving-agents-the-whole-computer</guid>
  <pubDate>Fri, 22 May 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-05-22T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">I’ve been thinking a lot about how much room we should actually give coding agents to work.  </p><p class="paragraph" style="text-align:left;">Qwen3.7-Max running for 35 hours with 1,000+ tool calls makes long-horizon agents feel a lot more real. But today’s npm compromise is the less fun side of the same story: attackers are now targeting Claude Code and Codex hooks directly.  </p><p class="paragraph" style="text-align:left;">So the takeaway is pretty simple. Agents are getting better at doing real work, but the workflows around them need stricter specs, better memory, and tighter boundaries before we hand them bigger jobs.</p><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Qwen3.7-Max: 35 hours, 1,000+ tool calls, zero human intervention</b></p></li><li><p class="paragraph" style="text-align:left;"><b>GitHub Spec Kit forces AI to spec before it codes</b></p></li><li><p class="paragraph" style="text-align:left;"><b>314 npm packages compromised to hijack your coding agent</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Google’s own version of hosted Hermes/OpenClaw</b></p></li><li><p class="paragraph" style="text-align:left;"><b>A design skill that refuses to look AI-generated</b></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/the-ultimate-guide-to-goal?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">The Ultimate Guide to /goal</a></b></p><p class="paragraph" style="text-align:left;">HTTP is a primitive. JSON is a primitive. /goal is becoming one for coding agents.</p><p class="paragraph" style="text-align:left;">A few weeks ago, OpenAI&#39;s Codex CLI added /goal as a way to give the coding worker a job with a defined done state. Claude Code added it this week.</p><p class="paragraph" style="text-align:left;">Hermes Agent, the orchestrator I run on a Mac Mini to coordinate work between coding workers, has had /goal built in for a while.</p><p class="paragraph" style="text-align:left;">This guide walks through what /goal actually is, the three roles in a multi-agent setup, a real end-to-end run, the verification rule, and how to run goals in parallel without workers stepping on each other. </p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/the-ultimate-guide-to-goal?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">Read The Ultimate Guide to /goal</a></b></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://qwen.ai/blog?id=qwen3.7&utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">Qwen3.7-Max: 35 Hours, 1,000+ Tool Calls, Zero Human Intervention</a></b></h3><div class="image"><a class="image__link" href="https://qwen.ai/blog?id=qwen3.7&utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/114869bf-286e-4721-a018-6db53a8bfce7/image.png?t=1779346463"/></a></div><p class="paragraph" style="text-align:left;">What’s the maximum number of steps and tool calls you’ve seen an LLM doing without you babysitting? 20? 50? Max 100? </p><p class="paragraph" style="text-align:left;">Qwen3.7-Max just ran a fully autonomous kernel optimization session for 35 hours straight, making over 1,000 tool calls.</p><p class="paragraph" style="text-align:left;">Alibaba&#39;s Qwen Team released their latest model Qwen 3.7-Max, specifically for the agent era. It tops SWE-Pro at 60.6% (vs Opus 4.6&#39;s 48.2%), leads TerminalBench, and takes the crown on MCP-Mark. It also tops all the benchmarks on pure reasoning; best-in-class!</p><p class="paragraph" style="text-align:left;">What makes it genuinely different is scaffold generalisation. You can plug it into Claude Code, OpenClaw, Hermes Agent, or Qwen Code and get consistent results without prompt gymnastics.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Reward hacking defense built in</b>: During 80+ hours of RL training on SWE tasks, the model&#39;s monitoring system autonomously caught 1,618 reward hacking attempts and generated 13 new heuristic rules to block them. The model is training itself to be honest.</p></li><li><p class="paragraph" style="text-align:left;"><b>1M token context, 65K output</b>: Scores 90.4% on MRCR-v2 128K, far ahead of every competitor on long-context retrieval.</p></li><li><p class="paragraph" style="text-align:left;"><b>48 languages natively</b>: Leads multilingual benchmarks across the board, including WMT24++ translation and MMLU-ProX.</p></li><li><p class="paragraph" style="text-align:left;"><b>Pricing</b>: Qwen 3.7 Max is roughly half the price of GPT-5.4 and less than a third of Claude Opus 4.6, while matching both on SWE-Pro and TerminalBench.</p></li><li><p class="paragraph" style="text-align:left;"><b>Closed source</b>: Qwen3.7-Max is proprietary and will be available via Alibaba Cloud Model Studio API. Open-weight variants at smaller sizes are expected to follow.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/github/spec-kit?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>GitHub Spec Kit forces AI to spec before it codes</b></a></h3><div class="image"><a class="image__link" href="https://github.com/github/spec-kit?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2759a7a6-014d-4025-b01d-dfa266471b23/Screenshot_2026-05-19_at_10.00.47_PM.png?t=1779253251"/></a></div><p class="paragraph" style="text-align:left;">Still throwing vague prompts at your coding agent and hoping it doesn&#39;t torch your project?</p><p class="paragraph" style="text-align:left;"><b>GitHub</b> just open-sourced <b>Spec Kit</b>, a toolkit that makes the AI create a structured specification before it writes a single line of code. The agent figures out what you want, asks clarifying questions, plans the architecture, generates a task list, then implements. All structured, all inspectable, all before any code exists.</p><p class="paragraph" style="text-align:left;">103K stars already. Works with 30+ coding agents out of the box: Claude Code, Cursor, Codex, Gemini CLI, Junie, and more. And it&#39;s completely stack-agnostic, so it doesn&#39;t care if you&#39;re writing Rust or Rails.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Structured before creative: T</b>he agent can&#39;t start coding until it&#39;s written a spec, asked clarifying questions, and planned the architecture. The sequence is enforced, not optional.</p></li><li><p class="paragraph" style="text-align:left;"><b>30+ agent integrations: </b>Works with Claude Code, Copilot, Gemini, Codex, Cursor, and pretty much every coding agent you&#39;re already using. Same spec, any agent.</p></li><li><p class="paragraph" style="text-align:left;"><b>Extensible via presets and extensions: </b>Customize the workflow with your own templates. Runtime resolution follows a priority chain: project-local overrides beat presets beat extensions beat core.</p></li><li><p class="paragraph" style="text-align:left;"><b>MIT-licensed:</b> Install via <code>uv tool install</code> from the GitHub repo. Full greenfield, exploration, and brownfield workflows supported out of the box.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://safedep.io/mini-shai-hulud-strikes-again-314-npm-packages-compromised/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>314 npm Packages Compromised to Hijack Your Coding Agent</b></a></h3><div class="image"><a class="image__link" href="https://safedep.io/mini-shai-hulud-strikes-again-314-npm-packages-compromised/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/791065cf-9863-4d51-aee0-950b135bde1d/Screenshot_2026-05-19_at_10.01.59_PM.png?t=1779253324"/></a></div><p class="paragraph" style="text-align:left;">22 minutes. That&#39;s how long it took an attacker to publish 637 malicious versions across 317 npm packages with a combined 11+ million monthly downloads.</p><p class="paragraph" style="text-align:left;">The compromised account &quot;atool&quot; pushed a payload from the &quot;Mini Shai-Hulud&quot; toolkit, the same one behind the SAP compromise three weeks ago. But the interesting bit is that the malware specifically targets AI coding agents. It injects Claude Code SessionStart hooks, Codex hooks, and VS Code &quot;runOn: folderOpen&quot; tasks. It harvests AWS credentials, Kubernetes tokens, SSH keys, GitHub PATs, and even 1Password and Bitwarden vaults. Exfiltration is disguised as OpenTelemetry traces to blend in with your existing observability stack.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Your coding agent is a vector</b>: The payload hooks into Claude Code and Codex session startup, silently piping every credential it finds to a C2 server. If you&#39;re running these agents in environments with cloud access, this is as bad as it sounds.</p></li><li><p class="paragraph" style="text-align:left;"><b>Packages you probably use</b>: size-sensor (4.2M downloads/month), echarts-for-react (3.8M), @antv/scale (2.2M), timeago.js (1.15M). Check your lockfile.</p></li><li><p class="paragraph" style="text-align:left;"><b>Persistent and stealthy</b>: A LaunchAgent/systemd service called &quot;kitty-monitor&quot; survives reboots and uses GitHub commit search as a dead-drop C2 channel, polling for RSA-PSS signed commands.</p></li><li><p class="paragraph" style="text-align:left;"><b>Full advisory and IoCs available</b>: SafeDep published the complete list of all 317 compromised packages with deobfuscated payloads and remediation steps.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/karpathy/status/2056753169888334312?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>Andrej Karpathy joins Anthropic</b></a><br>Yesterday, he announced he&#39;s joined Anthropic. &quot;The next few years at the frontier of LLMs will be especially formative,&quot; he wrote. He shaped the early GPT era at OpenAI, then left to build Eureka Labs for AI education. Now he&#39;s back in the lab at what might be the most interesting research org in the field right now. One to watch.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/Google/status/2056791134295273554?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>Google releases a managed personal 24/7 agent</b></a><br>Gemini App now comes with Spark, an always-on personal AI agent, running on Gemini 3.5 and built on Antigravity. It navigates your digital life and takes actions on your behalf, even when you close your laptop. You can set up cron jobs (schedule tasks), teach it new Skills, and create end-to-end workflows. Rolling out to trusted testers now, with beta access for Google AI Ultra subscribers next week.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://elevenlabs.io/speech-engine?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>Turn any chat agent into a voice agent with one prompt</b></a><br>ElevenLabs just shipped Speech Engine, and it’s pretty straightforward: keep your existing chat agent exactly as it is, add Speech Engine on top, and now it talks. You don’t need to rearchitect your LLM stack or swap out your RAG pipeline. It bundles speech-to-text, turn detection, interrupt handling, TTS, and audio orchestration into a single pipeline with ultra-low latency. Works with any LLM that produces text, has built-in stream extraction, and covers 70+ languages.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://deepmind.google/models/gemini-omni/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>Google releases the Nano Banana of video-gen</b></a><br>Gemini Omni is Google&#39;s new any-input-to-any-output model. Feed it images, text, video, audio, or any combination, and it generates or edits video through conversation. Multi-turn editing keeps scenes consistent across back-and-forth iterations, and it applies real-world physics to generated content. Available in the Gemini app, Google Flow, and YouTube Shorts. </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/warpdotdev/status/2056772856835453395?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>Multi-agent orchestration comes to Warp Oz</b></a><br>Warp just shipped multi-agent orchestration in Oz with support for Claude Code, Codex, and the Warp Agent. Use /orchestrate to delegate complex tasks across a team of agents running locally or in the cloud. If you&#39;re already in the Warp terminal, this makes it your control plane for parallel agent work.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/google-antigravity/antigravity-cli?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>Google Antigravity with Gemini 3.5 Flash now in your Terminal</b></a><br>Google Antigravity just shipped a CLI written in Go, powered by Gemini 3.5 Flash, and built for async workflows where agents run tasks in the background and report back when done. It shares the same tool and app server as Antigravity 2.0, so anything you build on the platform also works inside Google Search, where Antigravity powers the new agentic coding features: custom generative UIs, dashboards, and &quot;mini apps&quot; spun up from natural language.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://blog.google/products-and-platforms/products/search/search-io-2026/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow"><b>Google Search gets its biggest overhaul in 25 years</b></a><br>The search box itself is being rebuilt: AI-powered, dynamically expanding, with multimodal inputs (text, images, files, videos, even Chrome tabs). New &quot;search agents&quot; will monitor the web 24/7 for specific criteria like apartment listings or sneaker drops, and agentic booking lets you complete local service bookings right from Search. Launching for Google AI Pro and Ultra subscribers this summer.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/nutlope/hallmark?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">Hallmark</a></b>: Design skill by Hassan El Mghari that encodes anti-slop rules into Claude Code, Cursor, and Codex. Has four modes: build (generates pages that refuse to repeat the same structure twice), study (extracts a design&#39;s DNA from a URL or screenshot without copying pixels), audit (scores existing pages against its anti-pattern catalogue), and redesign (same content, deliberately different bones).</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/colbymchenry/codegraph?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">CodeGraph</a></b>: Pre-indexed code knowledge graph for Claude Code, Codex, Cursor, and OpenCode that cuts tool calls by 92% and speeds up tasks by 71%. 100% local, MIT-licensed.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/rohitg00/agentmemory?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">AgentMemory</a></b>: Persistent memory MCP server for coding agents with 4-tier consolidation inspired by how the brain organizes memory during sleep. Works with every major coding agent.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> (111k+ </b>🌟 <b>) </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=stop-giving-agents-the-whole-computer"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=adc964d4-941d-40c2-b580-ceeca7b5c3a2&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Vercel Built a Programming Language for AI Agents</title>
  <description>+ Garry Tan’s agent brain, Codex on mobile, and a $1.3M agent bill</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e8b919ad-5f47-49c0-b6f6-966b8b3fdf1a/Vercel_Built_a_Programming_Language_for_AI_Agents.png" length="1513867" type="image/png"/>
  <link>https://www.theunwindai.com/p/vercel-built-a-programming-language-for-ai-agents</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/vercel-built-a-programming-language-for-ai-agents</guid>
  <pubDate>Mon, 18 May 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-05-18T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">I was looking at today’s updates and kept coming back to the Vercel one.</p><p class="paragraph" style="text-align:left;">A programming language built with AI agents in mind sounds a little ridiculous at first. Like, do agents really need their own language now?</p><p class="paragraph" style="text-align:left;">But then you look at the rest of today’s news: Garry Tan’s knowledge brain for his personal agents, Codex on mobile, and Peter Steinberger’s $1.3M monthly token spend. </p><p class="paragraph" style="text-align:left;">Suddenly, it feels less ridiculous. </p><p class="paragraph" style="text-align:left;">If agents are going to do real work, people are going to build weird new infrastructure around them. Some of it will be overkill. Some of it will probably become the new normal.</p><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Vercel built a programming language for AI agents</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Garry Tan open-sourced the brain running his AI agents</b></p></li><li><p class="paragraph" style="text-align:left;"><b>LiteLLM launches sandboxes for agent fleets</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Peter Steinberger spent $1.3M in OpenAI tokens in 30 days</b></p></li><li><p class="paragraph" style="text-align:left;"><b>Codex crossed 4M weekly users and landed on mobile</b></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.theunwindai.com/p/the-ultimate-guide-to-goal?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>The Ultimate Guide to /goal</b></a></p><p class="paragraph" style="text-align:left;">HTTP is a primitive. JSON is a primitive. /goal is becoming one for coding agents.</p><p class="paragraph" style="text-align:left;">A few weeks ago, OpenAI&#39;s Codex CLI added /goal as a way to give the coding worker a job with a defined done state. Claude Code added it this week.</p><p class="paragraph" style="text-align:left;">Hermes Agent, the orchestrator I run on a Mac Mini to coordinate work between coding workers, has had /goal built in for a while.</p><p class="paragraph" style="text-align:left;">This guide walks through what /goal actually is, the three roles in a multi-agent setup, a real end-to-end run, the verification rule, and how to run goals in parallel without workers stepping on each other. </p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.theunwindai.com/p/the-ultimate-guide-to-goal?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Read The Ultimate Guide to /goal</b></a></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/garrytan/gbrain?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Garry Tan Open-Sources the AI Brain That Runs His Life</b></a></h3><div class="image"><a class="image__link" href="https://x.com/garrytan/status/2055670533451366479?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/71de464f-1f2c-488e-9511-a6cad27fcfe9/Screenshot_2026-05-17_at_8.42.38_PM.png?t=1779075763"/></a></div><p class="paragraph" style="text-align:left;">Garry Tan’s gstack made your AI agent ship code like a sprint team. Now, he open-sourced the knowledge system that powers his personal AI agents. </p><p class="paragraph" style="text-align:left;"><b>GBrain</b> is not a notes app or a RAG pipeline. It&#39;s a structured, self-maintaining knowledge brain that currently holds 17,000+ pages, tracks 4,000+ people, and runs 21 autonomous cron jobs. Garry built it in 12 days.</p><p class="paragraph" style="text-align:left;">The core idea is great: instead of re-deriving knowledge from scratch every query (like RAG does), GBrain pre-computes and maintains a &quot;compiled truth&quot; for every entity. Your agent gets richer context every time, and the whole thing compounds daily as it ingests meetings, emails, and calls. You wake up, and the brain is smarter than when you went to bed.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>MECE knowledge structure</b> - Everything is organized into clean categories: people, companies, deals, meetings, projects, concepts, originals, and media. Each page has a compiled summary on top and an append-only evidence trail below.</p></li><li><p class="paragraph" style="text-align:left;"><b>Self-wiring graph</b> - Every page-write automatically extracts entity references and creates typed links (attended, works_at, invested_in) with zero LLM calls. Ask &quot;who works at Acme AI?&quot; and get answers vector search alone can&#39;t reach.</p></li><li><p class="paragraph" style="text-align:left;"><b>34 built-in skills</b> - From signal detection to content ingestion to research synthesis. Intelligence lives in markdown skill files, not the runtime. This is how Garry&#39;s agents know how to do things, not just remember things.</p></li><li><p class="paragraph" style="text-align:left;"><b>Auto-enrichment tiers</b> - Mention someone once, they get a stub. Three mentions triggers web enrichment. Meet them in person or mention them 8+ times, and the full research pipeline kicks in. The brain decides how much attention someone deserves.</p></li><li><p class="paragraph" style="text-align:left;"><b>Usage</b> - Works as a standalone CLI, an MCP server for Claude Code and Cursor, or a one-click deploy on OpenClaw or Railway. Check it out at <a class="link" href="https://github.com/garrytan/gbrain?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">github.com/garrytan/gbrain</a>.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/vercel-labs/zero?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Vercel Built a Programming Language for AI Agents</b></a></h3><div class="image"><a class="image__link" href="https://github.com/vercel-labs/zero?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b73a5f20-ca0e-4d50-9da4-1974386c4459/image.png?t=1779075737"/></a></div><p class="paragraph" style="text-align:left;">Programming languages were designed for humans and then retrofitted for AI. Chris Tate from Vercel just changed that. </p><p class="paragraph" style="text-align:left;"><b>Zero</b> is a new systems language where AI agents are first-class users of the entire toolchain. Compiler errors come back as structured JSON with stable error codes and fix suggestions that agents can parse and act on programmatically.</p><p class="paragraph" style="text-align:left;">Think of it as a language where the compiler talks to your agent the same way a senior engineer talks to a junior one: here&#39;s what&#39;s wrong, here&#39;s the error code, here&#39;s exactly how to fix it. </p><p class="paragraph" style="text-align:left;">Still experimental, but the idea is genuinely novel. And the fact that it&#39;s coming from Vercel, not a random weekend project, means there&#39;s real conviction behind it.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Structured diagnostics</b> - Every compiler error returns JSON with stable codes, line locations, and repair metadata. No more regex-parsing error messages. Agents can read errors and fix code without scraping human-readable text.</p></li><li><p class="paragraph" style="text-align:left;"><b>Tiny output footprint</b> - Compiles down to extremely small native binaries with no garbage collector, event loop, and hidden runtime overhead. Think CLI tools and serverless functions.</p></li><li><p class="paragraph" style="text-align:left;"><b>Full CLI toolchain</b> - One command handles check, build, run, test, format, inspect, dependency graphs, and docs. Everything an agent needs in one place, all machine-readable output.</p></li><li><p class="paragraph" style="text-align:left;"><b>Human-readable too</b> - Despite being agent-first, the syntax is clean and readable. File extension is .0, which is a fun touch.</p></li><li><p class="paragraph" style="text-align:left;"><b>Try it now</b> - Zero is open-source (Apache 2.0) from Vercel. Check it out on <a class="link" href="http://github.com/vercel-labs/zero?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">GitHub</a> and <a class="link" href="https://zerolang.ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">zerolang.ai</a>.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/BerriAI/litellm-agent-platform?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>LiteLLM Launches an Agent Platform with K8s Sandboxes</b></a></h3><div class="image"><a class="image__link" href="https://github.com/BerriAI/litellm-agent-platform?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/eca3980c-80a6-4790-8d44-70b436e091e9/Screenshot_2026-05-17_at_8.46.29_PM.png?t=1779075995"/></a></div><p class="paragraph" style="text-align:left;">The team behind LiteLLM just shipped something bigger: a full platform for running fleets of coding agents in isolated Kubernetes sandboxes. </p><p class="paragraph" style="text-align:left;">Each agent session gets its own fresh pod. Your real API keys never touch agent code.</p><p class="paragraph" style="text-align:left;">A great feature is the credential vault. Agents only see stub tokens, and the platform transparently swaps in real secrets on every outbound connection. </p><p class="paragraph" style="text-align:left;">If you&#39;re scaling from one coding agent to a team of them, this makes it super safe.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Agent-agnostic sandboxes</b> - First-class support for Claude Code, Codex, and Hermes agents, all running in isolated pods that persist for 24 hours after you detach.</p></li><li><p class="paragraph" style="text-align:left;"><b>Credential vault</b> - Agents never see your real API keys. The vault injects real secrets transparently on outbound connections, so a rogue agent can&#39;t leak anything. This alone is worth the setup.</p></li><li><p class="paragraph" style="text-align:left;"><b>Terminal-first workflow</b> - The CLI lets you open a sandbox, attach your terminal, do your work, and Ctrl-D to detach. Simple and clean.</p></li><li><p class="paragraph" style="text-align:left;"><b>Full developer API</b> - Create agents, open sessions, send messages, and read replies, all via REST. Build your own orchestration on top.</p></li><li><p class="paragraph" style="text-align:left;"><b>Open source and self-hostable</b> - MIT licensed with local dev support and a production deploy path on AWS. Check it out at <a class="link" href="https://github.com/BerriAI/litellm-agent-platform?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">github.com/BerriAI/litellm-agent-platform</a>.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/steipete/status/2055346265869721905?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Peter Steinberger spent $1.3M in API tokens in 30 days</a></b><br>Peter Steinberger, creator of OpenClaw and now an OpenAI employee, casually revealed he burned through $1.3M worth of API tokens in a single month. That&#39;s roughly $20K per day, mostly on GPT-5.5 powering agents that manage the OpenClaw repo. Yes, he almost certainly has unlimited access as an OpenAI employee, so this isn&#39;t out of pocket. The number is wild, but the useful part is what it says about agent economics. Once agents run continuously, token spend becomes infrastructure, not just API usage</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://openai.com/index/work-with-codex-from-anywhere/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Codex hits 4M+ weekly users, now on Mobile</a></b><br>OpenAI&#39;s coding agent Codex just crossed 4 million weekly active users and is now available on iOS and Android through the ChatGPT app. You can kick off coding tasks, review diffs, and approve PRs from your phone. The &quot;desk-only&quot; constraint for AI-assisted coding is officially gone. 4M WAU also makes it one of the most widely adopted coding agents by a wide margin.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://interfaze.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">AI model built for deterministic dev tasks</a></b><br>Interfaze is a new model architecture that takes a fundamentally different approach. Instead of one autoregressive model doing everything, it breaks tasks into deterministic sub-steps like OCR, web search, classification, extraction, and then orchestrates purpose-built models for each. The result is structured output accuracy that beats GPT-5.4-Mini and matches Gemini-3-Flash. It&#39;s OpenAI SDK-compatible, comes with built-in web search from its own crawler, and runs tasks like audio transcription (1h 35m podcast in ~50 seconds) and document OCR natively. Free tier available at <a class="link" href="https://interfaze.ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">interfaze.ai</a>.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://openai.com/index/personal-finance-chatgpt/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">ChatGPT now wants to manage your personal finance</a></b><br>OpenAI launched a personal finance experience in ChatGPT for Pro users in the U.S. Connect your bank accounts via Plaid, and ChatGPT gives you a spending dashboard, subscription tracker, and financial guidance grounded in your transaction data. The before/after examples are genuinely compelling, going from generic &quot;save more money&quot; advice vs. &quot;cap dining at $450/month based on your Feb-May spend.&quot; OpenAI has also partnered with Intuit that signals the next step: actionable finance, like applying for credit cards and scheduling tax appointments directly from chat.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://clawpatch.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Clawpatch</a></b>: Code review for agent-written code. It maps a repo into semantic slices like routes, commands, packages, and tests, then reviews bounded contexts instead of isolated files. </p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="http://github.com/MinishLab/semble?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Semble</a></b>: Code search built specifically for AI coding agents. Instead of dumping entire files into context, Semble indexes your repo and returns only the relevant snippets, using 98% fewer tokens than grep-and-read. Drop-in MCP server for Claude Code, Codex, and Cursor.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="http://github.com/DrCatHicks/learning-opportunities?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Learning Opportunities</b></a>: A Claude Code and Codex plugin that pauses after significant coding work and offers short, evidence-based learning exercises. The idea: if an AI agent writes your code, you should still understand what it did and why. Turns AI-assisted coding into actual skill building.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> (111k+ </b>🌟 <b>) </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=vercel-built-a-programming-language-for-ai-agents"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=205d901f-4696-44b8-a2ea-a9c805534e59&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>The Ultimate Guide to /goal</title>
  <description>/goal isn’t a feature. It’s the new primitive.</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e7aab05b-a06c-4833-915c-c7d5ffbbce13/The_Ultimate_Guide_to_goal.png" length="2686872" type="image/png"/>
  <link>https://www.theunwindai.com/p/the-ultimate-guide-to-goal</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/the-ultimate-guide-to-goal</guid>
  <pubDate>Mon, 18 May 2026 04:32:06 +0000</pubDate>
  <atom:published>2026-05-18T04:32:06Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <category><![CDATA[Ai Blogs]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b>/goal is not a feature. It is a primitive. </b></p><p class="paragraph" style="text-align:left;">HTTP is a primitive. JSON is a primitive. /goal is becoming one for coding agents.</p><p class="paragraph" style="text-align:left;">A few weeks ago, OpenAI&#39;s Codex CLI added /goal as a way to give the coding worker a job with a defined done state. Claude Code added it this week. </p><p class="paragraph" style="text-align:left;">Hermes Agent, the orchestrator I run on a Mac Mini to coordinate work between coding workers, has had /goal built in for a while. </p><p class="paragraph" style="text-align:left;">So I now have a builder, a reviewer, and an orchestrator that all accept the same instruction format, even though they share nothing else.</p><p class="paragraph" style="text-align:left;">If you&#39;ve only seen /goal used as a fancier prompt, you&#39;ve missed what it changes.</p><h2 class="heading" style="text-align:left;"><b>What /goal actually is</b></h2><p class="paragraph" style="text-align:left;">A regular prompt asks an agent for the next response. You read what comes back, decide if it&#39;s right, and push the agent forward to the next step. You&#39;re steering every turn.</p><p class="paragraph" style="text-align:left;">/goal flips that. You write down what &quot;done&quot; looks like, submit it once, and the agent works toward it until it gets there. Here&#39;s a real one:</p><div class="codeblock"><pre><code>/goal Build the app described in SPEC.md. Done means tests pass, build passes, README is accurate, and git status only shows relevant project files.</code></pre></div><p class="paragraph" style="text-align:left;">The goal stays active until it&#39;s achieved, paused, blocked, cleared, or it runs out of budget.</p><p class="paragraph" style="text-align:left;">This is different from putting the word &quot;goal&quot; inside a normal one-shot command. If you write codex exec &#39;goal: build the app&#39;, that&#39;s still a prompt with a label. The real primitive lives inside an interactive worker session. You launch the CLI, you submit /goal, and you walk away.</p><p class="paragraph" style="text-align:left;">The shift is from prompting (you driving) to assigning (the agent driving toward a target you defined).</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/02460539-5ebc-46fd-abb3-6471754457b2/diag1.gif?t=1779078302"/></div><h2 class="heading" style="text-align:left;"><b>The three tools that currently speak /goal</b></h2><p class="paragraph" style="text-align:left;">The three tools accepting /goal aren&#39;t all the same kind of thing, so it&#39;s worth being specific.</p><p class="paragraph" style="text-align:left;"><b>Codex</b> is OpenAI&#39;s coding CLI. Strong at implementation, especially when given a clear spec. /goal is how you give it that spec.</p><p class="paragraph" style="text-align:left;"><b>Claude Code</b> is Anthropic&#39;s coding CLI. Strong at the inverse: finding what&#39;s wrong with code that looks right. Spec compliance, safety issues, error states, security holes. /goal is how you point it at code and ask for a review.</p><p class="paragraph" style="text-align:left;"><b>Hermes Agent</b> is a different kind of tool entirely. Not a coding worker, but an orchestrator that coordinates work between coding workers like the two above. /goal is how Hermes hands off tasks to whichever tool is right for the job, and also how I tell Hermes what I want in the first place.</p><p class="paragraph" style="text-align:left;">What matters isn&#39;t that any one of them shipped /goal. It&#39;s that three different teams converged on the same primitive, and that convergence is what makes it possible to compose them.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/444f1a13-913d-464d-ac28-4ec472960a59/diag2.gif?t=1779078344"/></div><h2 class="heading" style="text-align:left;"><b>Setting things up</b></h2><p class="paragraph" style="text-align:left;">The first time I needed Codex and Claude Code on the Mac Mini that runs Hermes, I didn&#39;t install them by hand. I sent Hermes a message asking it to install both and log me in. It handled the rest.</p><p class="paragraph" style="text-align:left;">That&#39;s the workflow now. You don&#39;t type install commands. Setup is just another goal.</p><p class="paragraph" style="text-align:left;">If you don&#39;t have an orchestrator running yet, the install pages for Codex and Claude Code are easy enough to follow. But once you do, you shouldn&#39;t set up another tool by hand. The point of having an orchestrator is that mechanical work stops being yours.</p><h2 class="heading" style="text-align:left;"><b>What Hermes adds on top of /goal</b></h2><p class="paragraph" style="text-align:left;">A raw /goal is useful on its own. But it leaves you with a coordination problem.</p><p class="paragraph" style="text-align:left;">If Codex is running in one terminal and Claude Code is running in another, you have to remember which process is doing what. You have to check logs. You have to manually pass review findings from one tool to the other.</p><p class="paragraph" style="text-align:left;">Hermes turns those loose runs into a workflow:</p><ol start="1"><li><p class="paragraph" style="text-align:left;">You message Hermes (in my case, over Telegram from my phone)</p></li><li><p class="paragraph" style="text-align:left;">Hermes creates goal cards on a Kanban board</p></li><li><p class="paragraph" style="text-align:left;">Hermes picks the right worker for each card</p></li><li><p class="paragraph" style="text-align:left;">The worker runs the goal in the background</p></li><li><p class="paragraph" style="text-align:left;">The card stores the process id, PID, repo, and done criteria</p></li><li><p class="paragraph" style="text-align:left;">When the build is ready, Hermes hands the repo to the reviewer</p></li><li><p class="paragraph" style="text-align:left;">If the review blocks, Hermes sends the findings back as a fix goal</p></li><li><p class="paragraph" style="text-align:left;">Hermes verifies the final output by inspecting the filesystem, tests, build, and git state</p></li></ol><p class="paragraph" style="text-align:left;">The board is what /goal becomes when there&#39;s an orchestrator on top of it. Every goal has a card, every card has a status, every handoff leaves a trail. Instead of hunting through terminals, you watch the work move across columns on your phone.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9012b167-870b-41b0-a16d-556c1b77d7fc/image.png?t=1779078169"/></div><h2 class="heading" style="text-align:left;"><b>The three roles</b></h2><p class="paragraph" style="text-align:left;">The tools change. The roles don&#39;t.</p><p class="paragraph" style="text-align:left;"><b>Orchestrator.</b> Owns the control loop. Task decomposition, worker selection, Kanban cards, background processes, dependencies, final verification, the user-facing summary. In my setup, Hermes.</p><p class="paragraph" style="text-align:left;"><b>Builder.</b> Takes a spec and produces working code. Implementation is the bottleneck this role solves. Codex tends to be strong here.</p><p class="paragraph" style="text-align:left;"><b>Reviewer.</b> Reads what the builder produced and finds what&#39;s wrong with it. Correctness is the bottleneck. Claude Code tends to be strong here.</p><h2 class="heading" style="text-align:left;"><b>A real run, end to end</b></h2><p class="paragraph" style="text-align:left;">I gave Hermes agent a goal to do this:</p><div class="codeblock"><pre><code>/goal Build a CLI tool that finds X mentions of me and pings me when something blows up.</code></pre></div><p class="paragraph" style="text-align:left;">Hermes broke the request into six cards.</p><blockquote align="center" class="twitter-tweet"><a href="https://twitter.com/Saboo_Shubham_/status/2054260705365475609?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=the-ultimate-guide-to-goal"><p> Twitter tweet </p></a></blockquote><p class="paragraph" style="text-align:left;"><b>Card 1: Spec.</b> Hermes wrote SPEC.md itself, capturing the stack, repo path, read-only constraints, mock mode requirements, tests, and verification commands. Owned by the PM role.</p><p class="paragraph" style="text-align:left;"><b>Card 2: Codex builds.</b> Codex ran /goal against SPEC.md. It created the project files, implemented the UI and backend, added tests, and got the app to a passing state. About 15 minutes. When it finished, npm test passed, npm run build passed, and git status showed only relevant new files.</p><p class="paragraph" style="text-align:left;"><b>Card 3: Claude Code reviews.</b> Claude Code ran /goal to review what Codex built. Checked spec compliance, read-only safety, API key handling, error states, tests, UI usefulness, bugs, and security issues. Result: PASS, no blocking issues.</p><p class="paragraph" style="text-align:left;"><b>Card 4: Codex fix loop.</b> Skipped, because the review passed. The card still matters when skipped. It shows Hermes can model conditional work. If Claude Code had blocked, Hermes would have handed the findings back to Codex as a new /goal.</p><p class="paragraph" style="text-align:left;"><b>Card 5: Claude Code final verification.</b> Skipped for the same reason.</p><p class="paragraph" style="text-align:left;"><b>Card 6: Hermes final summary.</b> Working app at the local path, UI and API both verified in mock mode. Codex built it with /goal. Claude Code reviewed it with /goal and returned PASS.</p><p class="paragraph" style="text-align:left;">All of that came from one message. Three different tools did the actual work, but I only ever talked to Hermes.</p><h2 class="heading" style="text-align:left;"><b>The verification rule</b></h2><p class="paragraph" style="text-align:left;">Hermes never trusted Codex&#39;s self-report. After Codex marked the build done, Hermes ran the commands itself:</p><div class="codeblock"><pre><code>npm test         # 17 tests passed
npm run build    # vite build passed</code></pre></div><p class="paragraph" style="text-align:left;">The verifier is what makes a /goal a contract instead of a promise. Don&#39;t trust the worker&#39;s self-report as final. Trust the verifier.</p><p class="paragraph" style="text-align:left;">Coding agents are confident. They&#39;ll tell you the build passes when the build was never run. They&#39;ll tell you tests pass when they wrote tests that never executed. The verifier closes that gap.</p><p class="paragraph" style="text-align:left;">Without verification, /goal is just a fancier prompt. With verification, it becomes a contract.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ae40a901-ecea-40df-acad-388ddca732fb/diagram3.gif?t=1779078360"/></div><h2 class="heading" style="text-align:left;"><b>Running multiple goals</b></h2><p class="paragraph" style="text-align:left;">You can run multiple /goals in parallel, but you can&#39;t point multiple coding workers at the same files without thinking about it first.</p><p class="paragraph" style="text-align:left;">My default is one main builder per repo. If I want parallelism, I add it across clear boundaries. Different repos, different branches, git worktrees, separate packages, docs vs code, tests vs implementation. Anywhere two workers can&#39;t step on each other.</p><p class="paragraph" style="text-align:left;">The bad pattern is three workers all editing the same file in the same repo. You get conflicts, partial overwrites, and one worker silently undoing another&#39;s work.</p><p class="paragraph" style="text-align:left;">The better pattern is one writer at a time on any given file. Builder writes, reviewer only reads, fix goals stay scoped to the fix. Or run three builders in three worktrees on three competing approaches and let the orchestrator pick the best one.</p><p class="paragraph" style="text-align:left;">The board is what makes this practical. Without it, parallel background workers become terminal chaos.</p><h2 class="heading" style="text-align:left;"><b>What changes for me</b></h2><p class="paragraph" style="text-align:left;">The useful framing here is not &quot;I can run agents in the background.&quot;</p><p class="paragraph" style="text-align:left;">It&#39;s that one message turns into a pipeline across three different coding tools, and I watch the whole thing move across one board.</p><p class="paragraph" style="text-align:left;">You stop sitting in a terminal waiting for one agent to finish, and start managing a queue of work with visible state.</p><p class="paragraph" style="text-align:left;">If Codex and Claude Code had each invented their own job-handoff format, no orchestrator could route between them. The board is impressive, but the primitive makes the board even more useful. </p><p class="paragraph" style="text-align:left;">The workers can change, but the primitive stays the same. The next coding tool that adopts /goal will join this pipeline without me changing anything. I&#39;ll just route work to it.</p><p class="paragraph" style="text-align:left;">That&#39;s what good primitives do.</p><p class="paragraph" style="text-align:left;">For more such cool tips and interesting ideas around Hermes, OpenClaw, Claude Code, Codex and other 24/7 agent teams.</p><p class="paragraph" style="text-align:left;"><b>Follow → </b><b><a class="link" href="https://x.com/@Saboo_Shubham_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=the-ultimate-guide-to-goal" target="_blank" rel="noopener noreferrer nofollow">@Saboo_Shubham_</a></b></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">We share in-depth blogs and tutorials like this 2-3 times a week, to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=the-ultimate-guide-to-goal" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;"><b>Don’t forget to share this tutorial on your social channels and tag Unwind AI (</b><span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=the-ultimate-guide-to-goal" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=the-ultimate-guide-to-goal" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=the-ultimate-guide-to-goal" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span><b>) to support us!</b></p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=the-ultimate-guide-to-goal"><span class="button__text" style=""> Subscribe now for FREE - Get instant access to more LLM, RAG & AI Agent tutorials </span></a></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=b28a5198-384f-4d50-8417-63c04b15ad27&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>/goal in Claude Code, Codex, and Hermes Agent</title>
  <description>+ OpenClaw creator open-sourced a Mac automation tool</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2f3866a1-c6df-456c-9f5b-3a0bf33b5260/goal_in_Claude_Code__Codex__and_Hermes_Agent.png" length="2255054" type="image/png"/>
  <link>https://www.theunwindai.com/p/goal-in-claude-code-codex-and-hermes-agent</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/goal-in-claude-code-codex-and-hermes-agent</guid>
  <pubDate>Tue, 12 May 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-05-12T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/openclaw/Peekaboo?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>OpenClaw creator open-sourced a native Mac automation tool</b></a></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/antirez/ds4?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Run DeepSeek V4 Flash locally on a 128GB Mac</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://openai.com/daybreak/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>OpenAI enters serious cyber defense with Daybreak</b></a></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/Saboo_Shubham_/status/2054017280376455265?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">/goal in Codex CLI, Hermes Agent, and Claude Code</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/GTG-Labs/sangria?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Let agents pay for your API in ~3 lines of code</b></a></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.theunwindai.com/p/build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Build a Multimodal Agentic RAG App with Gemini Embedding 2 and Google ADK</b></a></p><p class="paragraph" style="text-align:left;">In this tutorial, you&#39;ll build a fully-working <b>multimodal agentic RAG app</b> where text, URLs, PDFs, images, audio, and video all share a single 768-dimension embedding space, and a small Google Agent Development Kit (ADK) coordinator turns the retrieved evidence into a grounded, cited answer.</p><p class="paragraph" style="text-align:left;">The two pieces doing the heavy lifting are <b>Gemini Embedding 2</b>, which embeds every modality into the same vector space, and <b>Google ADK</b>, which wraps the retrieval call in an agent that inspects the workspace, calls the retrieval tool, and writes the answer.</p><p class="paragraph" style="text-align:left;">You&#39;ll see exactly how those two pieces compose without any extra orchestration framework.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/c9ccb7a4-e58d-4b21-bb2e-c0358b7a3e7e/Build_a_Multimodal_Agentic_RAG_App_with_Gemini_Embedding_2_and_Google_ADK.png?t=1778367901"/><div class="embed__content"><p class="embed__title"> Build a Multimodal Agentic RAG App with Gemini Embedding 2 and Google ADK </p><p class="embed__description"> (100% open source) </p></div></a></div><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this every week, designed to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://github.com/openclaw/Peekaboo?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">OpenClaw creator open-sourced a native Mac automation tool</a></b></h3><div class="image"><a class="image__link" href="https://github.com/openclaw/Peekaboo?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/61f25a9d-9e77-422a-ac35-11d42e471826/Screenshot_2026-05-11_at_10.29.37_PM.png?t=1778563783"/></a></div><p class="paragraph" style="text-align:left;">Still using computer use agents doing screenshot → click → reason → hallucinate → repeat?</p><p class="paragraph" style="text-align:left;">Peter Steinberger open-sourced <b>Peekaboo</b>, a MacOS-native toolkit that hands agents the accessibility tree directly, so clicks land on real elements with IDs, not pixel coordinates that drift every time the window moves. </p><p class="paragraph" style="text-align:left;">Beyond click and type, it covers the stuff every other agent fails at, like Spaces switching, Dock right-clicks, menu bar extras, file dialogs, drag-to-Trash.</p><p class="paragraph" style="text-align:left;">Use it with Claude Code, Codex, OpenClaw, Hermes Agent, or which agent harness you like, via CLI and MCP server. Bring any model like Claude, GPT-5.1, Grok 4-fast, or Ollama for fully local runs. MIT-licensed.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Native, not virtualized</b>: Runs as a real macOS process with Screen Recording + Accessibility permissions. It can drive any app you can, including ones that block automation inside browsers or VMs.</p></li><li><p class="paragraph" style="text-align:left;"><b>Structured menu discovery</b>: peekaboo menu returns the full menu tree as JSON, so agents navigate &quot;File → Export → PDF…&quot; by name instead of pattern-matching on screenshots.</p></li><li><p class="paragraph" style="text-align:left;"><b>Multi-screen and multi-Space aware</b>: First-class support for moving windows between Spaces, switching desktops, and targeting elements on specific displays.</p></li><li><p class="paragraph" style="text-align:left;"><b>Drop into OpenClaw</b>: Lives in the repo as skills/peekaboo-cli, so you can install as an OpenClaw skill alongside the 5,400+ others.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow">Voxtral TTS: Outperforms ElevenLabs on naturalness</a></b></h3><div class="image"><a class="image__link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/bbfd6994-a831-402b-9b4b-5ec87a49f5bb/image.png?t=1778564785"/></a></div><p class="paragraph" style="text-align:left;">When it comes to voice agents, naturalness is a key factor. <a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow">Voxtral TTS</a> outperforms ElevenLabs Flash v2.5 on naturalness and matches ElevenLabs v3 quality with emotion-steering support. Lightweight at 4B parameters, built for production.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Wins 58.3% of flagship voice preference tests</b>: In side-by-side human evaluations against ElevenLabs Flash v2.5, Voxtral TTS wins on naturalness across flagship voices and 68.4% on voice customization.</p></li><li><p class="paragraph" style="text-align:left;"><b>Emotion-aware output</b>: Contextual understanding (neutral, happy, sarcastic, and more) determines whether output sounds considered or robotic.</p></li><li><p class="paragraph" style="text-align:left;"><b>70ms model latency, ~9.7x real-time factor</b>: Streams natively and integrates into any existing STT and LLM stack.</p></li><li><p class="paragraph" style="text-align:left;"><b>Voice cloning from 3 seconds of audio</b>: Adapts to tone, personality, rhythm, and intonation. Zero-shot, no fine-tuning required.</p></li><li><p class="paragraph" style="text-align:left;"><b>Open weights under CC BY NC 4.0</b>: Deploy on your own infrastructure, extend to your own voice library.</p></li></ol><p class="paragraph" style="text-align:left;"><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow">Try it now!</a></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/antirez/ds4?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Run DeepSeek V4 Flash locally on a 128GB Mac</b></a></h3><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d8fb5cfa-8ad5-4a8e-b1d6-ab699c9c1a9d/image.png?t=1778564592"/></div><p class="paragraph" style="text-align:left;">The creator of Redis just built a dedicated inference engine for a quasi-frontier model. It might be the most opinionated piece of AI infrastructure released this year.</p><p class="paragraph" style="text-align:left;">Salvatore Sanfilippo (antirez) released <b>DwarfStar4</b>, a purpose-built C + Metal engine that runs DeepSeek V4 Flash, a 284B parameter open-source model with a 1M token context window, locally on a 128GB MacBook at ~27 tokens/second.</p><p class="paragraph" style="text-align:left;">No generic runtime or framework. Just raw C + Metal doing one thing really well. The engine uses an asymmetric 2-bit quantization that fits the entire model in ~81GB, ships with a disk-based KV cache that&#39;s a lifesaver for agent workflows, and built-in APIs that plug straight into agents like OpenClaw, Hermes, Claude Code, Opencode, and Pi. </p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>One Model, Maximum Optimization</b>: DS4 isn&#39;t a generic model runner. It&#39;s a dedicated engine built exclusively for DeepSeek V4 Flash, squeezing out performance that general-purpose tools can&#39;t match.</p></li><li><p class="paragraph" style="text-align:left;"><b>Runs on a MacBook</b>: A specialized 2-bit quantization compresses the 284B model to ~81GB while keeping code generation and tool calling quality intact, making it genuinely usable on 128GB Macs.</p></li><li><p class="paragraph" style="text-align:left;"><b>Disk KV Cache for Agents</b>: Saves session state to your SSD so agent clients that resend large system prompts every request can skip the expensive prefill after the first run. That’s a massive time saver.</p></li><li><p class="paragraph" style="text-align:left;"><b>Agent-Ready Out of the Box</b>: Ships with OpenAI and Anthropic-compatible server APIs plus ready-to-use configs for Claude Code, opencode, and Pi.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://openai.com/daybreak/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">OpenAI enters serious cyber defense with Daybreak</a></b><br>OpenAI just shipped Daybreak, a cyber defense stack built on GPT-5.5 and Codex Security.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/3550ffc9-abbb-4459-9efe-9ed5f576e4ae/image.png?t=1778565362"/><div class="image__source"><span class="image__source_text"><p>iykyk</p></span></div></div><p class="paragraph" style="text-align:left;">The idea is Codex ingests your repo, builds a threat model specific to your codebase, then maps attack paths and validates real vulnerabilities in sandboxed environments. It generates patches, runs them, and sends audit-ready evidence back into your existing security stack. Here’s how the access will work for now: standard GPT-5.5 stays general-purpose, Trusted Access unlocks for verified defenders doing vuln triage and malware analysis, and GPT-5.5-Cyber is for authorized red teaming, pen testing, and controlled validation.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://thinkingmachines.ai/blog/interaction-models/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Thinking Machines show what they’re building with $2B funding</b></a><br>Thinking Machines finally demoed what they’re working on: &quot;interaction models.&quot; At first glance, it feels a lot like the GPT-4o demo from 2 years ago: real-time, audio-video-text. The interesting part is underneath though: a 276B MoE “interaction model” (12B active, 0.40s latency) that handles the live conversation, and a separate background model runs reasoning, searches, and tool calls mid-chat, then feeds results back in. Full-duplex isn&#39;t new (hi Moshi from Kyutai Labs), but the architectural design is interesting, and the early benchmarks on latency and quality are solid. </p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://claude.com/blog/new-in-claude-managed-agents?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Claude Agents can now dream between sessions</b></a><br>Anthropic&#39;s Claude Managed Agents (their hosted agent runtime, launched last month) got a solid update: dreaming, outcomes, and multi-agent orchestration. Dreaming is the one worth paying attention to! It reviews past agent sessions between runs, surfaces recurring mistakes and workflow patterns, and folds them back into memory automatically. Outcomes lets you define a success rubric evaluated by a separate grader in its own context window, looping the agent back until output clears the bar. Multi-agent orchestration does what you&#39;d expect — lead agent delegates to specialist subagents, each with their own model and tools, running in parallel.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/NousResearch/status/2052140057222369541?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>What the Hermes Agent community is actually building</b></a><br>If you’re still wondering what people are using Hermes Agent for, here’s a wall of 200+ of them to inspire you! The team used Hermes Agent itself to scrape the entire <span style="color:#0f1419;font-family:TwitterChirp, -apple-system, &quot;system-ui&quot;, &quot;Segoe UI&quot;, Roboto, Helvetica, Arial, sans-serif;font-size:17px;">internet for these usecases and added them to their docs. And if you&#39;ve found an interesting use case, you can submit your own!</span></p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/Saboo_Shubham_/status/2054017280376455265?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>/goal in Codex CLI, Hermes Agent, and Claude Code</b></a><br>ICYMI and have been manually re-prompting your coding agents with &quot;keep going,&quot; that&#39;s now a solved problem across the board. <code>/goal</code> gives the agent a durable objective with a clear done-condition. It keeps looping - planning, editing, running, and verifying until that condition is actually met or you tell it to stop. Codex CLI shipped it first, Hermes Agent picked it up in v0.13.0, and Claude Code now has its own native version. And here’s an interesting workflow we discovered: use Hermes Agent as an orchestration layer to fire <code>/goal</code> across Codex CLI and Claude Code simultaneously, and track all the running objectives on Hermes&#39;s Kanban board.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/GTG-Labs/sangria?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Sangria</b></a>: An open-source SDK that lets you put a paywall on any API endpoint so AI agents can pay per request automatically via the x402 protocol and USDC on Base. Drops into Express/Fastify/Hono/FastAPI with minimal code.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://stack.cardor.dev/ahk?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>agent-harness-kit</b></a>: An open-source TypeScript CLI that scaffolds a structured multi-agent workflow into any codebase. One command sets up multiple agents (lead, explorer, builder, reviewer), a SQLite task backlog, and a health gate that runs before any agent can start or close work. It&#39;s provider-agnostic (Claude Code, OpenCode) and ships its own local MCP server.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/cursor_ai/status/2052432778743210127?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>/orchestrate</b></a>: Skill that decomposes large tasks into a tree of parallel cloud agents: planners, workers, and verifiers. It runs on the Cursor SDK&#39;s cloud runtime, so each agent gets an isolated VM, and the whole tree reconciles back through git and structured handoffs.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=goal-in-claude-code-codex-and-hermes-agent"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=b5fbfc2f-6b96-495d-800f-f93534024df9&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Build a Multimodal Agentic RAG App with Gemini Embedding 2 and Google ADK</title>
  <description>(100% open source)</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/c9ccb7a4-e58d-4b21-bb2e-c0358b7a3e7e/Build_a_Multimodal_Agentic_RAG_App_with_Gemini_Embedding_2_and_Google_ADK.png" length="2588693" type="image/png"/>
  <link>https://www.theunwindai.com/p/build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk</guid>
  <pubDate>Sat, 09 May 2026 23:07:20 +0000</pubDate>
  <atom:published>2026-05-09T23:07:20Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Ai Tutorial]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">If you have built a RAG app before, you know how quickly the &quot;just retrieve the right chunk&quot; problem fragments the moment your sources stop being plain text. Product PDFs, UI screenshots, recorded calls, demo videos, and support notes all carry the answer your user is asking for, but each lives in its own embedding silo. </p><p class="paragraph" style="text-align:left;">Stitching them together usually means three pipelines, two vector stores, and a glue layer you regret pretty soon.</p><p class="paragraph" style="text-align:left;">In this tutorial, you&#39;ll build a fully-working <b>multimodal agentic RAG app</b> where text, URLs, PDFs, images, audio, and video all share a single 768-dimension embedding space, and a small Google Agent Development Kit (ADK) coordinator turns the retrieved evidence into a grounded, cited answer. </p><p class="paragraph" style="text-align:left;">The two pieces doing the heavy lifting are <b>Gemini Embedding 2</b>, which embeds every modality into the same vector space, and <b>Google ADK</b>, which wraps the retrieval call in an agent that inspects the workspace, calls the retrieval tool, and writes the answer. </p><p class="paragraph" style="text-align:left;">You&#39;ll see exactly how those two pieces compose without any extra orchestration framework.</p><p class="paragraph" style="text-align:left;"><b>Don’t forget to share this tutorial on your social channels and tag Unwind AI (</b><span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.facebook.com/profile.php?id=61561355694033&utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Facebook</a></b></span><b>) to support us!</b></p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk"><span class="button__text" style=""> Subscribe now for FREE - Get instant access to more LLM, RAG & AI Agent tutorials </span></a></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>What We’re Building</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">A multimodal Agentic RAG demo where you can drop in any file or URL and ask questions across the whole index. The same retrieval packet powers both the answer and the citation panel, so the UI never disagrees with the model.</p><p class="paragraph" style="text-align:left;"><b>Key features:</b></p><ul><li><p class="paragraph" style="text-align:left;"><b>Truly multimodal index</b> — text, URLs, PDFs, images, audio, and video all live in one cosine-similarity space.</p></li><li><p class="paragraph" style="text-align:left;"><b>Gemini Embedding 2 with task prefixes</b> — separate prefixes for documents and queries to improve retrieval quality.</p></li><li><p class="paragraph" style="text-align:left;"><b>Google ADK agent</b> — coordinates <code>inspect_embedding_space</code> and <code>retrieve_relevant_context</code> tools, then synthesizes a grounded answer.</p></li><li><p class="paragraph" style="text-align:left;"><b>Single retrieval, two consumers</b> — <code>/ask</code> retrieves once, then passes the same packet to the agent and to the UI.</p></li><li><p class="paragraph" style="text-align:left;"><b>3D PCA embedding view</b> — every source is one point; ask a question and the query and cited sources light up in the same projection.</p></li><li><p class="paragraph" style="text-align:left;"><b>SSRF-safe URL ingestion</b> — private and loopback IPs blocked unless you opt in.</p></li></ul></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>How It Works</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">End-to-end, one question flows like this:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>You add sources.</b> Each source is chunked (text/URL) or uploaded once (PDF/image/audio/video). Every chunk gets a Gemini Embedding 2 vector with the <code>task: retrieval document</code> prefix. Files get a media vector blended with a text annotation vector, so titles still help retrieval.</p></li><li><p class="paragraph" style="text-align:left;"><b>You ask a question.</b> <code>/ask</code> embeds the query with the <code>task: question answering | query</code> prefix, scores every chunk by cosine similarity, keeps the best chunk per source, takes the top <i>k</i>, and projects everything into 3D using power-iteration PCA.</p></li><li><p class="paragraph" style="text-align:left;"><b>The agent runs.</b> <code>_run_adk_agent</code> builds a per-request agent whose <code>retrieve_relevant_context</code> tool is a closure over the already-computed retrieval packet. The agent calls <code>inspect_embedding_space</code>, then &quot;calls&quot; the retrieval tool, then writes a grounded answer with no inline citation IDs.</p></li><li><p class="paragraph" style="text-align:left;"><b>The UI renders.</b> The frontend shows the answer text, the citation panel (built from the same <code>matches</code>), the agent trace, and the updated 3D view with the query point and highlighted sources.</p></li></ol><p class="paragraph" style="text-align:left;">The architectural insight is the <b>single-retrieval contract</b>: one query embedding, one ranked list of matches, two consumers (the agent and the UI). That&#39;s what keeps citations honest.</p></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Prerequisites</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Before we begin, make sure you have the following:</p><ol start="1"><li><p class="paragraph" style="text-align:left;">Python installed on your machine (version 3.12 is recommended)</p></li><li><p class="paragraph" style="text-align:left;">Your <a class="link" href="https://aistudio.google.com/api-keys?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow">Gemini API key</a> for using Gemini Embedding 2</p></li><li><p class="paragraph" style="text-align:left;">A code editor of your choice</p></li><li><p class="paragraph" style="text-align:left;">Basic Python and FastAPI familiarity</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Code Walkthrough</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h4 class="heading" style="text-align:left;"><b>Setting Up the Environment</b></h4><p class="paragraph" style="text-align:left;">First, let&#39;s get our development environment ready:</p><ol start="1"><li><p class="paragraph" style="text-align:left;">Clone the GitHub repository:</p></li></ol><div class="codeblock"><pre><code>git clone https://github.com/Shubhamsaboo/awesome-llm-apps.git</code></pre></div><h6 class="heading" style="text-align:left;">🌟<b> </b><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow">Don&#39;t forget to star the opensource repo to show your support.</a></b></h6><ol start="2"><li><p class="paragraph" style="text-align:left;">Go to the <a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/rag_tutorials/multimodal_agentic_rag?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow"><b>multimodal_agentic_rag</b></a><b> </b>folder:</p></li></ol><div class="codeblock"><pre><code>cd rag_tutorials/multimodal_agentic_rag/backend</code></pre></div><ol start="3"><li><p class="paragraph" style="text-align:left;">Install the <a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/blob/main/voice_ai_agents/insurance_claim_live_agent_team/requirements.txt?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow">required dependencies</a>:</p></li></ol><div class="codeblock"><pre><code>pip install -r requirements.txt</code></pre></div><ol start="4"><li><p class="paragraph" style="text-align:left;">Grab your<a class="link" href="https://aistudio.google.com/api-keys?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow"> Gemini API key from Google AI Studio</a> and set it in your current session:</p></li></ol><div class="codeblock"><pre><code>export GOOGLE_API_KEY=&quot;your-google-ai-studio-key&quot;</code></pre></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h4 class="heading" style="text-align:left;"><b>Creating the App</b></h4><p class="paragraph" style="text-align:left;">Project structure:</p><div class="codeblock"><pre><code>rag_tutorials/multimodal_agentic_rag/
|-- README.md
|-- assets/
|   `-- multimodal-agentic-rag-architecture.png
|-- backend/
|   |-- app_state.py
|   |-- rag_store.py
|   |-- requirements.txt
|   |-- server.py
|   `-- agentic_rag_agent/
|       |-- __init__.py
|       `-- agent.py
`-- frontend/
    |-- index.html
    |-- package.json
    |-- src/
    |   |-- App.tsx
    |   |-- main.tsx
    |   `-- styles.css
    |-- tsconfig.json
    `-- vite.config.ts</code></pre></div><p class="paragraph" style="text-align:left;">We’ll skip the frontend code walkthrough and focus on the backend architecture.</p><h3 class="heading" style="text-align:left;"><b>1. The Shared Store (</b><code>app_state.py</code><b>)</b></h3><p class="paragraph" style="text-align:left;">A single line keeps the in-memory index addressable from both FastAPI and the ADK tools:</p><div class="codeblock"><pre><code>from rag_store import MultimodalRagStore

RAG_STORE = MultimodalRagStore()</code></pre></div><p class="paragraph" style="text-align:left;">Both <code>server.py</code> and the ADK tool functions import <code>RAG_STORE</code> from here, so the agent always sees the same sources you uploaded through the UI.</p><h3 class="heading" style="text-align:left;"><b>2. The Multimodal Store (</b><code>rag_store.py</code><b>)</b></h3><p class="paragraph" style="text-align:left;">This is where most of the interesting code lives. A few constants set the contract:</p><div class="codeblock"><pre><code>EMBED_MODEL = &quot;gemini-embedding-2&quot;
DEFAULT_DIMENSIONS = 768
CHUNK_WORDS = 170
CHUNK_OVERLAP = 35
INLINE_MEDIA_LIMIT_BYTES = 18 * 1024 * 1024</code></pre></div><p class="paragraph" style="text-align:left;">We chunk text into roughly 170-word windows with 35-word overlap and embed each chunk separately. Anything bigger than ~18 MB or any audio/video file goes through the <b>Gemini File API</b> instead of inline bytes.</p><h4 class="heading" style="text-align:left;"><b>Embedding text with task prefixes</b></h4><p class="paragraph" style="text-align:left;">Gemini Embedding 2 supports task prefixes — small instructions like <code>&quot;task: retrieval document&quot;</code> or <code>&quot;task: question answering | query&quot;</code> that tell the model how this content will be used. Documents and queries get different prefixes, which measurably improves retrieval:</p><div class="codeblock"><pre><code>def _embed_text(self, text: str, task_prefix: str) -&gt; list[float]:
    content = f&quot;&#123;task_prefix&#125;: &#123;text&#125;&quot;
    client = self._require_client()

    result = client.models.embed_content(
        model=EMBED_MODEL,
        contents=[content],
        config=types.EmbedContentConfig(output_dimensionality=self.dimensions),
    )
    return result.embeddings[0].values</code></pre></div><p class="paragraph" style="text-align:left;">The interesting bit is <code>output_dimensionality=768</code>: Gemini Embedding 2 supports truncating to smaller, latency-friendlier vectors right at the API call, so you don&#39;t have to pay for storage or cosine math on the full embedding width.</p><h4 class="heading" style="text-align:left;"><b>Embedding files (PDFs, images, audio, video)</b></h4><p class="paragraph" style="text-align:left;">Multimodal is where Gemini Embedding 2 earns its keep. Small images and PDFs go inline; large files and all media go through the File API:</p><div class="codeblock"><pre><code>def _embed_file(self, data, mime_type, title, notes):
    client = self._require_client()

    use_file_api = (
        len(data) &gt; INLINE_MEDIA_LIMIT_BYTES
        or mime_type.startswith(&quot;video/&quot;)
        or mime_type.startswith(&quot;audio/&quot;)
    )
    if use_file_api:
        return self._embed_uploaded_file(data, mime_type, title), &quot;gemini-file-api&quot;

    part = types.Part.from_bytes(data=data, mime_type=mime_type)
    result = client.models.embed_content(
        model=EMBED_MODEL,
        contents=[part],
        config=types.EmbedContentConfig(output_dimensionality=self.dimensions),
    )
    return result.embeddings[0].values, &quot;gemini-inline&quot;</code></pre></div><p class="paragraph" style="text-align:left;">The File API path uploads the file, polls until its state is <code>ACTIVE</code>/<code>SUCCEEDED</code>, embeds via <code>Part.from_uri</code>, and then <b>deletes the uploaded file</b> in a <code>finally</code> block — important so you don&#39;t leak storage on every upload.</p><p class="paragraph" style="text-align:left;">To make a PDF or image still findable by its title (e.g., &quot;the launch deck&quot;), we blend the media vector with a text vector of the title plus user-provided notes:</p><div class="codeblock"><pre><code>media_vector, embedding_path = self._embed_file(...)
annotation_vector = self._embed_text(f&quot;&#123;title&#125;. &#123;notes&#125;&quot;, &quot;task: retrieval document&quot;)
vector = _blend_vectors(media_vector, annotation_vector)  # 68% media / 32% text</code></pre></div><p class="paragraph" style="text-align:left;">This is a small but very effective trick: native multimodal embeddings are great at semantic content, but humans often search by the label they gave the file.</p><h4 class="heading" style="text-align:left;"><b>Search: cosine similarity per chunk, deduplicated per source</b></h4><div class="codeblock"><pre><code>def search(self, query: str, top_k: int = 6) -&gt; dict[str, Any]:
    query_vector = self._embed_text(query, &quot;task: question answering | query&quot;)
    source_vectors = self._source_vectors()
    projections = self._pca_projection(&#123;**source_vectors, query_id: query_vector&#125;)
    ...
    for chunk in self.chunks:
        score = round(_cosine(query_vector, chunk.vector), 4)
        current = source_matches.get(chunk.source_id)
        if not current or score &gt; current[&quot;score&quot;]:
            source_matches[chunk.source_id] = &#123; ... &#125;
    matches = sorted(source_matches.values(), key=lambda m: m[&quot;score&quot;], reverse=True)[:top_k]</code></pre></div><p class="paragraph" style="text-align:left;">Three subtle decisions here: we score every chunk but keep only the <b>best chunk per source</b>, we project source vectors and the query vector together so the 3D view shares the same basis, and we return a fully-formed <code>space</code> snapshot so the frontend never has to ask twice.</p><h4 class="heading" style="text-align:left;"><b>PCA projection in pure Python</b></h4><p class="paragraph" style="text-align:left;">The <code>_pca_projection</code> method runs power iteration to find the top three principal components and projects every vector into 3D — no NumPy, no scikit-learn. That keeps the dependency list short and the projection deterministic per request.</p><h4 class="heading" style="text-align:left;"><b>The retrieval payload</b></h4><p class="paragraph" style="text-align:left;">The agent doesn&#39;t see raw chunks; it sees a clean, model-friendly payload:</p><div class="codeblock"><pre><code>def retrieval_payload(self, results):
    return &#123;
        &quot;provider&quot;: self.embedding_provider,
        &quot;matches&quot;: [
            &#123;
                &quot;citation&quot;: m[&quot;id&quot;],
                &quot;source&quot;: m[&quot;title&quot;],
                &quot;modality&quot;: m[&quot;modality&quot;],
                &quot;similarity&quot;: m[&quot;score&quot;],
                &quot;evidence&quot;: m[&quot;text&quot;],
            &#125;
            for m in results[&quot;matches&quot;]
        ],
    &#125;</code></pre></div><p class="paragraph" style="text-align:left;">This is the exact same packet that <code>/ask</code> returns to the frontend, which is how we guarantee the answer and the citation panel never drift.</p><h3 class="heading" style="text-align:left;"><b>3. The ADK Agent (</b><code>agentic_rag_agent/agent.py</code><b>)</b></h3><p class="paragraph" style="text-align:left;">A short, sharp ADK agent with two tools and a focused instruction:</p><div class="codeblock"><pre><code>def retrieve_relevant_context(query: str, top_k: int = 5) -&gt; dict:
    &quot;&quot;&quot;Retrieve the most relevant multimodal source evidence for a user question.&quot;&quot;&quot;
    return RAG_STORE.retrieval_tool(query=query, top_k=top_k)


def inspect_embedding_space() -&gt; dict:
    &quot;&quot;&quot;Inspect current sources, modalities, dimensions, and embedding provider.&quot;&quot;&quot;
    return RAG_STORE.space_tool()


def build_agent(retrieval_tool=retrieve_relevant_context) -&gt; Agent:
    return Agent(
        name=&quot;multimodal_agentic_rag_agent&quot;,
        model=&quot;gemini-3-flash-preview&quot;,
        description=&quot;Agentic RAG coordinator for a multimodal Gemini Embedding 2 workspace.&quot;,
        instruction=&quot;&quot;&quot;
You are the Google ADK coordinator for a multimodal agentic RAG workspace.

For every user question:
1. Use inspect_embedding_space to understand the current workspace.
2. Use retrieve_relevant_context with the user&#39;s question before answering.
3. Ground the answer in the retrieved evidence. Do not invent facts...
4. Do not include raw citation ids, source ids, bracket citations...
5. Start with a clear direct answer in 2-3 sentences.
6. If helpful, add a short &quot;Key points:&quot; section with simple hyphen bullets.
&quot;&quot;&quot;,
        tools=[inspect_embedding_space, retrieval_tool],
        generate_content_config=genai_types.GenerateContentConfig(
            temperature=0.25,
            max_output_tokens=900,
        ),
    )</code></pre></div><p class="paragraph" style="text-align:left;"><code>build_agent</code> accepts an injectable <code>retrieval_tool</code>. That&#39;s how <code>server.py</code> swaps in a closure that returns the <b>already-computed</b> retrieval packet, instead of letting the agent embed the query a second time.</p><h3 class="heading" style="text-align:left;"><b>4. The FastAPI Server (</b><code>server.py</code><b>)</b></h3><p class="paragraph" style="text-align:left;">The endpoint surface is small and predictable:</p><div style="padding:14px 15px 14px;"><table class="bh__table" width="100%" style="border-collapse:collapse;"><tr class="bh__table_row"><th class="bh__table_header" width="33%"><p class="paragraph" style="text-align:left;">Method</p></th><th class="bh__table_header" width="33%"><p class="paragraph" style="text-align:left;">Endpoint</p></th><th class="bh__table_header" width="33%"><p class="paragraph" style="text-align:left;">What it does</p></th></tr><tr class="bh__table_row"><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>GET</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>/health</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;">Liveness, ADK availability, dimensions, source counts</p></td></tr><tr class="bh__table_row"><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>GET</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>/space</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;">Current sources, points, events, projection metadata</p></td></tr><tr class="bh__table_row"><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>POST</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>/sources/text</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;">Add a text source</p></td></tr><tr class="bh__table_row"><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>POST</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>/sources/url</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;">Fetch and index a public URL (SSRF-protected)</p></td></tr><tr class="bh__table_row"><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>POST</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>/sources/file</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;">Upload PDF, image, audio, or video</p></td></tr><tr class="bh__table_row"><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>DELETE</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>/sources/&#123;id&#125;</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;">Remove a source and its chunks</p></td></tr><tr class="bh__table_row"><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>POST</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;"><code>/ask</code></p></td><td class="bh__table_cell" width="33%"><p class="paragraph" style="text-align:left;">Retrieve once, run ADK answer flow, return citations</p></td></tr></table></div><p class="paragraph" style="text-align:left;">The key piece is <code>/ask</code>. It retrieves once, builds a clean payload, and injects a closure into the agent so it can&#39;t redo the embedding:</p><div class="codeblock"><pre><code>@app.post(&quot;/ask&quot;)
async def ask(req: AskRequest):
    retrieval = await run_in_threadpool(RAG_STORE.search, req.question, req.top_k)
    retrieval_payload = RAG_STORE.retrieval_payload(retrieval)
    answer = await _run_adk_agent(req.question, retrieval_payload)
    trace = [
        &#123;&quot;agent&quot;: &quot;space_inspector&quot;,   &quot;status&quot;: &quot;complete&quot;, &quot;detail&quot;: ...&#125;,
        &#123;&quot;agent&quot;: &quot;retrieval_tool&quot;,    &quot;status&quot;: &quot;complete&quot;, &quot;detail&quot;: ...&#125;,
        &#123;&quot;agent&quot;: &quot;answer_synthesizer&quot;,&quot;status&quot;: &quot;complete&quot;, &quot;detail&quot;: ...&#125;,
    ]
    return &#123;
        &quot;answer&quot;: answer,
        &quot;matches&quot;: retrieval[&quot;matches&quot;],
        &quot;query_point&quot;: retrieval[&quot;query_point&quot;],
        &quot;trace&quot;: trace,
        &quot;space&quot;: retrieval[&quot;space&quot;],
    &#125;</code></pre></div><p class="paragraph" style="text-align:left;">And the closure injection inside <code>_run_adk_agent</code>:</p><div class="codeblock"><pre><code>async def _run_adk_agent(question: str, retrieval: dict[str, Any]) -&gt; str:
    def retrieve_relevant_context(query: str, top_k: int = 6) -&gt; dict:
        &quot;&quot;&quot;Return the exact retrieval packet already embedded for this request.&quot;&quot;&quot;
        return retrieval

    request_agent = build_agent(retrieve_relevant_context)
    request_runner = Runner(agent=request_agent, app_name=APP_NAME, session_service=session_service)
    session = await session_service.create_session(app_name=APP_NAME, user_id=USER_ID)
    content = genai_types.Content(
        role=&quot;user&quot;,
        parts=[genai_types.Part(text=f&quot;Question: &#123;question&#125;\nUse the retrieval tool result for this exact question.&quot;)],
    )
    final_text = &quot;&quot;
    async for event in request_runner.run_async(user_id=USER_ID, session_id=session.id, new_message=content):
        text = _event_text(event)
        if text:
            final_text = text
    return final_text</code></pre></div><p class="paragraph" style="text-align:left;">The agent thinks it&#39;s calling a real retrieval tool. It is — the tool just returns a cached result. This is a clean way to keep agent semantics while skipping a redundant embedding round-trip.</p><p class="paragraph" style="text-align:left;">A couple of safety details worth highlighting:</p><ul><li><p class="paragraph" style="text-align:left;"><b>SSRF protection</b>: <code>_validate_fetch_url</code> rejects non-HTTP schemes and resolves the hostname; if any returned IP is private, loopback, link-local, or reserved, ingestion fails. Set <code>ALLOW_PRIVATE_URLS=true</code> only when you really need it.</p></li><li><p class="paragraph" style="text-align:left;"><b>Threadpool offloading</b>: every blocking call (text chunking, file reads, search, PCA) runs in <code>run_in_threadpool</code> so the FastAPI event loop stays responsive.</p></li><li><p class="paragraph" style="text-align:left;"><b>Configurable CORS</b>: <code>ALLOWED_ORIGINS</code> is read from the env, defaulting to the Vite dev server.</p></li></ul><h3 class="heading" style="text-align:left;"><b>5. The Frontend (very brief)</b></h3><p class="paragraph" style="text-align:left;">The frontend is a single React/Vite app (<code>frontend/src/App.tsx</code>) that wraps three panels: a <b>source manager</b> for adding text/URLs/files, a <b>Q&A panel</b> that calls <code>/ask</code> and renders the answer plus a separate citations list, and a <b>3D embedding view</b> built on Three.js that uses the projection coordinates returned by the backend. Every source is one colored point (color encodes modality), and after a question the query point and the cited sources are highlighted in the same PCA basis.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h4 class="heading" style="text-align:left;"><b>Running the App</b></h4><p class="paragraph" style="text-align:left;">With our code in place, it&#39;s time to launch the app.</p><p class="paragraph" style="text-align:left;">Start the backend</p><div class="codeblock"><pre><code>python server.py</code></pre></div><p class="paragraph" style="text-align:left;">The backend listens on <code>http://localhost:8897</code>.</p><p class="paragraph" style="text-align:left;">Start the frontend in a second terminal:</p><div class="codeblock"><pre><code>cd multimodal_agentic_rag/frontend
npm install
npm run dev -- --port 5177</code></pre></div><p class="paragraph" style="text-align:left;">If your backend lives on a different port, point the frontend at it:</p><div class="codeblock"><pre><code>VITE_API_URL=http://localhost:8897 npm run dev -- --port 5177</code></pre></div><ol start="1"><li><p class="paragraph" style="text-align:left;">Add a few sources — try a paragraph of text, a public URL, a PDF, and an image.</p></li><li><p class="paragraph" style="text-align:left;">Watch them appear as colored points in the embedding view.</p></li><li><p class="paragraph" style="text-align:left;">Ask a question in the Q&A panel.</p></li><li><p class="paragraph" style="text-align:left;">Inspect the answer, the cited sources, and the agent trace.</p></li><li><p class="paragraph" style="text-align:left;">Notice the orange query point land near the sources the agent cites.</p></li></ol><p class="paragraph" style="text-align:left;">A quick health check from the terminal:</p><div class="codeblock"><pre><code>curl http://localhost:8897/health</code></pre></div><p class="paragraph" style="text-align:left;">Expected response shape on a fresh start (the store begins empty):</p><div class="codeblock"><pre><code>&#123;
  &quot;status&quot;: &quot;ok&quot;,
  &quot;adk&quot;: true,
  &quot;setup_error&quot;: &quot;&quot;,
  &quot;sources&quot;: 0,
  &quot;chunks&quot;: 0,
  &quot;dimensions&quot;: 768,
  &quot;provider&quot;: &quot;gemini-embedding-2&quot;,
  &quot;modalities&quot;: &#123;&#125;,
  &quot;chunk_modalities&quot;: &#123;&#125;,
  &quot;projection&quot;: &quot;pca_3d&quot;</code></pre></div></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Working Application Demo</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/1rPZZIfekWs" width="100%"></iframe><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Conclusion</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">You&#39;ve now built a multimodal agentic RAG app that puts text, URLs, PDFs, images, audio, and video into a single Gemini Embedding 2 space, retrieves with cosine similarity over chunked vectors, and uses a tightly-scoped Google ADK agent to write grounded, citation-friendly answers, without a separate vector database, in a few hundred lines of Python.</p><p class="paragraph" style="text-align:left;">A few directions worth exploring from here:</p><ul><li><p class="paragraph" style="text-align:left;"><b>Swap the in-memory store</b> for a managed vector DB (pgvector, Qdrant, Vertex AI Vector Search) and persist the chunk metadata.</p></li><li><p class="paragraph" style="text-align:left;"><b>Add re-ranking</b> with a cross-encoder or a Gemini reranker between cosine retrieval and the agent.</p></li><li><p class="paragraph" style="text-align:left;"><b>Background ingestion</b> with a queue (Celery, RQ, or a simple async worker) so large videos don&#39;t block the API.</p></li><li><p class="paragraph" style="text-align:left;"><b>Evals:</b> wire a small eval set with question/answer pairs and track citation precision and answer faithfulness over changes.</p></li><li><p class="paragraph" style="text-align:left;"><b>Auth + multi-tenancy</b> so different users see different workspaces.</p></li><li><p class="paragraph" style="text-align:left;"><b>Observability</b>: log the retrieval packet alongside the final answer; the single-retrieval contract makes faithfulness audits straightforward.</p></li></ul><p class="paragraph" style="text-align:left;">Keep experimenting with different configurations and features to build more sophisticated AI applications.</p><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this 2-3 times a week, to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;"><b>Don’t forget to share this tutorial on your social channels and tag Unwind AI (</b><span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span><b>) to support us!</b></p></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-multimodal-agentic-rag-app-with-gemini-embedding-2-and-google-adk"><span class="button__text" style=""> Subscribe now for FREE - Get instant access to more LLM, RAG & AI Agent tutorials </span></a></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=a2e798cb-b59a-463c-a2c0-7bc5d0c5e8e9&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>LLM with 12M Context Window </title>
  <description>+ Free web Search and Fetch for your Hermes and Claws</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2d55902f-28ce-4278-9576-683e9fd76860/LLM_with_12M_Context_Window.png" length="2630931" type="image/png"/>
  <link>https://www.theunwindai.com/p/llm-with-12m-context-window</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/llm-with-12m-context-window</guid>
  <pubDate>Wed, 06 May 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-05-06T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/alex_whedon/status/2051663268704636937?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Attention Is All You Need - Just 1/1000th of It</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.tinyfish.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Free web Search and Fetch, for every dev and AI agent</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/ashpreetbedi/status/2049180168200106150?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow"><b>Stop RAG-ing, start Grepping your company knowledge</b></a></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/HeyGen/status/2051697813554405384?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Make your Hermes Agent your video editor with one Skill</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/Manavarya09/design-extract?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow"><b>Extract a website’s complete design system with one command</b></a></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/anatomy-of-agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Anatomy of Agent SKILLS</a></b></p><p class="paragraph" style="text-align:left;">Your agent has a 200k token context window. </p><p class="paragraph" style="text-align:left;">The 400 tokens of instructions it actually needs are buried under tool definitions, reference docs, and brand guides it never asked for. So it ignores them. </p><p class="paragraph" style="text-align:left;">This is the most common reason agents fail in production. It&#39;s not a model problem or a framework problem. </p><p class="paragraph" style="text-align:left;">In this blog, you&#39;ll learn the anatomy of Agent Skills: why the first two lines of SKILL.md are the most important writing you&#39;ll do, and how the LLM itself routes queries to the right skill without embeddings or retrieval layers. </p><p class="paragraph" style="text-align:left;">Read on to learn the five parts that make skills work, then pick one workflow you do every week and ship your first skill today.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/anatomy-of-agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/e45fcf50-862b-47e8-8f08-f74aa7cfc466/Anatomy_of_Agent_SKILLS.png?t=1777437345"/><div class="embed__content"><p class="embed__title"> Anatomy of Agent SKILLS </p><p class="embed__description"> Same agent, much less context </p></div></a></div><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this every week, designed to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://x.com/alex_whedon/status/2051663268704636937?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow"><b>Attention Is All You Need - Just 1/1000th of It</b></a></h3><div class="image"><a class="image__link" href="https://x.com/alex_whedon/status/2051663268704636937?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/911a23ed-5a23-4a31-a328-943daaa2ddaa/image.png?t=1778042931"/></a></div><p class="paragraph" style="text-align:left;">12 million tokens of context in a single pass, and it does it at roughly 1/1000th the attention compute of current frontier models. </p><p class="paragraph" style="text-align:left;">Meet <b>SubQ</b>, the first large language model built on a fully subquadratic sparse attention (SSA) architecture, where compute scales linearly with context length instead of quadratically.</p><p class="paragraph" style="text-align:left;">Transformers compare every token to every other token, which means doubling input length quadruples the compute. SubQ&#39;s architecture focuses only on the token relationships that actually matter, making million-token workloads fast and cheap enough to be practical. </p><p class="paragraph" style="text-align:left;">The model is entering private beta today with an API, a coding agent called SubQ Code, and a long-context search tool called SubQ Search.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Benchmark Performance</b>: SubQ 1M-Preview scores 95% on RULER 128K and 81.8 on SWE-Bench Verified, putting it on par with or ahead of Opus 4.6 and Deepseek V4 Pro on both long-context accuracy and code tasks.</p></li><li><p class="paragraph" style="text-align:left;"><b>Speed & Efficiency</b>: Its sparse attention runs 52x faster than FlashAttention at 1M tokens while requiring 63% less compute.</p></li><li><p class="paragraph" style="text-align:left;"><b>5% Cost of Opus 4.7</b>: Though the pricing is not out, the team claims it costs &lt;5% Opus&#39;s cost at scale, with RULER 128K running for $8 vs ~$2,600. Take this with a grain of salt for now!</p></li><li><p class="paragraph" style="text-align:left;"><b>Private Beta:</b> All three products (API, Code, Search) are available via early access at <a class="link" href="https://subq.ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">subq.ai</a>. </p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h2 class="heading" style="text-align:left;"><a class="link" href="https://www.tinyfish.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow"><b>Free web Search and Fetch, for every dev and AI agent</b></a></h2><div class="image"><a class="image__link" href="https://www.tinyfish.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/afe49e96-4af7-4032-bd7e-5147cb730982/Stand_By_We_re_Going_Live_Shortly__1_.png?t=1778043034"/></a></div><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.tinyfish.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">TinyFish</a></b> just made their Web Search and Fetch endpoints free with generous rate limits. Forever. No credit card or “7-day” trial”. Just sign up and grab your API key.</p><p class="paragraph" style="text-align:left;">Search returns structured JSON for agents. Fetch renders any URL in a real browser with full JavaScript, SPAs, anti-bot, all of it, strips the unnecessary content, and returns clean markdown.</p><p class="paragraph" style="text-align:left;">Everything runs on TinyFish&#39;s own custom Chromium fleet. Owning the stack end-to-end makes their Search and Fetch both free and fast.</p><ul><li><p class="paragraph" style="text-align:left;">Works with Claude Code, OpenClaw, Hermes Agent, Cursor, Codex, and any agent framework</p></li><li><p class="paragraph" style="text-align:left;">Available via API, MCP, Python + TypeScript SDKs, CLI, and Skills</p></li><li><p class="paragraph" style="text-align:left;">One API key. No credit card.</p></li></ul><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.tinyfish.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Grab your API key now!</a></b></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://x.com/ashpreetbedi/status/2049180168200106150?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Stop RAG-ing, start Grepping your company knowledge</a></b></h3><div class="image"><a class="image__link" href="https://x.com/ashpreetbedi/status/2049180168200106150?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/8731e9be-b714-4ba8-8a9c-283806cbc840/image.png?t=1778044571"/></a></div><p class="paragraph" style="text-align:left;">Your company&#39;s best knowledge is rotting in Slack threads nobody will ever search again. </p><p class="paragraph" style="text-align:left;">And no, RAG is not the solution. The index is always stale. The chunks land at the wrong boundaries.</p><p class="paragraph" style="text-align:left;">Turns out, coding agents already cracked this. They don&#39;t search, they <code>grep</code></p><p class="paragraph" style="text-align:left;"><b>Scout</b>, an open-source context agent from Agno, borrows the trick and <i>navigates</i> your information sources live. It connects to Slack, Google Drive, Linear, MCP servers, and more, walking each source&#39;s native API at query time to assemble real answers with real citations. </p><p class="paragraph" style="text-align:left;">As it works, it builds its own wiki and CRM. Say &quot;Josh from Anthropic shared a paper on RLMs&quot; and Scout files Josh as a contact, parses the paper into a wiki page, and links them together. </p><p class="paragraph" style="text-align:left;">The whole thing is open-source, ready to fork and customize.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Context Providers</b>: Instead of exposing dozens of API-specific tools to the main agent, Scout wraps each source behind a thin sub-agent layer. The main agent sees <code>query_slack</code>, not Slack&#39;s twelve endpoints, keeping context clean.</p></li><li><p class="paragraph" style="text-align:left;"><b>Navigation over search</b>: Scout queries live APIs at request time, so a Slack message sent thirty seconds ago is immediately available, and citations always point to real, openable paths.</p></li><li><p class="paragraph" style="text-align:left;"><b>Self-building CRM and wiki</b>: It populates a Postgres-backed CRM and a knowledge wiki as it learns. It even creates new database tables on demand.</p></li><li><p class="paragraph" style="text-align:left;"><b>Ready to clone and use</b>: Ships with Docker Compose, connects to Agno&#39;s AgentOS for multi-user sessions and scheduled tasks, and plugs into Slack with full thread history. More connectors coming soon!</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/OpenAIDevs/status/2050275713824211041?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow"><b>Codex gets a Tamagotchi</b></a><br>OpenAI shipped pets for Codex. These are animated companions that double as a persistent status overlay for your coding agent. The pet visually maps to whether Codex is actively working, waiting for input, or flagging something for review. Think of it as the most adorable process monitor you never asked for.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/claudeai/status/2051679629488865498?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Just another day of the Claude team shipping</a></b><br>Anthropic just dropped ten agent templates built specifically for finance work, like pitchbooks, KYC screening, month-end close, valuation checks, the whole grind. They plug into Cowork and Claude Code or run autonomously as Managed Agents, and they come wired to data sources like Moody&#39;s, Third Bridge, and S&P Capital IQ. Oh, and Claude now works inside Excel, PowerPoint, Word, and Outlook with context that follows you across apps — so yes, your comps model can become a deck without explaining everything twice. Install them as plugins in Cowork and Claude Code</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/HeyGen/status/2051697813554405384?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow"><b>Make your Hermes Agent your video editor with one Skill</b></a><br>Hermes Agents can now spin up full videos, courtesy HyperFrames Agent Skill by HeyGen. Just do <code>$ hermes skills install hyperframes</code>, and your agent becomes a video editor that treats HTML as the source of truth for video. Feed it an X post, a PDF, or a GitHub repo, and it&#39;ll script, animate with GSAP, lay captions over TTS narration, and render a finished MP4, all orchestrated end-to-end by the agent itself.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/Manavarya09/design-extract?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow"><b>Designlang</b></a>: Extract any website&#39;s complete design system with one command. It reads the design system off the live DOM, and emits 17+ files — DTCG tokens, Tailwind config, shadcn theme, Figma variables, motion tokens, typed component anatomy, brand voice, page-intent labels, and a paste-ready prompt pack for v0 / Lovable / Cursor / Claude Artifacts.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.interactlabs.ai/blog-article/introducing-interact-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Interact AI</a></b>: Replaces your static website with an adaptive, conversational interface that recomposes itself per visitor in real time. A founder sees compliance content, a CISO sees security controls, all generated on the fly from your data. It&#39;s not a chatbot widget in the corner; the conversation <i>is</i> the page, and everything the visitor says carries through into signup and product onboarding. </p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/yizhiyanhua-ai/fireworks-tech-graph?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow"><b>fireworks-tech-graph</b></a>: A skill that turns plain descriptions of your system into polished SVG + PNG technical diagrams. It ships with 5 visual styles, 8 diagram types, and built-in knowledge of AI/agent patterns like RAG pipelines, Mem0 memory layers, and multi-agent flows.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=llm-with-12m-context-window"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=189810ca-6424-4ee1-a562-fc8a1bfaf464&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Multi-Agent Kanban board in Hermes Agent</title>
  <description>+ Make Claude Code talk less and ship more</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5898b9bf-a5d1-45fe-8dc4-e84e22bfac0b/Multi-Agent_Kanban_board_in_Hermes_Agent.png" length="1501056" type="image/png"/>
  <link>https://www.theunwindai.com/p/multi-agent-kanban-board-in-hermes-agent</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/multi-agent-kanban-board-in-hermes-agent</guid>
  <pubDate>Mon, 04 May 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-05-04T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/JuliusBrussee/caveman?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Make Claude Code talk less and ship more</b></a></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow"><b>Mistral Vibe just introduced Remote Agents</b></a></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/browser-use/browser-harness?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>LLMs can operate your browser with just 600 LoC and a Skill</b></a></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/NousResearch/status/2050997692977844324?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Multi-agent task coordinator Kanban in Hermes Agent</b></a></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/aattaran/deepclaude?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Claude Code&#39;s agent loop with DeepSeek V4, 17x cheaper</a></b></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/anatomy-of-agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Anatomy of Agent SKILLS</a></b></p><p class="paragraph" style="text-align:left;">Your agent has a 200k token context window. </p><p class="paragraph" style="text-align:left;">The 400 tokens of instructions it actually needs are buried under tool definitions, reference docs, and brand guides it never asked for. So it ignores them. </p><p class="paragraph" style="text-align:left;">This is the most common reason agents fail in production. It&#39;s not a model problem or a framework problem. </p><p class="paragraph" style="text-align:left;">In this blog, you&#39;ll learn the anatomy of Agent Skills: why the first two lines of SKILL.md are the most important writing you&#39;ll do, and how the LLM itself routes queries to the right skill without embeddings or retrieval layers. </p><p class="paragraph" style="text-align:left;">Read on to learn the five parts that make skills work, then pick one workflow you do every week and ship your first skill today.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/anatomy-of-agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/e45fcf50-862b-47e8-8f08-f74aa7cfc466/Anatomy_of_Agent_SKILLS.png?t=1777437345"/><div class="embed__content"><p class="embed__title"> Anatomy of Agent SKILLS </p><p class="embed__description"> Same agent, much less context </p></div></a></div><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this every week, designed to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://github.com/JuliusBrussee/caveman?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Make Claude Code talk less and ship more</a></b></h3><div class="image"><a class="image__link" href="https://github.com/JuliusBrussee/caveman?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a24e0b55-4e9a-4556-9b3b-fd44358a49e8/Screenshot_2026-05-03_at_8.51.40_PM.png?t=1777866706"/></a></div><p class="paragraph" style="text-align:left;">Why use many token when few token do trick? </p><p class="paragraph" style="text-align:left;"><b>Caveman</b> is a Claude Code skill and plugin that makes Claude talk like a caveman. It cuts roughly 75% of output tokens while keeping every bit of technical accuracy intact. </p><p class="paragraph" style="text-align:left;">Claude’s responses and fillers like &quot;Sure, I&#39;d be happy to help you with that&quot; cost tokens and add zero value. With Caveman active, Claude drops articles, kills pleasantries, and eliminates hedging, but leaves code blocks, technical terms, and error messages completely untouched. The result is faster responses, lower API costs, and honestly, funnier code reviews. </p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>75% token reduction</b> – Strips filler words, hedging, and pleasantries from Claude&#39;s responses while preserving all technical content, code blocks, and exact error messages.</p></li><li><p class="paragraph" style="text-align:left;"><b>One-line install</b> – Run <code>claude install-skill JuliusBrussee/caveman</code> And you&#39;re set.</p></li><li><p class="paragraph" style="text-align:left;"><b>Toggle on demand</b> – Activate with <code>/caveman</code> or &quot;caveman mode&quot; plugin and switch back to normal Claude anytime with &quot;stop caveman&quot; or &quot;normal mode.&quot;</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow">Mistral Vibe just introduced Remote Agents</a></b></h3><div class="image"><a class="image__link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d2a92fb9-babd-4c5d-a1f6-e3fb5190ed09/image.png?t=1777869105"/></a></div><p class="paragraph" style="text-align:left;">Mistral just shipped <b>remote agents in Vibe</b> - async coding sessions that run in the cloud, execute in parallel, and ping you when they&#39;re done. </p><p class="paragraph" style="text-align:left;">Powered by <b>Mistral Medium 3.5</b>, a new 128B dense model built for long-horizon agentic tasks. The model merges instruction-following, reasoning, and coding into a single set of weights, scores 77.6% on SWE-Bench Verified, and is now the default across both Vibe CLI and Le Chat. </p><p class="paragraph" style="text-align:left;">ICYMI: Vibe&#39;s CLI is an open-source<span style="color:#222222;"> terminal-native coding agent</span>, working inside your codebase with slash-command skills, custom subagents, and MCP integrations.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>CLI or Le Chat</b> — Spawn remote coding agents from either surface.</p></li><li><p class="paragraph" style="text-align:left;"><b>Teleport local sessions to the cloud</b> — Move an ongoing CLI session to a remote sandbox mid-task, with history, state, and approvals carrying across. When done, the agent opens a PR on GitHub and notifies you.</p></li><li><p class="paragraph" style="text-align:left;"><b>Mistral Medium 3.5</b> — Open-weights released under modified MIT, self-hostable on as few as four GPUs, with configurable reasoning effort.</p></li><li><p class="paragraph" style="text-align:left;"><b>Work mode in Le Chat</b> — A new agentic mode that handles multi-step tasks like inbox triage, research synthesis, and cross-tool workflows, with connectors on by default and explicit approval for sensitive actions.</p></li></ol><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow">Try it today!</a></b></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://github.com/browser-use/browser-harness?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">LLMs can operate your browser with just 600 LoC and a Skill</a></b></h3><div class="image"><a class="image__link" href="https://github.com/browser-use/browser-harness?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ddd1bcf3-785c-42a8-be9f-639f4246a398/Screenshot_2026-05-03_at_9.03.38_PM.png?t=1777867443"/></a></div><p class="paragraph" style="text-align:left;">Browser Use is a tens-of-thousands-line browser framework. Now, the team turned around and shipped a 592-line alternative that lets the LLM be the framework.</p><p class="paragraph" style="text-align:left;"><b>Browser Harness</b> is a new open-source project that gives an LLM a raw CDP websocket connection to Chrome, a small set of starter helpers, and a skill file - that’s it!</p><p class="paragraph" style="text-align:left;">The agent sees exactly how its tools work at the protocol level, writes new ones when it needs them, and self-corrects when something breaks.</p><p class="paragraph" style="text-align:left;">In one interesting example shared by the team, the agent needed to upload a file and it didn’t have the tool, so it wrote one itself and moved on.</p><p class="paragraph" style="text-align:left;">And now, after Browser Harness blew up, the team just packaged it into <a class="link" href="https://github.com/browser-use/desktop-app?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Browser Use Desktop</b></a>, an open-source desktop app that lets Claude Code, Codex, or any coding agent control your browser, without trying to replace your Chrome.</p><p class="paragraph" style="text-align:left;">You can even trigger these agents by texting yourself <code>@BU</code> on WhatsApp.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/NousResearch/status/2050997692977844324?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Multi-agent task coordinator Kanban in Hermes Agent</a></b><br>The latest version of Hermes Agent ships a Kanban multi-agent system where agents, each with their own tools, skills, and personality profiles, claim tasks off a board, fan out work through linked dependencies, and pass files via shared workspaces or git worktrees. There&#39;s a live dashboard, per-task comment threads that both humans and agents write into, heartbeat monitoring, and the whole thing is SQLite-backed so it survives crashes and reboots.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/svpino/status/2050563493708112190?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>20 Claude Code tips you probably missed</b></a><br>This one’s a nice recap or fresher on the best hotkeys and commands to use Claude Code. It&#39;s one of those lists where you go in thinking you know most of it and come out humbled. We definitely didn’t know about 11 and 13 in the list. Worth bookmarking!</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://blog.cloudflare.com/agents-stripe-projects/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent#" target="_blank" rel="noopener noreferrer nofollow"><b>Agents can now create account, buy domain, and deploy</b></a><br>Cloudflare and Stripe just teamed up so coding agents can go from zero to deployed autonomously. This includes creating a Cloudflare account, buying a domain, starting a paid subscription, and shipping to production, all without any human intervention anywhere. Stripe acts as the identity provider and handles payment tokenization (with a $100/month spending cap per provider, so your agent doesn&#39;t go on a domain shopping spree), while Cloudflare auto-provisions accounts on the fly. Just install Stripe CLI + <span style="text-decoration:underline;"><a class="link" href="https://docs.stripe.com/projects?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Stripe Projects plugin</a></span><span style="text-decoration:underline;">,</span> and you can start! The team will also release an open protocol soon.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/aattaran/deepclaude?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Deepclaude</b></a> - A proxy that lets you swap Claude Code&#39;s backend for DeepSeek V4 Pro, OpenRouter, or any Anthropic-compatible API. You keep the same Claude Code agent loop and UX but at roughly 17x lower cost.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/lahfir/agent-desktop?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Agent-desktop</b></a> - A Rust-based CLI that gives AI agents native desktop control by reading OS accessibility trees and outputting structured JSON with stable element references. Think browser automation, but for any native app on your machine.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/heardlabs/heard?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow"><b>Heard</b></a> - A voice companion that reads your AI coding agent&#39;s terminal replies aloud so you don&#39;t have to keep watching the screen. Wispr handles what you say to your agent; Heard handles what it says back. It monitors agent output and uses TTS to keep you in the loop while you do other things.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=multi-agent-kanban-board-in-hermes-agent"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=9efc7ddf-852e-400c-8936-d119d891ec86&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Build a Voice-First Insurance Claim Live Agent Team</title>
  <description>Multi-agent voice-first FNOL app with Google ADK and Gemini Live (100% open source)</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/ea0d05ce-7e82-4576-b8f6-58417c1ea6a7/Build_a_Voice-First_Insurance_Claim_Live_Agent_Team.png" length="1674572" type="image/png"/>
  <link>https://www.theunwindai.com/p/build-a-voice-first-insurance-claim-live-agent-team</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/build-a-voice-first-insurance-claim-live-agent-team</guid>
  <pubDate>Sat, 02 May 2026 18:39:34 +0000</pubDate>
  <atom:published>2026-05-02T18:39:34Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Ai Tutorial]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Filing an insurance claim by phone is messy: the claimant tells an emotional, unstructured story, and an agent on the other end tries to translate it into a rigid form in real time. Voice AI is built for exactly this gap.</p><p class="paragraph" style="text-align:left;">In this tutorial, you will build a <b>voice-first FNOL (first notice of loss) app</b> where a claimant talks naturally and an agent assembles a structured claim packet live. The UI shows the transcript, extracted facts, missing items, routing, and an adjuster-ready handoff.</p><p class="paragraph" style="text-align:left;">The stack is Google ADK for the workflow graph and Gemini Live for the voice. </p><p class="paragraph" style="text-align:left;"><b>What is Google ADK?</b> Google&#39;s framework for building production-ready multi-agent systems. It provides model-agnostic agent orchestration, native tool integration (like Google Search), and a powerful Sequential Agent pattern that lets you chain specialized agents into sophisticated workflows. </p><p class="paragraph" style="text-align:left;"><b>Don’t forget to share this tutorial on your social channels and tag Unwind AI (</b><span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.facebook.com/profile.php?id=61561355694033&utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Facebook</a></b></span><b>) to support us!</b></p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team"><span class="button__text" style=""> Subscribe now for FREE - Get instant access to more LLM, RAG & AI Agent tutorials </span></a></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>What We’re Building</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">A voice and text FNOL intake app that:</p><ul><li><p class="paragraph" style="text-align:left;">Lets the claimant speak or type</p></li><li><p class="paragraph" style="text-align:left;">Streams audio responses back in real time via Gemini Live</p></li><li><p class="paragraph" style="text-align:left;">Extracts structured claim facts into Pydantic schemas</p></li><li><p class="paragraph" style="text-align:left;">Classifies claim type and severity</p></li><li><p class="paragraph" style="text-align:left;">Applies deterministic rules for missing fields, required documents, fraud signals, and safety escalations</p></li><li><p class="paragraph" style="text-align:left;">Builds a Markdown adjuster handoff packet during the call</p></li><li><p class="paragraph" style="text-align:left;">Avoids promising coverage, payment, or liability</p></li></ul></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>How It Works</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Every time the claimant speaks or types, the app does this in the background:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Listen.</b> Voice gets transcribed; text comes in directly.</p></li><li><p class="paragraph" style="text-align:left;"><b>Understand.</b> Gemini reads the full conversation so far and pulls out structured facts like name, policy, date, location, what happened, evidence, injuries.</p></li><li><p class="paragraph" style="text-align:left;"><b>Classify.</b> Gemini decides what kind of claim this is and how severe it looks.</p></li><li><p class="paragraph" style="text-align:left;"><b>Apply rules.</b> Python checks: are required fields missing? What documents will the adjuster need? Are there fraud or safety red flags?</p></li><li><p class="paragraph" style="text-align:left;"><b>Decide routing.</b> The rules pick one of four lanes: <code>ready_for_adjuster</code>, <code>needs_docs</code>, <code>special_investigation</code>, or <code>emergency_escalation</code>. A safety flag (injury, unsafe housing) always wins.</p></li><li><p class="paragraph" style="text-align:left;"><b>Build the packet.</b> A Markdown handoff packet is assembled, plus the next thing the agent should say.</p></li><li><p class="paragraph" style="text-align:left;"><b>Update the UI.</b> Fields, timeline, and packet refresh in the browser. The agent speaks back.</p></li></ol><p class="paragraph" style="text-align:left;">The key idea is the split of labor: <b>Gemini handles messy human language, Python handles decisions that have to be consistent.</b> You don&#39;t want an LLM deciding whether a claim has all its documents or whether to escalate for safety — those are exactly the decisions that need to be the same every time.</p></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Prerequisites</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Before we begin, make sure you have the following:</p><ol start="1"><li><p class="paragraph" style="text-align:left;">Python installed on your machine (version 3.12 is recommended)</p></li><li><p class="paragraph" style="text-align:left;">Your <a class="link" href="https://aistudio.google.com/api-keys?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow">Gemini API key</a> for using Gemini models</p></li><li><p class="paragraph" style="text-align:left;">A code editor of your choice</p></li><li><p class="paragraph" style="text-align:left;">Basic Python, FastAPI, and async familiarity</p></li><li><p class="paragraph" style="text-align:left;">A browser with microphone access for voice</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Code Walkthrough</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h4 class="heading" style="text-align:left;"><b>Setting Up the Environment</b></h4><p class="paragraph" style="text-align:left;">First, let&#39;s get our development environment ready:</p><ol start="1"><li><p class="paragraph" style="text-align:left;">Clone the GitHub repository:</p></li></ol><div class="codeblock"><pre><code>git clone https://github.com/Shubhamsaboo/awesome-llm-apps.git</code></pre></div><h6 class="heading" style="text-align:left;">🌟<b> </b><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow">Don&#39;t forget to star the opensource repo to show your support.</a></b></h6><ol start="2"><li><p class="paragraph" style="text-align:left;">Go to the <a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/voice_ai_agents/insurance_claim_live_agent_team?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow"><b>insurance_claim_live_agent_team</b></a><b> </b>folder:</p></li></ol><div class="codeblock"><pre><code>cd awesome-llm-apps/voice_ai_agents/insurance_claim_live_agent_team</code></pre></div><ol start="3"><li><p class="paragraph" style="text-align:left;">Install the <a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/blob/main/voice_ai_agents/insurance_claim_live_agent_team/requirements.txt?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow">required dependencies</a>:</p></li></ol><div class="codeblock"><pre><code>pip install -r requirements.txt</code></pre></div><ol start="4"><li><p class="paragraph" style="text-align:left;">Grab your<a class="link" href="https://aistudio.google.com/api-keys?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow"> Gemini API key from Google AI Studio</a>.<br></p></li><li><p class="paragraph" style="text-align:left;">Copy the env file and add your key:</p></li></ol><div class="codeblock"><pre><code>cp .env.example .env</code></pre></div><ol start="6"><li><p class="paragraph" style="text-align:left;">In <code>.env</code>:</p></li></ol><div class="codeblock"><pre><code>GOOGLE_GENAI_USE_VERTEXAI=False
GOOGLE_API_KEY=your-google-api-key</code></pre></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h4 class="heading" style="text-align:left;"><b>Creating the App</b></h4><p class="paragraph" style="text-align:left;">Project structure:</p><div class="codeblock"><pre><code>insurance_claim_live_agent_team/
├── __init__.py            # Exports root_agent for `adk web`
├── agent.py               # ADK graph + run_claim_workflow
├── schemas.py             # Pydantic data contracts
├── policies.py            # Deterministic insurance rules
├── examples.py            # Demo claimant prompts
├── requirements.txt
├── .env.example
├── live_demo/
│   ├── server.py          # FastAPI transport
│   ├── index.html         # Frontend
│   ├── styles.css         # Frontend
│   └── app.js             # Frontend
└── README.md</code></pre></div><p class="paragraph" style="text-align:left;">We skip the frontend code (<code>index.html</code>, <code>styles.css</code>, <code>app.js</code>). It&#39;s a static cockpit that talks to two endpoints (<code>/api/message</code>, <code>/api/audio</code>) and a WebSocket (<code>/ws/live</code>) — swap it for any UI you like as long as those three surfaces are honored.</p><h4 class="heading" style="text-align:left;"><code>schemas.py</code><b> — Data Contracts</b></h4><ol start="1"><li><p class="paragraph" style="text-align:left;">Pydantic models for everything passed between steps. Headline schema is <code>ClaimNarrative</code> - this is what the LLM extracts from a claimant&#39;s story:</p></li></ol><div class="codeblock"><pre><code>class ClaimNarrative(BaseModel):
    policyholder_name: str
    policy_number: str
    contact_method: str
    date_of_loss: str
    loss_location: str
    loss_description: str
    estimated_loss_usd: Optional[float] = None
    injuries_or_safety_concerns: list[str] = Field(default_factory=list)
    evidence_available: list[str] = Field(default_factory=list)
    documents_mentioned: list[str] = Field(default_factory=list)
    # ...</code></pre></div><p class="paragraph" style="text-align:left;">One schema per pipeline step: <code>FieldValidation</code>, <code>ClaimClassification</code>, <code>CoverageEvidenceDecision</code>, <code>DocumentChecklist</code>, <code>FraudSafetyGate</code>, <code>ClaimIntakePacket</code>. <code>Literal</code> types lock down values like <code>ClaimType</code> and <code>RoutingDecision</code> so the rules can switch on them safely.</p><h4 class="heading" style="text-align:left;"><code>policies.py</code><b> — Deterministic Insurance Rules</b></h4><ol start="1"><li><p class="paragraph" style="text-align:left;">The boring file, by design. Five public functions, each takes structured inputs and returns structured outputs:</p></li></ol><div class="codeblock"><pre><code>validate_required_claim_fields(claim)
apply_coverage_and_evidence_rules(claim, validation, classification)
generate_document_checklist(claim, classification, coverage)
fraud_signal_and_safety_gate(claim, validation, classification, coverage)
build_claim_intake_packet(...)</code></pre></div><ol start="2"><li><p class="paragraph" style="text-align:left;">Required documents are mapped per claim type:</p></li></ol><div class="codeblock"><pre><code>TYPE_REQUIRED_DOCS = &#123;
    &quot;home_water_damage&quot;: [
        (&quot;Photos or video of damaged areas before cleanup&quot;, &quot;...&quot;),
        (&quot;Mitigation or drying invoice&quot;, &quot;...&quot;),
        (&quot;Repair estimate or contractor assessment&quot;, &quot;...&quot;),
    ],
    &quot;auto_collision&quot;: [...],
    # ...
&#125;</code></pre></div><p class="paragraph" style="text-align:left;"><code>fraud_signal_and_safety_gate</code> is the most consequential — if it sees injury or unsafe-living mentions, it forces routing to <code>emergency_escalation</code> regardless of the rest of the pipeline. </p><h4 class="heading" style="text-align:left;"><code>agent.py</code><b> — The ADK Graph and the Workflow Bridge</b></h4><p class="paragraph" style="text-align:left;">Two jobs: define the <code>SequentialAgent</code> (<code>root_agent</code>), and expose <code>run_claim_workflow</code> for the server.</p><ol start="1"><li><p class="paragraph" style="text-align:left;">The graph is seven steps, alternating LLM and Python nodes:</p></li></ol><div class="codeblock"><pre><code>def create_workflow() -&gt; SequentialAgent:
    return SequentialAgent(
        name=&quot;insurance_claim_live_agent_team&quot;,
        sub_agents=[
            create_normalizer(),                            # LLM
            FunctionNode(name=&quot;ValidateRequiredClaimFields&quot;, ...),   # Python
            create_classifier(),                            # LLM
            FunctionNode(name=&quot;ApplyCoverageAndEvidenceRules&quot;, ...),
            FunctionNode(name=&quot;GenerateDocumentChecklist&quot;, ...),
            FunctionNode(name=&quot;FraudSignalAndSafetyGate&quot;, ...),
            FinalPacketNode(name=&quot;FinalClaimIntakePacket&quot;, ...),
        ],
    )

root_agent = create_workflow()</code></pre></div><ol start="2"><li><p class="paragraph" style="text-align:left;">LLM nodes are <code>LlmAgent</code> with a Pydantic <code>output_schema</code> — that&#39;s what guarantees structured output:</p></li></ol><div class="codeblock"><pre><code>def create_normalizer() -&gt; LlmAgent:
    return LlmAgent(
        name=&quot;NormalizeClaimNarrative&quot;,
        model=MODEL,  # &quot;gemini-3-flash-preview&quot;
        instruction=&quot;&quot;&quot;You are the intake specialist...
        Read the claim narrative and produce a structured ClaimNarrative.
        Do not invent policy numbers, contacts, dates, or evidence.&quot;&quot;&quot;,
        output_schema=ClaimNarrative,
        output_key=&quot;normalized_claim&quot;,
    )</code></pre></div><ol start="3"><li><p class="paragraph" style="text-align:left;">Python nodes wrap a handler that reads from and writes to ADK session state:</p></li></ol><div class="codeblock"><pre><code>class FunctionNode(BaseAgent):
    handler: Callable[[InvocationContext], dict[str, Any]]
    output_key: str

    @override
    async def _run_async_impl(self, ctx):
        result = self.handler(ctx)
        ctx.session.state[self.output_key] = result
        yield _state_event(self.name, self.summary, &#123;self.output_key: result&#125;)</code></pre></div><ol start="4"><li><p class="paragraph" style="text-align:left;">Handlers delegate straight to <code>policies.py</code> — no logic in the graph layer:</p></li></ol><div class="codeblock"><pre><code>def _validate_claim_handler(ctx):
    return validate_required_claim_fields(ctx.session.state.get(&quot;normalized_claim&quot;))</code></pre></div><ol start="5"><li><p class="paragraph" style="text-align:left;">The bridge that lets the live UI use this graph is <code>run_claim_workflow</code>:</p></li></ol><div class="codeblock"><pre><code>async def run_claim_workflow(claimant_transcript, *, session_id=None, user_id=&quot;live-ui&quot;):
    if not claimant_transcript.strip():
        return build_initial_workflow_state()

    session_service = InMemorySessionService()
    await session_service.create_session(app_name=APP_NAME, user_id=user_id, session_id=...)
    runner = Runner(app_name=APP_NAME, agent=root_agent, session_service=session_service)

    message = genai_types.Content(role=&quot;user&quot;, parts=[
        genai_types.Part(text=f&quot;Use this claimant transcript as source of truth.\n\n&#123;claimant_transcript&#125;&quot;)
    ])

    async for _ in runner.run_async(user_id=user_id, session_id=..., new_message=message):
        pass

    state = (await session_service.get_session(...)).state
    # validate each output through Pydantic, return as a dict</code></pre></div><h4 class="heading" style="text-align:left;"><code>live_demo/server.py</code><b> — FastAPI Transport</b></h4><ol start="1"><li><p class="paragraph" style="text-align:left;">After the refactor, <code>server.py</code> does no claim reasoning. It serves the frontend, manages per-claimant sessions, accepts text/audio/WebSocket, and calls <code>run_claim_workflow</code>:</p></li></ol><div class="codeblock"><pre><code>from agent import MODEL, blank_claim, build_initial_workflow_state, run_claim_workflow</code></pre></div><ol start="2"><li><p class="paragraph" style="text-align:left;">The bridge function is tiny:</p></li></ol><div class="codeblock"><pre><code>async def _process_with_adk_graph(session, *, add_claimant_facing_reply):
    workflow = await run_claim_workflow(_claimant_text(session), session_id=session.session_id)
    if add_claimant_facing_reply:
        packet = workflow[&quot;claim_intake_packet&quot;]
        session.transcript.append(&#123;&quot;speaker&quot;: &quot;Agent&quot;, &quot;text&quot;: packet[&quot;claimant_next_message&quot;]&#125;)
    return _state_from_workflow(session, workflow)</code></pre></div><p class="paragraph" style="text-align:left;">The agent&#39;s reply comes from the deterministic packet&#39;s <code>claimant_next_message</code> — no separate LLM call.</p><ol start="3"><li><p class="paragraph" style="text-align:left;">The three transport surfaces:</p></li></ol><div class="codeblock"><pre><code>@app.post(&quot;/api/message&quot;)        # text turn
@app.post(&quot;/api/audio&quot;)          # uploaded audio (transcribed, then same path)
@app.websocket(&quot;/ws/live&quot;)       # Gemini Live bidirectional voice</code></pre></div><p class="paragraph" style="text-align:left;">For <code>/ws/live</code>, the graph runs as a side task (<code>asyncio.create_task(...)</code>), so audio keeps streaming back while the structured packet updates in the background. The user never waits on the rules.</p><h4 class="heading" style="text-align:left;"><code>__init__.py</code><b> — Optional Surface for </b><code>adk web</code></h4><div class="codeblock"><pre><code>from .agent import root_agent
__all__ = [&quot;root_agent&quot;]</code></pre></div><p class="paragraph" style="text-align:left;">Two reasons it exists: it makes the folder a Python package (so the relative imports inside <code>agent.py</code> resolve), and it lets ADK&#39;s CLI tools — <code>adk web</code>, <code>adk run</code>, <code>adk api_server</code> — auto-discover <code>root_agent</code>. </p><p class="paragraph" style="text-align:left;">The live demo imports <code>run_claim_workflow</code> from <code>agent.py</code> directly; <code>adk web</code> exercises the same <code>root_agent</code> through ADK&#39;s dev UI. Both hit the same graph.</p><div class="blockquote"><blockquote class="blockquote__quote"></blockquote></div><h4 class="heading" style="text-align:left;"><code>examples.py</code><b> — Demo Prompts</b></h4><p class="paragraph" style="text-align:left;">Five canned claimant narratives like basement flood, car accident with injuries, stolen laptop without a police report, travel cancellation, and a deliberately vague claim. Useful for <code>adk web</code> testing and rule sanity-checks.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h4 class="heading" style="text-align:left;"><b>Running the App</b></h4><p class="paragraph" style="text-align:left;">With our code in place, it&#39;s time to launch the app.</p><p class="paragraph" style="text-align:left;">Start the backend and frontend with one command:</p><div class="codeblock"><pre><code>python -m uvicorn live_demo.server:app --reload --host 127.0.0.1 --port 4177</code></pre></div><p class="paragraph" style="text-align:left;">Open:</p><div class="codeblock"><pre><code>http://127.0.0.1:4177/index.html</code></pre></div><p class="paragraph" style="text-align:left;">Click the mic to start a live claim, or type. Watch the right panel populate fields and build the handoff packet as you talk. Try a prompt from <code>examples.py</code> for a quick demo.</p><p class="paragraph" style="text-align:left;">To inspect the graph step-by-step:</p><div class="codeblock"><pre><code>adk web</code></pre></div></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Working Application Demo</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><iframe allow="accelerometer; autoplay; clipboard-write; encrypted-media; gyroscope; picture-in-picture" allowfullscreen="true" class="youtube_embed" frameborder="0" height="100%" src="https://youtube.com/embed/h0FgCIgT7Mk" width="100%"></iframe><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Conclusion</b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">You&#39;ve built a voice-first FNOL intake app with a clean separation: schemas as contracts, policies as deterministic logic, an ADK graph as the workflow, and FastAPI as pure transport.</p><p class="paragraph" style="text-align:left;">A few directions to play with from here:</p><ul><li><p class="paragraph" style="text-align:left;">Support new claim types and see how far the pipeline carries you for free</p></li><li><p class="paragraph" style="text-align:left;">Persist conversations so claims survive a backend restart</p></li><li><p class="paragraph" style="text-align:left;">Expand the fraud and safety signals for richer routing</p></li><li><p class="paragraph" style="text-align:left;">Wire the final packet into your real claims system, CRM, or messaging tool</p></li><li><p class="paragraph" style="text-align:left;">Build an eval set to catch regressions when you tweak prompts or rules</p></li></ul><p class="paragraph" style="text-align:left;">The hybrid LLM-plus-rules pattern generalizes far beyond insurance!</p><p class="paragraph" style="text-align:left;">Keep experimenting with different configurations and features to build more sophisticated AI applications.</p><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this 2-3 times a week, to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;"><b>Don’t forget to share this tutorial on your social channels and tag Unwind AI (</b><span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span><b>) to support us!</b></p></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-a-voice-first-insurance-claim-live-agent-team"><span class="button__text" style=""> Subscribe now for FREE - Get instant access to more LLM, RAG & AI Agent tutorials </span></a></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=c3e1602a-4d25-404b-b07b-24fede5d6682&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Open-source Claude Design</title>
  <description>+ Cursor runtime, harness, and models now in Cursor SDK</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f23ac543-b082-4daa-8145-b63ff64b497c/Open-source_Claude_Design.png" length="1712748" type="image/png"/>
  <link>https://www.theunwindai.com/p/open-source-claude-design</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/open-source-claude-design</guid>
  <pubDate>Fri, 01 May 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-05-01T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/nexu-io/open-design?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Open-source replica of Claude Design</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://cursor.com/blog/typescript-sdk?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Cursor runtime, harness, and models now in Cursor SDK</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/supermemoryai/smfs?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Give your agent a filesystem, it gets a memory</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/warpdotdev/status/2049153766977421444?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Warp is now open-source</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/firecrawl/status/2049159615703654469?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Convert any document into MD or JSON for agents</a></b></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.theunwindai.com/p/anatomy-of-agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Anatomy of Agent SKILLS</b></a></p><p class="paragraph" style="text-align:left;">Your agent has a 200k token context window. </p><p class="paragraph" style="text-align:left;">The 400 tokens of instructions it actually needs are buried under tool definitions, reference docs, and brand guides it never asked for. So it ignores them. </p><p class="paragraph" style="text-align:left;">This is the most common reason agents fail in production. It&#39;s not a model problem or a framework problem. </p><p class="paragraph" style="text-align:left;">In this blog, you&#39;ll learn the anatomy of Agent Skills: why the first two lines of SKILL.md are the most important writing you&#39;ll do, and how the LLM itself routes queries to the right skill without embeddings or retrieval layers. </p><p class="paragraph" style="text-align:left;">You&#39;ll also see how skills enable independent team ownership and composition at scale, doing for agents what npm did for JavaScript. </p><p class="paragraph" style="text-align:left;">Read on to learn the five parts that make skills work, then pick one workflow you do every week and ship your first skill today.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/anatomy-of-agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/e45fcf50-862b-47e8-8f08-f74aa7cfc466/Anatomy_of_Agent_SKILLS.png?t=1777437345"/><div class="embed__content"><p class="embed__title"> Anatomy of Agent SKILLS </p><p class="embed__description"> Same agent, much less context </p></div></a></div><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this every week, designed to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/nexu-io/open-design?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Open-source Replica of Claude Design</b></a></h3><div class="image"><a class="image__link" href="https://github.com/nexu-io/open-design?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1dbf8dd8-378c-4836-959d-636af581ee85/Screenshot_2026-04-30_at_5.09.22_PM.png?t=1777594166"/></a></div><p class="paragraph" style="text-align:left;">When something good ships closed-source, the clock starts on the open-source version.</p><p class="paragraph" style="text-align:left;"><b>Open Design</b> is a local-first replica of Claude Design that ditches the lock-in entirely.</p><p class="paragraph" style="text-align:left;">The trick is that it doesn&#39;t ship an agent at all. The daemon scans your machine for Claude Code, Codex, Cursor Agent, Gemini CLI, OpenCode, or whichever it finds. That same agent gets driven by a stack of 19 skills, 71 brand-grade design systems, and a sandboxed iframe preview. </p><p class="paragraph" style="text-align:left;">Same artifact-first loop as Claude Design. None of the cloud, subscription, or model lock-in.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>BYO agent -</b> Whatever CLI is on your <code>PATH</code> becomes the design engine. No Anthropic key needed if you have Claude Code installed; no Claude Code needed if you have Codex or Gemini. </p></li><li><p class="paragraph" style="text-align:left;"><b>19 Skills + 71 design systems -</b> Prototypes, decks, dashboards, mobile apps, plus brand systems for Linear, Stripe, Vercel, Notion, Apple, Tesla, and 65 others.</p></li><li><p class="paragraph" style="text-align:left;"><b>Anti-slop, hard-coded -</b> Turn 1 is always a question form. Turn 2 picks one of five locked visual directions. The agent self-critiques 1-5 before emitting.</p></li><li><p class="paragraph" style="text-align:left;"><b>Skills are just folders -</b> Drop a <code>SKILL.md</code> into <code>skills/</code>, restart, it appears in the picker. Same convention as Claude Code.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://cursor.com/blog/typescript-sdk?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Cursor Runtime, Harness, and Models now in Cursor SDK</b></a></h3><div class="image"><a class="image__link" href="https://cursor.com/blog/typescript-sdk?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/183fbfd4-2289-4ffd-89a6-fb9a8fdf22bc/image.png?t=1777618844"/></a></div><p class="paragraph" style="text-align:left;">The agents that power your Cursor desktop app, CLI, and web app are now just a <code>npm install</code> away. </p><p class="paragraph" style="text-align:left;">Cursor has launched the <b>Cursor SDK</b> in public beta, giving you programmatic access to the same runtime, harness, and models that ship inside Cursor itself. Build custom coding agents in a few lines of TypeScript, run them locally or on Cursor&#39;s cloud against a dedicated VM, and pick any frontier model you want. </p><p class="paragraph" style="text-align:left;">You get codebase indexing, semantic search, MCP server support, skills auto-loaded from <code>.cursor/skills/</code>, hooks for controlling the agent loop, and subagents that delegate tasks with their own prompts and models.</p><p class="paragraph" style="text-align:left;">Cursor has also open-sourced starter projects, including a coding agent CLI, a prototyping tool, and an agent-powered kanban board.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/supermemoryai/smfs?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Give Your Agent a Filesystem, It Gets a Memory</b></a></h3><div class="image"><a class="image__link" href="https://github.com/supermemoryai/smfs?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/442acea0-dc33-4c90-b21a-789e2831b61c/Screenshot_2026-04-30_at_11.53.34_PM.png?t=1777618422"/></a></div><p class="paragraph" style="text-align:left;">“RAG is dead” and “filesystems are all you need” are both kind of right and kind of wrong.</p><p class="paragraph" style="text-align:left;">Agentic search on a codebase is genuinely beautiful - until you try it on folders with your notes, PDFs, transcripts, etc., and watch it fail.</p><p class="paragraph" style="text-align:left;">SMFS is what happens when you stop picking sides. From Supermemory, it’s a mountable filesystem that turns your memory container into a real local directory. Your agent navigates it with the same <code>ls</code>, <code>cat</code>, and <code>grep</code> it already knows, except grep is now semantic by default.</p><p class="paragraph" style="text-align:left;">Drop in PDFs, videos, audio, and docs - SMFS handles extraction, embeddings, and search automatically.</p><p class="paragraph" style="text-align:left;">It’s open source, written in Rust, and plays well with Daytona, E2B, Cloudflare, and Vercel sandboxes.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Semantic by default</b> - Inside a mount, plain grep hits Supermemory’s hybrid semantic index; any flag falls through to standard grep.</p></li><li><p class="paragraph" style="text-align:left;"><b>Multi-agent, shared state</b> - Multiple agents can mount the same container - one writes, the others see it on the next pull.</p></li><li><p class="paragraph" style="text-align:left;"><b>Live profile file</b> - A virtual profile.md at the mount root regenerates on read from every memory in the container, so the agent’s first useful action in a new directory - reading a file - just works.</p></li><li><p class="paragraph" style="text-align:left;"><b>Works where there’s no filesystem</b> - For Cloudflare Workers, edge runtimes, and browser agents, @supermemory/bash exposes the same Unix surface as a single run_bash tool the agent calls.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/warpdotdev/status/2049153766977421444?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Warp is now open-source</b></a><br>Warp is now open source, dropping its client code, roadmap, and a lightweight contributor workflow on GitHub under a mix of MIT (UI crates) and AGPL v3 (everything else). OpenAI is the founding sponsor, GPT models drive the new agent management flows, and Warp still plays nicely with Claude Code, Codex, and Gemini CLI if you&#39;d rather BYO agent. The cherry on top: <a class="link" href="https://build.warp.dev?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">build.warp.dev</a> lets you spectate fleets of Oz agents triaging issues and shipping PRs, then jump into a live web-compiled Warp terminal to follow along.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/GoogleCloudTech/status/2047567704807346675?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>A2A + MCP: Five Ways They Fit Together</b></a><br>This one is a solid playbook on how A2A and MCP fit together. Five multi-agent patterns covering everything from agent discovery to ambient event meshes that react to Pub/Sub streams in the background. The TL;DR: A2A handles agent-to-agent, MCP handles agent-to-tool, and if you&#39;re not using both you&#39;re going to be rewriting things in twelve months. </p><p class="paragraph" style="text-align:left;"><i>The five patterns:</i></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Agent Card Discovery</b> — every A2A agent publishes a JSON doc at a well-known URL describing its capabilities, auth, and rate limits; Agent Registry makes them findable across your org</p></li><li><p class="paragraph" style="text-align:left;"><b>Delegated Specialization</b> — a Python coordinator hands off to a Go security agent, a Java risk agent, a TypeScript marketing agent; each team ships independently and the coordinator never knows the difference</p></li><li><p class="paragraph" style="text-align:left;"><b>Tool Bridge via MCP</b> — one protocol for every data source and API, with 60+ ready integrations and Apigee API Hub auto-converting existing REST APIs into agent-accessible tools</p></li><li><p class="paragraph" style="text-align:left;"><b>Cross-Organization Federation</b> — Agent Gallery&#39;s 100+ validated partner agents (Adobe, ServiceNow, Workday, Salesforce) collaborate with yours while each side enforces its own governance</p></li><li><p class="paragraph" style="text-align:left;"><b>Ambient Event Mesh</b> — agents listen continuously on BigQuery tables and Pub/Sub streams, processing events in the background and escalating to humans via Mission Control when needed</p></li></ol><p class="paragraph" style="text-align:left;">Worth a read if you&#39;re stitching agents together right now and tired of building custom connectors for every API.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://openai.com/index/open-source-codex-orchestration-symphony/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Open-source spec for Codex orchestration</b></a><br>OpenAI just open-sourced Symphony, a spec that turns your Linear board into a control plane for Codex agents — every open ticket gets its own agent workspace, runs continuously, and shows up as a PR for review. Internally, some teams saw landed PRs jump 500% in three weeks, though the more interesting bit is the mental shift: engineers stop babysitting sessions and start managing work. It&#39;s shipped as a <a class="link" href="https://Spec.md?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Spec.md</a> you point at your coding agent of choice, with an Elixir reference implementation — fitting, given the whole thing was built by Codex itself.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/firecrawl/status/2049159615703654469?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Firecrawl /parse</b></a>: New endpoint that converts documents like PDFs, DOCX, and XLSX into clean Markdown or JSON that AI agents can actually work with. It runs on a Rust-based engine that&#39;s up to 5x faster than before, handles files up to 50 MB, and preserves reading order and tables.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/ammaarreshi/gemma-chat?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Gemma Chat</b></a>: Vibe code without internet. This open-source Electron app runs Gemma 4 natively on Apple Silicon. You describe what you want to build, and it writes the code with a live preview that updates as the model types. No internet connection needed after the initial model download.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://skills.sh/browserbase/skills/browser-trace?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>/browser-trace skill</b></a>: Attaches a read-only observer to your agent&#39;s browser session, capturing network requests, DOM snapshots, screenshots, and CDP logs into a filesystem. It pairs with Stagehand, Playwright, or whatever&#39;s already driving the page so you can debug failed runs, attach mid-flight to live automations, or just keep tabs on what your agent is actually doing.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://helmor.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow"><b>Helmor</b></a>: A local-first macOS IDE built around orchestrating coding agents rather than writing code yourself. Built for agent orchestration, review, testing, merge, and everything around the code.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=open-source-claude-design"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=657b9ecb-0a42-4532-b9cd-7cd9878c1b6a&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Anatomy of Agent SKILLS</title>
  <description>Same agent, much less context</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/e45fcf50-862b-47e8-8f08-f74aa7cfc466/Anatomy_of_Agent_SKILLS.png" length="1570774" type="image/png"/>
  <link>https://www.theunwindai.com/p/anatomy-of-agent-skills</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/anatomy-of-agent-skills</guid>
  <pubDate>Wed, 29 Apr 2026 04:37:24 +0000</pubDate>
  <atom:published>2026-04-29T04:37:24Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <category><![CDATA[Ai Blogs]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Your agent has a 200k token context window. The instruction it actually needed was 400 tokens long. It ignored them anyway.</p><p class="paragraph" style="text-align:left;">Those 400 tokens were buried at position 142k under six tool definitions, four reference docs, and a brand guide nobody asked it to read.</p><p class="paragraph" style="text-align:left;">This is the most common reason agents fail in production. It&#39;s not become of the model or the framework. <b>The prompt got too big and the right thing got buried.</b></p><p class="paragraph" style="text-align:left;">Skills are the cleanest fix for this problem. Not a bigger model. Not a bigger window. Not a smarter retriever. Just a small set of design decisions about where context lives and when it loads.</p><p class="paragraph" style="text-align:left;">Five parts make the whole thing work. Here is how each one fits.</p><h2 class="heading" style="text-align:left;"><b>1. A skill is a folder</b></h2><p class="paragraph" style="text-align:left;">A skill is not a Python class or a registered tool. It is a folder on disk with a markdown file inside it.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d678150a-2be2-470e-a6b1-838625abe129/skill1.gif?t=1777436718"/></div><p class="paragraph" style="text-align:left;">SKILL.md is the only required file. References hold docs the agent reads on demand. Assets hold templates and brand files. Scripts hold code the agent can execute. Everything except SKILL.md is opt-in.</p><p class="paragraph" style="text-align:left;">Because a skill is just files, you version it in git. You diff it in pull requests. You copy it between projects. You publish it on GitHub. The format is the contract.</p><p class="paragraph" style="text-align:left;">Same SKILL.md works across Claude Code, Codex, Gemini CLI, Cursor, Agent Development Kit, LangChain, and a growing list of agent tools and frameworks. One folder, many runtimes.</p><h2 class="heading" style="text-align:left;"><b>2. The first two lines are the search index</b></h2><p class="paragraph" style="text-align:left;">Open any SKILL.md and the first thing you see is YAML frontmatter with two fields. These two fields are not just metadata. <b>They are the search index.</b></p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/281dc009-24a3-4a9d-9985-90e298204be6/skill2.gif?t=1777436810"/></div><p class="paragraph" style="text-align:left;">At session start, the agent loads the name and description of every installed skill. Roughly 100 tokens per skill. The body, the references, the scripts, all of it stays on disk.</p><p class="paragraph" style="text-align:left;">When a request comes in, the model reads its own catalog and decides which skill to open. The description is what it matches against. Write a vague description and the skill never fires. Write a sharp one with concrete trigger words and the skill activates exactly when it should.</p><p class="paragraph" style="text-align:left;">This single line is the most important piece of writing in the whole skill. People spend hours on the body and ten seconds on the description, then wonder why their skill never gets used.</p><p class="paragraph" style="text-align:left;">Flip that ratio.</p><h2 class="heading" style="text-align:left;"><b>3. Progressive disclosure is the whole trick</b></h2><p class="paragraph" style="text-align:left;">A single skill can hold tens of thousands of tokens of instructions and reference material. An agent with twenty skills could carry hundreds of thousands of tokens. Multiple full context windows of dead weight before the user has even typed.</p><p class="paragraph" style="text-align:left;">Progressive disclosure prevents this with three loading tiers</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2bdde0a3-17ec-4180-ae10-25def2d23626/skill3.gif?t=1777436831"/></div><ul><li><p class="paragraph" style="text-align:left;"><b>L1 metadata.</b> Name and description. Always loaded at session start. Roughly 100 tokens per skill.</p></li><li><p class="paragraph" style="text-align:left;"><b>L2 instructions.</b> The body of SKILL.md. Loaded only when the description matches a user task. Usually a few thousand tokens.</p></li><li><p class="paragraph" style="text-align:left;"><b>L3 references.</b> Files in references/, assets/, and scripts/. Loaded only when the L2 instructions explicitly point the agent there.</p></li></ul><p class="paragraph" style="text-align:left;">An agent with twenty skills installed pays the same upfront cost as an agent with one. Add a 21st skill tomorrow and yesterday&#39;s tasks cost what they cost yesterday.</p><p class="paragraph" style="text-align:left;">The catch is that progressive disclosure only saves tokens if you actually use the tiers. Stuff every example into SKILL.md and the body balloons to 10K tokens. Now every task that triggers the skill pays that cost.</p><p class="paragraph" style="text-align:left;">Keep SKILL.md short. Push edge cases, long examples, and reference tables into references/. The agent pulls them in only when it needs them.</p><h2 class="heading" style="text-align:left;"><b>4. Agent routes the query</b></h2><p class="paragraph" style="text-align:left;">When a request comes in, the model does what you would do looking at a toolbox. Read the labels. Pick the right one. Open it.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9444af0b-a658-4907-ba54-94dd3918a1de/skill4.gif?t=1777436857"/></div><p class="paragraph" style="text-align:left;">The user says <i>&quot;clean up this messy CSV and dedupe rows.&quot;</i> The model scans the catalog of descriptions. <i><b>pdf-forms</b></i>, low match. brand-voice, low match. <i><b>data-clean: CSV cleanup, dedup, nulls</b></i>, strong match. The body of <i><b>data-clean</b></i> loads. Work begins.</p><p class="paragraph" style="text-align:left;">Two details matter here.</p><p class="paragraph" style="text-align:left;"><b>The match is not vector retrieval.</b> The model decides directly from descriptions inside its own context. No embedding step. No similarity score. No separate routing layer. The LLM is the router.</p><p class="paragraph" style="text-align:left;"><b>The match is also exclusive.</b> Only one skill activates per task. The others stay at L1. Their bodies never enter the context window. The cost of skills you do not need is essentially zero.</p><p class="paragraph" style="text-align:left;">This is what makes skills feel different from MCP tools or function calls. Tools are always loaded, always visible, always paid for. Skills load only when relevant.</p><h2 class="heading" style="text-align:left;"><b>5. Composition without context bloat</b></h2><p class="paragraph" style="text-align:left;">Scale this up. One agent, eight skills installed. Three different tasks come in over the course of a session.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/dc38fc4f-ecb2-45b2-b3b5-19494f517994/skill5.gif?t=1777436886"/></div><p class="paragraph" style="text-align:left;">The skills the agent did not use stay at L1. Roughly 100 tokens each, no body, no references. The body cost is paid only on the tasks that need it.</p><p class="paragraph" style="text-align:left;">The pattern matters beyond context economics.</p><p class="paragraph" style="text-align:left;"><b>Teams can ship skills independently.</b> The data team owns <i><b>data-clean</b></i> and <i><b>sql-runner</b></i>. The design team owns brand-voice and deck-build. The platform team wires up the agent. Nobody coordinates. Nobody merges prompts. Nobody rebuilds the system prompt every time a new capability lands.</p><p class="paragraph" style="text-align:left;">Skills are doing for agents what npm did for JavaScript. Small, focused, composable units behind a clear interface.</p><p class="paragraph" style="text-align:left;">The package manager won JavaScript. The same shape is going to win agents.</p><h2 class="heading" style="text-align:left;"><b>With Agent skills vs without</b></h2><p class="paragraph" style="text-align:left;">Put the five parts together and the gap between an agent built with skills and one built without is sharp enough to draw on a single page.</p><div class="image"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/aee728a4-5188-4a34-894c-36bd79e093e2/withwithout.gif?t=1777436941"/></div><p class="paragraph" style="text-align:left;">If you build agents and you have not written a skill yet, pick one workflow you do every week. Write a skill for it. One folder, one SKILL.md, version it in git. Watch the agent activate it.</p><p class="paragraph" style="text-align:left;">The format is just files. The leverage is enormous.</p><p class="paragraph" style="text-align:left;">Start today. One skill. One workflow. See what changes.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">We share in-depth blogs and tutorials like this 2-3 times a week, to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=anatomy-of-agent-skills" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;"><b>Don’t forget to share this tutorial on your social channels and tag Unwind AI (</b><span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=anatomy-of-agent-skills" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=anatomy-of-agent-skills" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=anatomy-of-agent-skills" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span><b>) to support us!</b></p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=anatomy-of-agent-skills"><span class="button__text" style=""> Subscribe now for FREE - Get instant access to more LLM, RAG & AI Agent tutorials </span></a></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=65789665-b2f4-47c2-be48-c003fa7beb11&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Slack for AI Employees </title>
  <description>+ OpenClaw Clawsweeper for PRs and Issues</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9f3de46f-c239-4d42-8a81-eec9ae6f62ac/Slack_for_AI_Employees.jpg" length="240629" type="image/jpeg"/>
  <link>https://www.theunwindai.com/p/slack-for-ai-employees</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/slack-for-ai-employees</guid>
  <pubDate>Mon, 27 Apr 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-04-27T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/nex-crm/wuphf?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">Slack for AI Employees with Karpathy’s Wiki</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow">Own the full audio stack: transcript, live STT, and voice synthesis</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/openclaw/clawsweeper?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">Clawsweeper runs 50 Codex in parallel to scan and close PRs/issues</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/deepseek_ai/status/2047516922263285776?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">Open-source DeepSeek V4 is a bigger deal than R1</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/liu8in/status/2048157838497992926?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>Turn Claude Design into a Motion Video Designer</b></a></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/how-to-run-a-24-7-ai-agent-that-grows-with-you?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">How to Run a 24/7 AI Agent that Grows with You</a></b></p><p class="paragraph" style="text-align:left;">I’ve been running a bunch of agents every day for months.</p><p class="paragraph" style="text-align:left;">The problem is that I was the one constantly tweaking and learning while they just held onto context.</p><p class="paragraph" style="text-align:left;">I tested it out by putting the same Monica agent on Hermes at the same time.</p><p class="paragraph" style="text-align:left;">She started creating her own playbook from my edits and keeps getting better without me.</p><p class="paragraph" style="text-align:left;">In this blog, I detailed how I did it. You’ll also understand the difference between an agent you manage and one that actually grows with you.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/how-to-run-a-24-7-ai-agent-that-grows-with-you?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/dfbcddab-2347-4afa-8972-4c6855a0cc2a/How_to_Run_a_247_AI_Agent_that_Grows_with_You.png?t=1776233672"/><div class="embed__content"><p class="embed__title"> How to Run a 24/7 AI Agent that Grows with You </p><p class="embed__description"> My agent started teaching itself while I slept </p></div></a></div><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this every week, designed to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/nex-crm/wuphf?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>Slack for AI Employees with Karpathy’s Wiki</b></a></h3><div class="image"><a class="image__link" href="https://github.com/nex-crm/wuphf?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/b979da93-4d0b-406b-99f3-b3b16a6d4c0e/Screenshot_2026-04-26_at_10.05.01_PM.png?t=1777266305"/></a></div><p class="paragraph" style="text-align:left;">This is Slack reimagined, where your AI employees communicate with each other and have a shared memory.</p><p class="paragraph" style="text-align:left;">WUPHF is an open-source working version of this with a very fun structure: a multi-agent &quot;office&quot; where a CEO, PM, engineers, designer, and CRO coordinate with each other and ship work. </p><p class="paragraph" style="text-align:left;">One command (<code>npx wuphf</code>) opens the office in your browser, complete with a <code>#general</code> channel and the team already inside, claiming tasks. </p><p class="paragraph" style="text-align:left;">Underneath it all sits the Karpathy-style wiki with a plain markdown + git for storage, BM25 + SQLite for search, and that’s it! Agents draft in private notebooks, the durable stuff gets promoted to a shared team wiki, and a &quot;Pam the Archivist&quot; git identity signs every wiki commit. </p><p class="paragraph" style="text-align:left;">It&#39;s MIT, self-hosted, BYO API keys, and the whole brain lives locally.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;">One command launches a tmux/web office with a CEO, PM, engineers, designer, CMO, and CRO already inside a shared channel.</p></li><li><p class="paragraph" style="text-align:left;">The shared brain is the Karpathy markdown wiki: per-agent notebooks for rough drafts, with reviewed entries promoted to the team wiki.</p></li><li><p class="paragraph" style="text-align:left;">Fresh sessions per turn keep input flat at ~87k tokens with a 97% cache hit rate. Accumulated-session orchestrators climb to ~484k over the same window.</p></li><li><p class="paragraph" style="text-align:left;">Mix runtimes freely - Claude Code, Codex, and OpenClaw agents can share the same office and the same wiki.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow"><b>Own the full audio stack: transcription, live STT, and voice synthesis</b></a></h3><div class="image"><a class="image__link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/bb014717-5e8e-470e-b1be-523aa807733b/image.png?t=1777265349"/></a></div><p class="paragraph" style="text-align:left;"><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow"><b>Mistral Speech</b></a> gives you three open-weights models in one coherent stack: batch transcription, live STT, and natural voice synthesis. Build and deploy a full audio pipeline via API, or self-host on-prem and at the edge.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Voxtral Mini Transcribe 2</b> - ~4% word error rate on FLEURS at $0.003/min. Speaker diarization, context biasing for up to 100 terms, and recordings up to 3 hours in a single request.</p></li><li><p class="paragraph" style="text-align:left;"><b>Voxtral Realtime</b> - Live transcription with latency configurable down to sub-200ms. At 480ms, within 1-2% WER: the voice agent sweet spot. Runs on phones, laptops, and smartwatches.</p></li><li><p class="paragraph" style="text-align:left;"><b>Voxtral TTS</b> - 70ms model latency, ~9.7x real-time factor, voice cloning from 3 seconds of audio across 9 languages. Emotion-steering built in.</p></li><li><p class="paragraph" style="text-align:left;"><b>Deployable anywhere</b> - On-prem, private cloud, serverless, or Mistral Compute. GDPR and HIPAA-compliant.</p></li></ol><p class="paragraph" style="text-align:left;"><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow"><b>Try it now!</b></a></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/openclaw/clawsweeper?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>Clawsweeper Runs 50 Codex in Parallel to Scan and Close PRs/Issues</b></a></h3><div class="image"><a class="image__link" href="https://github.com/openclaw/clawsweeper?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d422832f-1581-4a41-8f21-80db6afb22c4/Screenshot_2026-04-26_at_10.04.25_PM.png?t=1777266269"/></a></div><p class="paragraph" style="text-align:left;">4,000 GitHub issues closed in a single day, by a bot, not a burnout. </p><p class="paragraph" style="text-align:left;">The creator of OpenClaw just shipped <b>ClawSweeper</b>, an automated maintenance bot that runs 50 Codex instances in parallel to scan, triage, and close stale or already-resolved issues and PRs across the OpenClaw repo.</p><p class="paragraph" style="text-align:left;">Here&#39;s the context: OpenClaw&#39;s explosive growth left the repo sitting on 13,500+ open items. That’s an impossible backlog for any human team. ClawSweeper tackles this by running parallel review shards powered by GPT-5.4, each checking out the repo at <code>main</code> and deeply analyzing whether an issue has already been fixed, can&#39;t be reproduced, or is just too incoherent to act on. It only closes when the evidence is strong, and maintainer-authored items are never touched.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;">ClawSweeper only closes for five specific reasons: already implemented, unreproducible, belongs as a skill/plugin, incoherent, or stale past 60 days. Everything else stays open, untouched.</p></li><li><p class="paragraph" style="text-align:left;">Codex runs without GitHub write tokens and against a read-only checkout. If it leaves even a single untracked file behind, the review fails.</p></li><li><p class="paragraph" style="text-align:left;">Every proposed close includes a hash of the issue&#39;s state at review time. If anything changes between proposal and apply, the closure is automatically skipped.</p></li><li><p class="paragraph" style="text-align:left;">A planner scans all open items and prioritizes by activity: active issues first, then PRs, then recent issues, then older weekly reviews. The most relevant work gets handled first.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://claude.com/blog/building-agents-that-reach-production-systems-with-mcp?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>Building agents that reach production systems with MCP</b></a><br>Anthropic just dropped a guide on wiring up agents to production systems via MCP. It leans into hard-won patterns from running 200+ MCP servers at scale. If you&#39;re building agents that need to talk to real infrastructure behind auth, this is the playbook. Some of our takeaways:</p><ul><li><p class="paragraph" style="text-align:left;">Don&#39;t mirror your API 1:1 into MCP tools. Group tools around what the agent is actually trying to do. One <code>create_issue_from_thread</code> beats four chained primitives every time.</p></li><li><p class="paragraph" style="text-align:left;">If your service has hundreds of operations (think AWS, K8s), skip the mega-toolset. Expose a thin tool surface that accepts code: the agent writes a short script, your server runs it in a sandbox against your API, and only the result returns. Cloudflare’s MCP inspired by Code Mode is a great example.</p></li><li><p class="paragraph" style="text-align:left;">Instead of stuffing all tool definitions into context upfront, lazy-load them at runtime. Pair that with programmatic tool calling for another ~37% token reduction.</p></li><li><p class="paragraph" style="text-align:left;">MCP gives agents access to tools; skills teach them <i>how</i> to use those tools well. Bundling both as plugins gives the best of both.</p></li></ul><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/deepseek_ai/status/2047516922263285776?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>Open-source DeepSeek V4 is a bigger deal than R1</b></a><br>DeepSeek V4 is a 1.6T-param MoE model with 1M native context, MIT-licensed, scoring within a hair of Opus 4.6 on SWE-bench (80.6% vs 80.8%), and it costs $3.48 per million output tokens. For reference, GPT-5.4 charges $30. That&#39;s a different economic model entirely, and it&#39;s open-weight so you can self-host it. Add in that it was trained on Huawei Ascend 950 chips (not Nvidia). This a serious inflection point: the best open-source model in the world now sits at ~95% of frontier performance for ~12% of the price.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://openai.com/index/introducing-openai-privacy-filter/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>OpenAI open-sources a local-first PII redaction model</b></a><br>OpenAI dropped Privacy Filter, an open-weight 1.5B-parameter model (only ~50M active) that detects and redacts PII across eight categories like names, addresses, API keys, and the works in a single pass with a 128k context window. It runs locally on a laptop or even in a browser, so your sensitive text never has to leave your machine, and it ships under Apache 2.0, so you can fine-tune and commercialize freely.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/liu8in/status/2048157838497992926?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>HyperFrames</b></a>: Open-source framework by HeyGen that turns HTML into rendered video. It ships as a Skill you can add to Claude (Code or Design) that teaches it how to write video compositions in HTML. Just prompt what you want, Claude authors the scenes with proper timelines and animations, and you render to MP4 locally.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/stainlu/hermes-labyrinth?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">Hermes Labyrinth</a></b>: An observability plugin for Hermes Agent. It turns agent activity into a map of crossings: prompts, tool calls, tool results, failures, model switches, subagents, approvals, memory hits, redactions, context compression, cron runs, and reportable evidence.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/davila7/claude-code-templates/tree/main/cli-tool/components/hooks/monitoring?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>Claude Code Hook - Context Timeline</b></a>: A monitoring hook that visualizes your main agent&#39;s context window and all spawned subagents as a live timeline from session start. It makes it way easier to track what&#39;s happening across parallel contexts than squinting at console output.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/FarzaTV/status/2048203459976188261?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow"><b>Clicky</b></a>: A free Mac menu bar app by Farza that lets you talk to AI and spin up agents that can build apps, do research, and interact with native Apple apps like Notes, Calendar, and Reminders. It&#39;s designed for consumers with zero setup. Just install and start talking.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=slack-for-ai-employees"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=dbcfd539-5bd6-419f-bd68-04ddce602bbc&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>GPT 5.5 &amp; Kimi K2.6 Built for Autonomous AI Agents </title>
  <description>+ Google Agent CLI to build and deploy ADK agents</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/4eb67d4f-4708-4eea-9725-5a9d521d6a5c/GPT_5.5___Kimi_K2.6_Built_for_Autonomous_AI_Agents.png" length="2850798" type="image/png"/>
  <link>https://www.theunwindai.com/p/gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents</guid>
  <pubDate>Fri, 24 Apr 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-04-24T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://blog.cloudflare.com/introducing-agent-memory/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Cloudflare&#39;s persistent managed memory for AI agents</b></a></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow"><b>Talk to your coding agent, have it talk back</b></a></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.kimi.com/blog/kimi-k2-6?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">New Kimi model for your Hermes and OpenClaw</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://openai.com/index/introducing-gpt-5-5/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">OpenAI drops its new frontier model GPT 5.5</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/google/agents-cli?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Google Agent CLI to build and deploy ADK agents from terminal</b></a></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/how-to-run-a-24-7-ai-agent-that-grows-with-you?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">How to Run a 24/7 AI Agent that Grows with You</a></b></p><p class="paragraph" style="text-align:left;">I’ve been running a bunch of agents every day for months.</p><p class="paragraph" style="text-align:left;">The problem is that I was the one constantly tweaking and learning while they just held onto context.</p><p class="paragraph" style="text-align:left;">I tested it out by putting the same Monica agent on Hermes at the same time.</p><p class="paragraph" style="text-align:left;">She started creating her own playbook from my edits and keeps getting better without me.</p><p class="paragraph" style="text-align:left;">In this blog, I detailed how I did it. You’ll also understand the difference between an agent you manage and one that actually grows with you.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/how-to-run-a-24-7-ai-agent-that-grows-with-you?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/dfbcddab-2347-4afa-8972-4c6855a0cc2a/How_to_Run_a_247_AI_Agent_that_Grows_with_You.png?t=1776233672"/><div class="embed__content"><p class="embed__title"> How to Run a 24/7 AI Agent that Grows with You </p><p class="embed__description"> My agent started teaching itself while I slept </p></div></a></div><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this every week, designed to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://blog.cloudflare.com/introducing-agent-memory/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Cloudflare&#39;s Managed Persistent Memory for AI Agents</b></a></h3><div class="image"><a class="image__link" href="https://blog.cloudflare.com/introducing-agent-memory/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/fe206412-ab2a-4559-8002-ce4dbfedb8ad/Screenshot_2026-04-23_at_10.28.21_PM.png?t=1777008506"/></a></div><p class="paragraph" style="text-align:left;">There&#39;s a weird tension at the heart of every long-running agent: keep everything in context and watch quality rot, or prune and lose important things. </p><p class="paragraph" style="text-align:left;">Cloudflare&#39;s new <b>Agent Memory</b>, out in private beta, picks a third option - extract the important stuff from conversations at compaction time and make it retrievable on demand. The API is deliberately small: ingest a conversation, remember something specific, recall what you need, list, or forget. </p><p class="paragraph" style="text-align:left;">Unlike local-first setups like OpenClaw where memory is md files the model writes itself, this is a managed architecture built for memory at scale, where ingestion quality and retrieval sophistication actually matter. You can call it from a Worker via binding or hit the REST API from anywhere, and it plugs directly into the Cloudflare Agents SDK.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>What sticks, what fades -</b> Memories get sorted into facts, events, instructions, and tasks. Facts update in place instead of piling up, and tasks are ephemeral by design.</p></li><li><p class="paragraph" style="text-align:left;"><b>Retrieval that actually holds up -</b> Five search channels run in parallel and get fused together, so queries work whether you phrase them as keywords or questions.</p></li><li><p class="paragraph" style="text-align:left;"><b>Smaller models often win -</b> Llama 4 Scout handles the structured work, Nemotron 3 does synthesis. Bigger wasn&#39;t better for most stages.</p></li><li><p class="paragraph" style="text-align:left;"><b>Already running inside Cloudflare -</b> It’s already powering as an OpenCode plugin with shared team memory, an agentic code reviewer, and an internal chatbot. Early access is open via waitlist if you&#39;re building agents on Cloudflare.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow"><b>Talk to Your Coding Agent, Have it Talk Back</b></a></h3><div class="image"><a class="image__link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/cb27dba4-88db-4734-8c36-f1e924633b68/image.png?t=1777007152"/></a></div><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow"><b>Mistral Vibe</b></a> is a terminal-native coding agent powered by Mistral&#39;s models that lets you explore, modify, and interact with your codebase through natural language. The latest update adds voice mode: toggle it on to give instructions by voice instead of typing, and enable text-to-speech to have Vibe read its output back to you.</p><p class="paragraph" style="text-align:left;"><b>Key highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Voice input</b> - Toggle voice mode on with /voice to talk to Vibe instead of typing.</p></li><li><p class="paragraph" style="text-align:left;"><b>Text-to-speech readouts</b> - Vibe can read its output back to you, keeping you in flow without having to read through long responses.</p></li><li><p class="paragraph" style="text-align:left;"><b>Chat rewind</b> - /rewind lets you navigate back through your conversation and fork from any point. Useful for exploring a different approach without starting a new session.</p></li><li><p class="paragraph" style="text-align:left;"><b>Parallel tool execution</b> - Vibe now reads multiple files, runs searches, and calls several sub-agents simultaneously to speed up sessions.</p></li><li><p class="paragraph" style="text-align:left;"><b>Session resume</b> - Pick up any previous session with /resume and full context intact. Or launch with vibe --continue to jump straight back into your last session.</p></li></ol><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow">Get started now!</a></b></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://www.kimi.com/blog/kimi-k2-6?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>New Kimi Model for Your Hermes and OpenClaw</b></a><b> </b>🦞</h3><div class="image"><a class="image__link" href="https://www.kimi.com/blog/kimi-k2-6?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/c2a2750f-ba8f-44e7-8259-be0a4d2a16d6/image.png?t=1777000404"/></a></div><p class="paragraph" style="text-align:left;">This open-source model ran autonomously for 5 days straight, managing incidents, running code, handling system ops with zero human babysitting. </p><p class="paragraph" style="text-align:left;">In another run, it spent 13 hours rewriting an 8-year-old financial matching engine, made 1,000+ tool calls, and more than doubled throughput. </p><p class="paragraph" style="text-align:left;">That’s <b>Kimi K2.6</b> by you for by Moonshot AI, a 1T MoE model (32B active) built for long-horizon coding, proactive 24/7 agents, and multi-agent orchestration. </p><p class="paragraph" style="text-align:left;">The benchmarks hold up: 58.6 on SWE-Bench Pro, 80.2 on SWE-Bench Verified, 66.7 on Terminal-Bench 2.0, and 54.0 on HLE with tools, sitting alongside Claude Opus 4.6 and GPT-5.4 while being fully open-weight. </p><p class="paragraph" style="text-align:left;">Beyond text and code, Moonshot has clearly leaned hard into design this time — K2.6 turns a single prompt into Awwwards-level frontends with hero sections, scroll animations, and even generated image/video assets. </p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Agent Swarm baked into the model</b> - K2.6 can decompose one prompt into up to 300 specialist sub-agents running 1000s of coordinated steps in parallel, with the model itself acting as orchestrator. </p></li><li><p class="paragraph" style="text-align:left;"><b>Use with OpenClaw and Hermes </b>- K2.6 is explicitly tuned to power proactive always-on agent frameworks like OpenClaw and Hermes. Definitely worth trying if you’re still struggling with the models.</p></li><li><p class="paragraph" style="text-align:left;"><b>Claw Groups for team-style multi-agent work</b> - A research preview where you and your teammates can bring your own agents, running on different devices, using different models, with their own tools and memory, into one shared workspace. K2.6 sits at the center as coordinator, assigning work to whichever agent is best suited and reassigning tasks if one fails. </p></li></ol><p class="paragraph" style="text-align:left;">The model is OpenAI/Anthropic-API compatible and available now via <a class="link" href="https://kimi.com?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">kimi.com</a>, the Moonshot API, Kimi Code CLI, and Hugging Face.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://openai.com/index/codex-for-almost-everything/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Codex goes from coding agent to desktop Copilot</b></a><br>The new Codex takes its &quot;(almost) everything&quot; tagline pretty seriously: background computer use lets multiple agents click around your Mac without stepping on your own work, a native browser turns frontend iteration into a point-and-annotate loop instead of a prompt-and-hope one. Add image generation inside the agent loop, a memory preview, and cross-session automations that resume tasks across days, and OpenAI is quietly reshaping Codex into a desktop control surface rather than a coding sidekick. Rolling out now to Codex desktop app users.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://blog.cloudflare.com/email-for-agents/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Email inbox for your AI agents</b></a><br>Every agent can now have their own inbox with Cloudflare&#39;s Email Service in public beta. Besides just inbox, it enables a bunch of features like asynchronous work, persistent state across replies, and a channel users already know how to use. It&#39;s a neat counterpoint to the default assumption that agents need bespoke chat interfaces. Also shipping: an Email MCP server, CLI tooling, and an open-source Agentic Inbox you can deploy in one click.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://openai.com/index/introducing-gpt-5-5/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>OpenAI drops its new frontier model GPT 5.5</b></a><br>OpenAI shipped GPT-5.5 today (codename &quot;Spud&quot;), six weeks after 5.4 and one week after Opus 4.7. Interesting to see the release cadence now measured in &quot;days since last frontier model.&quot; 😅 It takes the top spot on Terminal-Bench 2.0 (82.7%), FrontierMath Tier 4 (35.4%), and long-context retrieval at 1M tokens, while Opus 4.7 hangs onto SWE-Bench Pro and MCP Atlas. API access is &quot;coming very soon&quot; at $5/$30 per million tokens, with the Pro variant at a spicy $30/$180.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><b>OpenAI and Google shipped their enterprise agent platforms on the same day</b><br><a class="link" href="https://cloud.google.com/blog/products/ai-machine-learning/introducing-gemini-enterprise-agent-platform?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Google&#39;s Gemini Enterprise Agent Platform</b></a> absorbs Vertex AI. All future roadmap ships through Agent Platform, which bundles Agent Studio (low-code), an upgraded ADK, and a proper governance stack: Agent Identity, Registry, Gateway, plus runtime that supports multi-day workflows with persistent Memory Bank. It gives a single platform to build agents, scale to production, establish centralized control, and optimize with full traces.<br><br>A few hours later, <a class="link" href="https://openai.com/index/introducing-workspace-agents-in-chatgpt/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>OpenAI released Workspace Agents in ChatGPT</b></a>. These are Codex-powered team agents that keep running while you&#39;re offline, plug into Slack, Drive, Salesforce, and Notion, and can be scheduled or triggered by events. Build one in the sidebar by describing a workflow in plain English, then let it handle month-end close or inbound lead routing while your team does literally anything else. Free until May 6, after which credit-based pricing kicks in.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/google/agents-cli?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Agents CLI in Agent Platform</b></a>: Google released an open-source CLI that gives AI agents like Claude Code, Cursor, and Gemini CLI a machine-readable path into Google Cloud&#39;s agent stack. Describe the agent in a simple prompt, and they use CLI + the bundled skills to scaffold an ADK project, run evals, and deploy to Agent Runtime, Cloud Run, or GKE.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/cosmicstack-labs/mercury-agent?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Mercury</b></a>: Open-source CLI and Telegram AI agent that asks before it acts, like shell blocklists, folder-scoped file access, and explicit approval for sensitive commands. Personality lives in <code>soul.md</code> and <code>persona.md</code>, and daily token budgets keep it from burning through your API credits.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.kimi.com/blog/kimi-vendor-verifier?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Kimi Vendor Verifier</b></a>: Moonshot’s open-source toolkit that checks whether third-party inference providers are actually running Kimi K2 correctly. It runs six benchmarks to separate real model issues from sloppy engineering at the vendor layer.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/trycua/status/2047383200348221632?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow"><b>Cua Driver</b></a>: An open-source macOS driver that lets any agent like Claude Code, Codex, or your own click and type inside real Mac apps in the background, without stealing your cursor or pulling focus across Spaces. It works by tapping into a few of Apple&#39;s private, undocumented system APIs to deliver input directly to a target app, so the app responds while your foreground window keeps everything it had.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=gpt-5-5-kimi-k2-6-built-for-autonomous-ai-agents"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=961d7379-518f-47cf-89ed-0f443b53ed4c&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Build AI Coding Agents that Run ∞ in the Cloud </title>
  <description>+ Search, Fetch, Browser, and Web agent in one CLI + Skill</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/d1923d9d-ee0a-4188-85f8-ea2e9dc6cc9b/How_to_Run_a_247_AI_Agent_that_Grows_with_You__1_.png" length="2581684" type="image/png"/>
  <link>https://www.theunwindai.com/p/build-ai-coding-agents-that-run-in-the-cloud</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/build-ai-coding-agents-that-run-in-the-cloud</guid>
  <pubDate>Wed, 15 Apr 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-04-15T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/tinyfish-io/tinyfish-cookbook/blob/main/skills/use-tinyfish/SKILL.md?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">One npm to give your AI Agents the entire live web</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow">AI coding agent that adapts to your workflow</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://open-agents.dev/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>Build your own agents that run infinitely in the cloud</b></a></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://claude.com/blog/introducing-routines-in-claude-code?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">Run your Claude Code Routines on autopilot</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/awesome_agent_skills/self-improving-agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>Self-improving agent skills with built-in eval loop</b></a></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.theunwindai.com/p/how-to-run-a-24-7-ai-agent-that-grows-with-you?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>How to Run a 24/7 AI Agent that Grows with You</b></a></p><p class="paragraph" style="text-align:left;">I’ve been running a bunch of agents every day for months.</p><p class="paragraph" style="text-align:left;">The problem is that I was the one constantly tweaking and learning while they just held onto context.</p><p class="paragraph" style="text-align:left;">I tested it out by putting the same Monica agent on Hermes at the same time.</p><p class="paragraph" style="text-align:left;">She started creating her own playbook from my edits and keeps getting better without me.</p><p class="paragraph" style="text-align:left;">In this blog, I detailed how I did it. You’ll also understand the difference between an agent you manage and one that actually grows with you.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/how-to-run-a-24-7-ai-agent-that-grows-with-you?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/dfbcddab-2347-4afa-8972-4c6855a0cc2a/How_to_Run_a_247_AI_Agent_that_Grows_with_You.png?t=1776233672"/><div class="embed__content"><p class="embed__title"> How to Run a 24/7 AI Agent that Grows with You </p><p class="embed__description"> My agent started teaching itself while I slept </p></div></a></div><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this every week, designed to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://github.com/tinyfish-io/tinyfish-cookbook/blob/main/skills/use-tinyfish/SKILL.md?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>One npm to Give Your AI Agents the Entire Live Web</b></a></h3><div class="image"><a class="image__link" href="https://github.com/tinyfish-io/tinyfish-cookbook/blob/main/skills/use-tinyfish/SKILL.md?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/a9446bf4-9f75-42a6-950a-155bf70da551/Newsletter_Design_Presentation__1_.png?t=1776230466"/></a></div><p class="paragraph" style="text-align:left;">One npm install and a Skill markdown file - that&#39;s all it takes to give your coding agent full access to the live web.</p><p class="paragraph" style="text-align:left;"><a class="link" href="https://www.tinyfish.ai/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>TinyFish</b></a> just released their CLI + Agent Skill that lets Claude Code, Codex, Cursor, OpenClaw, and other agents autonomously call four web primitives directly from the terminal: Search, Fetch, Browser, and Agent. </p><p class="paragraph" style="text-align:left;">Everything runs on their infrastructure, so you can automate workflows at scale.</p><p class="paragraph" style="text-align:left;">You install the skill file once, and your agent figures out the rest: which endpoint to hit, how to structure the call, where to write the output. No integration code needed. </p><p class="paragraph" style="text-align:left;">All four primitives are built in-house on TinyFish&#39;s custom Chromium engine with 28 anti-bot mechanisms at the C++ level. The platform already ranks #1 on Mind2Web (89.9% accuracy) with 40M+ operations running for clients like Google, DoorDash, and Volkswagen.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Skills over MCP for execution</b>: The skill approach uses 87% fewer tokens per operation, writes output to the filesystem instead of the context window, and delivers 2x higher task completion on complex tasks.</p></li><li><p class="paragraph" style="text-align:left;"><b>End-to-end signal loop</b>: Because TinyFish owns search, fetch, browser, and agent, every run teaches the entire stack to improve - something you can&#39;t build by assembling third-party APIs.</p></li><li><p class="paragraph" style="text-align:left;"><b>One platform, one session</b>: Same IP, same fingerprint, same cookies across your whole workflow. Sites see one client, not five disconnected tools making separate requests.</p></li><li><p class="paragraph" style="text-align:left;"><b>Free tier available</b>: 500 steps, no credit card required, with an <a class="link" href="https://github.com/tinyfish-io/tinyfish-cookbook?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">open-source cookbook and skills on GitHub</a>.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow"><b>Mistral Vibe: An AI Coding Agent That Adapts to Your Workflow</b></a></h3><div class="image"><a class="image__link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/f5807d05-c007-4573-89fd-3f4d448d73b2/image.png?t=1776232729"/></a></div><p class="paragraph" style="text-align:left;"><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow"><b>Mistral Vibe</b></a> gives you a configurable agent runtime: composable skills, custom subagents, unified modes, MCP integrations, and fine-grained control over what the agent is allowed to do.</p><p class="paragraph" style="text-align:left;">For developers who want their tools to work the way they work, this is worth a close look.</p><p class="paragraph" style="text-align:left;"><b>Key highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Unified agent modes</b>: Configure custom modes that combine tools, permissions, and behaviors. Switch contexts without switching tools - from a read-only exploration mode to a fully autonomous execution mode, all managed through a single config.</p></li><li><p class="paragraph" style="text-align:left;"><b>Custom subagents on demand</b>: Define specialized agents for recurring tasks: deploy scripts, PR reviews, test generation. Invoke them when needed; they run independently and return results asynchronously.</p></li><li><p class="paragraph" style="text-align:left;"><b>Skills with .agents/skills/ standard support</b>: Skills extend agents with new tools, slash commands, and specialized behaviors. Vibe now supports the .agents/skills/ standard, improving interoperability across agent tooling. </p></li><li><p class="paragraph" style="text-align:left;"><b>Web search, notifications, and session resume</b>: The latest updates add web search mid-session, ping notifications when the agent needs attention, and /resume to continue any previous session with full context.</p></li><li><p class="paragraph" style="text-align:left;"><b>MCP in two steps</b>: Add your preferred tool to config.toml and it&#39;s connected. No wrappers, no boilerplate beyond the config tweak and an AGENTS.md file in the workspace.</p></li></ol><p class="paragraph" style="text-align:left;"><a class="link" href="https://console.mistral.ai/codestral/cli?utm_source=unwindai&utm_medium=newsletter&utm_campaign=vibe" target="_blank" rel="noopener noreferrer nofollow"><b>Try it today!</b></a></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><a class="link" href="https://open-agents.dev/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>Build Your Own Agents that Run Infinitely in the Cloud</b></a></h3><div class="image"><a class="image__link" href="https://open-agents.dev/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/1249e33c-4d31-4c3c-816c-57bc91cec083/Screenshot_2026-04-14_at_10.09.25_PM.png?t=1776229770"/></a></div><p class="paragraph" style="text-align:left;">Vercel looked at how Stripe, Spotify, and Block build internal coding agents, and decided to open source the entire playbook. </p><p class="paragraph" style="text-align:left;"><b>Open Agents</b> is a ready-to-fork platform for spinning up cloud coding agents that go from a chat prompt to committed code changes, no laptop required. </p><p class="paragraph" style="text-align:left;">The platform separates the agent brain from its execution sandbox, which turns out to be the key architectural decision. It means your model choices, sandbox lifecycle, and agent logic can all move independently. It ships with a chat UI, durable workflow runtime, isolated VMs with git baked in, and multi-model support through AI Gateway. </p><p class="paragraph" style="text-align:left;">Everything deploys on Vercel infrastructure and the repo is explicitly built to be customized, not consumed.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Three-Layer Design</b>: Web app handles auth and UI, the agent runs as a durable workflow outside the sandbox, and the VM provides filesystem, shell, and git access. Clean separation means each layer scales on its own.</p></li><li><p class="paragraph" style="text-align:left;"><b>Workflows That Don&#39;t Die</b>: Agent runs survive restarts, retry failures, and auto-checkpoint. You can disconnect and reconnect to a running session from anywhere.</p></li><li><p class="paragraph" style="text-align:left;"><b>Sandbox-Per-Session</b>: Every coding session gets an isolated environment with repo cloning, branch management, auto-commit, and optional PR creation. Sandboxes hibernate on inactivity and resume from snapshots.</p></li><li><p class="paragraph" style="text-align:left;"><b>Built to Fork</b>: Not a SaaS product, it&#39;s a template. Deploy your own instance, wire up your GitHub App, swap in your preferred models, and build your team&#39;s custom coding agent platform.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://claude.com/blog/introducing-routines-in-claude-code?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>Run your Claude Code Routines on autopilot</b></a><br>Anthropic just shipped &quot;Routines&quot; for Claude Code. Configure a prompt, point it at a repo and your connectors, then let it run on a schedule, via API webhook, or in response to GitHub events, all on Anthropic&#39;s cloud so your laptop can stay shut. Think nightly issue triage, auto-porting PRs across SDKs, or Datadog alerts that arrive pre-triaged with a draft fix. Available today across all paid plans, ​​with Claude Code on the web enabled.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/theo/status/2043611205856837680?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>What really is an Agent Harness?</b></a><br>Theo&#39;s latest video is the most fun explanation of AI coding harnesses you&#39;ll find — he builds one in ~60 lines of Python, then proceeds to lie to the model about what its tools do (it just believes you), swap the entire toolset for a single bash command (the model figures out file I/O on its own), and show why the same Opus model jumps from 77% to 93% just by moving it into Cursor. His core argument: the harness <i>is</i> the product, and whoever obsessively tweaks their system prompts and tool descriptions per model wins. Do watch this.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://blog.google/products-and-platforms/products/chrome/skills-in-chrome/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>Turn your AI prompts in Chrome into reusable Skills</b></a><br>Chrome now lets you turn your best Gemini prompts into saved Skills - reusable, one-click AI workflows you can fire on any page or across multiple tabs. Great abstraction: instead of re-typing the same prompt every time you hit a recipe or a product page, you save it once, customize it, and trigger it with <code>/</code> going forward. Google&#39;s also shipping a curated Skills library with ready-made options, and your saved Skills sync across all signed-in Chrome desktops.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/claudeai/status/2044131493966909862?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>Claude Code&#39;s new desktop app is built for parallel agents</b></a><br>Anthropic just dropped a major redesign of the Claude Code desktop app, and it&#39;s clearly built for the multi-agent workflow most of us have been using - spinning up parallel sessions across repos, each doing its own thing while you play air traffic controller. The new sidebar tracks all active sessions, there&#39;s a drag-and-drop layout with an integrated terminal and file editor, and side chats let you branch off quick questions without polluting the main task context. Available now on Pro, Max, Team, and Enterprise.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps/tree/main/awesome_agent_skills/self-improving-agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>Self-Improving Agent Skills</b></a>: Automatically optimize your agent skills using a multi-agent system built with Google ADK (Agent Development Kit) and Gemini. Upload a skill, let the agents generate test scenarios and evaluation criteria, then watch as three specialized ADK agents collaborate to improve your skill through iterative optimization.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/kontext-security/kontext-cli?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">Kontext CLI</a></b>: Open-source tool that wraps AI coding agents with enterprise-grade identity, credential management, and governance. Instead of copy-pasting long-lived API keys, you declare dependencies, and the CLI handles OIDC auth, short-lived token exchange, and automatic expiry.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/BayramAnnakov/claude-reflect?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow"><b>Claude Reflect</b></a>: A Claude Code plugin that builds persistent memory by automatically capturing your corrections and preferences during coding sessions, then syncing them to CLAUDE.md so Claude doesn&#39;t repeat the same mistakes. It also analyzes your session history to spot repeating workflows and generates reusable slash commands from them.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Hot Takes </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><span style="color:inherit;font-family:inherit;font-size:inherit;">Markdown is the new Python</span></p><p class="paragraph" style="text-align:left;">~ <b><a class="link" href="https://x.com/theo/status/2043819374889554261?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">Theo - t3.gg</a></b><br></p></li><li><p class="paragraph" style="text-align:left;"><span style="color:inherit;font-family:inherit;font-size:inherit;">the ai labs, in competing with each other, are burning huge amounts of the commons on public trust in ai to win minor points against the others. their lobbyists, pr machines, lawsuits. it’s the very opposite of what marxist class struggle analysis would tell you</span><br><span style="color:inherit;font-family:inherit;font-size:inherit;">~ </span><b><a class="link" href="https://x.com/tszzl/status/2044190488408785282?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">roon</a></b></p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=build-ai-coding-agents-that-run-in-the-cloud"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=f5cb07ed-e7d1-4494-9d14-ba9b584736ad&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>How to Run a 24/7 AI Agent that Grows with You</title>
  <description>My agent started teaching itself while I slept</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/dfbcddab-2347-4afa-8972-4c6855a0cc2a/How_to_Run_a_247_AI_Agent_that_Grows_with_You.png" length="2695268" type="image/png"/>
  <link>https://www.theunwindai.com/p/how-to-run-a-24-7-ai-agent-that-grows-with-you</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/how-to-run-a-24-7-ai-agent-that-grows-with-you</guid>
  <pubDate>Wed, 15 Apr 2026 06:14:54 +0000</pubDate>
  <atom:published>2026-04-15T06:14:54Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <category><![CDATA[Ai Blogs]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">I&#39;ve been running six agents on OpenClaw for months. They work. But the longer I ran them, the more I noticed a tax: I was maintaining the system more than the system was improving itself.</p><p class="paragraph" style="text-align:left;">I kept updating SOUL.md files, pruning stale memory, and repeating the same corrections across agents. The agents were not getting better on their own. I was getting better at managing them.</p><p class="paragraph" style="text-align:left;">Most agents can store context. Far fewer can turn completed work into reusable method. That is the difference I wanted to test.</p><p class="paragraph" style="text-align:left;">If you haven&#39;t read my previous article about OpenClaw, read that here: </p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/how-i-built-an-autonomous-ai-agent-team-that-runs-24-7?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=how-to-run-a-24-7-ai-agent-that-grows-with-you" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/021464df-88a1-4a8d-8905-a1e70c3f2ac4/How_I_Built_an_Autonomous_AI_Agent_Team_That_Runs_247.png?t=1770938302"/><div class="embed__content"><p class="embed__title"> How I Built an Autonomous AI Agent Team That Runs 24/7 </p><p class="embed__description"> Full step-by-step tutorial to make your own Agent Team with OpenClaw </p></div></a></div><h2 class="heading" style="text-align:left;"><b>The correction loop</b></h2><p class="paragraph" style="text-align:left;">I started calling my approach corrective prompt-engineering.</p><p class="paragraph" style="text-align:left;">You watch the agent, notice a failure, explain the fix, update memory or instructions, and wait for the behavior to stick. Kelly&#39;s early drafts were full of emojis. I told her: no emojis, no hashtags, short punchy sentences. After a week, she got it. Dwight captured too much noise. I told him to focus on signal. He adjusted.</p><p class="paragraph" style="text-align:left;">That loop works. But it has a hidden assumption: <b>you are the learning system.</b></p><p class="paragraph" style="text-align:left;">Every improvement depends on you noticing the problem, diagnosing it, and writing down the correction. The agents do not get better while you sleep. They get better when you pay attention.</p><h2 class="heading" style="text-align:left;"><b>The experiment</b></h2><p class="paragraph" style="text-align:left;">I kept the OpenClaw squad running, but set up a second Monica on Hermes Agent. Same job. Same Mac Mini. Same access to Dwight&#39;s intel files. Not a replacement. Just a parallel test.</p><p class="paragraph" style="text-align:left;">Hermes Agent is open-source from Nous Research. I was skeptical. Every agent tool claims some version of &quot;memory&quot; and &quot;learning.&quot; Most of them mean &quot;we store your chat history.&quot; That is not learning. That is hoarding.</p><p class="paragraph" style="text-align:left;">Setup was simple. One command:</p><div class="codeblock"><pre><code>curl -fsSL https://raw.githubusercontent.com/NousResearch/hermes-agent/main/scripts/install.sh | bash</code></pre></div><p class="paragraph" style="text-align:left;">Then <i><b>&#39;hermes setup&#39;</b></i>. It detected OpenClaw on the Mac Mini and offered to import settings, memories, and API keys. For Telegram, it walked me through BotFather setup. I sent the token and it connected. One tip: use a frontier model. The learning loop needs strong reasoning to work well.</p><h2 class="heading" style="text-align:left;"><b>The first thing that felt different</b></h2><p class="paragraph" style="text-align:left;">The first time I checked &#39;<b>~/.hermes/skills/&#39;</b> and found files I did not create, something clicked.</p><p class="paragraph" style="text-align:left;">The one that stood out was &#39;<b>local-writing-canon-analysis/SKILL.md&#39;</b>. Monica had written a procedure for reading my published articles before drafting in my voice. Not &quot;I know your style&quot; in the abstract. More like: read the shipped work, then infer the voice from what actually survived editing.</p><p class="paragraph" style="text-align:left;">The file captured concrete editorial rules: keep the compounding thesis, avoid framing new topics as product comparisons, center the piece on outcomes like memory and reuse rather than features.</p><p class="paragraph" style="text-align:left;">I did not write it. She did. She watched what I kept, what I cut, and wrote down the pattern. Other skills showed up too: operational playbook, workflow procedures, reusable task patterns. Some were useful. Some were generic. But the default behavior, turning completed work into reusable procedure, was new.</p><p class="paragraph" style="text-align:left;">The other thing that stood out was recall. When I searched for telegram OR gateway OR restart OR stuck, Hermes agent surfaced the exact troubleshooting arc from a previous session: the polling conflict, the lesson that gateway status alone is not proof of health, and the practical fix path. All from a conversation weeks earlier that nobody had manually promoted into memory.</p><p class="paragraph" style="text-align:left;">The work leaves residue. Not just a memory of what happened, but a method for next time.</p><h2 class="heading" style="text-align:left;"><b>Who owns the correction loop</b></h2><p class="paragraph" style="text-align:left;"><b>OpenClaw Monica</b> improves when I notice a problem and teach the fix. I notice a problem, give feedback, and she stores the correction for next time. The improvement depends on both of us.</p><p class="paragraph" style="text-align:left;"><b>Hermes Monica</b> closes more of that loop herself. After a complex task, she evaluates what happened, decides what is worth keeping, and writes it into a skill file. I can inspect it, edit it, or delete it. But I did not have to initiate it.</p><p class="paragraph" style="text-align:left;">That is the shift that matters. <b>OpenClaw mostly stores what I explicitly teach. Hermes tries to turn finished work into future procedure.</b></p><p class="paragraph" style="text-align:left;">The difference I cared about was not recall architecture. It was who owned the improvement loop.</p><h2 class="heading" style="text-align:left;"><b>What still breaks</b></h2><p class="paragraph" style="text-align:left;">Agents still hallucinate confidence. Self-improvement is not self-verification.</p><p class="paragraph" style="text-align:left;">During a migration session, Monica treated the Hermes migration archive as if it contained everything from OpenClaw. It did not. The actual source of truth was the live &#39;<i><b>~/.openclaw/</b></i><i>&#39; </i>workspace. She inferred completeness from partial data and moved forward with confidence she had not earned.</p><p class="paragraph" style="text-align:left;">The Telegram gateway was the other one. After a restart, the gateway reported as loaded, but the bot was not responding. launchd said the service was running. The live responder was dead. Status screens flatter you. Logs tell the truth.</p><p class="paragraph" style="text-align:left;">Monica turned that failure into a recovery skill on her own. The file sits at <i><b>~/.hermes/skills/autonomous-ai-agents/hermes-telegram-gateway-recovery/SKILL.md</b></i>. It includes a diagnostic sequence: check gateway status, inspect logs, run the safe restart script, verify a real process plus Telegram reconnect before declaring success.</p><p class="paragraph" style="text-align:left;">That skill was not prompted. She hit the problem, solved it, and wrote the playbook (aka skill)</p><p class="paragraph" style="text-align:left;"><b>Mistake goes in. Procedure comes out.</b></p><p class="paragraph" style="text-align:left;">This is not automatic perfection. Some generated skills were too generic. Some captured the wrong level of abstraction. I still had to inspect, prune, and occasionally delete them. The difference was not that supervision disappeared. The difference was that useful procedure started showing up without me authoring every improvement by hand.</p><h2 class="heading" style="text-align:left;"><b>Running both</b></h2><p class="paragraph" style="text-align:left;">The OpenClaw squad handles research, content drafting, code review, and the newsletter on cron schedules. Hermes Monica runs as the chief of staff experiment on the same Mac Mini. She reads the same intel files Dwight produces on OpenClaw. Dwight writes DAILY-INTEL.md. Monica reads it. The handoff is a markdown file on disk.</p><p class="paragraph" style="text-align:left;">OpenClaw gives me manual control and predictability. Hermes gives Monica a more self-maintaining improvement loop.</p><p class="paragraph" style="text-align:left;">Both support the <span style="text-decoration:underline;"><a class="link" href="https://agentskills.io?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=how-to-run-a-24-7-ai-agent-that-grows-with-you" target="_blank" rel="noopener noreferrer nofollow">agentskills.io</a></span> standard, so skills are portable across OpenClaw, Hermes, Claude Code, and Cursor. Nothing is locked to one platform.</p><p class="paragraph" style="text-align:left;">I did not plan to run both long-term. But they complement each other. The agents I want full control over stay on OpenClaw. The agent I want to watch improve more autonomously runs on Hermes.</p><h2 class="heading" style="text-align:left;"><b>How to test this yourself</b></h2><p class="paragraph" style="text-align:left;">Start the boring way: one agent, one job, one recurring workflow. Install Hermes with one command. Run &#39;<b>hermes setup</b>&#39;. Pick a frontier model. Connect Telegram.</p><p class="paragraph" style="text-align:left;">The question is not whether day one looks good. The question is whether day fourteen is better without you hand-tuning every session.</p><p class="paragraph" style="text-align:left;">If you are already running OpenClaw, run Hermes alongside it. During setup, Hermes detects OpenClaw and offers to import your settings. Pick one agent. See what happens.</p><h2 class="heading" style="text-align:left;"><b>Day 30 versus day 1</b></h2><p class="paragraph" style="text-align:left;"><b>The system that maintains itself compounds faster than the one that depends on you.</b></p><p class="paragraph" style="text-align:left;">OpenClaw is an agent you actively shape. Hermes is an agent that turns more of its own work into future procedure and improve itself. Not universally better. A different learning loop.</p><p class="paragraph" style="text-align:left;">The best 24/7 agent is not the one that gives the best first answer. It is the one that gives a better answer on day 30 than day 1. Without you having to teach it every lesson by hand.</p><p class="paragraph" style="text-align:left;">Start with one agent. One job. Let it run. Then check whether day 30 is better than day 1.</p><p class="paragraph" style="text-align:left;">That is the test.</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">We share in-depth blogs and tutorials like this 2-3 times a week, to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=how-to-run-a-24-7-ai-agent-that-grows-with-you" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;"><b>Don’t forget to share this tutorial on your social channels and tag Unwind AI (</b><span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=how-to-run-a-24-7-ai-agent-that-grows-with-you" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=how-to-run-a-24-7-ai-agent-that-grows-with-you" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b>, </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=how-to-run-a-24-7-ai-agent-that-grows-with-you" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span><b>) to support us!</b></p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=how-to-run-a-24-7-ai-agent-that-grows-with-you"><span class="button__text" style=""> Subscribe now for FREE - Get instant access to more LLM, RAG & AI Agent tutorials </span></a></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=85509c5c-d1d0-4717-b228-5bfc7cc373d7&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

      <item>
  <title>Karpathy&#39;s AI Coding Agent Rant in a Claude.md File</title>
  <description>+ Google&#39;s best engineering practices in 20 Skills</description>
      <enclosure url="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/2268565b-9bab-45de-a203-48362dc887f9/Karpathy_s_AI_Coding_Agent_Rant_in_a_Claude.md_File.png" length="2772613" type="image/png"/>
  <link>https://www.theunwindai.com/p/karpathy-s-ai-coding-agent-rant-in-a-claude-md-file</link>
  <guid isPermaLink="true">https://www.theunwindai.com/p/karpathy-s-ai-coding-agent-rant-in-a-claude-md-file</guid>
  <pubDate>Mon, 13 Apr 2026 12:30:00 +0000</pubDate>
  <atom:published>2026-04-13T12:30:00Z</atom:published>
    <dc:creator>Shubham Saboo</dc:creator>
    <dc:creator>Gargi Gupta</dc:creator>
    <category><![CDATA[Daily Unwind]]></category>
  <content:encoded><![CDATA[
    <div class='beehiiv'><style>
  .bh__table, .bh__table_header, .bh__table_cell { border: 1px solid #6553a2; }
  .bh__table_cell { padding: 5px; background-color: #ffffff; }
  .bh__table_cell p { color: #030712; font-family: 'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
  .bh__table_header { padding: 5px; background-color:#d0c7e2; }
  .bh__table_header p { color: #6553a2; font-family:'Open Sans','Segoe UI','Apple SD Gothic Neo','Lucida Grande','Lucida Sans Unicode',sans-serif !important; overflow-wrap: break-word; }
</style><div class='beehiiv__body'><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">Today’s top AI Highlights:</p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow">Ship production-ready voice agents with open-weight TTS</a></b></p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/addyosmani/agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow"><b>Google&#39;s best engineering practices in 20 Agent Skills</b></a></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/forrestchang/andrej-karpathy-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Karpathy-inspired CLAUDE.md file to improve Claude Code behavior</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/garrytan/status/2042925773300908103?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Get 100x from your AI Agents with Thin Harness and Fat Skills</a></b></p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://x.com/hwchase17/status/2042978500567609738?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Own your memory for your agent harness</a></b></p></li></ol><p class="paragraph" style="text-align:start;">& so much more!</p><p class="paragraph" style="text-align:start;"><i><b>Read time: 3 mins</b></i></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>AI Tutorial </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://www.theunwindai.com/p/how-i-built-an-autonomous-ai-agent-team-that-runs-24-7?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">How I Built an Autonomous AI Agent Team That Runs 24/7</a></b></p><p class="paragraph" style="text-align:left;">Six AI agents run my entire life while I sleep.</p><p class="paragraph" style="text-align:left;">Not a demo. Not a weekend project.</p><p class="paragraph" style="text-align:left;">A real team that works 24/7, making sure I&#39;m never behind. Research done. Content drafted. Code reviewed. Newsletter ready. By the time I open Telegram in the morning, they&#39;ve already put in a full shift.</p><p class="paragraph" style="text-align:left;">By the end of this, you will understand exactly how to build an autonomous AI agent team that runs while you sleep.</p><div class="embed"><a class="embed__url" href="https://www.theunwindai.com/p/how-i-built-an-autonomous-ai-agent-team-that-runs-24-7?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank"><img class="embed__image embed__image--left" src="https://beehiiv-images-production.s3.amazonaws.com/uploads/asset/file/021464df-88a1-4a8d-8905-a1e70c3f2ac4/How_I_Built_an_Autonomous_AI_Agent_Team_That_Runs_247.png?t=1770938302"/><div class="embed__content"><p class="embed__title"> How I Built an Autonomous AI Agent Team That Runs 24/7 </p><p class="embed__description"> Full step-by-step tutorial to make your own Agent Team with OpenClaw </p></div></a></div><p class="paragraph" style="text-align:left;">We share hands-on tutorials like this every week, designed to help you stay ahead in the world of AI. <span style="text-decoration:underline;"><b><a class="link" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">If you&#39;re serious about leveling up your AI skills and staying ahead of the curve, subscribe now and be the first to access our latest tutorials.</a></b></span></p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b>Unwind AI</b> (<b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">X</a></b><b>, </b><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=referral&utm_campaign=last-week-in-ai-a-weekly-unwind" target="_blank" rel="noopener noreferrer nofollow">LinkedIn</a></b><b>, </b><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Threads</a></b>) to support us!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Latest Developments </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow">Ship Production-Ready Voice Agents with Open-weights TTS</a></b></h3><div class="image"><a class="image__link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" rel="noopener" target="_blank"><img alt="" class="image__image" style="border-radius:0px 0px 0px 0px;border-style:solid;border-width:0px 0px 0px 0px;box-sizing:border-box;border-color:#E5E7EB;" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/dc575688-6afa-4aee-b4ba-18e84298dd4e/image.png?t=1776052252"/></a></div><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow">Voxtral TTS</a></b> is a lightweight 4B parameter text-to-speech model with state-of-the-art performance in multilingual voice generation. It goes beyond reading text aloud, capturing a speaker&#39;s personality: natural pauses, rhythm, intonation, and emotional dexterity. Available now via API, Mistral Studio, and open weights on Hugging Face.</p><p class="paragraph" style="text-align:left;"><b>Key highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Emotion-aware speech generation</b> - Contextual understanding (neutral, happy, sarcastic, and more) determines whether output sounds natural or robotic. Voxtral TTS models both meaning and personality, not just pronunciation.</p></li><li><p class="paragraph" style="text-align:left;"><b>Zero-shot custom voice context in 9 languages</b> - Provide a 5-25 second voice prompt and the model adapts to accent, inflection, rhythm, and even disfluencies. Supports English, French, German, Spanish, Dutch, Portuguese, Italian, Hindi, and Arabic.</p></li><li><p class="paragraph" style="text-align:left;"><b>Cross-lingual voice adaptation</b> - Generate speech in one language using a voice prompt in another. A French voice prompt with English text produces natural French-accented English, useful for cascaded speech-to-speech translation.</p></li><li><p class="paragraph" style="text-align:left;"><b>70ms model latency, ~9.7x real-time factor</b> - Built for voice agent applications. Streams natively and integrates into any existing STT and LLM stack.</p></li><li><p class="paragraph" style="text-align:left;"><b>Beats ElevenLabs Flash v2.5 on naturalness</b> - Human evaluations show Voxtral TTS wins on flagship voices (58.3% vs 41.7%) and voice customization (68.4% vs 31.6%).</p></li></ol><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://console.mistral.ai/build/audio/text-to-speech/?utm_source=unwindai&utm_medium=newsletter&utm_campaign=audio" target="_blank" rel="noopener noreferrer nofollow">Try it now!</a></b></p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://github.com/addyosmani/agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Google&#39;s Best Engineering Practices in 20 Agent Skills</a></b></h3><div class="image"><a class="image__link" href="https://github.com/addyosmani/agent-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/c97140bf-5a35-4816-bf73-df5e958007ab/Screenshot_2026-04-12_at_9.27.53_PM.png?t=1776054478"/></a></div><p class="paragraph" style="text-align:left;">Google&#39;s Addy Osmani just packaged 14+ years of senior engineering judgment into Markdown files, and now your AI coding agent can follow them. </p><p class="paragraph" style="text-align:left;">He open-sourced <b>Agent Skills,</b> a collection of 20 structured engineering workflows as Skills and 7 slash commands that cover the full software development lifecycle, from writing specs to shipping code. Each skill is a step-by-step process with verification gates, defined inputs/outputs, and anti-rationalization tables that prevent your agent from taking shortcuts like skipping tests or jumping straight to code. </p><p class="paragraph" style="text-align:left;">The skills map to six development phases (Define → Plan → Build → Verify → Review → Ship), and they bake in real engineering principles like Hyrum&#39;s Law, Chesterton&#39;s Fence, and trunk-based development. </p><p class="paragraph" style="text-align:left;">The repo has already crossed 14K stars already in just a couple of days already!</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Workflows, Not Prompts</b> - Each of the 20 skills has defined inputs, outputs, and verification gates. The agent can&#39;t skip to the next step without completing the current one.</p></li><li><p class="paragraph" style="text-align:left;"><b>7 Slash Commands</b> - <code>/spec</code>, <code>/plan</code>, <code>/build</code>, <code>/test</code>, <code>/review</code>, <code>/code-simplify</code>, and <code>/ship</code> map to the full development lifecycle, making it easy to invoke the right process at the right time.</p></li><li><p class="paragraph" style="text-align:left;"><b>Three Specialist Personas</b> - Ships with code-reviewer, test-engineer, and security-auditor agent personas you can load for targeted review workflows.</p></li></ol><p class="paragraph" style="text-align:left;">You need to check it out now!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><h3 class="heading" style="text-align:left;"><b><a class="link" href="https://github.com/forrestchang/andrej-karpathy-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Karpathy-inspired CLAUDE.md File to Improve Claude Code Behavior</a></b></h3><div class="image"><a class="image__link" href="https://github.com/forrestchang/andrej-karpathy-skills?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/9fb772cf-6f4e-46b9-9161-ad966d5bf33f/Screenshot_2026-04-12_at_9.36.29_PM.png?t=1776054993"/></a></div><p class="paragraph" style="text-align:left;">Andrej Karpathy diagnosed exactly how LLMs fail at coding. And now there&#39;s a one-file fix you can install in 10 seconds. </p><p class="paragraph" style="text-align:left;"><b>andrej-karpathy-skills</b> is an open-source <code>CLAUDE.md</code> file by developer Forrest Chang that encodes four behavioral principles directly into Claude Code, all derived from Karpathy&#39;s observations with building using AI coding agents. </p><p class="paragraph" style="text-align:left;">Back in January, Karpathy documented how he shifted from mostly manual coding to 80% agent coding in a matter of weeks. But he was equally blunt about the failure modes: agents making wrong assumptions without checking, over-complicating everything, and quietly messing with unrelated code as side effects. </p><p class="paragraph" style="text-align:left;">This project turns those observations into guardrails. Drop the file in your project root, and Claude Code automatically picks up the constraints for cleaner diffs, fewer rewrites, and agents that pause to ask before assuming.</p><p class="paragraph" style="text-align:left;"><b>Key Highlights:</b></p><ol start="1"><li><p class="paragraph" style="text-align:left;"><b>Assumptions get surfaced, not buried</b> - The file forces the agent to explicitly state what it&#39;s assuming and ask for clarification when multiple interpretations exist, instead of silently picking one.</p></li><li><p class="paragraph" style="text-align:left;"><b>Simplicity is enforced, not suggested</b> - If 200 lines could be 50, rewrite. No speculative flexibility, no abstractions for single-use code, no features nobody asked for.</p></li><li><p class="paragraph" style="text-align:left;"><b>Surgical precision on edits</b> - The agent only modifies what&#39;s directly relevant to the task — no touching adjacent comments, no reformatting, no surprise refactors.</p></li><li><p class="paragraph" style="text-align:left;"><b>Success criteria over step-by-step hand-holding</b> - Inspired by Karpathy&#39;s observation that LLMs are best when looping toward verifiable goals: write the test first, then make it pass.</p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Quick Bites </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/openclaw/status/2043132528094036332?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow"><b>OpenClaw’s latest update might make your Claws better with GPT-5.4</b></a><br>A huge release dropped by the OpenClaw team today with a formal GPT-5.4 vs Opus 4.6 agentic parity gate baked into the release pipeline. It’s a clear signal that the project is treating GPT-5.4 as a first-class citizen post-Anthropic lockout. The update also ships ChatGPT memory import ingestion, richer video generation modes, and a pile of fixes for the OAuth/transport bugs that plagued early 5.4 adopters. If you rage-quit Claude after the April 4 subscription crackdown and found GPT-5.4 underwhelming, this is the release worth retesting with retuned bootstrap files.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/garrytan/status/2042925773300908103?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow"><b>Get 100x from your AI Agents with Thin Harness and Fat Skills</b></a><br>If you&#39;ve been thinking about how to structure AI agents, Garry Tan&#39;s &quot;thin harness, fat skills&quot; essay is a really satisfying read. The gist: stop cramming logic into bloated orchestration layers with 40 MCP tools. Instead, put judgment into rich markdown skill files and push execution into fast, deterministic tooling underneath. He walks through a real system at YC where matching skills self-improve by analyzing mediocre event feedback, which is exactly the kind of compounding loop that makes you go &quot;oh, right, that&#39;s how this should work.&quot; Do give it a read.</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/hwchase17/status/2042978500567609738?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow"><b>Own your memory for your agent harness</b></a><br>The thing that makes your AI agent actually good over time isn&#39;t the model, it&#39;s all the context it&#39;s built up about you, your work, and your preferences. Harrison Chase connects this nicely to why owning your agent harness matters: when you use a closed agent harness, you&#39;re giving away your agent&#39;s memory. And memory is really just context! Model providers know this, which is why they&#39;re quietly pulling more state behind their APIs. His story about accidentally losing his email agent&#39;s memory and having to reteach it everything from scratch is a perfect example of why this matters. Another interesting read coming on X!</p><p class="paragraph" style="text-align:left;"></p><p class="paragraph" style="text-align:left;"><a class="link" href="https://x.com/marmaduke091/status/2043382991901147158?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow"><b>Leaked screenshots from Anthropic show Full-Stack App Builder</b></a><br>Another (rumoured) leak from Anthropic showing a Lovable-like interface where you can describe apps in natural language, pick templates for chatbots, games, or photo albums, and it spins up full-stack apps. Powered by Claude Opus 4.6, it offers live previews, one-click publishing, and more. Not sure how true this is, but we won’t be surprised!</p></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#ffffff;"><b>Tools of the Trade </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/multica-ai/multica?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow"><b>Multica</b></a>: Turns coding agents into real teammates. Assign issues to an agent like you&#39;d assign to a colleague, and they&#39;ll pick up the work, write code, report blockers, and update statuses autonomously. Your agents show up on the board, participate in conversations, and compound reusable skills over time. Works with Claude Code, Codex, OpenClaw, and OpenCode.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/RunanywhereAI/rcli?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow"><b>RCLI</b></a>: A fully on-device voice AI for macOS that runs a complete STT → LLM → TTS pipeline natively on Apple Silicon, no API keys needed. It can control 38 macOS actions by voice (Spotify, volume, screenshots, etc.), do local RAG over your documents with ~4ms retrieval, and claims up to 550 tok/s on M3+.</p></li><li><p class="paragraph" style="text-align:left;"><a class="link" href="https://github.com/vercel-labs/just-bash?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow"><b>just-bash</b></a>: A sandboxed bash implementation in TypeScript from Vercel, designed to be the execution environment AI agents use when they need to run shell commands. It provides an in-memory filesystem, broad Unix command coverage, and configurable execution limits to prevent runaway scripts.</p></li><li><p class="paragraph" style="text-align:left;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Awesome LLM Apps</a></b><b> </b>- A curated collection of LLM apps with RAG, AI Agents, multi-agent teams, MCP, voice agents, and more. The apps use models from OpenAI, Anthropic, Google, and open-source models like DeepSeek, Qwen, and Llama that you can run locally on your computer. <br><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">(Now accepting GitHub sponsorships)</a></p></li></ol><div class="image"><a class="image__link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" rel="noopener" target="_blank"><img alt="" class="image__image" style="" src="https://media.beehiiv.com/cdn-cgi/image/fit=scale-down,format=auto,onerror=redirect,quality=80/uploads/asset/file/5842cecc-c30d-48e9-a805-783f55950a3e/image.png?t=1755755385"/></a></div></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:#6553a2;border-radius:5px;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><h2 class="heading" style="text-align:center;"><span style="color:#FFFFFF;"><b>Hot Takes </b></span></h2></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:0.0px 0.0px 0.0px 0.0px;padding:5.0px 5.0px 5.0px 5.0px;"><ol start="1"><li><p class="paragraph" style="text-align:left;">vibecoding is fentanyl for people with an iq over 120</p><p class="paragraph" style="text-align:left;">~ <b><a class="link" href="https://x.com/runaway_vol/status/2043028468259041762?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Blair Dulder CPA™</a></b><br><br></p></li><li><p class="paragraph" style="text-align:left;">OpenAI is starting to look like Google in 2002 & Anthropic is starting to look like Yahoo in 2002.</p><p class="paragraph" style="text-align:left;">~ <b><a class="link" href="https://x.com/yacineMTB/status/2043127335184900417?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">kache</a></b></p></li></ol></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;border-color:#6553a2;border-radius:5px;border-style:solid;border-width:1px;margin:5.0px 5.0px 5.0px 5.0px;padding:5.0px 5.0px 5.0px 5.0px;"><p class="paragraph" style="text-align:left;">That’s all for today! See you tomorrow with more such AI-filled content.</p><p class="paragraph" style="text-align:left;">Don’t forget to share this newsletter on your social channels and tag <b><a class="link" href="https://www.theunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow">Unwind AI</a></b> to support us!</p><p class="paragraph" style="text-align:start;"><b>Unwind AI</b> - <span style="text-decoration:underline;"><b><a class="link" href="https://x.com/unwind_ai_?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">X</a></b></span> | <span style="text-decoration:underline;"><b><a class="link" href="https://www.linkedin.com/company/unwind-ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">LinkedIn</a></b></span><b> </b>|<b> </b><span style="text-decoration:underline;"><b><a class="link" href="https://www.threads.net/@unwind_ai?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Threads</a></b></span></p><p class="paragraph" style="text-align:left;"><span style="text-decoration:underline;"><b><a class="link" href="https://github.com/Shubhamsaboo/awesome-llm-apps?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Awesome LLM Apps</a></b></span><b> | </b><span style="text-decoration:underline;"><b><a class="link" href="https://sponsorunwindai.com/?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file" target="_blank" rel="noopener noreferrer nofollow" style="color: #6553a2">Sponsor Us</a></b></span></p><p class="paragraph" style="text-align:start;"><b>PS:</b> We curate this AI newsletter every day for FREE, your support is what keeps us going. If you find value in what you read, share it with at least one, two (or 20) of your friends 😉 </p></div><p class="paragraph" style="text-align:left;"></p><div class="button" style="text-align:center;"><a target="_blank" rel="noopener nofollow noreferrer" class="button__link" style="" href="https://www.theunwindai.com/subscribe?utm_source=www.theunwindai.com&utm_medium=newsletter&utm_campaign=karpathy-s-ai-coding-agent-rant-in-a-claude-md-file"><span class="button__text" style=""> Subscribe now for FREE! </span></a></div><p class="paragraph" style="text-align:left;"></p><div class="section" style="background-color:transparent;margin:0.0px 0.0px 0.0px 0.0px;padding:0.0px 0.0px 0.0px 0.0px;"><p class="paragraph" style="text-align:left;"></p></div></div><div class='beehiiv__footer'><br class='beehiiv__footer__break'><hr class='beehiiv__footer__line'><a target="_blank" class="beehiiv__footer_link" style="text-align: center;" href="https://www.beehiiv.com/?utm_campaign=69ecc6b1-695e-4ef7-b47f-fdfbc839c04c&utm_medium=post_rss&utm_source=unwind_ai">Powered by beehiiv</a></div></div>
  ]]></content:encoded>
</item>

  </channel>
</rss>
