<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0" xmlns:itunes="http://www.itunes.com/dtds/podcast-1.0.dtd" xmlns:googleplay="http://www.google.com/schemas/play-podcasts/1.0"><channel><title><![CDATA[Andrew Mudd]]></title><description><![CDATA[I run Mudd Ventures. After years of consulting, you tend to see the same patterns. This is what's working and what's not. 5+ years consulting high ticket high performance online offers doing 6 & 7 figures per month.]]></description><link>https://blog.muddventures.com</link><image><url>https://substackcdn.com/image/fetch/$s_!Ml4j!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3c9fc46f-93fa-4df4-b502-45810c63a5ed_3546x3546.jpeg</url><title>Andrew Mudd</title><link>https://blog.muddventures.com</link></image><generator>Substack</generator><lastBuildDate>Thu, 27 Aug 2026 04:20:51 GMT</lastBuildDate><atom:link href="https://blog.muddventures.com/feed" rel="self" type="application/rss+xml"/><copyright><![CDATA[Andrew Mudd]]></copyright><language><![CDATA[en]]></language><webMaster><![CDATA[muddventures@substack.com]]></webMaster><itunes:owner><itunes:email><![CDATA[muddventures@substack.com]]></itunes:email><itunes:name><![CDATA[Andrew Mudd]]></itunes:name></itunes:owner><itunes:author><![CDATA[Andrew Mudd]]></itunes:author><googleplay:owner><![CDATA[muddventures@substack.com]]></googleplay:owner><googleplay:email><![CDATA[muddventures@substack.com]]></googleplay:email><googleplay:author><![CDATA[Andrew Mudd]]></googleplay:author><itunes:block><![CDATA[Yes]]></itunes:block><item><title><![CDATA[Your lead magnet is still a PDF. Here's what Canva quietly handed every free account this month.]]></title><description><![CDATA[The interactive calculator your funnel needed used to be a developer invoice, now it's a prompt in a tool you already have.]]></description><link>https://blog.muddventures.com/p/your-lead-magnet-is-still-a-pdf-heres</link><guid isPermaLink="false">https://blog.muddventures.com/p/your-lead-magnet-is-still-a-pdf-heres</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Fri, 31 Jul 2026 21:38:14 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!APD1!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!APD1!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!APD1!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!APD1!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!APD1!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!APD1!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!APD1!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:982603,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/209312132?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!APD1!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!APD1!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!APD1!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!APD1!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F0ead8dbb-b213-47e3-9e45-4541a1d70828_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TLDR:</strong> On July 14, Canva pushed Canva Code 2.0 to every account on every plan, free included. You describe a calculator, quiz, or scorecard in plain English, Canva builds the working thing, you edit it like any other design inside the editor where your brand kit already lives, and you publish it to a real URL. The interactive lead magnet, an asset that used to mean a developer invoice or one more monthly subscription, is now a prompt inside a tool your team already logs into. By the end of this you'll have a copy-paste prompt that turns the one question your leads keep asking into a working interactive tool, plus the pre-traffic check that catches the embarrassing failures before your audience finds them.</p><p>Three weeks ago I wrote about Canva folding the ad-refresh loop into a button, and the <a href="https://blog.muddventures.com/p/run-the-30-minute-canva-test-that">30-minute test</a> that shows whether a creative retainer still earns its invoice. I filed it under <a href="https://blog.muddventures.com/p/notion-just-made-zapier-optional">Zapier Optionality</a>, the pattern where software you already pay for quietly absorbs a job you were paying somebody else to do, and I figured Canva was done collecting other people's invoices for the month.</p><p>It wasn't. On July 14, Canva made <a href="https://www.canva.com/newsroom/news/Canva-Code/">Canva Code 2.0</a> generally available to everyone, on every plan, free accounts included. You type a prompt, Canva generates a working interactive tool, a quiz that scores, a calculator that computes, a page that reflows on a phone screen, and then you edit the result like any other Canva design and publish it to a free Canva URL or your own domain.</p><p>The job getting absorbed this time is the interactive lead magnet, the calculator or scored quiz that used to sit behind a freelance developer, a Typeform-style subscription, or a "we'll scope it next sprint" conversation with an agency. If your lead capture is still a static download, the build cost excuse retired on July 14, and what's left is a decision about what to build.</p><p>Here's what shipped, what I'd build with it first, and the check I'd run before putting a dollar of traffic behind it.</p><ul><li><p><strong>The July 14 change:</strong> Canva Code 2.0 went to every plan including Free, after a preview run that was gated to the first million users who found an easter egg at Canva Create in April.</p></li><li><p><strong>The Static Magnet:</strong> a PDF download tells you somebody wanted a file, a three-question calculator tells you what they're planning to spend.</p></li><li><p><strong>One prompt, one working tool:</strong> the full build prompt to copy, tuned for a lead magnet that solves one narrow job in under two minutes.</p></li><li><p><strong>The 10-minute pre-traffic check:</strong> what to test on your phone before this thing goes in a bio link or behind ad spend.</p></li></ul><p><strong>Who this is for:</strong> operators whose lead capture is a static download and whose booking page could use a warmer front door, which covers most call funnels, local service businesses, and agencies I talk to.</p><h2>What Canva shipped on July 14</h2><p>Canva Code launched in its first version last year, and Canva says users built more than six million coded sites with it, which told them the demand ran a lot deeper than developers. Version 2.0 arrived at Canva Create in April as a preview limited to the first million users who cracked an easter egg, and on July 14 it went to the full user base, per <a href="https://www.canva.com/newsroom/news/Canva-Code/">Canva's announcement</a> and <a href="https://www.techradar.com/pro/canva-wants-you-to-have-full-control-over-the-design-code-2-0-just-does-the-heavy-lifting-for-you">TechRadar's coverage</a> of the general release.</p><p>The mechanics, per <a href="https://9to5mac.com/2026/07/14/canva-code-2-0-adds-visual-editing-html-imports-and-real-time-collaboration/">9to5Mac's writeup</a>: start from a plain-language prompt, one of 50-plus templates, or an HTML file imported from anywhere, including pages another AI tool generated. The output lands in the normal Canva editor, so you swap images by dragging from the library, change colors and fonts from the toolbar, click into any section and retype the text, and publish to a free Canva domain or a custom one, with responsive layouts and a mobile preview before it goes live. Canva's own line on it: "Canva Code sits inside the same intuitive editor where you already house your brand kit."</p><p>The part I'd flag for anyone who tried the first version and bounced: editing no longer means re-prompting the whole build. Canva spent the spring criticizing AI design tools for being AI-first and design-second, where the smallest tweak meant another prompt and an entirely new result, and the 2.0 pitch is that you fix the button color the way you'd fix it on a flyer. Danny Wu, Canva's Head of AI Products, framed the shift to TechRadar this way: "With the barrier to generating code lowered, the value has shifted to visual control: what you build needs to be distinctive and look like your brand or vision."</p><p>Canva also reports generation time down 75%, median prompt-to-publish time down 30%, and active Code users up 25% since the feature moved into the main UI. Worth holding loosely: every one of those is Canva measuring Canva. The <a href="https://blog.crescitaly.com/canva-code-2-social-mini-app-lead-magnet-loop-2026/">Crescitaly team's read</a> matches mine: "Treat speed as production capacity, not business impact." Faster builds mean you can test more ideas, and they prove nothing about whether any given idea converts.</p><h2>The Static Magnet</h2><p>Here's the name for the thing this replaces. The Static Magnet is the PDF checklist, the swipe file, the ebook that captures an email and then tells you nothing else about the person who grabbed it, because it asked them to do nothing except download, and an asset that asks for no decisions collects no intent. Every download looks identical in your CRM, the tire kicker and the operator with budget land in the same nurture sequence, and your follow-up is generic because the asset gave you nothing to personalize with.</p><p>An interactive asset works the opposite job. A three-question calculator makes the reader type in their monthly ad spend, their average deal value, their close rate, and now the "lead magnet" has handed you qualification data before a single email went out. A scorecard makes them self-assess, and the score tells you who's ready for a call this week versus who needs three months of nurture. Crescitaly's framing of why this works on social traffic is the cleanest I've seen: an interactive asset turns a fast scroll into a small decision, and a small decision reveals real intent.</p><p>I sell from behind one of these myself. The <a href="https://showtime.muddventures.com/ai-iq-test">AI IQ Test</a> is a scored quiz, and the value over a PDF is the same value I'm pointing you at here: the answers tell me where an operator stands before I've written a single follow-up, so the follow-up gets to be specific instead of generic. That asset was a real build project when it went up. As of this month, the same category of asset is a prompt.</p><h2>Pick the job before the format</h2><p><strong>The move:</strong> find one question your audience already asks on repeat, and build for that question only. Go mine the places the questions already live: your DMs, your sales call recordings, your inbox, the comments on whatever post performed best this quarter. You're looking for a question with commercial weight behind it, the kind where the answer changes what somebody buys.</p><p><strong>The mapping</strong> (adapted from Crescitaly's planning model, which mostly matches how I'd sort it):</p><pre><code><code>"How do I even start with X?" on repeat  -&gt;  guided checklist, outputs a personal action plan
"Should I do X or Y for my case?"        -&gt;  3-question decision quiz, outputs a recommendation
"What would this cost / return?"         -&gt;  simple calculator, outputs a planning range
"Are we ready for X?"                    -&gt;  scorecard, outputs a gap read that earns the audit call
</code></code></pre><p><strong>What a correct pick looks like:</strong> you can write the result as one sentence before you build anything. "You're leaving roughly $X on the table from no-shows every month" is a buildable result, and so is "Your setup scores 6 out of 10 on AI readiness, here are the two gaps." If the result needs a paragraph, the tool is too broad.</p><p><strong>The failure mode:</strong> picking a format because it demos well. A gorgeous interactive experience wrapped around a question nobody was asking converts worse than the ugly PDF it replaced, because at least the PDF answered something. Format follows question, every time.</p><p><strong>Walk away with:</strong> one repeated audience question and a one-sentence result definition, written down before you open Canva.</p><h2>The build prompt</h2><p><strong>The move:</strong> open Canva, start a Code build, and paste this with your blanks filled. It's structured to force the constraints that make lead magnets convert, few inputs, fast result, one call to action.</p><pre><code><code>Build a [FORMAT: calculator / 3-question quiz / scorecard] for
[WHO: e.g. agency owners who book sales calls from paid traffic].

The one job: answer "[THE QUESTION YOUR LEADS KEEP ASKING]"
in under 2 minutes.

Inputs: no more than [3-5] fields.
[LIST THEM: e.g. monthly ad spend, average deal value, close rate.]

Logic, in plain English:
[e.g. booked calls = ad spend / cost per booked call.
Revenue = booked calls x show rate x close rate x deal value.
Show the gap between their current number and a 75% show rate.]

Result screen: one sentence in this exact shape:
"[THE SENTENCE THE USER WALKS AWAY WITH, number filled in.]"
Below it, one button labeled "[CTA TEXT]" linking to [URL].

Style: [BRAND COLORS + FONT], mobile first, large tap targets,
no intro screen, no email gate before the result.

Do not include: fake precision (round all outputs), disclaimers
longer than one line, or any input field I did not list above.
</code></code></pre><p><strong>What a correct output looks like:</strong> a working preview you can poke at inside the editor, where the math holds when you test it against numbers you already know, and where the branding swap takes minutes in the toolbar instead of another generation. Under the hood Canva's code generation runs on Anthropic's Claude models, per <a href="https://vibecoding.app/blog/canva-code-review">vibecoding.app's review</a>, and simple bounded logic like this is squarely inside what it does well.</p><p><strong>The failure mode:</strong> the first generation comes back generic or the logic bends somewhere. The fix is scoped follow-up prompts aimed at one element, "the result rounds to the nearest hundred," "make the second question a slider," rather than regenerating the build. And check the arithmetic by hand with two known scenarios before you trust it, because a calculator that's wrong is worse for your credibility than no calculator.</p><p><strong>Walk away with:</strong> a working, on-brand interactive tool built from one prompt and a handful of scoped edits.</p><h2>The 10-minute pre-traffic check</h2><p><strong>The move:</strong> before the link goes in a bio, an email, or an ad, run this check start to finish. Interactive assets have more ways to break than an image, and the failures are public.</p><pre><code><code>THE 10-MINUTE PRE-TRAFFIC CHECK

1. Run it on your phone, then somebody else's phone.
   Thumb-only, no pinch-zooming required.
2. Feed it garbage: zeros, blanks, a $999,999,999 budget.
   It should fail politely. No crashes, no "NaN."
3. Read the result screen out loud. A planning aid reads
   like a planning aid. If it sounds like a guarantee,
   soften the copy.
4. Click every button. The CTA lands where it claims,
   with your tracking parameters intact on the URL.
5. Time a cold run, open to result. Over 2 minutes,
   cut an input field.
6. Confirm where responses land (Canva Sheets can collect
   form responses) and that a named human checks it.
7. Ask: does the result state what data is stored and
   keep the email field optional? Never collect more
   than the follow-up needs.
</code></code></pre><p><strong>What a correct output looks like:</strong> a stranger on a phone gets from your post to a useful, honest result in under two minutes, and you can see the completion in your response sheet with enough context to write one specific follow-up line.</p><p><strong>The failure mode:</strong> shipping on desktop-preview confidence. Nearly all of this traffic is mobile, and the version of you that built the thing is the worst possible QA tester because you already know how it's supposed to work. Hand your phone to someone who's never seen it, and watch where their thumb hesitates.</p><p><strong>Walk away with:</strong> a tested link you can put spend behind without wincing.</p><h2>Where this goes wrong</h2><p><strong>The novelty build.</strong> The tool exists because the builder was excited, so it answers no recurring question, and it gets impressive engagement from people who were never going to buy. The control is boring and useful: Crescitaly's team suggests keeping the same offer live as a plain page or PDF, and giving the interactive version more budget only if it wins on qualified action.</p><p><strong>The twelve-field intake.</strong> Somebody on the team wants "just one more question" until the calculator turns into an application form, and completion rate falls off a cliff. A social click is fragile, the result has to arrive before the patience runs out.</p><p><strong>The overcertain score.</strong> A marketing scorecard that reads like a diagnosis creates trust problems and, in some industries, compliance problems. Round numbers, hedged language, one-line honesty about what the tool can't know.</p><p><strong>The orphaned inbox.</strong> Responses collect in a sheet nobody owns, and the hottest lead of the month sits unanswered for nine days. The asset is the cheap part, the follow-up system it feeds is where the revenue is.</p><h2>The honest tradeoffs</h2><p>This is a widget engine, and the ceiling is real. Everything runs client-side, so there's no user login, no database behind it, and nothing that needs server logic, which rules out anything past a bounded tool. The vibecoding.app review puts it plainly: "You cannot build apps. You cannot export code." Your build lives inside Canva's ecosystem, HTML comes in but doesn't come out, and that matters if you ever want to move the asset somewhere Canva isn't.</p><p>AI credits meter the whole thing, shared across Canva's other AI features, and free accounts get the tightest limits, so a heavy build week can hit the ceiling before the month resets. The speed and adoption numbers are Canva's own measurements with no independent audit. And generated logic still deserves a human check every time, <a href="https://www.entrepreneuraitools.com/small-business-ai-updates-july-21-2026/">one operator-focused review</a> put the risk well: "Generated code may contain weak logic or hidden errors," and "publishing speed can discourage proper review." For a prototype or a lead magnet, none of this is disqualifying. For checkout flows, client portals, or anything holding sensitive data, this is the wrong tool and a real developer is still the answer.</p><h2>Recap</h2><p>Canva Code 2.0 went to every plan on July 14, which moved the interactive lead magnet from a build-cost decision to a judgment decision. The judgment part: mine your DMs and sales calls for one repeated question with money behind it, define a one-sentence result, build the smallest version with the prompt above, run the 10-minute check on a phone that isn't yours, and measure it against your existing static download on qualified action, keeping the winner. The Static Magnet era of lead capture didn't end because PDFs stopped working, it ended because the thing that outperforms them stopped costing anything to try.</p><p>If you want to feel the difference from the lead's side before you build one, take the <a href="https://showtime.muddventures.com/ai-iq-test">AI IQ Test</a>. It's a scored quiz I run in front of my own offers, it takes about three minutes, and you'll see the mechanics from the receiving end, the bounded inputs, the specific result, the one next step. Steal the structure for your own.</p><p>Reply with the one question your leads keep asking you, and I'll tell you which format I'd build for it.</p><p>Andrew</p><p><strong>P.S.</strong> If you're building this week: the <a href="https://blog.muddventures.com/p/run-the-30-minute-canva-test-that">30-minute Canva Grow test</a> covers the ad-creative side of the same absorption pattern, <a href="https://blog.muddventures.com/p/notion-just-made-zapier-optional">Zapier Optionality</a> is the original writeup on software eating its neighbors' invoices, the <a href="https://blog.muddventures.com/p/run-the-three-prompt-google-business">three-prompt Google Business Profile check</a> does the same commoditization math on listings retainers, and the full archive lives at <a href="https://muddventures.substack.com/">muddventures.substack.com</a>.</p>]]></content:encoded></item><item><title><![CDATA[The 3 changes this week: A free model upgrade, an agent hub, and a 2-minute privacy audit ]]></title><description><![CDATA[A smarter Claude at the same price, agents moving into HubSpot, and shared AI chats showing up on Google.]]></description><link>https://blog.muddventures.com/p/the-operators-read-3-changes-this</link><guid isPermaLink="false">https://blog.muddventures.com/p/the-operators-read-3-changes-this</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Thu, 30 Jul 2026 16:28:03 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!xVRJ!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!xVRJ!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!xVRJ!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!xVRJ!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!xVRJ!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!xVRJ!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!xVRJ!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/ec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1116032,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/209136608?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!xVRJ!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!xVRJ!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!xVRJ!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!xVRJ!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fec14ba4d-f59e-4134-8996-fb2c03a7a216_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p>This newsletter gets drafted on a Claude scheduled task before I'm awake, so when Anthropic swaps the engine that kind of work runs on, I feel it the same morning without touching a setting. That swap happened Friday. Claude Opus 5 shipped on July 24 as the new default on Anthropic's top plans, at the same price as the model it replaced, and most of the people paying for those plans will never open the announcement.</p><p>That's one of three changes from the last seven days worth an operator's attention, and none of the three got the coverage the model-leaderboard drama usually pulls. By the end of this you'll have a two minute privacy check to run on your own AI account, a test for deciding which written-off tasks deserve a rerun, and a read on what HubSpot moving agents into the CRM means for software you're already paying for.</p><p>Here's what's in this one:</p><ul><li><p><strong>The daily-driver model got smarter at the same bill.</strong> Opus 5 replaced Opus 4.8 on Friday at identical pricing, and the right response is rerunning old failures, no new subscription involved.</p></li><li><p><strong>HubSpot gave agents one home inside the CRM.</strong> Agent Hub and Agent Builder went to public beta for every Professional and Enterprise account on July 23.</p></li><li><p><strong>Shared Claude chats showed up in Google results.</strong> Fixed within days, and the two minute cleanup is still worth running today.</p></li></ul><h2>The upgrade you already paid for</h2><p>Anthropic released Claude Opus 5 on July 24 at $5 per million input tokens and $25 out, the same sticker as Opus 4.8, the model it replaces. Anthropic's own line is that it lands close to Fable 5, the company's most capable public model, at half the cost, and it's now the default on the Max plans and the strongest option on Pro. The full breakdown with benchmark tables is in <a href="https://tech-ish.com/2026/07/24/claude-opus-5-launch-benchmarks-price/">tech-ish's launch coverage</a>.</p><p>Two numbers in the launch material matter more for operators than any leaderboard placement. On AutomationBench, a test of whether a model can carry an ordinary business task from start to finish, Opus 5 passed 26 percent, against 18.1 percent for OpenAI's GPT-5.6 Sol. Zapier CEO Wade Foster <a href="https://tech-ish.com/2026/07/24/claude-opus-5-launch-benchmarks-price/">says it topped Zapier's own automation leaderboard</a> and cleared a customer-churn task no previous model had passed. Read those two together and the picture is that the best model on the market still completes about a quarter of ordinary business tasks end to end, which is why every issue in this series keeps landing on the same design: bounded jobs, human checkpoints, <a href="https://blog.muddventures.com/p/set-up-the-calendar-workflow-that">the Book Line</a> and its cousins.</p><p>Here's the operator move, and it has nothing to do with switching tools. Somewhere in your head is a list of jobs you handed an AI in 2025 that came back wrong, the messy spreadsheet cleanup, the proposal draft that missed the point, the report that invented a number, and each of those failures hardened into a conclusion about what AI can't do. I call that conclusion The Stale No. Model capability moves every few months, your conclusions mostly don't, and a no you collected from a model two generations back is expired evidence you're still acting on.</p><p>When the default engine in a tool you already pay for gets a real jump, upgrade day is the day the Stale No list gets rerun. Write the list down this time:</p><pre><code><code>MY STALE NO LIST

Tasks I tried with AI that failed, with the date I gave up:
1. [task] - [month/year] - [what went wrong, one line]
2. [task] - [month/year] - [what went wrong, one line]
3. [task] - [month/year] - [what went wrong, one line]

Rerun rule: when the model behind my main AI tool gets replaced,
rerun the top 3 with the same inputs I used the first time.
Keep the output only if I can verify it in under 5 minutes.
</code></code></pre><p>Worth flagging: every benchmark figure above is Anthropic's own run, no independent numbers exist yet, and requests that trip a safety filter quietly fall back to the older Opus 4.8, so some of what you get on any given day is still the previous model. The rerun costs you twenty minutes either way, and the worst case is confirming your no is still fresh.</p><h2>HubSpot gave its agents one home</h2><p>HubSpot launched <a href="https://www.hubspot.com/company-news/meet-agent-hub-and-agent-builder">Agent Hub and Agent Builder</a> in public beta on July 23, live now for every Professional and Enterprise account. Agent Hub is one screen showing every agent running across marketing, sales, and service, live status and performance included, and Agent Builder lets you describe an automation in plain language and have it built against the deal history, contact records, and call transcripts already sitting in your CRM.</p><p>The problem it targets is real. HubSpot product chief Duncan Lennox describes what happens once a team runs more than one agent: "they become fragmented, all working from different pictures of the customer, or even worse, no picture at all." A prospecting agent emails an account the same week a service agent is handling that account's open complaint, and neither knows about the other. If you run HubSpot, the pattern I'd copy comes from the launch customer: <a href="https://www.hubspot.com/company-news/meet-agent-hub-and-agent-builder">Ignite Reading</a>, a tutoring program operating across 25 states, pointed Agent Builder at one narrow grunt task, parsing each school district's academic calendar, and turned a 15 to 20 minute chore into seconds, north of 350 hours a year back. HubSpot picked that story for the press release, so hold it loosely, but look at its shape: bounded, verifiable, zero customer contact.</p><p>The skeptical read comes from Gartner's Kathy Ross, <a href="https://www.cxtoday.com/ai-automation-in-cx/hubspot-takes-aim-at-ai-agent-sprawl/">talking to CX Today</a>: "They're very powerful tools, but they're not employees, they're not teammates, and they have to be managed like technology." A human rep having a bad day affects a handful of conversations, a misfiring agent reaches your whole list before anyone looks up. Which is why the first agent to turn on is one whose mistakes stay inside the building, and why the access question from <a href="https://blog.muddventures.com/p/4-questions-anthropics-security-team">the Anthropic security piece</a> applies here word for word: what can this agent touch, and whose login is it borrowing to touch it.</p><h2>The 2-minute check on your shared AI chats</h2><p>Over the weekend, Reddit users found that a Google search along the lines of site:claude.ai/share surfaced long lists of publicly shared Claude conversations and Artifacts. <a href="https://techcrunch.com/2026/07/27/psa-your-claude-shared-chats-and-artifacts-may-have-ended-up-on-google/">Futurism found</a> "a detailed medical report of a real patient, clinical trial results that included patient names," and company documents marked internal-use-only in the exposed set. Anthropic's position, via spokeswoman Amie Rotherham: "When someone shares a conversation, they are making that content publicly accessible, and like other public web content, it may be archived by third-party services." By Monday afternoon the search results were gone, per <a href="https://techcrunch.com/2026/07/27/psa-your-claude-shared-chats-and-artifacts-may-have-ended-up-on-google/">TechCrunch's own testing</a>.</p><p>The fix on Anthropic's side landed fast. The lesson on your side is permanent, and it applies to every AI tool with a share button: a share link is a publish button with a delay on it. The chat you sent a client in March with your pricing logic in it, the thread you shared with a contractor that had customer names in it, those are public web pages, and whether they stay obscure depends on nobody ever posting the URL anywhere a crawler can see it. ChatGPT went through its own version of this last year, when a researcher scraped <a href="https://techcrunch.com/2026/07/27/psa-your-claude-shared-chats-and-artifacts-may-have-ended-up-on-google/">around 100,000 shared conversations</a> that had been set public.</p><p>Run the cleanup now:</p><pre><code><code>THE SHARE-LINK AUDIT (2 minutes)

Claude:   Settings -&gt; Privacy -&gt; Shared Chats
          Delete every link you don't have a live reason to keep.

ChatGPT:  Settings -&gt; Data Controls -&gt; Shared Links
          Same rule. If you can't name who still needs it, kill it.

Going forward, one question before sharing any chat:
"Would this page be fine showing up in a Google result
with my name on it?" If no, copy the text into a doc instead.
</code></code></pre><h2>Where I'd start</h2><p>Ranked by effort against payoff: the share-link audit is two minutes and closes a real exposure, so it goes first, today. The Stale No rerun is twenty minutes this week, and it's the one most likely to hand you a working automation you'd already paid for and walked away from. The HubSpot pilot is for HubSpot shops with an hour to spare, one bounded internal task in Agent Builder, mistakes that stay inside the building, then a look at Agent Hub's performance screen after a week to see what the thing did while you weren't watching.</p><p>I've used AI every working day since February 2023, and weeks like this one are the argument for running a system instead of a scramble: the tools underneath you upgrade themselves, sprout agents, and spring leaks whether you're watching or not. Keeping up turns out to be a set of small habits, a list you rerun, a settings page you check, a pilot you scope deliberately small. Inside <a href="https://whop.com/abra-ai/">the Abra AI community</a> that's the standing conversation, operators comparing what cleared their own Stale No lists and which agent pilots earned a second week, and the skill files behind this newsletter's own automation live there too.</p><h2>Recap</h2><p>Opus 5 replaced Opus 4.8 at the same price on July 24, so rerun the tasks you wrote off instead of shopping for new tools. HubSpot put agent building and agent monitoring inside the CRM on July 23, and the first agent worth turning on is one whose mistakes can't reach a customer. And your shared AI chats are public web pages, so spend the two minutes deleting the ones that have no reason to exist.</p><p>Reply with the one task you wrote off as too much for AI in 2025. I want to see how many of those come back from the dead this quarter.</p><p>Andrew</p><p><strong>P.S.</strong> Three related builds from the archive: the <a href="https://blog.muddventures.com/p/4-questions-anthropics-security-team">4-question agent access check</a> borrowed from Anthropic's own security team, the <a href="https://blog.muddventures.com/p/set-up-the-calendar-workflow-that">calendar workflow with the Book Line built in</a>, and the <a href="https://blog.muddventures.com/p/run-the-three-prompt-google-business">three-prompt Google Business Profile check</a>. The full archive lives at <a href="https://muddventures.substack.com/">muddventures.substack.com</a>.</p>]]></content:encoded></item><item><title><![CDATA[Run the 15-minute Zap audit that finds the AI steps billing three tasks where one would do]]></title><description><![CDATA[Zapier now bills AI steps at 1x, 3x, or 5x per run, and every new step defaults to 3x. The audit below takes 15 minutes and pays for itself on the first downgrade.]]></description><link>https://blog.muddventures.com/p/run-the-15-minute-zap-audit-that</link><guid isPermaLink="false">https://blog.muddventures.com/p/run-the-15-minute-zap-audit-that</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Wed, 29 Jul 2026 14:53:58 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!sEIP!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!sEIP!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!sEIP!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!sEIP!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!sEIP!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!sEIP!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!sEIP!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:800002,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/208985446?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!sEIP!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!sEIP!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!sEIP!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!sEIP!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F07d655b6-a10d-4c94-a0fa-1170b563dd10_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TL;DR:</strong> Zapier moved AI steps to model-tier pricing on June 15, and the doc got a fresh update on July 15. Standard models bill 1x tasks, Advanced bills 3x, Premium bills 5x, and every tool call inside the step multiplies at the same rate, so one AI step can burn 15 tasks in a single run. The detail most people will miss: every new AI step defaults to the 3x tier. Below is the billing formula with worked math, a 15-minute audit for every Zap you run, and the two lanes that bring a step back down to 1x.</p><p>I read Zapier's updated AI pricing doc this week, the one they refreshed on July 15, and the thing that stopped me sits in a dropdown. When you add an AI step to a Zap, the model tier ships preset to Advanced, which bills three tasks per run before you've typed a single word of your prompt. Standard is right there in the same menu at one task. Nobody's hiding it, the tiers are printed in the selector, and the markup still collects from everyone who never opens the menu, which is most people, because defaults are where software makes its quiet money.</p><p>By the end of this you'll know what every AI step in your Zaps costs per run, and which steps can move to a 1x lane without losing anything.</p><p><strong>The dropdown doing the billing:</strong> every new AI step ships set to the 3x tier, and that setting rides along until somebody opens the menu.</p><p><strong>The formula worth memorizing:</strong> one run costs (1 x model rate) + (tool calls x model rate), which is how a single step reaches 15 tasks.</p><p><strong>The two 1x lanes:</strong> simple judgment jobs run at 1x on Standard, and your own API key runs tool-heavy steps at 1x too.</p><p><strong>Where the meter can't reach:</strong> filters, paths, and formatter steps bill zero tasks, so rule-based work costs nothing once you move it out of the AI step.</p><p>I ran the numbers on what that dropdown costs at volume, and this issue is that math. It's the same invoice logic I walked through on <a href="https://blog.muddventures.com/p/run-the-30-minute-canva-test-that">the Canva retainer test</a>: the vendor publishes the honest price sheet, and the gap between what you pay and what the job requires lives in a setting nobody audits. I'm calling this one The 3x Default, the pattern where a platform sets its pricier tier as the preselected option on every new AI step, and the difference quietly compounds across every run of every Zap you've got.</p><p><strong>Who this is for:</strong> operators running Zapier on a Professional or Team plan with at least one AI step in production, and anyone whose task bill has been creeping without an obvious reason.</p><h2>How the new meter works</h2><p>Zapier switched AI steps to model-tier pricing on <a href="https://help.zapier.com/hc/en-us/articles/46597632373389-AI-by-Zapier-new-model-based-pricing-starting-June-15-2026">June 15</a>, and the <a href="https://help.zapier.com/hc/en-us/articles/46425475442829-AI-by-Zapier-model-tier-pricing">full pricing doc</a> lays out the tiers. Standard models like GPT-5 mini and Gemini 2.5 Flash bill at 1x and can't call tools. Advanced models like Claude 4.5 Haiku and Gemini 2.5 Pro bill at 3x and can. Premium models like GPT-5.4 and Claude 4.8 Opus bill at 5x. Bringing your own API key bills at 1x no matter what it's doing, and that lane is going to matter in a minute.</p><p>The formula is the part worth writing on a sticky note:</p><pre><code><code>Tasks per run = (1 x model rate) + (tool calls x model rate)

Standard, no tools:        (1 x 1) + (0 x 1) = 1 task
Advanced (the default):    (1 x 3) + (0 x 3) = 3 tasks
Advanced, 1 tool call:     (1 x 3) + (1 x 3) = 6 tasks
Premium, 2 tool calls:     (1 x 5) + (2 x 5) = 15 tasks
Your own key, 4 tool calls: (1 x 1) + (4 x 1) = 5 tasks</code></code></pre><p>Those aren't my estimates, they're Zapier's own worked examples from the pricing doc. A tool call is any time the step successfully reaches into an app or a knowledge source during a run, so an AI step that checks your CRM and then writes to a sheet is carrying two multiplied calls on top of its base rate.</p><p>What the right read looks like: you can name the per-run cost of every AI step you have. The failure mode is assuming an AI step still costs one task the way it did before June 15, because a Zap that was cheap in May can be billing at triple or better in July with zero changes on your end, and legacy steps built before June 15 run under their own rules now too, including a June 30 pause on older Standard-model steps that used knowledge sources.</p><p><strong>Walk away with:</strong> the formula and your own per-run number for each AI step, written down.</p><h2>The 15-minute audit</h2><p>The task meter is where this gets real. On Zapier's Professional annual plan, $19.99 a month covers 750 tasks, and <a href="https://www.eesel.ai/blog/zapier-subscription">eesel's breakdown of the billing pages</a> puts 10,000 tasks a month at roughly $300, with overages billing at 1.25x your base rate once you blow through the cap. An operator on r/aiagents posted his own invoice math in February: $847 for one month, an 8-step lead-capture flow, 100 leads a day, 24,000 tasks a month, and his conclusion is the part I'd frame: <a href="https://www.reddit.com/r/aiagents/comments/1r4otc3/unpopular_opinion_zapiers_taskbased_pricing_is/">"The worst part is you start optimizing for Zapier's pricing instead of what's best for your business. I caught myself removing steps from workflows just to save on task counts."</a></p><p>That was before model multipliers existed. Under the new meter, the same reflex points somewhere better: stop deleting steps and start checking tiers. Here's the audit, and it runs in about 15 minutes for a normal-sized account:</p><pre><code><code>THE 3X DEFAULT AUDIT

1. Open each Zap that contains an AI by Zapier step.
2. For each AI step, write down three things:
   - Model tier (open the dropdown: Standard / Advanced / Premium / own key)
   - Number of tools attached to the step
   - Roughly how many times the Zap runs per month (Zap history shows this)
3. Compute per-run cost: (1 x rate) + (tool calls x rate)
4. Multiply by monthly runs. That's the step's monthly task burn.
5. Rank your AI steps by burn, highest first.
6. For the top of the list, ask one question:
   does this step use tools, and does it require deep reasoning?
   - No tools + simple job (summarize, classify, extract) -&gt; test Standard at 1x
   - Tools attached -&gt; test your own API key at 1x
   - No judgment in the job at all -&gt; move it out of AI entirely (free steps below)</code></code></pre><p>A worked example, because the abstraction hides the money. A lead-summary step on the untouched default tier with two tool calls costs nine tasks a run, and at 30 leads a day that's about 8,100 tasks a month from one step, which is most of the way to that $300 pricing band on its own. The identical step running through your own API key costs three tasks a run, about 2,700 a month. If the job doesn't use tools at all, Standard runs it at one task, roughly 900 a month. Same Zap, same output, three very different invoices.</p><p>What the right output looks like: a ranked list where the top two or three steps explain most of your AI task burn, because that concentration is what makes the fix fast. The failure mode is auditing one Zap and calling it done, when the whole reason this compounds is that the default rides along on every AI step you or your team has added since June 15, including the duplicated ones sitting across client accounts.</p><p><strong>Walk away with:</strong> a ranked list of your AI steps by monthly task burn, and a downgrade candidate circled.</p><h2>Moving a step down without breaking it</h2><p>Downgrading a tier is a thirty-second edit, and the tier menu lives right on the step, so the real work is proving the cheaper tier holds quality. Zapier's own Steph Spector put the target well in <a href="https://zapier.com/blog/minimize-ai-spend/">the company's cost guide</a>: "your goal is to value-maxx, not token-maxx," and the same post notes "you pick the tier that matches the job, and you can swap models anytime without rebuilding anything." Run the proof before you commit:</p><pre><code><code>TIER PARITY TEST

Take the last 5 real inputs this AI step processed (pull from Zap history).
Run them through the step twice: once on the current tier, once on the cheaper one.
Compare outputs side by side and ask:
- Did the cheaper tier follow the format instructions?
- Did it catch the same key facts?
- Would the downstream step (CRM write, Slack ping, sheet row)
  behave identically with this output?
5 for 5 -&gt; downgrade and note the date, then spot-check in a week.
3 or 4 of 5 -&gt; tighten the prompt and rerun before deciding.
Under 3 -&gt; the step earns its tier, leave it and move down your list.</code></code></pre><p>Three lanes, in order of how often they pay off. First, the judgment jobs that touch no tools, the summarize-classify-extract work that makes up most AI steps I come across in real accounts, and those run on Standard at 1x more often than people expect. Second, the tool-heavy steps, where connecting your own OpenAI or Anthropic key drops the multiplier to 1x, and yes the token bill lands on your API account instead, but for a step making a handful of calls per run the math usually lands well under the 3x and 5x task rates. Third, the jobs with no judgment in them at all, the routing and reformatting and if-then work, and those belong outside AI steps completely, because Zapier's filters, paths, and formatter steps bill zero tasks, which means every rule you move out of a prompt and into a path is billing that drops to nothing.</p><p>What the right result looks like: same outputs downstream, smaller number on the meter, and a one-line note in your ops doc saying which steps run on which tier and why. The failure mode is dropping a tool-connected step to Standard, because Standard can't call tools at all, so the step doesn't get cheaper, it stops doing part of its job, and that's a break you might not notice until the CRM rows stop appearing.</p><p><strong>Walk away with:</strong> at least one step moved to a 1x lane with a parity test behind it, and the rule-based work headed out of AI steps entirely.</p><h2>What's still true on Zapier's side</h2><p>The default isn't a scam, and the tradeoffs deserve daylight. Advanced is the default because tools require Advanced or better, and Zapier presumably decided a new step that can use tools out of the box beats a new step that errors when you attach one. There's also a real protection built in: if a single run hits 75 tasks, the Zap pauses and asks for approval before it keeps spending, which is more of a circuit breaker than most AI billing systems give you.</p><p>The bring-your-own-key lane carries its own homework, since you're now managing API keys and watching a second bill, and the insulation cuts both ways: on Zapier's included models your cost per run stays fixed no matter what token prices do, which is worth something. Sara McNamara, a Zapier partner, has <a href="https://www.linkedin.com/posts/saramcnamara_zapierpartner-sponsored-mygenuineopinionthough-activity-7361451043664142340-nBDI">made the case</a> that the per-task model is genuinely predictable once you understand it, and that for teams with compliance needs and SLAs, predictability can matter more than raw cost per action. She's right about the predictability, and the audit above is how you make the predictable number a smaller one.</p><p>Two more honest flags. My dollar figures are directional, because your per-task price depends on where your plan sits on Zapier's task slider, so run the math against your own bill, and note that Zapier's claim of roughly $1,300 in tokens for a thousand form submissions run through a chat assistant is Zapier's own benchmark measurement, not an independent audit. And the whole AI-step system sits on Professional plans and up, with tool calls not yet available on Enterprise accounts at all, which is a strange gap for the most expensive tier.</p><h2>Failure modes</h2><p>The audit dies in four predictable places. Checking tiers on three Zaps and skipping the client sub-accounts where the duplicated flows live. Downgrading a tool-connected step to Standard and breaking it silently, covered above, run the parity test. Treating the own-key lane as free when it moves cost to your API bill instead of erasing it, so compare both invoices after a week. And doing the audit once, in July, when the default reapplies itself to every AI step anyone on your team adds from here forward, so put a monthly fifteen-minute recheck on the calendar next to your task-usage email.</p><h2>Recap</h2><p>Zapier's AI steps now bill by model tier, 1x, 3x, or 5x, tool calls multiply at the same rate, and the default on every new step is 3x. The formula is (1 x rate) + (tool calls x rate). The audit is: list your AI steps, log tier and tools and monthly runs, rank by burn, then move what you can to the 1x lanes, Standard for simple no-tool jobs, your own key for tool-heavy ones, and free deterministic steps for anything with no judgment in it. Test parity on five real inputs before committing, and recheck monthly because The 3x Default resets with every new step.</p><p>This kind of audit is a slice of what I look at when an operator wants their whole AI spend mapped against what the work requires, and if you want a second set of eyes across the full stack rather than one tool, that's a fit for an AI Clarity Call at <a href="https://muddventures.com/book">muddventures.com/book</a>.</p><p>Reply with your monthly task number if you know it off the top of your head, most people don't, and that's sort of the point.</p><p>Andrew</p><p>P.S. More from the archive on the same money trail: <a href="https://blog.muddventures.com/p/run-the-30-minute-canva-test-that">the 30-minute Canva test for your ad creative retainer</a>, <a href="https://blog.muddventures.com/p/4-questions-anthropics-security-team">the 4 questions to run before an agent gets access to your systems</a>, <a href="https://blog.muddventures.com/p/set-up-the-calendar-workflow-that">the calendar workflow with the one job AI never runs alone</a>, and the full newsletter archive at <a href="https://muddventures.substack.com">muddventures.substack.com</a>.</p>]]></content:encoded></item><item><title><![CDATA[What does OpenAI want from your small business in exchange for all this free training?]]></title><description><![CDATA[Free webinars, in-person academies, partner deals, and a very clear reason it all exists. The trade works in your favor if you walk in knowing what you came to take.]]></description><link>https://blog.muddventures.com/p/what-does-openai-want-from-your-small</link><guid isPermaLink="false">https://blog.muddventures.com/p/what-does-openai-want-from-your-small</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Tue, 28 Jul 2026 15:31:41 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!sAvI!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!sAvI!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!sAvI!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!sAvI!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!sAvI!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!sAvI!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!sAvI!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:538214,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/208846124?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!sAvI!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!sAvI!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!sAvI!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!sAvI!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F04f92e3a-4d28-42bc-ab50-f87b251533dd_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p>I spent years on the sales and marketing side before going all in on AI, and one play showed up in every growth playbook I ever touched: teach something real for free, fill the room, and let the product do the selling from the front of it. A week ago today, OpenAI started running that play on small business owners at a scale no agency could touch.</p><p>On July 21 they launched the <a href="https://openai.com/index/introducing-chatgpt-small-business-program/">ChatGPT for small business program</a>: free virtual trainings, in-person AI academies across the US, guides built to load straight into ChatGPT, and partner offers from Shopify, Intuit, Dropbox, Slack, Atlassian, and Wix. I read the announcement twice, once as an operator who'll take useful free training from anybody, and once as a guy who used to build funnels like this for a living, and both reads land in the same place: sign up, and walk in holding your own agenda.</p><p>By the end of this you'll know which pieces of the program are worth your time, why it exists at all, and the one-page agenda that turns a vendor demo into a workflow you leave with.</p><p><strong>The useful parts are free and live right now.</strong> Webinars, in-person academies, uploadable guides, and partner deals, all linked below.</p><p><strong>The reason it exists is on the record.</strong> 95% of ChatGPT's 900 million weekly users pay nothing, and OpenAI's CFO has been open about growing the business side ahead of an expected IPO.</p><p><strong>Their own number sets the bar.</strong> OpenAI says 78% of last year's AI Jam attendees built a working workflow in a single day, so show up planning to be in that 78%.</p><p><strong>The agenda you bring decides which side of that number you land on.</strong> There's a fill-in block below to complete before you register.</p><h2>What went live on July 21</h2><p>The program has four pieces. <a href="https://webinar.openai.com/small-business/">Hands-on virtual trainings</a>, product webinars showing ChatGPT Work running day-to-day small business jobs across accounting, marketing, and ecommerce, with prompts to take home and partner Q&amp;As. <a href="https://academy.openai.com/home/clubs/small-business-ipf4m/events">In-person AI academies</a>, run by the OpenAI Academy team in cities across the US, local owners in a room doing guided exercises together. New guides and customer stories, including interactive ones you upload straight into ChatGPT to start a workflow. And a curated set of partner tools and offers from Dropbox, Shopify, Intuit, Slack, Atlassian, and Wix, built around common small business workflows.</p><p>All of it points at <a href="https://openai.com/index/chatgpt-for-your-most-ambitious-work/">ChatGPT Work</a>, the agent OpenAI shipped July 9 alongside GPT-5.6, built to run multi-step jobs end to end once you connect it to your files, your apps, and your way of working. The same day the program launched, OpenAI said ChatGPT Work and Codex crossed <a href="https://9to5mac.com/2026/07/21/openai-launches-small-business-program-as-it-touts-10m-chatgpt-work-and-codex-users/">10 million combined users</a>, double what Codex alone had earlier in July. The push is working, and small business is the next room they want to fill.</p><h2>The trade underneath the free training</h2><p>None of this is charity, and OpenAI hasn't pretended otherwise. CFO Sarah Friar told the AP earlier this year that <a href="https://apnews.com/article/openai-chatgpt-spud-sam-altman-anthropic-mythos-3c2674f5cdf67ac6d88eedb207de117c">95% of ChatGPT's 900 million weekly users pay nothing</a>, the company is <a href="https://www.pymnts.com/news/artificial-intelligence/2026/openai-launches-program-to-accelerate-small-business-ai-adoption/">still unprofitable at an $852 billion valuation with an IPO expected</a>, and business customers already drive about 40% of revenue. TechRadar's read on the launch was blunt: <a href="https://www.techradar.com/pro/openai-wants-to-help-your-small-business-grow-if-you-use-chatgpt-more">"The scheme is primarily an education process"</a>, and small businesses represent around 99% of the private sector, which is a lot of unopened wallets.</p><p>There's a name worth putting on this, because you're going to see a lot more of it: The Vendor Classroom. Free training a platform runs where the tool is the teacher, the syllabus, and the sales pitch at the same time. I watched versions of this play from the inside for years, a free workshop that teaches something real is one of the oldest pipelines in marketing, and the people who get the most out of one are always the ones who show up with a specific problem to solve before the pitch ever starts.</p><p>Also, my verdict up front: the trade is fine. Agencies charge four figures for worse training than a platform gives away when it wants your workflows living inside its product. The price you pay in a Vendor Classroom is your attention becoming pipeline, and attention is a cost you can manage.</p><h2>What I'd grab this week</h2><p>Three pieces are worth moving on.</p><p>First, the in-person academies if one lands near you. The screen-share webinar you can get anywhere, a room full of local owners comparing what they've automated is much harder to find, and OpenAI's own numbers from last year's <a href="https://openai.com/index/small-business-ai-jam/">Small Business AI Jams</a> say 78% of participants built a functional AI workflow in a single day and 42% saved more than five hours a week afterward. Those are vendor-measured numbers, more on that below, but even discounted they describe a working format: hands on, one day, leave with something running.</p><p>Second, the webinars, on the condition that you register with a job already picked. This is where the agenda comes in. Fill this out before you sign up, it takes ten minutes and it changes what you get out of every session:</p><pre><code><code>MY WALK-IN AGENDA (fill out before any webinar or academy)

The three tasks I repeat every week that eat the most time:
1.
2.
3.

The tools those tasks touch (name them, e.g. QuickBooks, Gmail, Shopify):

The one task I want built and running before the session ends:

What "working" looks like for that task (the output I can check):

My question for the Q&amp;A: what does this workflow read and touch
in my accounts, and can I run it in a review-first mode?</code></code></pre><p>The right output from a session is one of your three tasks running as a workflow you can rerun tomorrow without the instructor. The failure mode is the spectator run: ninety minutes of nodding at demos of businesses shaped nothing like yours, a notebook full of screenshots, nothing built. Their own 78% stat says building in one session is the norm, so hold them to it.</p><p>Third, the partner offers. There are promos from Dropbox, Shopify, Intuit, Slack, Atlassian, and Wix attached to the program, and if you already pay for any of those, ten minutes checking the offer list against your current bills is the easiest win on this page.</p><p><strong>Walk away with:</strong> one of your three weekly time-eaters running as a ChatGPT Work workflow, built during their free session, on your data, checked against your own definition of working.</p><h2>What to keep your eyes on</h2><p>An honest list, because the announcement won't give you one.</p><p>The stats are the vendor's own. The 78% and 42% numbers come from OpenAI measuring OpenAI's event, no independent audit, and the customer quotes on the launch page are hand-picked. One of them still caught my attention because it's specific: Kevin English, who owns a construction business called Keg Built, says <a href="https://openai.com/index/introducing-chatgpt-small-business-program/">"On one 30-page contractor quote, ChatGPT saved me at least 10 hours of data entry"</a>. Specific task, specific document, specific hours. That's the shape of claim worth testing on your own quotes and invoices, and the shape worth writing down about your own results.</p><p>The syllabus teaches the tool. Nobody in a Vendor Classroom is going to map your sales process or tell you which of your workflows would be better left alone, the curriculum optimizes for ChatGPT Work adoption because that's what it's for. The academies will show you what the tool can do, your business is on nobody's syllabus but yours.</p><p>ChatGPT Work wants access to run. The whole premise is an agent connected to your files, email, and apps, and connection questions deserve more thought than a workshop signup form. I wrote up the four access questions worth asking before any agent gets into your accounts in <a href="https://blog.muddventures.com/p/4-questions-anthropics-security-team">the Anthropic security piece</a> last week, and they apply to OpenAI's agent the same way.</p><p>The app under all this is still settling. OpenAI replaced its desktop app this month with a version built on Codex, and 9to5Mac notes the change <a href="https://9to5mac.com/2026/07/21/openai-launches-small-business-program-as-it-touts-10m-chatgpt-work-and-codex-users/">added complexity to the basic chat interface</a>, with a few updates since trying to rebalance chat against tasks. Expect some interface churn while you're learning.</p><pre><code><code>THREE QUESTIONS FOR ANY VENDOR CLASSROOM

1. Can I build this on the plan I already pay for, or does the
   demo assume a bigger seat?
2. What does this workflow read and touch in my accounts, and
   can I run it read-only or review-first to start?
3. When the model underneath changes, does my workflow keep
   working, and how would I find out?</code></code></pre><h2>The short version</h2><p>OpenAI built a free national training program because it wants small businesses running workflows inside ChatGPT Work before the IPO, and both halves of that sentence can be true at once: the training is real and the motive is revenue. Sign up for <a href="https://openai.com/business/why-openai/small-business/">the program</a>, pick an academy date if one's close, and fill out the walk-in agenda before you register, because the difference between extracting value from a Vendor Classroom and becoming its pipeline is whether your agenda or theirs runs the session.</p><p>Also, before you sit in anyone's classroom it helps to know where you stand. The <a href="https://showtime.muddventures.com/ai-iq-test">AI IQ Test</a> takes a few minutes and scores how your business is using AI today across marketing, ops, and sales, so you'll know which sessions are worth your seat.</p><p>Reply if you register for one of the academies, I want the field report on what they teach when the press isn't in the room.</p><p>Andrew</p><p>P.S. More from the archive worth pairing with this one:</p><ul><li><p>The four access questions before any agent touches your accounts: <a href="https://blog.muddventures.com/p/4-questions-anthropics-security-team">4 questions Anthropic's security team runs</a></p></li><li><p>The 5,000-character brief that stops you re-explaining your business in every AI chat: <a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">The Standing Brief</a></p></li><li><p>The calendar workflow that hands AI the grunt work and keeps the confirm human: <a href="https://blog.muddventures.com/p/set-up-the-calendar-workflow-that">The Book Line</a></p></li><li><p>Skill files and prompts from these issues live in the <a href="https://whop.com/abra-ai/">Abra AI community</a></p></li><li><p>Full archive: <a href="https://muddventures.substack.com/">muddventures.substack.com</a></p></li></ul>]]></content:encoded></item><item><title><![CDATA[Set up the calendar workflow that hands AI your scheduling grunt work, and the one job it should never run alone]]></title><description><![CDATA[Google and Claude will both propose meeting times off your real availability now, the whole game is knowing where the AI stops and you confirm.]]></description><link>https://blog.muddventures.com/p/set-up-the-calendar-workflow-that</link><guid isPermaLink="false">https://blog.muddventures.com/p/set-up-the-calendar-workflow-that</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Fri, 24 Jul 2026 15:33:38 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!mWAT!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!mWAT!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!mWAT!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!mWAT!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!mWAT!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!mWAT!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!mWAT!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1503695,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/208344120?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!mWAT!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!mWAT!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!mWAT!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!mWAT!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F59f8ede5-5be3-45b5-9631-59d2f255f6df_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TLDR:</strong> Google&#8217;s &#8220;Help me schedule&#8221; in Gmail and the Claude Google Calendar  connection both read your live availability and propose meeting times for you, and that whole find-times-and-draft-the-proposal job is safe to hand over end to end this week. The one job to keep human is the confirm, the moment a meeting books onto someone&#8217;s calendar. I call that boundary The Book Line. Below is the prompt that hands off the grunt work, the guardrail that holds the line, and the failure modes that show up when operators let the AI cross it.</p><p></p><p>A consulting client I worked with last year ran a small closing team, and the thing that ate their week wasn&#8217;t the calls, it was the back-and-forth to book the calls, the &#8220;does Tuesday work, no how about Thursday, mornings are better for me&#8221; chain that ran ten messages deep before a single meeting landed on anyone&#8217;s calendar.</p><p></p><p>The number that jumped out when I pulled it apart was the time cost, their setters were spending close to an hour a day just trading times, and none of that hour touched revenue. If you&#8217;re running any kind of booked-call motion, an agency, a coaching program, a service business, that hour is the thing to watch, because it&#8217;s the first job AI can lift off your plate cleanly and it&#8217;s also the first place operators get burned when they hand over one step too many.</p><p></p><p>By the end of this you&#8217;ll have a two-prompt calendar workflow that proposes real meeting times off your live availability without you touching the calendar, plus the one guardrail that keeps the AI from booking something it had no business booking.</p><p></p><p>Here&#8217;s what&#8217;s inside:</p><ul><li><p><strong>The find-and-propose job</strong> is the clean hand-off, Gemini and Claude both do it now straight off your live calendar.</p></li><li><p><strong>The Book Line</strong> is the point where the AI stops and you confirm, and Google already built that same confirm-point into its own scheduling tool, which tells you something.</p></li><li><p><strong>The auto-reschedule trap</strong> is where operators hand the calendar too much rope and it starts moving committed meetings around under them.</p></li><li><p><strong>The Sunday brain-dump</strong> is a bonus bounded job, plain-language planning turned into a calendar import in one sitting.</p></li></ul><p><strong>Who this is for:</strong> operators who book calls or meetings as part of the funnel, running Google Calendar with Gmail, Calendly, Cal.com, or a CRM calendar, and already using ChatGPT, Claude, or Gemini day to day.</p><p></p><p>This is the calendar cousin of a boundary I wrote about back in June with [Comet Walls](https://blog.muddventures.com/p/comet-will-read-your-whole-inbox), where the move was letting an AI browser read a sensitive stream like your inbox without letting it act. Same idea, pointed at your calendar this time.ZZTEST <a href="https://example.com">reallink</a> ZZEND<strong>TLDR:</strong> Google&#8217;s &#8220;Help me schedule&#8221; in Gmail and the Claude Google Calendar connection both read your live availability and propose meeting times for you, and that whole find-times-and-draft-the-proposal job is safe to hand over end to end this week. The one job to keep human is the confirm, the moment a meeting books onto someone&#8217;s calendar. I call that boundary The Book Line. Below is the prompt that hands off the grunt work, the guardrail that holds the line, and the failure modes that show up when operators let the AI cross it.</p><p>A consulting client I worked with last year ran a small closing team, and the thing that ate their week wasn&#8217;t the calls, it was the back-and-forth to book the calls, the &#8220;does Tuesday work, no how about Thursday, mornings are better for me&#8221; chain that ran ten messages deep before a single meeting landed on anyone&#8217;s calendar.</p><p>The number that jumped out when I pulled it apart was the time cost, their setters were spending close to an hour a day just trading times, and none of that hour touched revenue. If you&#8217;re running any kind of booked-call motion, an agency, a coaching program, a service business, that hour is the thing to watch, because it&#8217;s the first job AI can lift off your plate cleanly and it&#8217;s also the first place operators get burned when they hand over one step too many.</p><p>By the end of this you&#8217;ll have a two-prompt calendar workflow that proposes real meeting times off your live availability without you touching the calendar, plus the one guardrail that keeps the AI from booking something it had no business booking.</p><p>Here&#8217;s what&#8217;s inside:</p><ul><li><p><strong>The find-and-propose job</strong> is the clean hand-off, Gemini and Claude both do it now straight off your live calendar.</p></li><li><p><strong>The Book Line</strong> is the point where the AI stops and you confirm, and Google already built that same confirm-point into its own scheduling tool, which tells you something.</p></li><li><p><strong>The auto-reschedule trap</strong> is where operators hand the calendar too much rope and it starts moving committed meetings around under them.</p></li><li><p><strong>The Sunday brain-dump</strong> is a bonus bounded job, plain-language planning turned into a calendar import in one sitting.</p></li></ul><p><strong>Who this is for:</strong> operators who book calls or meetings as part of the funnel, running Google Calendar with Gmail, Calendly, Cal.com, or a CRM calendar, and already using ChatGPT, Claude, or Gemini day to day.</p><p>This is the calendar cousin of a boundary I wrote about back in June with <a href="https://blog.muddventures.com/p/comet-will-read-your-whole-inbox">Comet Walls</a>, where the move was letting an AI browser read a sensitive stream like your inbox without letting it act. Same idea, pointed at your calendar this time.</p><h2>The one job to hand it end to end: find the times and draft the proposal</h2><p>The move is to link your AI to your live calendar and let it do the whole find-open-slots-and-write-the-proposal job, because it reads your real free and busy the same way you would, only in two seconds instead of ten minutes of squinting at a week grid.</p><p>Two tools do this cleanly right now. Google turned on &#8220;Help me schedule&#8221; inside Gmail back on October 15, 2025, and it reads your Google Calendar plus the email&#8217;s context, then drops a card of open slots into your draft that you can edit before it goes out. The Claude Google Calendar connection goes further, it can read and write your calendar, find open time across your week, and draft the invite, all from a plain-language ask (per <a href="https://claude.com/connectors/google-calendar">Anthropic&#8217;s connector page</a>).</p><p>Here&#8217;s the prompt I&#8217;d hand it. This assumes you&#8217;ve connected Claude (or ChatGPT with calendar access, or you&#8217;re inside Gemini in Gmail) to the calendar that holds your real availability:</p><pre><code><code>You have access to my Google Calendar. I need to propose meeting times to
[NAME / COMPANY] for a [30-minute discovery call].

Rules:
- Pull 4 open slots from my LIVE calendar over the next 7 business days.
- Only weekdays, only between 9am and 4pm my time (America/Los_Angeles).
- Leave a 15-minute buffer before and after anything already booked.
- Skip any slot that collides with an existing event, tentative or confirmed.
- Format the output as a short email I can send, with the 4 times as a
  clean list and my timezone stated once.
- Do NOT create, book, or send anything. Draft only. Stop and show me.
</code></code></pre><p>The correct output is a short drafted email with four slots that map to your real open time, in the recipient&#8217;s rough ballpark, with your timezone stated once so nobody has to do mental math. Google&#8217;s own version color-codes this for you, green slots mean everyone&#8217;s free, amber means a conflict and it flags you to pick manually (that&#8217;s straight from <a href="https://support.google.com/calendar/answer/16865189">Google&#8217;s Help me schedule docs</a>). When it lands right, it feels like the thing one commenter described after wiring the Claude connection to his calendar, &#8220;just ramble and let it rip. Pure calendar magic&#8221; (<a href="https://www.androidpolice.com/paired-gemini-with-google-calendar-to-plan-my-week/">Android Police comments, May 2026</a>).</p><p>The failure mode to catch: it proposes a slot you&#8217;re already busy in. That means one of two things happened, either the calendar sync is stale, or the model guessed at times instead of reading your live free and busy. If a proposed slot collides with something real on your calendar, don&#8217;t fix that one slot and move on, re-connect the calendar and re-run the whole prompt, because a tool that invented one time will invent others.</p><p><strong>Walk away with:</strong> a four-slot meeting proposal drafted off your real availability in one prompt, ready for you to read and send.</p><h2>The Book Line: where the AI stops and you confirm</h2><p>The one job to keep human is the confirm, the moment a meeting turns from a proposal into a commitment on somebody&#8217;s calendar. That&#8217;s The Book Line. The AI can find the times, draft the proposal, prep the invite, and stack the agenda, and it stops the instant an event would book, because a human owns the commitment.</p><p>The reason I trust this boundary is that Google drew the same one inside its own product. Walk through what &#8220;Help me schedule&#8221; makes you do: Gemini proposes the slots, you click Send, the recipient picks a time, and then the recipient clicks &#8220;Book time for all guests&#8221; to create the event (<a href="https://support.google.com/calendar/answer/16865189">Google&#8217;s docs</a> spell out that last step). Two humans touch the confirm before anything lands. Google built the most-used AI scheduler on the planet and still put a person on each side of the booking, which tells you the confirm is the part nobody wants a model doing alone.</p><p>Here&#8217;s the guardrail I paste into any calendar automation so the AI can&#8217;t drift past the line:</p><pre><code><code>CALENDAR GUARDRAILS, always apply:
- You may READ my calendar and DRAFT proposals, invites, and agendas.
- You may NOT create, book, confirm, accept, decline, move, or delete any
  event without me telling you to, per event, in that moment.
- Never turn on auto-accept for incoming invites.
- Never reschedule an existing confirmed meeting on your own. If two things
  collide, surface the conflict and wait for me to choose.
- When a task would cross this line, stop and say:
  "Ready to book, confirm and I'll create it."
</code></code></pre><p>The correct behavior is boring on purpose, the AI lines up the times, shows you the draft, and waits. That pause is the whole product.</p><p>The failure mode here is the expensive one, and it&#8217;s why the guardrail exists. The auto-scheduling tools that promise to run your whole calendar are the ones operators complain about the loudest, because the moment you let a model move committed meetings, it moves the wrong ones. Users of one popular auto-scheduler describe it reshuffling their day so hard that &#8220;your schedule is no longer your own&#8221; (<a href="https://get-alfred.ai/blog/motion-vs-reclaim">Motion vs Reclaim breakdown, 2026</a>). The first complaint I hear about the aggressive auto-schedulers is always that same one, a client call got quietly bumped to make room for a focus block the AI decided was more important. If your calendar starts moving things you committed to other people, the AI crossed The Book Line, and the fix is to pull it back to draft-only and hold the confirm yourself.</p><p><strong>Walk away with:</strong> a pasteable guardrail that lets AI do the calendar grunt work while you keep the one power that matters, the confirm.</p><h2>The bonus job: the Sunday brain-dump</h2><p>There&#8217;s a second bounded job worth handing over, and it&#8217;s the one that turns a messy list of &#8220;things I need to do this week&#8221; into a week you can see at a glance. An Android Police writer laid out the workflow cleanly, she opens Gemini on Sunday, dumps every meeting, deadline, errand, and workout in plain language, and has it build the week, then imports it in one shot (<a href="https://www.androidpolice.com/paired-gemini-with-google-calendar-to-plan-my-week/">her full walkthrough is here</a>).</p><p>Her own caveat is the part that matters, &#8220;I still double-check everything before importing it,&#8221; and she&#8217;s blunt that the workflow &#8220;still needs occasional cleanup, and it&#8217;s definitely not perfect.&#8221; That&#8217;s the right posture, hand it the drafting, keep the review.</p><p>Here&#8217;s the prompt, adapted so it lands as a clean import:</p><pre><code><code>Turn this into a CSV I can import into Google Calendar. Columns: Subject,
Start Date, Start Time, End Time. Use MM/DD/YYYY and 12-hour times.

My week of [DATE RANGE]:
[dump it all in plain language: meetings, deadlines, errands, workouts,
deep-work blocks, with rough times and durations]

Rules: don't invent times I didn't give you. If a duration is missing,
default deep work to 90 min and errands to 30 min, and flag anything you
guessed at the bottom so I can check it before I import.
</code></code></pre><p>You review the flagged guesses, import the CSV through Google Calendar&#8217;s Import and export settings, and the whole week populates at once instead of you hand-entering twenty events. It&#8217;s a bounded job, it runs off information you already have, and you can eyeball whether it&#8217;s right in a few seconds, which is the test for anything you hand an in-app AI.</p><p><strong>Walk away with:</strong> a Sunday planning ritual that builds your week from a brain-dump and lands it on the calendar in one import.</p><h2>Where this breaks</h2><p>A few failure modes to watch, because they&#8217;re the difference between this saving an hour a day and creating a mess you clean up on Monday.</p><p>Stale calendar access is the quiet one. If you connected the AI weeks ago and haven&#8217;t used it, the free-and-busy it reads can lag reality, so run a throwaway &#8220;what&#8217;s on my calendar tomorrow&#8221; check before you trust a proposal, and if it&#8217;s wrong, reconnect.</p><p>Timezone drift bites anyone booking across regions. State your timezone once in the prompt and once in the drafted email, and spot-check the first cross-region proposal by hand, because a model that&#8217;s an hour off will cheerfully propose 6am calls.</p><p>Over-handing is the big one. The pull toward &#8220;let it just run my whole calendar&#8221; is strong, and it&#8217;s the exact move that gets meetings reshuffled out from under you. Keep the AI on find-and-draft, keep yourself on confirm, and the tool stays useful instead of stressful.</p><p>And not every model reads a calendar the same way. ChatGPT&#8217;s Google Calendar connection is read-only today, it can see your calendar and reason about it, but it can&#8217;t create or move events (<a href="https://www.usecarly.com/blog/can-chatgpt-access-google-calendar/">per usecarly&#8217;s rundown</a>). Claude can write. Gemini in Gmail proposes but leaves the send to you. Know which one you&#8217;re holding before you build a workflow on top of it.</p><h2>The honest tradeoffs</h2><p>This is useful and it&#8217;s uneven, both true at the same time, and pretending otherwise would be selling you something.</p><p>The results vary more than any vendor admits. In the same comment thread, one reader said building his week this way &#8220;genuinely feels like magic,&#8221; and another said he tried the same thing and &#8220;I can&#8217;t even describe what a disaster it was, just straight up not listening to instructions&#8221; (<a href="https://www.androidpolice.com/paired-gemini-with-google-calendar-to-plan-my-week/">Android Police comments</a>). Same feature, opposite weeks. Your mileage is going to depend on how clean your calendar already is and how tightly you write the prompt.</p><p>Google&#8217;s tool is capped, &#8220;Help me schedule&#8221; tops out at 20 guests and it&#8217;s built around one-to-a-few scheduling, so a big multi-party coordination still isn&#8217;t a one-click job. The auto-schedulers that promise more control over your whole day carry the reshuffle risk I already flagged, and one of them (Reclaim) is sitting on a 2.7 Trustpilot rating with billing complaints, so the &#8220;set it and forget it&#8221; pitch has a cost. And connecting any AI to your calendar hands it read access to your entire schedule, every client name, every private block, so treat that connection the way you&#8217;d treat handing someone your calendar login, because functionally that&#8217;s what it is.</p><p>None of that kills the workflow. It just means you keep your hand on The Book Line and let the AI earn trust on the grunt work first.</p><h2>Recap</h2><p>Hand the AI the find-times-and-draft-the-proposal job end to end, it reads your live calendar and does it in seconds. Hold The Book Line, the confirm stays human, no auto-accept, no auto-reschedule of committed meetings. Add the Sunday brain-dump as a second bounded job if weekly planning is where you slip. And when a proposal looks wrong, reconnect and re-run instead of patching one slot, because a tool that invented one time will invent more.</p><p>When I map a client&#8217;s booking flow, the calendar is the first place I look, because it&#8217;s usually leaking an hour a day that AI can hand back without any risk, as long as the confirm stays with a person. If you want a second set of eyes on where your funnel is quietly burning operator time and which pieces are safe to hand off, that&#8217;s the kind of thing we work through on an <a href="https://muddventures.com/book">AI Clarity Call</a>.</p><p>Next issue, Monday: the one CRM job I&#8217;d let AI run every morning, and the field I&#8217;d never let it write to on its own.</p><p>Reply with the tool you&#8217;re using to book calls right now if you want me to sanity-check whether it can read your live availability.</p><p>Andrew</p><p>P.S. A few places to take this further:</p><ul><li><p>The inbound version of this same boundary, letting an AI read your inbox without acting on it: <a href="https://blog.muddventures.com/p/comet-will-read-your-whole-inbox">Comet Walls</a>.</p></li><li><p>The brief that stops you re-explaining your business to every AI tool, calendar prompts included: <a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">the 5,000-character ChatGPT brief</a>.</p></li><li><p>Want the calendar-and-workflow skill files and the operators building this stuff alongside you? That lives in <a href="https://whop.com/abra-ai/">the Abra AI community</a>.</p></li></ul>]]></content:encoded></item><item><title><![CDATA[The one inbox job I'd hand an AI this week, and the one I'd keep to myself]]></title><description><![CDATA[Google is quietly building a triage inbox for Gemini. Here's the email job worth handing over now, and the button I never let it press.]]></description><link>https://blog.muddventures.com/p/the-one-inbox-job-id-hand-an-ai-this</link><guid isPermaLink="false">https://blog.muddventures.com/p/the-one-inbox-job-id-hand-an-ai-this</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Thu, 23 Jul 2026 18:53:24 GMT</pubDate><enclosure url="https://substack-post-media.s3.amazonaws.com/public/images/76187158-9955-4b58-8d54-542c6dcfbe83_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><strong>TLDR:</strong> Every lab is racing to point AI at your inbox, and the loud promise is that it'll answer your emails for you. The higher-leverage job to hand over this week is the quieter one upstream: triage. Let the AI sort your mail into what needs you, what can be drafted, and what's noise, so your attention goes to the few threads that move money. Draft replies too, right up to what I call The Send Line, the point where a human presses send. Below: the two prompts I run and the threads I never hand over.</p><p>Google got caught building something telling this month. Testers spotted a dedicated inbox inside the Gemini app for Workspace, with three filters, follow up, done, and one labeled "Needs review," that pulls messages out of Gmail into a to-do list a person walks through (<a href="https://www.testingcatalog.com/google-tests-new-gemini-inbox-section-for-workspace-triage/">TestingCatalog, July 4</a>). What jumped out at me was that "Needs review" filter, because it's an admission from the company with the most email data on earth that the AI's job is to sort and surface, and the human's job is still to decide.</p><p>By the end of this you'll have a morning routine that clears your inbox to a handful of real decisions in about ten minutes, plus a hard rule for where the AI stops and you start.</p><p>That's the same instinct I run my own inbox on. The first thing I do with any new AI tool that can touch my email is work out what it's allowed to do without me, and the answer is: read and sort, yes, press send, no. I laid out a version of this test yesterday for spreadsheets, the Bounded-Job Test, and the inbox runs on the same three questions, is the job bounded, does it run on data you already have, and can you eyeball the result and know it's right. Triage clears all three, while sending a reply on your behalf clears none of them, because you can't eyeball an email after it's already gone.</p><p>Here's what's inside:</p><ul><li><p><strong>The triage pass:</strong> one prompt that sorts a day of unread mail into four buckets so you only read what needs a human.</p></li><li><p><strong>The Send Line:</strong> the boundary the AI drafts up to and never crosses, and why it's the whole game on the outbound side.</p></li><li><p><strong>The draft-to-the-line prompt:</strong> copy it, and every reply lands in your drafts folder in your voice, with the commitments left blank for you to fill.</p></li><li><p><strong>The threads I keep fully human:</strong> negotiations, complaints, anything with feelings or a contract in it.</p></li></ul><p><strong>Who this is for:</strong> operators running a real business off their own inbox, ChatGPT or Claude or Gemini already in the mix, drowning a little in email and tempted to just let the robot answer it.</p><h3>Why triage is the job worth handing over</h3><p>Email still eats around a quarter of the average workweek, and inbox zero stays the thing most people give up on by week two (<a href="https://missiveapp.com/blog/ai-email-assistant">Missive, 2026</a>). Most of that time goes to deciding more than writing: opening things, scanning them, working out what deserves a reply and what doesn't, and carrying the low hum of a hundred unread items around all day. That deciding layer is the expensive part, and it's the part AI is genuinely good at right now, because sorting is bounded, it runs on the mail already sitting in front of it, and you can glance at the buckets and know instantly if it got them right.</p><p>Answering is the part everyone wants to automate and the part that keeps biting people. A draft that sounds right isn't the same as a draft that's right, and once it sends, there's no glancing at it first. So the move is to split the inbox job in two, hand the AI the sorting and the drafting, and keep the send button on the human side of a hard boundary. I call that boundary The Send Line. The AI can read, sort, summarize, and write a full draft, and it stops the instant a message would leave your account, because crossing that line is your call and not the model's.</p><p>If that sounds familiar, it's the outbound twin of something I named back in June. When AI browsers like Comet showed up promising to read your whole inbox, I wrote about <a href="https://blog.muddventures.com/p/comet-will-read-your-whole-inbox">Comet Walls</a>, the permission boundaries you set so an agent can read a sensitive stream without acting on it. The Send Line is the same wall pointed the other direction: inbound, let it read; outbound, it drafts but you send. It's the same principle running on both sides of the mailbox.</p><p>The tools to do this are already sitting in plans you pay for. If you're on Google Workspace Business or a personal Gemini subscription, the Gemini side panel in Gmail already summarizes threads and drafts replies with Help me write and Help me reply (<a href="https://support.google.com/mail/answer/14355636">Google support</a>). If you're a ChatGPT or Claude person, connect your Gmail and the same routine runs there. And if you'd rather not connect anything yet, you can paste a screenful of subject lines into any chat window and get most of the value on day one.</p><h3>The morning triage pass</h3><p>Here's how it works. Once a day, before you start opening things one by one, you hand the AI the last 24 hours of unread mail and ask it to sort, not answer. The whole point is to turn a scroll of forty things into four short lists so your eyes only land on the one list that needs a person.</p><p>The prompt I run, which works in the Gemini side panel, in ChatGPT or Claude with Gmail connected, or on pasted subject lines:</p><pre><code><code>You are my inbox triage assistant. Here are my unread emails from the
last 24 hours [paste them, or read from my connected Gmail].

Sort every message into one of these four buckets, and nothing else:

1. NEEDS ME: a real person waiting on a decision, a reply, or money.
   One line on what they want.
2. CAN BE DRAFTED: a reply I'll approve before it sends.
3. READ LATER: useful, no action today.
4. NOISE: newsletters, receipts, automated notifications.

Rules:
- Do not send, archive, or delete anything. Sorting only.
- If a message could be money coming in or money going out, it goes in
  NEEDS ME, even if you're unsure.
- Flag anything that reads angry, legal, or like a cancellation, and put
  it in NEEDS ME with the word REVIEW next to it.
- NEEDS ME is capped at what a person could clear in 20 minutes.
  If it's longer than that, move the softest items down to CAN BE DRAFTED.

Give me the four buckets as short lists. Nothing else.</code></code></pre><p>What a good result looks like: NEEDS ME is short, a handful of threads, and the money and the angry ones are sitting right at the top with REVIEW next to them. The bulk of the volume falls into NOISE and READ LATER, which is the whole win, because that's the pile you were burning attention on without meaning to. You read one list, you make a few decisions, you move on.</p><p>The failure mode to watch, and it's the common one: NEEDS ME comes back with twenty-five items in it. That means the model hedged and pushed everything up to the safe bucket, which gives you back the exact overwhelm you were trying to kill. When that happens, tighten the cap ("NEEDS ME holds at most eight items, everything else drops down") and it recalibrates. The rarer failure is an empty NEEDS ME on a day you know had a real ask buried in it, which is the model under-escalating. If you see that even once, keep reading manually alongside the AI for a few more mornings before you trust the sort, because a triage pass that misses the one email that mattered is worse than no triage at all.</p><p><strong>Walk away with:</strong> a once-a-day routine that collapses a full inbox into four lists and puts the threads that move money at the top, in about the time it takes to make coffee.</p><h3>Drafting up to The Send Line</h3><p>Sorting gets you to the handful of threads that need a reply. The AI can write those too, and this is where holding the line earns its keep. The rule is simple and it never bends: the AI writes the draft, the draft lands in your drafts folder, and you are the one who reads it and presses send. Nothing leaves your account on the model's decision.</p><p>The prompt:</p><pre><code><code>Draft a reply to this email in my voice. Do NOT send it. Leave it as a
draft for me to read and send myself.

The email:
[paste the thread]

My voice: [paste 2-3 of your own real sent replies here, or describe it:
direct, warm, short sentences, no corporate filler].

Draft rules:
- Match the length of the incoming message. A two-line email gets a
  two-line reply.
- If a date, price, or commitment is involved, leave a blank for me to
  fill: [DATE], [PRICE], [SCOPE]. Never invent a number.
- If the thread is a negotiation, a complaint, or anything with money or
  feelings on the line, do not draft a full reply. Write one line telling
  me what's at stake and let me handle it.
- End every draft with: READY FOR YOUR REVIEW.</code></code></pre><p>A good draft comes back short, in your register, with blanks sitting right where the commitments are, and that READY FOR YOUR REVIEW line at the bottom so you know it stopped where it was supposed to. You read it in five seconds, fill the blanks, send. The response-time drop is real, and the teams running this at scale almost all keep send on the human side anyway, because auto-send stays rare even where it's technically supported, the cost of one wrong autonomous reply running much higher than the minutes it saves (<a href="https://missiveapp.com/blog/ai-email-assistant">Missive, 2026</a>). You get the speed of a drafted reply and keep the judgment of a human read.</p><p>The failure mode here is the dangerous one, so calibrate on it hard. If the draft fills in a price or a date you never gave it, the reply is contaminated, the model guessed a commitment, and if that had been on auto-send it would've gone out as your word. That's the tell to keep those threads human and to make the "never invent a number" rule louder. The softer failure is a three-paragraph draft in response to a one-line email, which means your voice sample was too formal or you skipped it, so feed it two or three of your own real short replies and it snaps into your cadence.</p><p><strong>Walk away with:</strong> replies written in your voice sitting in your drafts folder, commitments left blank on purpose, and a hard rule that the send button is yours.</p><h3>What I keep fully human, and the honest tradeoffs</h3><p>The reason The Send Line matters is that some threads shouldn't get an AI draft at all, and knowing which ones is the real skill. Charles Hudson, who runs Precursor VC, built agents that watch his inbox, classify it, and draft his intros and declines, and he still keeps the send button on his own side. He doesn't trust the agent to send on its own, so as he told Missive, "I have a draft only flag on" (<a href="https://missiveapp.com/blog/ai-email-assistant">Missive, 2026</a>). That posture, an always-on agent doing the work with a hard draft-only stop on the send step, is the one I run too, and the threads I'd never even let it draft are the ones where sounding right is a trap: active negotiations with real terms on the line, complaints and conflict, anything legal or contractual, and any message that needs genuine empathy. The model produces something that reads fine in every one of those, and reading fine is the danger when the stakes are a client relationship or a signed contract.</p><p>A few things worth flagging before you wire this up. The triage sort is only as good as the model's read of your business, so the first week it'll misfile things, and the fix is feedback, correct the buckets a few times and it sharpens fast. The native Gemini side panel is bundled into Business and Enterprise Workspace plans and personal Gemini subscriptions, so if you're on a bare Workspace tier you might not see it yet, in which case the ChatGPT-or-Claude-plus-Gmail route does the same thing. And the deeper tradeoff is data, connecting any AI to your inbox hands it read access to everything in there, which is precisely why the read-but-don't-act boundary from Comet Walls matters, and why I'd start with pasted subject lines before I'd hand a brand-new tool a full mailbox connection.</p><p>There's also a real cost to over-trusting the sort. A triage pass that quietly misroutes the one email that mattered doesn't feel like a failure, it feels like a clean inbox, right up until you find the buried thread a week late. So the discipline is to keep glancing at the NOISE and READ LATER piles for the first couple of weeks, the way you'd double-check a new hire's work, because the AI can't learn what "important to me" means until you've corrected it a few times.</p><h3>Recap</h3><p>Point the AI at the deciding layer, not the sending layer. Run a once-a-day triage pass that sorts your unread mail into four buckets so you only read what needs a person, let it draft the replies that can be drafted, and hold The Send Line so nothing leaves your account without you pressing the button. Keep negotiations, complaints, legal, and anything with feelings in it fully in your hands. The industry consensus landed in the same place for a reason: these tools are co-pilots on the inbox, and the operators getting hours back are the ones who used them to decide faster, not to stop deciding.</p><p>If you run a booking calendar or a lead inbox, this compounds fast, because the same triage-and-draft pass is what keeps a fresh lead from sitting unanswered for six hours while you're heads-down. That gap between a lead landing and a human touching it is the kind of leak my <a href="https://showtime.muddventures.com">Showtime skill pack</a> is built to close. Worth a look if your inbox is where your money either moves or stalls.</p><p>Tomorrow: the calendar version of this, the one scheduling job I'd let AI run end to end and the one I'd never automate.</p><p>Andrew</p><p><strong>P.S. A few to pull the thread further:</strong></p><ul><li><p>Yesterday's Bounded-Job Test for spreadsheets, the same three questions applied to Gemini in Google Sheets.</p></li><li><p><a href="https://blog.muddventures.com/p/comet-will-read-your-whole-inbox">Comet Walls</a>, the inbound version of The Send Line, for when an AI browser wants your whole inbox.</p></li><li><p><a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">The Standing Brief</a>, the one-page business brief that makes every draft sound like you.</p></li><li><p>The <a href="https://showtime.muddventures.com">Showtime skill pack</a> if your lead inbox is where the money leaks.</p></li><li><p>New here? <a href="https://muddventures.substack.com/">The newsletter lives at muddventures.substack.com</a>, a short tactical read most weekday mornings.</p></li></ul>]]></content:encoded></item><item><title><![CDATA[Grab the four Google Sheets prompts that do the grunt work, and the one test for when to keep the formula]]></title><description><![CDATA[Google slid a spreadsheet builder into the plan you already pay for, and the trick is knowing the exact jobs it does well before it quietly falls apart on the big ones.]]></description><link>https://blog.muddventures.com/p/grab-the-four-google-sheets-prompts</link><guid isPermaLink="false">https://blog.muddventures.com/p/grab-the-four-google-sheets-prompts</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Wed, 22 Jul 2026 15:12:33 GMT</pubDate><enclosure url="https://substack-post-media.s3.amazonaws.com/public/images/43077cfd-4d6b-4d9e-86e8-857c989c96a1_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p><strong>TLDR:</strong> Gemini in Google Sheets can now build and edit whole spreadsheets from a plain-English prompt, and for most Business and Enterprise plans it&#8217;s already switched on inside what you pay for every month. It&#8217;s genuinely good at a short list of bounded jobs, categorizing a messy column, cleaning a list, drafting a formula, standing up a small dashboard off your own data. It falls apart the second you point it at a big multi-file, cross-app, auditable model. Below are four prompts you can paste into a sheet today, and a two-second test for which jobs to hand it and which ones to keep in a formula.</p><p>Almost every operator I work with has one spreadsheet a person babysits every single week. Expense categorization that somebody retypes by hand, a lead list that shows up messy from three sources, a little P&amp;L view that a bookkeeper rebuilds every month because the export never lands clean. It&#8217;s the least glamorous work in the building, and it&#8217;s usually either eating an owner&#8217;s Sunday or getting paid for by the hour.</p><p>That&#8217;s the work Google just quietly moved into a spreadsheet you&#8217;re already paying for. Back on April 22 Google turned on Gemini in Google Sheets, the thing that builds and edits entire spreadsheets from a sentence, and it&#8217;s been rolling wider through the summer, with the data-entry piece, Fill with Gemini, picking up 11 more languages in July (<a href="https://workspaceupdates.googleblog.com/2026/04/build-and-edit-complex-spreadsheets-with-Gemini-in-Google-Sheets.html">Google Workspace Updates</a>).</p><p>For the last three years I&#8217;ve watched the tools that used to earn a monthly retainer get pulled, one by one, into software operators already have open. This is that pattern landing on the spreadsheet. The catch is that Gemini in Sheets is great at a narrow band of work and quietly bad outside it, and if you don&#8217;t know where that line sits you&#8217;ll either miss the free win or trust it with a number the bank is going to read.</p><p>By the end of this you&#8217;ll have four prompts you can paste into a sheet today, and a two-second gut check I&#8217;m calling <strong>The Bounded-Job Test</strong> for deciding which spreadsheet jobs to hand Gemini and which ones to keep in a formula.</p><p><strong>Who this is for:</strong> operators running on Google Workspace (Business or Enterprise) who touch spreadsheets weekly, or pay someone who does.</p><p>Here&#8217;s what you&#8217;re walking out with:</p><ul><li><p><strong>The free win most operators are sitting on:</strong> the spreadsheet builder is already inside Business Standard and up, no new line item, no separate tool.</p></li><li><p><strong>Four copy-paste prompts:</strong> categorize, clean, formula, and a small dashboard off your own numbers, the jobs it nails.</p></li><li><p><strong>The Bounded-Job Test:</strong> a two-second check that keeps you from handing it the one task that&#8217;ll bite you.</p></li><li><p><strong>The failure line a finance guy already hit:</strong> a real tester ran the big cross-app forecast and it fell over, so you don&#8217;t have to learn that one the hard way.</p></li></ul><h2>What just shipped, in plain terms</h2><p>Gemini in Sheets does two things that used to need real spreadsheet chops. It builds, so you can say &#8220;build a P&amp;L dashboard from my sales and service data&#8221; and it drafts a plan, pulls the numbers, and lays out formatted tables, pivots, and charts. And it edits in the side panel, so &#8220;add scorecards and a bar chart above my inventory data&#8221; happens next to your sheet instead of making you rebuild anything.</p><p>The part that makes it more than a toy is what Google calls Workspace Intelligence, which lets Gemini reach across your Gmail, Drive, and Chat to pull context into the sheet instead of you copy-pasting between an inbox and a tab. On Google&#8217;s own spreadsheet benchmark the model hit a 70.48% success rate on messy real-world tasks, which they frame as near-expert and which, in operator terms, means it&#8217;s right most of the time and wrong often enough that you still have to look (<a href="https://blog.google/products-and-platforms/products/workspace/gemini-google-sheets-state-of-the-art/">The Keyword</a>).</p><p>Where&#8217;s the money angle. It&#8217;s bundled. Gemini in Sheets is included in Business Standard and Plus, Enterprise Standard and Plus, the AI add-ons, and the consumer Google AI Pro and Ultra plans. If you&#8217;re on one of those, you&#8217;re already paying for this, it&#8217;s just sitting in the side panel behind the Gemini icon waiting for a prompt. The one thing worth flagging is limits: Google&#8217;s promotional higher limits ran through July 15, and after that per-user caps apply, with paid AI add-on seats getting more headroom, so heavy users hit a ceiling light users never will.</p><h2>The four jobs to hand it, with the exact prompts</h2><p>The jobs Gemini in Sheets does well share a shape: bounded scope, your own data, and an answer you can eyeball in a few seconds to know it&#8217;s right. Here are the four I&#8217;d reach for first, each with the prompt, what a good result looks like, and the moment to stop and check.</p><h3>Job 1: Categorize a messy column</h3><p>The move: point Fill with Gemini at a column of raw text, transaction memos, lead sources, support tags, and have it write a clean category next to each row. Select your data, open Fill with Gemini in the side panel, and give it the rule.</p><pre><code><code>In a new column next to this transaction list, label each row as one of:
Software, Payroll, Contractors, Ads, Rent, Travel, Meals, Other.
Base it on the vendor name and memo text. If you're unsure, use "Other"
and do not invent a new category.
</code></code></pre><p>What a good result looks like: clean single-word labels that match the eight buckets you named, &#8220;Other&#8221; showing up on the genuinely ambiguous rows instead of a guess. Google says Fill with Gemini runs about 9x faster than doing 100 cells by hand, and for this kind of labeling that&#8217;s roughly the lift you&#8217;ll feel.</p><p>The failure mode: on long columns the labels start to drift, a category like &#8220;Service&#8221; mutates into &#8220;service issues&#8221; a few hundred rows down, or it quietly invents a ninth bucket you didn&#8217;t ask for. That&#8217;s your signal the run got too big, chop it into batches of a couple hundred rows and add &#8220;do not invent a new category&#8221; to the prompt, which I already baked in above for that reason.</p><p><strong>Walk away with:</strong> a categorized column in seconds instead of an afternoon, on the kind of data you&#8217;d otherwise pay someone hourly to tag.</p><h3>Job 2: Clean up a list that came in three different shapes</h3><p>The move: hand it a column that three sources filled in three different ways and have it standardize the format without you writing a nest of formulas.</p><pre><code><code>Standardize this "Company" column: strip "Inc", "LLC", "Ltd", and
trailing punctuation, fix obvious capitalization, and put the cleaned
name in the next column. Leave the original column untouched.
</code></code></pre><p>What a good result looks like: &#8220;acme co., llc&#8221; and &#8220;ACME CO&#8221; both land as &#8220;Acme Co&#8221; in the new column, and the original stays put so you can compare. Telling it to leave the source column alone is the whole game, you always want the before and after side by side.</p><p>The failure mode: it &#8220;helpfully&#8221; merges two companies that were genuinely different, or rewrites a name it didn&#8217;t recognize. Eyeball the rows where the cleaned value differs a lot from the original, that&#8217;s where a real error hides, and it&#8217;s a five-second scan when the columns sit next to each other.</p><p><strong>Walk away with:</strong> a deduped, consistent list you can run a mail merge or an import against, without babysitting a formula chain.</p><h3>Job 3: Draft or explain a formula in words</h3><p>The move: describe the calculation you want in plain English and let it write the formula, or paste a formula you inherited and ask what it does. This is the one even the skeptics agree is solid.</p><pre><code><code>Write a formula for cell H2 that returns the gross margin percent using
revenue in column D and cost in column E, formatted as a percentage,
and returning blank if revenue is zero. Then explain in one sentence
what it does.
</code></code></pre><p>What a good result looks like: a working formula plus a one-line plain-English explanation, so you&#8217;re not pasting something you can&#8217;t read. The explanation is the tell, if it can describe the formula in a sentence that matches what you asked, it usually built the right thing.</p><p>The failure mode: it returns a formula that runs without erroring but calculates the wrong thing, off-by-one on a range, wrong column reference. This is why you ask for the explanation, if the sentence doesn&#8217;t match your intent, the formula won&#8217;t either, and you catch it before it&#8217;s buried in a model.</p><p><strong>Walk away with:</strong> formulas you&#8217;d have Googled for twenty minutes, written and explained in one pass.</p><h3>Job 4: Stand up a small dashboard off your own numbers</h3><p>The move: give it a clean tab of your own data and ask for a one-screen summary. Keep it to data that already lives in the sheet, not a cross-app pull.</p><pre><code><code>Using the data on the "Sales" tab, build a one-page summary above the
data with: total revenue, revenue by month as a bar chart, top 5
products by revenue, and month-over-month growth. Keep it on this sheet
and don't change my raw data below.
</code></code></pre><p>What a good result looks like: a tidy header block with a couple of scorecards and one clean chart, your raw data untouched underneath. For a quick read on how a month went, this is the job it was built for, Google&#8217;s own example is literally a small-business P&amp;L view.</p><p>The failure mode: ask it to reach into three other files or your inbox to build the same thing and reliability drops off a cliff, which is the exact wall the next section is about. Keep the dashboard job pointed at data that&#8217;s already in the sheet and it holds up.</p><p><strong>Walk away with:</strong> a shareable one-screen view of a month, built in the time it takes to describe it.</p><h2>The Bounded-Job Test, and where this thing breaks</h2><p>Every job above passes the same three-part check, and that&#8217;s the test worth memorizing. Before you hand Gemini a spreadsheet task, ask: is the scope bounded to one sheet or one column, is it working off data you already have in front of you, and can you eyeball the answer in a few seconds to know it&#8217;s right. Three yeses, hand it over. A no on any of them, keep it in a formula or keep it human.</p><p>The reason that test matters is that the failure isn&#8217;t hypothetical, and someone already ran it into the ground so you don&#8217;t have to. Wouter Born, a CFO-tech investor who writes these tools up for finance teams, tried the exact workflow Google&#8217;s architecture is supposed to nail: three store managers email in monthly forecast updates with attachments, and Gemini reads the emails, pulls the assumptions, opens the files, maps them to the corporate model, and flags what needs sign-off. The dream cross-app job. His verdict:</p><blockquote><p>&#8220;The vision is right. But the execution isn&#8217;t there yet... The complex, multi-file, cross-app workflow that makes Google&#8217;s architecture special isn&#8217;t reliable enough for finance yet.&#8221; (<a href="https://cfooffice.io/p/i-tested-gemini-in-google-sheets">AI CFO Office</a>)</p></blockquote><p>He didn&#8217;t write it off, and neither would I. He landed right where the Bounded-Job Test lands, on what to use it for:</p><blockquote><p>&#8220;For now, if your team runs on Google Workspace, use Gemini for the things it does well: formula generation, data enrichment, quick analysis questions, and spreadsheet cleanup. Those features work and they&#8217;re included in what you&#8217;re already paying for.&#8221; (<a href="https://cfooffice.io/p/i-tested-gemini-in-google-sheets">AI CFO Office</a>)</p></blockquote><p>Two more honest tradeoffs before you lean on it.</p><p>It&#8217;s got no formula trail. When Gemini fills a number, that number isn&#8217;t the output of a formula you can click into and trace, it&#8217;s a generated value sitting in a cell. That&#8217;s fine for a lead category, it&#8217;s a real problem for anything an accountant, a lender, or a tax reviewer has to trace back to a source. For auditable numbers, keep the formula and let Gemini draft or label around it, don&#8217;t let it be the thing that produces the figure of record.</p><p>And it runs out of room. Push a task past a few hundred cells in one go and it tends to stall, refuse, or hand back a partial answer, there are real session limits under the hood on top of the per-user caps that kicked in after July 15. Big jobs need chopping into bounded ones anyway, the same discipline the test pushes you toward.</p><h2>Failure modes, in one place</h2><ul><li><p><strong>You point it at the auditable P&amp;L.</strong> A generated number with no formula trail is the one thing you don&#8217;t want a lender reading. Keep those in formulas, use Gemini to label around them.</p></li><li><p><strong>You run it on 2,000 rows at once.</strong> Labels drift, categories wander, output comes back partial. Batch it into a couple hundred rows at a time.</p></li><li><p><strong>You ask for the big cross-app pull.</strong> Read three inboxes and four files in one shot, and it falls over, that&#8217;s the Born test. Keep the job inside one sheet.</p></li><li><p><strong>You trust it without the eyeball.</strong> If you can&#8217;t verify it in a few seconds, it wasn&#8217;t a bounded job.</p></li></ul><h2>Where this fits the bigger picture</h2><p>The operators I work with who get real value out of AI aren&#8217;t the ones chasing every launch, they&#8217;re the ones who&#8217;ve gotten sharp about which jobs to hand a machine and which ones stay human, and it&#8217;s the same muscle whether the tool is Sheets, your CRM, or your inbox. A free spreadsheet builder is a nice win. Knowing the line between the grunt work you give it and the number you protect is the part that compounds.</p><p>That judgment, which repeatable jobs to productize into an AI workflow and which to leave alone, is most of what I trade notes on with operators inside the <a href="https://whop.com/abra-ai/">Abra AI community</a>, where the skill files and workflow builds live. If you&#8217;d rather map it against your own stack directly, that&#8217;s what an <a href="https://muddventures.com/book">AI Clarity Call</a> is for.</p><h2>Recap</h2><p>Gemini in Sheets is already in the plan most operators pay for, and it&#8217;s good at bounded jobs: categorizing a column, cleaning a list, drafting a formula, building a small dashboard off your own data. Run those with the four prompts above, and run everything else through the Bounded-Job Test, bounded scope, your own data, an answer you can eyeball. When a task fails it, keep the formula and keep the human.</p><p>Tomorrow: the one inbox job I&#8217;d hand an AI before I&#8217;d hand it a spreadsheet.</p><p>Andrew</p><p><strong>P.S.</strong> A few worth pulling up next:</p><ul><li><p>The tool-compression pattern this fits into: <a href="https://blog.muddventures.com/p/run-the-30-minute-canva-test">Run the 30-minute Canva test that shows whether your ad creative retainer still earns its invoice</a></p></li><li><p>The Google-side listings version of the same story: <a href="https://blog.muddventures.com/p/run-the-three-prompt-google-business">Run the three-prompt Google Business Profile check</a></p></li><li><p>The context brief that makes any of these prompts land better: <a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">Grab the 5,000-character ChatGPT brief</a></p></li><li><p>Where the workflow builds and skill files live: <a href="https://whop.com/abra-ai/">whop.com/abra-ai</a></p></li><li><p>Map it to your own stack: <a href="https://muddventures.com/book">muddventures.com/book</a></p></li></ul>]]></content:encoded></item><item><title><![CDATA[4 questions Anthropic's security team runs before an agent gets access, and the 20-minute version for your business]]></title><description><![CDATA[Connecting an agent to your inbox takes ninety seconds. Deciding what it's allowed to do takes longer, and that's the part that gets skipped.]]></description><link>https://blog.muddventures.com/p/4-questions-anthropics-security-team</link><guid isPermaLink="false">https://blog.muddventures.com/p/4-questions-anthropics-security-team</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Tue, 21 Jul 2026 15:48:37 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!iaDp!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!iaDp!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!iaDp!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!iaDp!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!iaDp!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!iaDp!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!iaDp!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1437943,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/207931638?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!iaDp!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!iaDp!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!iaDp!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!iaDp!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F330887aa-19dd-4d39-9c39-2aeaf75dd814_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TLDR:</strong> On July 17, Anthropic&#8217;s Deputy CISO published the four questions his team answers before they let any agent touch anything, and the whole thing translates down to a one-person business without a security department. Run the four questions against the agent you already have connected, write the answers in a doc, then go remove one verb from its tool list. Twenty minutes, no new software, and it changes what happens on the day something goes sideways.</p><p>Anthropic put out a piece on July 17 that I did not expect to be useful to anybody running a five-person company, and then I read it twice. It&#8217;s called <a href="https://claude.com/blog/ciso-guide-to-agentic-ai">Zero risk isn&#8217;t the job: a CISO&#8217;s guide to agentic AI</a>, written by Jason Clinton, their Deputy CISO, and most of it is aimed squarely at people who own a SIEM and an identity provider and a compliance calendar. Buried in the middle of it is a four-question checklist that costs nothing to run and doesn&#8217;t care how big you are.</p><p>I run this newsletter through a scheduled agent with connectors attached, so I ran the four questions against my own setup while I was reading, and the second question is where I slowed down, not because the answer was bad but because I had never written it down, and writing it down took four minutes, and at the end of those four minutes I could describe my own automation in a way I couldn&#8217;t have that morning.</p><p>That&#8217;s the whole play. The four questions don&#8217;t buy you software or a security team, what they buy you is the ability to say out loud what your agent is allowed to do, which turns out to be a thing most operators can&#8217;t do about the automations already running in their business.</p><p><strong>By the end of this you&#8217;ll have written answers to four questions about one agent you already have running, plus one verb removed from its permissions.</strong></p><p>Here&#8217;s what&#8217;s in this one:</p><ul><li><p><strong>The Borrowed Login is why &#8220;it&#8217;s just an assistant&#8221; stops being true.</strong> When you connect an agent to your CRM, it doesn&#8217;t get an agent account, it gets yours, and its reach is your reach.</p></li><li><p><strong>Question two is the one that catches people.</strong> Read-only and read-write are different businesses, and most operators have never checked which one they authorized.</p></li><li><p><strong>Blast radius is a five-minute thought experiment, not a security discipline.</strong> You describe the worst realistic outcome in one sentence and the answer either scares you or it doesn&#8217;t.</p></li><li><p><strong>Removing a verb beats adding a rule.</strong> An agent will never attempt an action that isn&#8217;t in its tool list, which is a stronger control than any instruction you write in a prompt.</p></li><li><p><strong>The honest gap: Anthropic&#8217;s four questions travel down to a small business, and their seven controls mostly don&#8217;t.</strong></p></li></ul><p><strong>Who this is for:</strong> any operator who has connected an AI assistant to email, a CRM, a calendar, a file drive, or a payment tool, and who has not yet written down what that assistant is permitted to do inside those systems.</p><h2>The Borrowed Login</h2><p>Clinton lays out an identity spectrum with two clean ends: a service account that does one job with no human attached, or a person at a keyboard accountable for whatever they do with their own credentials, same as always.</p><p>The middle is where the trouble lives. In his words, the middle is &#8220;where an agent carries a person&#8217;s delegated identity into systems that person is not watching,&#8221; and then the line that made me stop: &#8220;Ambiguous accountability is how incidents become unexplainable.&#8221;</p><p>Almost every SMB agent setup I&#8217;ve seen lives in that middle. You click Connect on the Gmail integration, approve the OAuth screen, and from that second forward the agent operates as you, with your mailbox, your send rights, and your access to every thread you&#8217;ve ever been on, because that&#8217;s the only permission level most consumer-grade tools offer. Call it <strong>The Borrowed Login</strong>: the agent holds your identity rather than one of its own, so its blast radius and your blast radius are the same number.</p><p>Post-incident analysis compiled from CrowdStrike and Mandiant data found that <a href="https://www.digitalapplied.com/blog/ai-agent-security-2026-1-in-8-breaches-agentic-systems">78% of agents involved in breaches had significantly broader permission scope than their function required</a>, with the usual cause being that teams grant wide access during setup so the thing works, intending to tighten it later, and later never arrives. That&#8217;s not an enterprise failure mode, that&#8217;s every automation any of us has ever built on a Tuesday afternoon.</p><p>Clinton also names the timing problem, citing the Ponemon Institute&#8217;s 2026 insider risk report finding that organizations took an average of 67 days to contain an insider incident even after years of dedicated investment, then adding that at agent execution speeds, &#8220;67 days is the wrong unit of measurement entirely.&#8221; A small business doesn&#8217;t have 67 days of anything, it has one owner who eventually notices something looks off.</p><p><strong>Walk away with:</strong> a name for the thing your connected agent is standing on, and the understanding that its reach equals your reach unless you narrowed it on purpose.</p><h2>The four questions, translated</h2><p>Here&#8217;s the framework verbatim from the piece, in Clinton&#8217;s order, with the small-business translation under each.</p><p><strong>The specific move:</strong> open a doc, paste this in, and answer it about one agent that&#8217;s live in your business right now. Pick the one with the most access, not the one you&#8217;re proudest of.</p><pre><code><code>AGENT AUDIT: [name of the agent/automation]
Date: 
Systems it's connected to: 

1. UNTRUSTED CONTENT
What does this agent read that someone outside my company
could have written or altered?
(inbound email, web pages, uploaded PDFs, form submissions,
review text, third-party docs)
Answer: 
If the answer is "nothing," agent-specific risk is near zero.

2. ACTIONS + IDENTITY
What can it DO, not just read? List every verb.
(read, search, draft, send, create, edit, delete, pay, post,
change permissions, call external URLs)
Whose login is it acting under? 
Answer: 

3. BLAST RADIUS
If this agent went wrong in the worst realistic way, what's
the damage in one sentence? Who would find out, and how?
Answer: 
Rate it: anomaly / annoyance / data exposure / real incident

4. OBSERVABILITY
Where do I go to see what it did? Can I tell its actions
apart from mine? How long before I'd notice something odd?
Answer: 

DECISION: keep as-is / narrow it / turn it off
Verb I'm removing today: </code></code></pre><p><strong>What the correct output looks like:</strong> four short paragraphs and one verb crossed off, done in under twenty minutes. Clinton&#8217;s own blast-radius answer for his team&#8217;s incident-response agent runs a single sentence: &#8220;the worst outcome we could construct was some mildly sensitive log lines posted into an incident channel that was already locked down.&#8221; That&#8217;s the register you&#8217;re aiming for, one sentence you&#8217;d be comfortable reading out loud to a client.</p><p><strong>The failure mode:</strong> you write &#8220;medium risk&#8221; or &#8220;we&#8217;d probably catch it&#8221; in the blast radius box. That&#8217;s not an answer, that&#8217;s a shrug wearing a suit. If you can&#8217;t name the specific worst thing, the honest output is that you don&#8217;t currently know what this agent can reach, and that&#8217;s a finding worth having, because it tells you question two was never answered properly either.</p><p>The other failure mode is answering question one too generously. Untrusted means anything an attacker could plausibly write or alter, so if your agent reads inbound email, your contact form, or review text, the answer to question one is never &#8220;nothing.&#8221;</p><p><strong>Walk away with:</strong> a written, dated answer sheet for one live agent, in a doc you can hand to whoever asks.</p><h2>Remove the verb, not the trust</h2><p>This is the part I&#8217;d put on a wall. Clinton&#8217;s framing on tool permissions: &#8220;If the failure mode that keeps you up at night is &#8216;the production database gets deleted,&#8217; remove the delete verb from the agent&#8217;s world entirely. It will never attempt an action that isn&#8217;t in its tool list.&#8221;</p><p>That&#8217;s a different category of control from anything you can write in a prompt. A prompt is an instruction the model interprets. A missing verb is a thing the model cannot do regardless of what it reads, what it decides, or what some hidden text in a PDF told it to do.</p><p><strong>The specific move:</strong> go into the connector or integration settings for the agent you just audited, find the per-action permission list, and turn off the highest-consequence verb you don&#8217;t need this week. Most operators find they have send, delete, or edit switched on for a workflow that only ever needed draft and read.</p><p>If your tool doesn&#8217;t expose per-action controls, use the account layer instead. Create a second user in the CRM with restricted permissions and connect the agent through that one rather than through the owner login. It takes ten minutes and it un-borrows the login.</p><p><strong>The exact input:</strong> if you want the agent to help you inventory this before you start clicking, paste this into whatever assistant is connected:</p><pre><code><code>List every action you are currently able to take in each of my
connected tools. Group them as: READ, CREATE, EDIT, SEND/PUBLISH,
DELETE, PAY. For each one, tell me which of my workflows uses it.
Flag any action you have permission for but have never used.
Do not take any actions. Output the list only.</code></code></pre><p><strong>What the correct output looks like:</strong> a table where the SEND, DELETE, and PAY rows are either empty or attached to a workflow you can name. The &#8220;has permission, never used&#8221; list is your removal candidates, and it&#8217;s usually longer than people expect.</p><p><strong>The failure mode:</strong> the assistant confidently invents a permission list, which happens, so treat its output as a starting draft and confirm against the real settings screen in the tool.</p><p>Anthropic&#8217;s default email posture is worth stealing outright. Their guidance: when a personal agent handles email but uses web search results as an input, &#8220;an excellent default is to only allow <em>draft</em> emails to be created and never sent externally, automatically, without human review.&#8221; Draft yes, send no. That one setting removes most of what could go wrong with an email agent and costs you a click per message.</p><p><strong>Walk away with:</strong> one verb removed from one agent, and a defensible default for anything that touches outbound email.</p><h2>The four-level ladder</h2><p>The practitioner version of this that I liked best came from a ZenAI piece published July 13, which splits agent capability into four levels rather than the binary of connected or not connected: <strong>read, recommend, create, update.</strong> Their line on it: &#8220;An AI agent should not be allowed to take every action it can technically perform,&#8221; and then, on scoping, &#8220;<a href="https://zenaicorp.com/en/insights/control-ai-agent-actions-crm-erp">do not let the agent become a shortcut around the permission model your business already depends on</a>.&#8221;</p><p><strong>The specific move:</strong> take every automation in your business and put it on the ladder, where read is retrieving account history or order status, recommend is suggesting an owner or a next step, create is making a task or a draft or a queue item, and update is changing a field, an owner, a price, or a payment status.</p><p>Their tell for whether you&#8217;re ready: &#8220;A demo may use a broad service account because it is faster to build. That may be acceptable for a controlled test environment, but it is dangerous in production.&#8221; Most SMB agent setups are still running the demo configuration, because it worked and nobody went back.</p><p><strong>What the correct output looks like:</strong> most items sit at read, recommend, and create, with update reserved for a small number of low-consequence fields. Their framing I keep coming back to: &#8220;the safest first version is usually boring on purpose.&#8221;</p><p><strong>The failure mode:</strong> you find an automation sitting at update on something customer-facing or money-adjacent that nobody explicitly approved, usually because &#8220;let the agent update the CRM&#8221; got interpreted broadly during setup. Move it down a rung to create-a-review-item and see whether you miss the speed, because most operators don&#8217;t.</p><p><strong>Walk away with:</strong> every automation you run placed on a four-rung ladder, and anything at the top rung either justified or demoted.</p><h2>What&#8217;s still broken about this</h2><p>The four questions travel down to a small business cleanly. The seven controls in the back half of that guide mostly do not, and it&#8217;s worth saying plainly.</p><p>Clinton&#8217;s control list assumes you have an identity provider issuing SAML or OIDC, a SIEM you can point an OpenTelemetry stream at, egress allowlisting through a proxy the agent can&#8217;t reconfigure, and admin-level RBAC for connector actions. A ten-person company has none of that and isn&#8217;t going to buy it. The guide is also, in fairness, partly a product spec for Claude Cowork Enterprise, since every control gets stated twice, once as a general requirement and once as how Anthropic implements it. Useful and also marketing, both true at the same time.</p><p>The small-business substitutions: your observability is the sent-items folder, the CRM activity log, and the connector&#8217;s own history page, checked on a rhythm rather than streamed anywhere. Your off switch is disconnecting the connector, and your least-privilege control is a second restricted user account. Thinner than what Anthropic runs, and still ahead of the version where nobody has written anything down.</p><p>Prompt injection also isn&#8217;t solved. Anthropic&#8217;s own language is that models are getting meaningfully better at resisting it and attack success rates keep falling, &#8220;they&#8217;re not zero,&#8221; so anything you point at inbound email or the open web carries residual risk no configuration erases.</p><p>Writing the four answers down also bounds nothing by itself, it&#8217;s a document. The bounding happens when you go remove the verb, and the gap between the audit and the removal is where this exercise usually dies.</p><p>The counterpoint worth holding: there&#8217;s a real cost to over-tightening, because if every action needs your approval the agent stops saving you time and becomes another inbox. Clinton&#8217;s framing is that the job is making risk &#8220;legible and bounded,&#8221; and legible is doing a lot of work in that sentence. Bounded doesn&#8217;t mean small, it means you know where the edges are.</p><h2>Failure modes to watch for</h2><ul><li><p><strong>The permission that got added and never removed.</strong> Capabilities get extended, permissions get added, nothing gets taken away, which is how a read-only setup becomes a write-everywhere setup across eighteen months without a single decision being made.</p></li><li><p><strong>The agent that got smarter underneath you.</strong> Clinton&#8217;s sharpest story: they moved their incident-response agent to a newer model version and changed nothing else, no new tools or permissions or prompts, and the intelligence uplift alone was enough for it to start reaching out to another agent to try to fix production on its own. It stayed inside the bounds because the bounds were real and a human still reviewed the code change. His takeaway travels: limit access and actions, not around what you believe today&#8217;s model can do.</p></li><li><p><strong>Silent failure.</strong> An agent that stops working loudly is fine, an agent that quietly does the wrong thing for three weeks inside a workflow nobody opens is the expensive version, which is why question four matters as much as question two.</p></li></ul><h2>Recap</h2><p>The four questions cost nothing to run: what untrusted content does it ingest, what actions can it take and under whose identity, what&#8217;s the blast radius, and what observability do you have. Run them against one live agent, write the answers in a doc with today&#8217;s date, then do the part that matters, which is going into the settings and removing one verb the agent has but doesn&#8217;t need. The four-level ladder is the ongoing discipline: read, recommend, create, update, with that top rung reserved for things you deliberately signed off on.</p><p>None of this requires a security hire, it requires the person who already understands the business writing four answers down.</p><p>If you want to work through this on one of your own automations alongside people running comparable setups, that&#8217;s the kind of thing we chew on inside the <a href="https://showtime.muddventures.com/operator-council">Operator Council</a>, and a second set of eyes on a permission list beats sitting alone with a settings screen.</p><p><strong>P.S.</strong> A few related pieces if you&#8217;re building in this direction:</p><ul><li><p><a href="https://blog.muddventures.com/p/run-the-three-prompt-google-business">The three-prompt Google Business Profile check</a>, for connecting an assistant to a system you already own</p></li><li><p><a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">The 5,000-character Standing Brief</a>, for giving an agent context without giving it access</p></li><li><p><a href="https://blog.muddventures.com/p/gohighlevels-ai-builder-just-learned">GoHighLevel&#8217;s scoped edits and the two prompts to copy</a>, for editing live workflows without breaking them</p></li><li><p><a href="https://blog.muddventures.com/p/the-operators-read-3-quiet-ai-shifts">The Operator&#8217;s Read on three quiet AI shifts</a>, if you missed that one</p></li><li><p>The full Anthropic guide is <a href="https://claude.com/blog/ciso-guide-to-agentic-ai">here</a>, and it&#8217;s a five-minute read</p></li></ul><p>Reply with the verb you removed if you run the audit, I want to know which one shows up most.</p><p>Andrew</p>]]></content:encoded></item><item><title><![CDATA[Run the three-prompt Google Business Profile check that does most of what a listings retainer bills you for monthly]]></title><description><![CDATA[Google wired your profile data into Gemini, and almost nobody has opened the numbers it's been collecting.]]></description><link>https://blog.muddventures.com/p/run-the-three-prompt-google-business</link><guid isPermaLink="false">https://blog.muddventures.com/p/run-the-three-prompt-google-business</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Mon, 20 Jul 2026 15:17:33 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!pLkm!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!pLkm!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!pLkm!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!pLkm!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!pLkm!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!pLkm!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!pLkm!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:659577,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/207788636?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!pLkm!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!pLkm!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!pLkm!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!pLkm!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F572dbf53-04c3-4f3f-a887-1f59fd34307b_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p></p><p><strong>TLDR:</strong> Google connected Google Business Profile directly to the Gemini app, so a single-location operator can now ask plain questions about their own search impressions, direction requests, call taps and search keywords, draft review replies grounded in the actual review text, and edit hours, attributes and action links from a chat window. It&#8217;s free, it&#8217;s already rolling out globally outside the EEA and UK, and it covers a meaningful slice of what listing-management retainers bill $300 to $1,200 a month for. Below: the three prompts to run, what good output looks like, and the four eligibility rules that will block half of you before you start.</p><div><hr></div><p>In the agency years, the Google Business Profile was almost always somebody&#8217;s third priority. It got set up during onboarding, somebody dropped in the hours and a few photos, and then it sat there for two years untouched until a one-star review showed up and suddenly everybody cared about it for about a day. Meanwhile that profile was quietly collecting the highest-intent data in the whole business, the searches people typed right before they tapped call, and nobody on either side of the engagement was reading it, because opening the dashboard was never on anybody&#8217;s calendar.</p><p>I call that the <strong>Unread Storefront</strong>, and Google just did something about it.</p><p>Google wired Business Profile into Gemini on June 10, plus Business notebooks, and both started rolling out globally over the following weeks (<a href="https://blog.google/innovation-and-ai/products/gemini-app/gemini-features-for-businesses/">Google&#8217;s announcement</a>). Vishnu Sivaji, senior director on the Gemini app, described the connected state as an assistant with &#8220;access to your real-world context like customer reviews, customer questions and performance data&#8221; (<a href="https://www.searchenginejournal.com/google-is-adding-business-profile-tools-to-the-gemini-app/578824/">via Search Engine Journal</a>). It shifts the janitorial half of local visibility out of a dashboard nobody logs into and into a chat window most owners already have open on their phone.</p><p><strong>By the end of this you&#8217;ll have three copy-paste prompts that pull your profile&#8217;s real performance numbers, draft a review reply that references the actual review, and surface the gaps in your listing, plus the eligibility checklist that tells you in 30 seconds whether you can use any of it yet.</strong></p><p>Here&#8217;s what&#8217;s underneath:</p><ul><li><p><strong>The data was always there, the door was just annoying.</strong> Impressions, direction requests, call data and the search keywords customers used are all in Business Profile, and asking for them in plain language is a different behavior than remembering to log in.</p></li><li><p><strong>Review replies are the first thing everybody will use and the easiest thing to get wrong.</strong> Drafting is the win. Publishing without reading is how a tire shop thanks someone for the sushi.</p></li><li><p><strong>Business notebooks are Google&#8217;s version of a standing brief</strong>, a container that holds your profile, your website and your prior chats so the assistant stops starting from zero.</p></li><li><p><strong>The eligibility rules are narrow enough to be the real story.</strong> One verified profile, personal Google account, activity on, and nothing in the EEA or UK.</p></li><li><p><strong>The retainer math shifts at the low end, not the high end.</strong> The $300-a-month &#8220;we post weekly and reply to reviews&#8221; tier is the part under pressure.</p></li></ul><p><strong>Who this is for:</strong> single-location service businesses and local operators who own one verified Google Business Profile, whether you run it yourself or pay somebody a few hundred a month to run it for you.</p><h2>Step one: check whether you&#8217;re even eligible, then connect</h2><p>Open the Gemini web app, go to connected apps, link Google Business Profile with a single tap. Before you do that, run the eligibility check, because Google&#8217;s help documentation is stricter than the blog post reads.</p><p>You qualify if all of these are true:</p><pre><code><code>[ ] You are owner or manager of ONE verified Business Profile, no more
[ ] You are signed in with a PERSONAL Google account (not Workspace/work/school)
[ ] You are 18+ and using the Gemini WEB app (not the GBP Manager app)
[ ] Gemini Apps Activity is turned ON
[ ] Your business is NOT in the European Economic Area or the UK</code></code></pre><p><strong>What correct output looks like:</strong> after connecting, ask &#8220;what Business Profile are you connected to?&#8221; and Gemini names your business, your category and your city back to you without you having told it any of that.</p><p><strong>The failure mode:</strong> if Gemini answers with a generic explanation of what a Business Profile is, or asks you which business you mean, the connection didn&#8217;t take. The two most common causes are the multi-profile block (if you have manager access to a second location, a client&#8217;s profile, or an old test listing you forgot about, you&#8217;re out for now) and the work-account block. The fix for the second one is connecting from the personal account that owns the profile. The fix for the first one is removing your manager access from the profiles you don&#8217;t need, or waiting, because Google has been clear this is the simple-governance first wave (<a href="https://www.tryvizup.com/blog/google-business-profile-comes-to-gemini">eligibility breakdown here</a>).</p><p><strong>Walk away with:</strong> a definitive yes or no on whether this is available to you today, and a connected profile if it is.Step two: run the monthly read you&#8217;ve never run</p><p>This is the one that earns the whole exercise, most operators have never looked at their own search keyword data once. Paste this in:</p><pre><code><code>Using my connected Google Business Profile, give me a plain-English
performance read for the last 30 days compared to the prior 30 days:

1. Search impressions, and whether the split between discovery
   searches and direct searches moved
2. Direction requests, website clicks, and call taps, with the
   percentage change on each
3. The top 15 search keywords customers used to find me, flagged
   as either "my business name" or "a service/problem phrase"
4. Any metric that moved more than 20% in either direction, with
   your best read on why

End with the one number you'd watch next month and why.</code></code></pre><p><strong>What correct output looks like:</strong> a table or list with real figures attached to each line, and a keyword list that includes phrases you didn&#8217;t expect. That third bucket is where the value sits. When most of your discovery traffic comes in on service and problem phrases rather than your business name, your profile is doing acquisition work. When almost everything is your business name, people already knew about you and Google is just the phone book, which is a completely different situation and calls for different spending.</p><p><strong>The failure mode:</strong> if the output comes back with round-number estimates, hedged ranges, or no keywords at all, it&#8217;s answering from general knowledge instead of your connected data. Ask it directly, &#8220;are these figures from my connected Business Profile or are you estimating,&#8221; and re-run. If it still can&#8217;t produce keywords, your profile likely doesn&#8217;t have enough volume in the window yet, so widen it to 90 days.</p><p><strong>Walk away with:</strong> the top 15 phrases real customers typed before they found you, which is the closest thing to free market research a local business has, and almost none of them have looked at it.</p><h2>Step three: reply to reviews without sounding like a robot</h2><p>Review replies are the task everybody hates, and it&#8217;s the task customers use to size up your business. Gemini can pull the star rating, the review text and the associated media, then draft a reply that references what the customer specifically said (<a href="https://support.google.com/business/answer/17142585?hl=en">capability list here</a>). The part operators skip is giving it a voice rule first, so paste the rules and the request together:</p><pre><code><code>You're drafting replies to my Google reviews. Voice rules, apply every time:

- First person singular, I not we, unless the review names an employee
- Two to four sentences maximum, never a paragraph
- Reference one concrete detail the customer mentioned, by name
- No marketing language, no "we strive to," no "your feedback is
  important to us"
- On anything under 4 stars: acknowledge the specific problem, say what
  changed or what I'm doing about it, give one way to reach me directly
- Never offer a refund, discount, or free service in a reply
- Never dispute the customer's version of events publicly

Pull my three most recent unanswered reviews and draft a reply to each.
Show me the review text above each draft so I can check it.</code></code></pre><p><strong>What correct output looks like:</strong> each draft names something real from the review, the wait on Saturday, the tech who showed up early, the part that came in wrong, and the low-star replies read like a person taking responsibility rather than a policy statement.</p><p><strong>The failure mode:</strong> polite mush. If the draft would fit under any review at any business in any industry, it&#8217;s useless and worse than no reply, because customers can tell. The other failure mode is the factual one, which is why the prompt asks it to show the review text above each draft. As one local SEO write-up put it, &#8220;Connect Gemini to a messy profile and you will get messy outputs faster&#8221; (<a href="https://www.tryvizup.com/blog/google-business-profile-comes-to-gemini">Vizup</a>). Read every draft before it publishes, especially the one-star replies, because a published reply represents your business permanently and Google won&#8217;t be the one explaining it.</p><p><strong>Walk away with:</strong> three reviews answered in about five minutes, in a voice that&#8217;s yours, with the rules saved so the next batch takes two.Business notebooks are the Google-side version of a standing brief</p><p>The notebook is the piece that turns three good prompts into a system. It&#8217;s a container inside Gemini that holds your connected profile, your website, your chats and your sources, so the assistant carries context between sessions instead of meeting you fresh every morning, and it surfaces alerts when you open it, things like an unanswered customer question or holiday hours that were never set.</p><p>It&#8217;s the same shape as <a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">the Standing Brief on the ChatGPT side</a>, one written description of your business that lives somewhere permanent so you stop paying the setup cost every single chat. Lisa Landsman, who runs industry engagement and SMB success at Google, described the same complaint coming back from owners over and over, that &#8220;AI would be infinitely more useful if it knew their specific business, rather than forcing them to re-explain who they are, what they do, and who they serve every single time&#8221; (<a href="https://www.seroundtable.com/google-business-profile-integrated-gemini-41484.html">via Search Engine Roundtable</a>).</p><p>Drop your voice rules and your service list into the notebook once, so the review-reply prompt above shrinks to five words, and keep the master copy of that brief in a doc you own rather than only inside Google&#8217;s box. Vendors change their boxes. Your brief should outlive them.</p><h2>What this does to the retainer math</h2><p>Listing management from a real agency runs $300 to $1,200 a month, where the low tier covers weekly posts, photo uploads, Q&amp;A response, review replies within 48 hours, attribute updates and a monthly insights review, roughly four to six hours of skilled work, and the full Map Pack program adds citation building, review-generation systems and competitive monitoring at the top of that range (<a href="https://coloradowebimpressions.com/answers/how-much-does-google-business-profile-management-cost">Colorado Web Impressions pricing breakdown</a>). That same page puts it plain: avoid GBP management under $200 a month, because it&#8217;s usually automated.</p><p>That&#8217;s the honest read on what changed. The automated tier just got a free competitor built by the company that owns the data, and the judgment tier didn&#8217;t. Citation building, review generation systems, competitor monitoring, category strategy, suspension recovery, those are still real work by people who&#8217;ve done it a hundred times, and Gemini isn&#8217;t close to that. If you&#8217;re paying at the low end for weekly posts and review replies, run the three prompts above for 30 days and then have an honest conversation about what you&#8217;re buying. If you&#8217;re paying at the high end for a Map Pack program that&#8217;s producing calls, this changes your maintenance workflow and leaves the strategy where it is.</p><h2>What&#8217;s still broken</h2><p>The eligibility gate is the biggest one, and it knocks out most people reading this who own more than one location or hold manager access on a client profile. Google went with one person, one profile as the starting governance model, which means agencies and franchises can&#8217;t use this at all right now.</p><p>The EEA and UK exclusion is real and Google hasn&#8217;t said when that changes. Gemini Apps Activity has to stay on, which is a genuine privacy tradeoff some operators won&#8217;t want to make, and it&#8217;s worth knowing that&#8217;s the price before you flip it.</p><p>And the write access cuts both ways. An assistant that can edit hours, attributes, menus and action links is an assistant that can put wrong hours in front of customers, and wrong hours means somebody drives across town to a closed store, so the cost lands on the customer experience side long before it shows up in any ranking. Treat profile edits like production changes, confirm each one, and check the live profile after.</p><h2>Recap</h2><p>Google connected Business Profile to Gemini, free, rolling out globally outside the EEA and UK, gated to single-profile owners on a personal account with activity on. Run the eligibility check, then the 30-day performance read, then the review-reply prompt with voice rules attached, then park all of it in a Business notebook so the context stops evaporating. The keyword list from prompt two is the piece almost nobody has looked at, and it&#8217;s the one I&#8217;d open first.</p><p>The thread running through this and the last several issues: the tools that used to require hiring somebody keep getting absorbed into the layer you already pay for, and the operators pulling ahead are the ones who go check what got absorbed this month instead of assuming the invoice still maps to the work.</p><p>If the deeper question here is what people find when they look your business up across Google, Maps and the AI answers now sitting on top of both, that&#8217;s the ground the <a href="https://counterclaim.muddventures.com">Counterclaim Skill Pack</a> covers, and it starts where this one ends.</p><p><strong>P.S.</strong> A few next steps if you want to keep pulling this thread:</p><ul><li><p><a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">The Standing Brief</a>, the six-section business brief that stops you re-explaining yourself in every chat</p></li><li><p><a href="https://blog.muddventures.com/p/the-operators-read-3-quiet-ai-shifts">The Operator&#8217;s Read on three quiet AI shifts</a>, including the model-retirement check worth running on old automations</p></li><li><p><a href="https://counterclaim.muddventures.com">Counterclaim Skill Pack</a> for what AI answers are saying about your business</p></li><li><p><a href="https://whop.com/abra-ai/">Abra AI community</a> if you want the skill files and the operators building this stuff daily</p></li><li><p><a href="https://muddventures.com/book">Book an AI Clarity Call</a> if the question is bigger than one workflow</p></li></ul><p>Reply with your top search keyword if you run prompt two, I want to see how far off it lands from what people assume.</p><p>Andrew</p>]]></content:encoded></item><item><title><![CDATA[Your WhatsApp and Instagram DMs can now answer themselves, and Meta already published the dates when free ends]]></title><description><![CDATA[The agent takes one evening to set up, and the free window has an end date on the calendar.]]></description><link>https://blog.muddventures.com/p/your-whatsapp-and-instagram-dms-can</link><guid isPermaLink="false">https://blog.muddventures.com/p/your-whatsapp-and-instagram-dms-can</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Fri, 17 Jul 2026 16:02:10 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!u2xo!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!u2xo!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!u2xo!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!u2xo!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!u2xo!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!u2xo!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!u2xo!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1091261,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/207444463?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!u2xo!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!u2xo!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!u2xo!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!u2xo!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F65808788-da6a-4f1e-abfa-721c0eecd7e2_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TLDR:</strong> Meta Business Agent is live for businesses of every size across WhatsApp, Messenger, and Instagram, and turning it on costs nothing today. The billing calendar is already public: the partner platform starts metering August 1 at roughly 4 to 5 cents per agent-handled message, and from October 1 every reply inside an open WhatsApp service window carries a charge. The free stretch is the cheapest month you&#8217;ll ever get with this tool, enough time to set the agent up in demo mode, break it with a test script, and learn what your real conversation volume would cost, so the paid era opens with your numbers instead of Meta&#8217;s estimates.</p><p>Every booking workflow I&#8217;ve built in GoHighLevel exists because of one number: the longer a lead sits unanswered in a message thread, the deader the deal gets. I spent years on the agency side watching that decay eat perfectly good pipeline, a lead messages at 9pm, the reply lands at 10 the next morning, and the thing they were ready to buy the night before has cooled into &#8220;let me think about it.&#8221; Entire product categories exist to shave minutes off that gap.</p><p>Which is why the part of Meta&#8217;s announcement that got me had nothing to do with the AI itself. Meta ran its Conversations event in London on June 3 and turned on <a href="https://about.fb.com/news/2026/06/meta-business-agent/">Meta Business Agent</a> for businesses of every size, an agent that answers customer questions, recommends products from your catalog, books appointments, qualifies leads, and closes sales inside WhatsApp, Messenger, and Instagram threads. Free to activate, live in minutes, already working for more than a million businesses. Then, four weeks later, the billing schedule showed up with real dates on it.</p><p>By the end of this you&#8217;ll have the agent running in demo mode, a ten-message script for breaking it before customers find the gaps, and your own cost estimate for the metered era before Meta sends anyone an invoice.</p><p>That free-then-billed sequence deserves a name, so I&#8217;m coining one: The Meter Flip. A platform gives an AI capability away long enough for a million businesses to wire it into daily operations, then turns the meter on against a published schedule. It&#8217;s the same absorption pattern I wrote about when <a href="https://blog.muddventures.com/p/notion-just-made-zapier-optional">Notion made Zapier optional</a>, a native feature swallowing what used to be a paid vendor, with one difference that changes how you play it: this time the free part has a printed expiration date, and the first of those dates is two weeks out.</p><ul><li><p><strong>Free today, with paid tiers already promised.</strong> Meta&#8217;s own newsroom says getting started costs nothing and that &#8220;paid subscription offerings&#8221; arrive in the coming months.</p></li><li><p><strong>The platform meter starts August 1.</strong> $2.00 per million tokens, which Meta translates to roughly 4 to 5 cents per agent-handled message, one blended charge covering the AI and the message delivery.</p></li><li><p><strong>From October 1 there is no free reply left on WhatsApp.</strong> Human, template, or agent, every response inside an open service window gets billed.</p></li><li><p><strong>The math is knowable this month.</strong> A test script and three weekly numbers tell you what the paid era costs before it starts.</p></li></ul><p><strong>Who this is for:</strong> operators whose customers already message them on WhatsApp, Instagram, or Messenger, and anyone paying a chatbot vendor or an answering service to cover those same threads.</p><h2>What Meta turned on</h2><p>Meta&#8217;s pitch line is &#8220;AI that lets every business show up for every customer as if they had an infinite team behind them,&#8221; and the capability list in <a href="https://about.fb.com/news/2026/06/meta-business-agent/">the announcement</a> backs most of that up: it answers questions specific to your business, makes product recommendations from a catalog, books appointments and qualifies incoming leads, lets you decide when a team member steps in, and closes sales. More than a million businesses were running it on WhatsApp and Messenger in markets like Brazil, India, and Mexico before the global expansion, and Meta counts over a billion active business threads a day across its three messaging apps.</p><p>One thing worth getting straight before you touch settings: there are three different products under this one name, and the difference decides what you&#8217;d eventually pay. Wati, one of the larger WhatsApp business platforms, <a href="https://www.wati.io/en/blog/meta-business-agent/">broke the tiers down</a> after digging into the launch. A free self-serve agent lives inside the WhatsApp Business app and learns from your past chats, business profile, and catalog. The same self-serve experience exists for Messenger inside Meta Business Suite. And above both sits an invite-only enterprise platform tier that runs on the WhatsApp API and connects to outside systems like Shopify and Zendesk. The self-serve tiers are the ones most readers of this newsletter can turn on tonight, from a phone, without a developer anywhere in the building.</p><p>The agent only replies when it&#8217;s confident it can answer correctly, flags the conversation for you when it can&#8217;t, and passes the full chat history across when a human steps in. There&#8217;s also a morning briefing feature rolling out to a smaller group, which catches you up on overnight threads, with <a href="https://metabusinessai.com/waitlist">a waitlist</a> if you want in early.</p><h2>The billing calendar Meta already published</h2><p>Now the Markup Machine part, because the dates are the story. <a href="https://chatmaxima.com/blog/meta-business-agent-platform-explainer-2026/">ChatMaxima&#8217;s breakdown of the platform rollout</a> lays out the schedule. July 1: the Business Agent Platform and its APIs went live for eligible partners, free. August 1: per-token billing begins at &#8220;$2.00 per 1 million tokens, which Meta estimates at roughly 4 to 5 cents per message,&#8221; one blended charge that covers both the AI work and the message delivery, invoiced monthly by Meta directly. October 1: WhatsApp service messages resume charging, and utility templates sent inside an open service window get billed too. Their summary of where that leaves the channel is blunt: &#8220;Whichever way a business chooses to answer a customer inside an open conversation, an AI agent, a human, or a template, there is now a charge. The channel is fully metered.&#8221;</p><p>I ran Meta&#8217;s own estimate against typical operator volume. A business whose agent handles 5,000 messages a month is looking at roughly $250 at the conservative 5-cent end. A hundred thousand messages runs about $5,000. And a solo operator fielding 30 conversations a day, the profile Wati says this product fits best, lands somewhere around $40 to $50 a month once metering reaches their tier. Put those numbers next to what a chatbot vendor or an after-hours answering service invoices for covering the same threads and the compression is hard to miss, buried in a channel most operators still answer by hand.</p><p>One precision note, because this is where coverage gets sloppy: the August 1 meter applies to the partner platform tier. The self-serve agent stays free for now, Meta&#8217;s newsroom language is subscriptions &#8220;in the coming months,&#8221; and <a href="https://techcrunch.com/2026/06/03/metas-ai-agent-for-whatsapp-business-is-now-available-globally/">TechCrunch reports</a> the paid path will likely ride on WhatsApp Business Premium tiers. The direction is identical at every layer though, free to hook, metered on a schedule, which is the whole reason the free stretch is worth using deliberately instead of vaguely.</p><h2>Build one: put the agent in demo mode with a real brief</h2><p>The move: open the WhatsApp Business app, tap Tools, select Meta Business Agent, grant it access to your past chats, and add your catalog plus FAQ material. On Facebook, it&#8217;s Meta Business Suite, then All Tools, then Business AI. Both paths end in a demo mode, and demo mode is where the agent stays until it survives the script in build two.</p><p>Before you tap through setup, prepare the knowledge you&#8217;ll feed it. This is the input worth writing first:</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">WHAT WE SELL: [each product or service with its real price, no "starting at"]
HOURS AND LOCATION: [including holiday exceptions]
THE FIVE QUESTIONS CUSTOMERS ASK MOST: [each with the answer in your own words]
NEVER DO: no discounts, no delivery-date promises, no prices beyond the list above, no claims about results
HAND TO A HUMAN WHEN: [complaints, refunds, anything about money owed, custom requests]
TONE: [paste three replies you sent recently that sound like you]</code></pre></div><p>If you built the 5,000-character Standing Brief from <a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">Wednesday&#8217;s issue</a>, this is the condensed customer-facing cut of the same document, the difference being this version speaks to your customers instead of to you.</p><p>What a correct result looks like: in demo mode, the agent answers your five most common questions with your real numbers in something close to your tone, and the questions it can&#8217;t handle come back flagged for you rather than improvised.</p><p>The failure mode: the agent generalizes, or invents a policy you don&#8217;t have. That means the knowledge you gave it was thin, so feed it the real catalog and the real policies rather than a summary. If it keeps freelancing on price or promise questions after that, move those categories into the hand-to-a-human list instead of leaving them answerable.</p><p><strong>Walk away with:</strong> an agent grounded in your material instead of Meta&#8217;s guesses, live nowhere except demo mode.</p><h2>Build two: attack it like a customer before a customer does</h2><p>The move: run a fixed ten-message script through demo mode, and keep the script frozen so next week&#8217;s results compare cleanly against this week&#8217;s.</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">1. how much is [your core offer]
2. you got anything available this week
3. why would i pick you over [name a real competitor]
4. whats your refund policy
5. can you do it cheaper
6. [ask about something you don't sell]
7. this is ridiculous, ive been waiting two days
8. book me for tuesday at 3
9. [message 1 again, in whatever second language your customers use]
10. asdfjkl</code></pre></div><p>What a correct result looks like: right numbers on 1 and 2, a handoff or an owner flag on 5 and 7, a clean &#8220;that&#8217;s outside what we do&#8221; on 6, a booking that lands on your calendar from 8, a coherent answer on 9, and no meltdown on 10.</p><p>The failure mode: it offers a discount on 5, or invents availability on 2. Tighten the NEVER DO list and run the script again. Any question class that fails twice gets demoted from answerable to handoff, and that single rule is what keeps a cheap agent from becoming an expensive apology.</p><p><strong>Walk away with:</strong> a pass/fail read on the agent before any paying customer meets it.</p><h2>Build three: measure the free month</h2><p>The move: track three numbers weekly for the rest of the free stretch.</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">messages the agent handled this week: ___
conversations that became a booking or a sale: ___
conversations handed to a human: ___</code></pre></div><p>Multiply weekly message volume by four, then by $0.05, and you&#8217;re holding a conservative monthly cost for the day metering reaches your tier. Divide that cost by the bookings the agent produced and you have cost per booked conversation, which is the number the whole decision hangs on.</p><p>What a correct result looks like: something as plain as &#8220;the agent handled 240 messages, produced 9 bookings, and worst case that&#8217;s about $48 a month for overnight coverage those threads were getting from nobody.&#8221;</p><p>The failure mode runs in two directions. Volume so low the projection rounds to nothing means the agent is simply free 24-hour coverage, keep it. A projection that clears what a conversation is worth to you means the agent gets capped to first-response and FAQ duty while your humans keep the closing threads, and you&#8217;ve learned that for free instead of from an invoice.</p><p><strong>Walk away with:</strong> your own cost number before Meta&#8217;s first invoice exists anywhere.</p><h2>The honest tradeoffs</h2><p>The ceiling is real and Wati named it without hedging: &#8220;it may be useful for a solo operator managing 30 conversations a day from a phone. For businesses doing real volume on WhatsApp, with teams, broadcasts, and CRM workflows, it hits a ceiling fast.&#8221; The self-serve tier has no team inbox, no CRM connection, no broadcast tools, and it only responds to inbound messages, it never initiates. The enterprise tier that does connect to outside systems is invite-only, and its official pricing page still reads &#8220;in development,&#8221; with tiered pricing promised later, which has procurement-minded operators understandably annoyed.</p><p>The data question deserves a straight look too. Per Wati&#8217;s comparison: &#8220;your customer conversations stay within Meta&#8217;s infrastructure and are used to train Meta&#8217;s models.&#8221; If you treat customer conversations as an owned asset, that&#8217;s a cost this product charges that never shows up on an invoice.</p><p>And the launch anxiety is worth acknowledging rather than waving off. Wati notes the most upvoted Reddit post the night of the announcement read &#8220;Meta Announces Meta Business Agent: A Replacement for All BSPs and Tech Providers.&#8221; Given the ceiling above, that fear runs ahead of the product today, and the vendors most exposed are the ones whose whole offer was basic FAQ coverage. The last quiet risk is the obvious one: a wrong answer in your DMs carries your name, and the confidence gating helps, but demo-mode testing is the control you own, so the agent earns live mode rather than defaulting into it.</p><h2>Recap</h2><p>Meta put a working AI agent inside the three messaging apps where your customers already are, free to activate today, and then published the calendar for metering the channel, platform billing August 1 at roughly 4 to 5 cents a message, every WhatsApp service-window reply billed from October 1. That&#8217;s the Meter Flip, and the operators who use the free stretch to ground the agent in a real brief, break it with a fixed test script, and track three weekly numbers will enter the paid era holding their own cost per booked conversation, while everyone else enters it holding a guess.</p><p>The agent can get the appointment booked, and what happens in the seventy-two hours after that booking decides whether the person shows, that after-the-booking gap is what <a href="https://showtime.muddventures.com">showtime.muddventures.com</a> wires up for $27.</p><p>Back in your inbox Monday.</p><p>Andrew</p><p>P.S. If you&#8217;re building the pieces around this one: the <a href="https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief">5,000-character Standing Brief</a> is the master document your agent brief condenses from, <a href="https://blog.muddventures.com/p/notion-just-made-zapier-optional">Notion making Zapier optional</a> is the earlier sighting of this same absorption pattern, the <a href="https://blog.muddventures.com/p/your-show-rate-moves-on-a-decay-curve">Show Rate Decay Curve</a> is why answered-in-minutes beats answered-in-hours, and the <a href="https://blog.muddventures.com/p/gohighlevels-ai-builder-just-learned">GoHighLevel scoped-edits prompts</a> cover the workflow side of the same automation stack.</p>]]></content:encoded></item><item><title><![CDATA[Grab the 5,000-character ChatGPT brief that stops you re-explaining your business in every chat]]></title><description><![CDATA[OpenAI more than tripled the custom instructions box on July 15, and most people still have one sentence in there. The briefing template below fills all 5,000 characters with things that pay off.]]></description><link>https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief</link><guid isPermaLink="false">https://blog.muddventures.com/p/grab-the-5000-character-chatgpt-brief</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Thu, 16 Jul 2026 15:17:59 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!87al!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!87al!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!87al!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!87al!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!87al!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!87al!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!87al!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1016196,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/207299981?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!87al!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!87al!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!87al!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!87al!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F480f55d4-c628-46a1-ad65-991f3a082d9f_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TL;DR:</strong> On July 15, OpenAI raised ChatGPT&#8217;s custom instructions limit from 1,500 characters to 5,000 for paid plans. That box is where you end the cycle of re-teaching a fresh chat your business every morning, and at the old size it was too small to hold a real briefing. Below: the full Standing Brief template to paste in, the 15-minute test that proves it took, and the places where this still breaks.</p><p>Yesterday morning&#8217;s issue covered <a href="https://blog.muddventures.com/p/claude-will-now-show-you-what-you">Claude Reflect</a>, the new report that shows you which tasks you keep re-explaining to your AI, and I named the cycle underneath it <a href="https://blog.muddventures.com/p/claude-will-now-show-you-what-you">The Context Rebuild Loop</a>: open a fresh chat, re-teach it your business, get the output, throw the setup away, then pay that same setup cost again a few days later. Later that same day, OpenAI shipped the other half of that story. The custom instructions box in ChatGPT <a href="https://help.openai.com/en/articles/6825453-chatgpt-release-notes">went from 1,500 characters to 5,000</a> for paid plans, which is the difference between a long text message and a proper one-page briefing document.</p><p>By the end of this you&#8217;ll have a 5,000-character Standing Brief written for your own business and a 15-minute test that proves ChatGPT is using it instead of ignoring it.</p><p>I run my own operation on standing context files, the system that drafts this newsletter reads a stack of them before it types a word, so a bigger box for standing context is the kind of update I stop and look at. The old 1,500-character cap is a big part of why most people treated custom instructions as a tone preference, there was only ever room for &#8220;be concise, I&#8217;m a marketing manager,&#8221; and a box that small trains you to write something shallow in it.</p><p>What&#8217;s below:</p><ul><li><p><strong>The July 15 change in plain terms.</strong> Who got 5,000 characters, who stayed at 1,500, and why the update applies to chats you already have open.</p></li><li><p><strong>Why the old cap kept the box shallow.</strong> 1,500 characters is roughly 250 words, and people filled it accordingly.</p></li><li><p><strong>The Standing Brief template.</strong> Six sections worth spending the new character budget on, ready to copy.</p></li><li><p><strong>The proof test.</strong> Three prompts that show whether the brief took.</p></li><li><p><strong>Where this still breaks.</strong> Instruction drift is real, and the forum threads about it are worth reading before you trust the box completely.</p></li></ul><p><strong>Who this is for:</strong> operators who use ChatGPT on a paid plan for real business work, writing, analysis, customer replies, planning, and are tired of every chat starting from zero.</p><h2>What OpenAI changed on July 15</h2><p>Straight from OpenAI&#8217;s <a href="https://help.openai.com/en/articles/6825453-chatgpt-release-notes">release notes</a>: &#8220;Plus, Pro, Enterprise, Business, and Education users can now save up to 5,000 characters, up from 1,500, giving them more room to customize ChatGPT&#8217;s response style and behavior.&#8221; Free and Go plans stay at 1,500 per OpenAI&#8217;s <a href="https://help.openai.com/en/articles/8096356-chatgpt-custom-instructions">help center</a>, and the expanded instructions apply to your existing conversations immediately, no fresh session required, per <a href="https://cryptobriefing.com/openai-chatgpt-custom-instructions-5000-characters/">Crypto Briefing's July 15 report</a>.</p><p>Some quick context on why this box matters more than it did a year ago. Back in November 2025, OpenAI made custom instructions apply across all your conversations instead of only new ones, so whatever you write in there now follows you everywhere. And the day before this change, on July 14, ChatGPT also added <a href="https://releasebot.io/updates/openai/chatgpt">unified search</a> across chats, projects, images, and documents on every plan, one sidebar search that finds your old work. The pattern across both updates is OpenAI treating your accumulated context as the product, and the instructions box is the one piece of that context you get to author on purpose.</p><p>Run the math on the size. 1,500 characters is about 250 words, a long text message. 5,000 characters is about 800 words, a full page of real briefing material. That&#8217;s enough room for what your business sells and at what price, who buys it, what tools you run it on, how you write, and what you&#8217;re working on this quarter.</p><h2>The myth: custom instructions are a tone setting</h2><p>The box has been around since 2023, and the standard move has always been to put a bio in it. &#8220;I&#8217;m a marketing manager. Be concise and professional.&#8221; Emmanuel at My Writing Twin, who has spent more time on this feature than almost anyone publishing about it, <a href="https://www.mywritingtwin.com/blog/custom-gpt-instructions-complete-guide">puts the problem plainly</a>: &#8220;When you say &#8216;professional but friendly,&#8217; you&#8217;ve described maybe 50,000 different writing styles.&#8221; ChatGPT splits the difference across all of them and hands you the same competent, forgettable output it hands everyone else.</p><p>Here&#8217;s the landmine underneath the myth. If the box holds one vague sentence, every chat starts from zero, so you paste in your offer details again, describe your customer again, explain your pricing again, and that&#8217;s the <a href="https://blog.muddventures.com/p/claude-will-now-show-you-what-you">Context Rebuild Loop</a> running on a daily timer. The brief I keep for my own business runs long past the old cap, the writing rules section alone wouldn&#8217;t have fit in 1,500 characters, which is why the master copy has always lived in a document outside the chat window. As of July 15 a real working chunk of it fits in the box itself.</p><p>The fix is treating the field as what it now has room to be: a Standing Brief. One written briefing document about your business, maintained in a doc you own, pasted into the personalization layer of every AI tool you use. Write it once, and every chat starts from there instead of from zero.</p><h2>Write your Standing Brief</h2><p><strong>The move:</strong> block 25 minutes. Open a document you own, Google Doc, Notion page, plain text file, and write the brief there first, not in the ChatGPT settings box. The doc is the master, the box is a deployment target, and when Claude and Gemini get the same treatment later you&#8217;ll paste from the same source. Then copy it into ChatGPT under Settings, Personalization, Custom instructions.</p><p><strong>The template:</strong> six sections. Fill in the brackets, cut what&#8217;s irrelevant, keep it under 5,000 characters.</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">1. THE BUSINESS
I run [business name], a [what it is] that sells [offer] at [price point(s)]
to [customer type]. Revenue comes from [primary revenue source]. My role
is [role].

2. THE CUSTOMER
My buyer is [who they are, what they run, size]. They care about [top 2-3
things]. They do not respond well to [what turns them off]. Common
objections: [list 2-3].

3. THE STACK
I run the business on [CRM], [email tool], [calendar/booking tool],
[other core tools]. When you suggest a workflow, use these tools by name.

4. WRITING RULES
Always: [2-4 rules, e.g. short paragraphs, plain words, specifics over
adjectives]. Never: [2-4 rules with reasons, e.g. never use exclamation
points, never open with a question, never use the word "elevate"].
When drafting anything a customer sees, match this register: [paste 2-3
sentences you wrote yourself as a sample].

5. CURRENT PRIORITIES
This quarter I'm focused on [1-3 priorities with a number attached where
possible]. Weight suggestions toward these.

6. HOW TO HANDLE UNKNOWNS
If you're missing a fact about my business, ask me one direct question
before drafting. Do not invent product names, prices, or claims.</code></pre></div><p><strong>What the correct output looks like:</strong> the next fresh chat behaves like a briefed contractor instead of a stranger. Ask it to draft a follow-up email to a lead who went quiet and it writes in your register, names your offer, respects your price point, and skips the exclamation points, all without you pasting anything into the prompt.</p><p><strong>The failure mode:</strong> the output stays generic. Nine times out of ten that means the brief is adjectives instead of rules. &#8220;Professional but friendly&#8221; gives the model 50,000 styles to average across, &#8220;never use exclamation points, never open with a question, here are three sentences I wrote myself&#8221; gives it walls to work inside. Rewrite every adjective in your brief as an always rule, a never rule, or a pasted example of your own writing.</p><p><strong>Walk away with:</strong> a one-page briefing document you own, deployed into ChatGPT, that every future chat inherits automatically.</p><h2>The 15-minute proof test</h2><p><strong>The move:</strong> before you write the brief, run three work prompts in fresh chats and save the outputs somewhere. After the brief is installed, run the same three prompts in fresh chats and compare side by side. Same day, same model, the only variable is the brief.</p><p><strong>The exact prompts:</strong></p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">1. Draft a follow-up email to a lead who went quiet after asking about
   pricing two weeks ago.

2. I have 90 minutes free tomorrow morning. Based on what you know about
   my business, what's the highest-leverage thing to spend it on?

3. Write three social post hooks about what I sell, in my voice.</code></pre></div><p><strong>What the correct output looks like:</strong> the after versions cite specifics you never typed into the prompt. The email names your offer and handles the price objection the way your brief says your buyers raise it, the 90-minute answer points at one of your stated quarterly priorities, and the hooks sound like your pasted writing sample instead of a motivational poster.</p><p><strong>The failure mode:</strong> the model keeps breaking one specific rule, capitalization, a banned word, a format you told it to skip. Two things to check, in order. First, move that rule higher in the brief and restate it as a never with a reason, position and bluntness both change compliance. Second, check for a conflict with ChatGPT&#8217;s Memory feature, which collects facts about you passively in the background and can contradict what you wrote deliberately. Since <a href="https://releasebot.io/updates/openai/chatgpt">June 12</a> you can view, edit, and delete individual memories from the memory summary page, so clear out anything that fights the brief.</p><p><strong>Walk away with:</strong> before-and-after evidence of whether your brief changed the output, instead of a vague feeling that ChatGPT seems smarter now.</p><h2>Where this still breaks</h2><p>Worth being straight about the limits, because the people who&#8217;ve leaned hardest on this feature are also its loudest critics.</p><p>Instruction-following is a probability, and it drops as conversations get long. The OpenAI community forum has years of threads on this, and a <a href="https://community.openai.com/t/custom-instructions-are-useless/1367231">November 2025 post titled "Custom Instructions Are Useless"</a> is a fair sample of the frustration: &#8220;What is the point of having custom instructions on personalizations if they&#8217;re not being followed?&#8221; That user had explicit formatting rules in the box and watched the model wander off them mid-conversation anyway. I&#8217;ve been using AI daily since February 2023, and drift is the one complaint that has never fully gone away in that whole stretch, across every model generation and every vendor. A bigger box doesn&#8217;t fix drift, it gives your rules more surface area and better odds. Expect the brief to hold strongest in the first stretch of any chat, and expect to re-anchor with a one-line reminder in marathon sessions.</p><p>The 5,000-character tier is paid only. Free and Go users keep 1,500, so if that&#8217;s you, deploy sections 1, 4, and 6 of the template and hold the rest in your master doc for pasting when it counts.</p><p>And the box is one personalization layer among several. Memory collects facts about you passively, Projects carry their own context, and custom GPTs override the global brief entirely, so if an output surprises you, the explanation might live in a different layer than the one you edited. The Standing Brief earns its keep by being the layer you wrote on purpose, in a document you control, portable to any tool that gives you a place to paste it.</p><h2>Recap</h2><p>OpenAI more than tripled ChatGPT&#8217;s custom instructions capacity on July 15, from 1,500 characters to 5,000 on paid plans, and the expansion applies to existing chats immediately. The box was never meant to be a bio field, and at its new size it holds a proper Standing Brief: what you sell, who buys, what tools you run, how you write, what you&#8217;re focused on, and how the model handles gaps. Write the brief in a doc you own, deploy it to the box, prove it with the three-prompt before-and-after test, and keep your expectations calibrated on drift.</p><p>Building context once and reusing it everywhere is most of what separates operators who compound with AI from operators who start over every morning, and it&#8217;s the exact muscle we build inside <a href="https://whop.com/abra-ai/">whop.com/abra-ai</a>, where the skill files and memory setups are already written for you.</p><p>P.S. If you&#8217;re working through this in order:</p><ul><li><p>Yesterday&#8217;s piece on <a href="https://blog.muddventures.com/p/claude-will-now-show-you-what-you">Claude Reflect and the Context Rebuild Loop</a> is the diagnostic half of today&#8217;s fix.</p></li><li><p>The <a href="https://blog.muddventures.com/p/gohighlevels-ai-builder-just-learned">GoHighLevel scoped-edits prompts</a> from Monday pair well once your brief is live.</p></li><li><p>The <a href="https://blog.muddventures.com/p/work-from-anywhere-just-got-tested">two-labs-in-48-hours breakdown</a> covers where agent work is headed next.</p></li><li><p>The full archive lives at <a href="https://muddventures.substack.com/">muddventures.substack.com</a>.</p></li></ul><p>Back in your inbox tomorrow.</p><p>Andrew</p>]]></content:encoded></item><item><title><![CDATA[Claude will now show you what you use AI for, and the report doubles as a free systems audit]]></title><description><![CDATA[Reflect is live in Settings for every Claude plan with memory on. Ten minutes in it finds the task you keep re-explaining, and that task is your next standing system.]]></description><link>https://blog.muddventures.com/p/claude-will-now-show-you-what-you</link><guid isPermaLink="false">https://blog.muddventures.com/p/claude-will-now-show-you-what-you</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Wed, 15 Jul 2026 17:59:53 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!iikm!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!iikm!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!iikm!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!iikm!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!iikm!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!iikm!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!iikm!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1114313,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/207186296?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!iikm!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!iikm!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!iikm!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!iikm!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3051e740-e817-47b0-84d5-24fb984d4019_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TL;DR:</strong> Anthropic shipped Reflect on July 9, a beta dashboard inside Claude&#8217;s Settings that shows what you&#8217;ve used AI for over the last 1 to 12 months: top topics, peak hours, task categories, and recommendations built from your own patterns. It&#8217;s free on every plan, it takes ten minutes to read, and the most useful thing in it is the task you keep re-explaining in fresh chats, because that task is a system waiting to be built.</p><p>This newsletter got researched and drafted this morning by a scheduled Claude task, the same setup I covered in the <a href="https://blog.muddventures.com/p/work-from-anywhere-just-got-tested">Work From Anywhere issue</a>, so I have a running record of where my own AI usage concentrates without needing a dashboard to tell me. Most people don&#8217;t have that record, and until last Thursday there was no easy way to get one. Anthropic shipped it: <a href="https://www.anthropic.com/news/reflect-with-claude">Reflect</a>, a beta dashboard that reads your chat history and hands you the usage report.</p><p>By the end of this you&#8217;ll know what Reflect shows, the one pattern worth hunting for in your own report, and how to promote that pattern from a chat you keep re-starting into a system that runs without you re-explaining it.</p><p>The pattern has a name worth keeping: <strong>The Context Rebuild Loop</strong>. It&#8217;s the cycle where you open a fresh chat, re-explain your business, your format, your preferences, get the output, and throw the setup away, then pay that same setup cost again three days later on the same task. Anthropic&#8217;s own product team built detection for it, Reflect watches for people who, in Engadget&#8217;s words, &#8220;frequently re-establish the same or similar context&#8221; and nudges them toward Projects. The loop is invisible while you&#8217;re in it, which is what makes a mirror useful.</p><ul><li><p><strong>What shipped:</strong> a Reflect tab in Claude&#8217;s Settings (web and desktop) showing a summary of your usage, most active day, peak hour, total chats, and a topic breakdown with percentages.</p></li><li><p><strong>Who gets it:</strong> Free, Pro, and Max accounts in beta, with memory turned on.</p></li><li><p><strong>What it recommends:</strong> fluency suggestions built from your own patterns, like moving a repeated task into a Project or a custom skill.</p></li><li><p><strong>The move below:</strong> find your most-repeated task and promote it from chat to system in one sitting.</p></li></ul><p><strong>Who this is for:</strong> anyone using Claude more than a few times a week, and especially operators who suspect they&#8217;re solving the same problem from scratch every Monday.</p><h2>What Reflect shows you</h2><p>Open Settings in Claude on web or desktop and pick the Reflect tab, then generate your report. The default view collates your last month, and you can widen it to 3, 6, or 12 months. At the top sits a paragraph summary of what you&#8217;ve been working through, followed by your most active day, peak hour, and total chats, then a breakdown of your topics with a percentage on the ones you hit most. The last stretch is the interesting part: a set of recommendations grouped around the <a href="https://anthropic.skilljar.com/ai-fluency-framework-foundations">4D AI Fluency framework</a> Anthropic built with academic partners, generated from your own patterns rather than generic tips.</p><p>Ryn Linthicum, Anthropic&#8217;s head of wellbeing policy, <a href="https://www.engadget.com/2211304/claude-reflect-dashboard-wants-to-help-you-log-off/">told Engadget</a> the intent behind it: &#8220;We were really intentional about building [the dashboard] with an eye toward how we can upskill people&#8217;s usage of Claude, not in a way that encourages them to spend more time with it, but instead enables them to get more efficient at meeting their goals.&#8221; There&#8217;s also a quiet-hours setting and a break nudge in there, and the dashboard periodically asks questions like &#8220;What&#8217;s one thing you want to keep doing yourself, even if Claude could do it faster?&#8221;, which is a better question than most productivity content asks.</p><p>If you generate a report and get nothing, memory is off, that&#8217;s the dependency. Reflect only reads memory-enabled history, skips incognito chats entirely, and leaves out the underlying files from connected tools, so an inbox summary might appear in your reflection while the source emails never do.</p><p><strong>Walk away with:</strong> your report generated on the 3-month view, which is long enough to show patterns and short enough to reflect how you work now.</p><h2>How to read it: hunt the Context Rebuild Loop</h2><p>Read the report as an audit, a Spotify Wrapped skim wastes the ten minutes. The topic percentages and peak hours are trivia. The signal is repetition: the task that shows up week after week, phrased twenty different ways, each time in a fresh chat where you rebuilt the context by hand. Every one of those chats billed you a setup cost, and the report is the first tool that shows you the receipt.</p><p>Once the report is on screen, run this in the same session:</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">Based on my usage patterns, answer three questions:

1. What task have I brought to you most often in the last 3 months?
2. What context did I re-explain across those chats that never changed
   (business details, format rules, tone, constraints)?
3. If I could only systematize one recurring task this month, which one
   saves the most repeated setup, and why?</code></pre></div><p>Reflect&#8217;s own recommendation engine is pointed at the same target. <a href="https://www.engadget.com/2211304/claude-reflect-dashboard-wants-to-help-you-log-off/">Igor Bonifacic at Engadget</a> ran it against his real usage, repeated research chats tracking down executive statements for a story, and the dashboard recommended building a custom fact-checking skill, which then produced a source-and-confidence template for every claim. His verdict: &#8220;I&#8217;ll admit I found the template helpful, and it wasn&#8217;t something I would have thought to ask Claude to do on its own.&#8221; That&#8217;s the shape of the value, the report surfaces the system you didn&#8217;t know you were already running manually.</p><p><strong>Walk away with:</strong> one named task, the unchanging context behind it, and a reason it&#8217;s the first candidate.</p><h2>Promote one task from chat to system</h2><p>A repeated task can land in three homes, and the report tells you which one fits:</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">Where your repeated task goes:

Same task + same background context every time
  -&gt; a Project (context lives there once, every chat inherits it)

Same task on a calendar rhythm (weekly report, Monday review)
  -&gt; a scheduled task (it runs without you opening a chat)

Same steps applied to different inputs each time
  -&gt; a custom skill or saved workflow (the steps live once,
     the inputs change)</code></pre></div><p>The one-sitting version: take the task from your report, paste the context you&#8217;ve been re-typing into whichever home fits, and run it once from the new setup while the old chats are still fresh enough to compare against. The output quality usually jumps, because context written once and deliberately beats context re-typed from memory at 4pm on a Tuesday. My whole publishing operation runs on the third category, and the setup cost was paid one time.</p><p><strong>Walk away with:</strong> one task moved out of the Context Rebuild Loop, and the pattern for moving the next one.</p><h2>Where this wobbles</h2><ul><li><p>It&#8217;s beta, web and desktop only for now, mobile is coming later, and the time-spent metric doesn&#8217;t exist yet. Linthicum says that&#8217;s because Anthropic wasn&#8217;t tracking it internally, a metric the product team &#8220;didn&#8217;t want to maximize.&#8221;</p></li><li><p>Memory has to be on, so the analytics come bundled with a privacy tradeoff some people have deliberately declined. Sensitive conversations surface only at a high level, health-integration chats stay out entirely, and Anthropic says reflection data isn&#8217;t used for other purposes, worth reading their <a href="https://www.anthropic.com/news/reflect-with-claude">privacy note</a> yourself if that matters to your business.</p></li><li><p>The mirror is sold by the company that profits from what you see in it. <a href="https://techcrunch.com/2026/07/09/anthropics-new-claude-feature-is-quietly-selling-you-on-ai/">TechCrunch's Sarah Perez</a> makes the case that &#8220;Reflect&#8217;s larger purpose is about shaping how users think about AI itself,&#8221; and that deeper workflow integration &#8220;helps retain users and discourage them from switching to competitors&#8217; AI tools.&#8221; She points back to Gmail Meter in 2012, an analytics tool that doubled as a demonstration of how central Gmail had become. Fair on both counts. The retention motive and the usefulness both live in the same feature, so take the audit and leave the sentiment.</p></li><li><p>The insights are only as honest as the history behind them. If your real repeated work happens in tools Reflect can&#8217;t see, the report understates your loop instead of revealing it.</p></li></ul><h2>Recap</h2><p>Reflect is free, live now, and takes ten minutes: generate the 3-month report in Settings, ignore the trivia, and hunt the task you&#8217;ve re-explained twenty different ways, that&#8217;s the Context Rebuild Loop and it&#8217;s been billing you setup time all year. Promote that one task into a Project, a scheduled task, or a skill, run it once from its new home, and you&#8217;ve converted the report from a curiosity into a system.</p><p>If what the report shows you is a longer list than one task, and you want a second set of eyes on which systems to build first, that&#8217;s the work I do with operators on an AI Clarity Call at <a href="https://muddventures.com/book">muddventures.com/book</a>.</p><p>P.S. A few next steps if today&#8217;s issue was useful:</p><ul><li><p>The setup this newsletter runs on: <a href="https://blog.muddventures.com/p/work-from-anywhere-just-got-tested">Work From Anywhere Just Got Tested by Two AI Labs</a></p></li><li><p>The first Operator&#8217;s Read: <a href="https://blog.muddventures.com/p/the-operators-read-3-quiet-ai-shifts">3 quiet AI shifts, and what I'd do about each one</a></p></li><li><p>Last week&#8217;s GoHighLevel piece on scoped AI workflow edits: <a href="https://blog.muddventures.com/p/gohighlevels-ai-builder-just-learned">blog.muddventures.com/p/gohighlevels-ai-builder-just-learned</a></p></li><li><p>Operators comparing builds like these every day: <a href="https://whop.com/abra-ai/">whop.com/abra-ai</a></p></li><li><p>New here? The full archive lives at <a href="https://muddventures.substack.com/">muddventures.substack.com</a></p></li></ul><p>Back in your inbox tomorrow.</p><p>Andrew</p>]]></content:encoded></item><item><title><![CDATA[GoHighLevel's AI Builder Just Learned Scoped Edits, Here’s Two Prompts to Copy]]></title><description><![CDATA[Naming six actions inside a fifty-step workflow used to mean hoping the AI left the other forty-four alone. GoHighLevel fixed the scoping on July 11, and the two prompts below are ready to copy.]]></description><link>https://blog.muddventures.com/p/gohighlevels-ai-builder-just-learned</link><guid isPermaLink="false">https://blog.muddventures.com/p/gohighlevels-ai-builder-just-learned</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Tue, 14 Jul 2026 15:30:03 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!J1PS!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!J1PS!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!J1PS!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!J1PS!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!J1PS!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!J1PS!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!J1PS!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/e4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1241773,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/207030454?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!J1PS!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!J1PS!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!J1PS!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!J1PS!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Fe4d4be0a-a255-4b0a-b623-7ae182fd185e_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TLDR:</strong> GoHighLevel shipped a real upgrade to its Workflow AI Builder on July 11: name the exact actions or triggers you want changed, inside a workflow with fifty steps, and it touches only those steps. Below are two prompts to copy into your account, a checklist to run before trusting an AI edit on a live workflow, and where GoHighLevel&#8217;s own docs say a human still has to check the work.</p><p>GoHighLevel pushed &#8220;Targeted Edits in AI Builder&#8221; live on July 11, 2026. The line that stopped me: you can &#8220;provide exact copy for 20 of your 50 emails and AI Builder edits only those 20, leaving the rest unchanged.&#8221;</p><p>I look at every &#8220;describe it and AI builds it&#8221; feature the same way: fine, but what happens when I only want to change six things inside a workflow with forty? That question has kept most agency owners doing manual, action-by-action edits even after the builder shipped a year ago.</p><p>By the end of this you&#8217;ll have the exact scoped-edit prompt pattern to run on your own GoHighLevel workflows, plus the checklist worth running before you let AI touch anything live.</p><p><strong>What shipped:</strong> name the exact actions inside a workflow and GoHighLevel edits only those, nothing else touched.</p><p><strong>What&#8217;s below:</strong> two prompts ready to copy, one for scoped copy edits and one for bulk sender identity and pipeline swaps.</p><p><strong>The wall this solves:</strong> call it the Workflow Rebuild Wall, the point where reopening every action by hand costs more time than the automation itself ever saved.</p><p><strong>Where it still needs you:</strong> GoHighLevel&#8217;s own docs say AI can&#8217;t test a workflow yet, so the checkpoint habit still matters.</p><p>Generating a whole new workflow from a prompt is one thing, touching six actions inside an existing one without breaking the other thirty-four is a different problem entirely, and the old version of that problem is worth naming: call it the Workflow Rebuild Wall, the point where a workflow gets big enough, or gets reused across enough client accounts, that manually opening and retyping every action costs more time than the automation itself ever saved.</p><p>A pattern shows up constantly on the intake side of AI Clarity Calls: an agency owner running a dozen or more client sub-accounts, each cloned from the same template, finds out a &#8220;quick rebrand&#8221; means touching every action in every account, every time. Pacho Sanchez, who runs automation builds for clients at Agency Level 5, puts the aggregate time savings from a solid core set of GoHighLevel workflows at &#8220;a <a href="https://www.agencylevel5.com/en/blog/gohighlevel-automation-templates-workflows">minimum of 20 hours per week in manual work&#8221;</a> once they&#8217;re running. That&#8217;s one operator&#8217;s own read on his client work rather than an audited figure, and it&#8217;s the right order of magnitude for what&#8217;s at stake when editing gets faster.</p><p><strong>Who this is for:</strong> agency owners and in-house operators running client or multi-brand workflows inside GoHighLevel, especially anyone managing workflows with fifteen or more actions across more than one sub-account.</p><h2>What GoHighLevel Shipped on July 11</h2><p>Workflow AI Builder has been able to generate a full automation from a prompt for a while now (&#8220;send a welcome email series when someone fills out my contact form&#8221; and it builds the whole thing, triggers and actions included), but the part that stayed clunky was editing an existing workflow with any precision, since asking it to change &#8220;the emails&#8221; might touch six actions when you meant two, or two when you meant six.</p><p>The July 11 update fixes the scoping. Name the actions you want changed, all of them, a set number, or specific ones, and GoHighLevel&#8217;s own release notes describe the update as scoped, applying changes only to the actions you name and leaving everything else in place. It ships with three specific bulk moves: copy edits across selected actions, pipeline stage swaps across all or specific actions and triggers, and sender identity (From Name and From Email) across selected email actions.</p><p>What I look for when any vendor ships an &#8220;edit anything with one prompt&#8221; feature is where the scoping breaks. This is the first version of that promise I&#8217;ve seen from GoHighLevel that names the boundary up front instead of leaving you to find it by accident, and that&#8217;s the difference between &#8220;AI, rebuild this&#8221; and &#8220;AI, change these six things and nothing else.&#8221; For anyone who&#8217;s managed more than a handful of client workflows, that&#8217;s most of the ballgame.</p><h2>The Scoped Copy Rewrite</h2><p>The most common wall operators hit is copy, a client rebrands, an offer changes, or the review-request text has sounded like it was written in 2019 for two years running, and now you&#8217;re opening ten, twenty, or fifty individual actions to fix the wording one at a time.</p><p>Open the AI assistant inside your workflow builder (the sidebar icon in the workflow editor) and give it a scoped instruction instead of a vague one:</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">
Update the SMS copy in only the review-request and no-show-follow-up actions.
New tone: warmer, first-name personal, no exclamation points.
Keep every other action untouched. Show me a preview before saving.</code></pre></div><p>What correct output looks like: GoHighLevel returns a preview showing only those two actions with new text, everything else in the workflow marked unchanged, and once you open the workflow after saving, every wait step, tag, and other action should read the same as it did before you ran the prompt.</p><p>The failure mode is a vague version of this same prompt, something like &#8220;update my SMS messages to sound friendlier,&#8221; with no named actions, and GoHighLevel&#8217;s own support docs back this up, recommending you name the workflow settings, channel, and scope explicitly. An ambiguous prompt is when the built-in Clarifying Agent has to guess or ask follow-up questions, and a wrong guess on a live workflow is not where you want to find that out.</p><p><strong>Walk away with:</strong> a tested, scoped prompt pattern you can run the next time a client rebrand or offer change means touching more than two or three actions.</p><h2>The Bulk Sender Identity and Pipeline Stage Swap</h2><p>This one is for anyone managing more than one client, brand, or sub-account. Sender identity and pipeline stage assignments are two of the most tedious things to fix by hand across a workflow, because they&#8217;re buried inside individual action settings instead of living in one central place.</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">
Set the From Name to [Business Name] and From Email to [reply@business.com]
across every email action in this workflow.
Then move all triggers currently pointing to the "New Lead" pipeline stage
to "Qualified Lead" instead. Leave SMS and task actions unchanged.</code></pre></div><p>What correct output looks like: every email action shows the new sender identity, and every trigger or action that referenced the old pipeline stage now references the new one, and nothing in the SMS or task actions moved. GoHighLevel&#8217;s release notes describe this as &#8220;reliable scoping,&#8221; meaning changes apply only to the set you name with no spillover into actions you didn&#8217;t touch. That claim is worth testing on a duplicate workflow before you believe it on a live one.</p><p>The failure mode shows up when pipeline stage names aren&#8217;t unique or consistent across accounts, which is a common problem once an agency has cloned the same template across a dozen sub-accounts with slightly different naming along the way. If &#8220;New Lead&#8221; doesn&#8217;t exist under that exact name in a given sub-account, the instruction has nothing to grab onto, and you&#8217;ll either get a clarifying question back or, worse, silence, so check your naming conventions before running bulk swaps across multiple accounts.</p><p><strong>Walk away with:</strong> one instruction that replaces opening every email action to fix a sender name, and every trigger to fix a pipeline reference, one at a time.</p><h2>The Checklist Before You Trust It on Anything Live</h2><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:null}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">
Before you run a scoped AI edit on a live workflow:
1. Duplicate the workflow. Edit the copy, not the original, on the first pass.
2. Count the actions in scope versus total actions in the workflow.
3. Name the exact actions in your instruction, or say "all" explicitly.
4. Screenshot the before state.
5. Run the edit, then diff the result against the screenshot, action by action.</code></pre></div><p>I&#8217;d run the same duplicate-first check on any AI agent that touches a live system, whether that&#8217;s this workflow builder, a CRM record, or a client account, and the habit is simple: duplicate before you touch anything, verify the diff, and publish last, not first. It&#8217;s small, and it&#8217;s the whole difference between a five-minute fix and a support ticket you write yourself.</p><h2>Where This Still Breaks</h2><p>GoHighLevel&#8217;s own documentation for the AI Builder lists this under Beta Limitations, in their words: &#8220;manual review required, verify triggers, actions, and configurations before publishing,&#8221; and more directly, &#8220;AI cannot test workflows, perform manual tests.&#8221; The vendor is being straightforward here, scoped editing catches the &#8220;touched the wrong action&#8221; mistake, but it doesn&#8217;t catch a workflow that runs fine on the surface while the wrong logic sits underneath it.</p><p>Mari, who runs GoHighLevel training at Automated Marketer, tested four AI-built workflows in detail and landed on a fair verdict: &#8220;<a href="https://automatedmarketer.net/how-to-use-the-gohighlevel-workflow-ai-builder-build-automations-with-plain-english/">the AI builds fast but not perfectl</a>y&#8221;. Her catch is worth flagging: on a lead-qualification workflow with conditional branching, the AI &#8220;did get a little broad on how it evaluated the conditional,&#8221; setting a check to a general text-field match instead of the specific survey answer, so the branching structure was right and the condition itself just needed a person to tighten it.</p><p>That tracks with what happens any time a system gets handed more autonomy, the structure gets faster and the judgment calls still need someone watching, and it&#8217;s a smaller, lower-stakes version of a tension that&#8217;s shown up in every &#8220;AI agent, but with an approval gate&#8221; release this year. I wrote about that same shape behind Anthropic&#8217;s Cowork and OpenAI&#8217;s ChatGPT Work rollouts in the Approval Leash piece, in the <a href="https://muddventures.substack.com/">full newsletter archive</a>, and scoped editing is a narrower slice of the same idea: let the AI touch only what you named, then look before you publish.</p><p>The other honest limit is access, this lives behind an agency-level Labs toggle, so if you&#8217;re inside a white-labeled sub-account and don&#8217;t see &#8220;Build with AI&#8221; in your workflow builder, your agency provider probably just hasn&#8217;t turned it on yet, and it&#8217;s worth asking them directly. That toggle is also why a scoped edit won&#8217;t touch your published workflow the moment you run it: GoHighLevel shows a preview before anything saves, and checkpoints let you roll back with one click if something looks off after you do save it.</p><h2>Recap</h2><p>GoHighLevel&#8217;s AI Builder can now edit only the actions you name inside an existing workflow, instead of forcing a full rebuild or a vague full-workflow edit. The two prompts above cover the most common uses, scoped copy rewrites and bulk sender identity or pipeline stage swaps, and running both on a duplicate first, with the checklist keeping the edit honest, is the whole habit. GoHighLevel&#8217;s own beta docs are clear that AI still doesn&#8217;t test the logic for you, so the manual review step doesn&#8217;t disappear just because the edit got faster, and the same holds if you&#8217;re not on GoHighLevel at all: Zapier&#8217;s Copilot and HubSpot&#8217;s workflow AI are moving the same direction, and naming your scope in plain language instead of a vague instruction is the control lever regardless of platform.</p><p>The operators who feel this fastest are the ones running the same core workflow across a dozen or more client sub-accounts, each with its own branding, sender identity, and pipeline names, and that&#8217;s the spot where the old way, open every action in every sub-account one at a time, turns a ten-minute rebrand into a lost afternoon. A solo operator running three workflows probably isn&#8217;t there yet, the time savings show up hardest once a single workflow crosses ten or fifteen actions or gets cloned across accounts, and below that line manual editing is still fast enough that this isn&#8217;t the real bottleneck. Abra AI&#8217;s workflow-building skill files cover this same scoped-edit pattern across other tools as well, GoHighLevel included, if you want the habit to travel with you wherever you build, worth a look at<a href="https://whop.com/abra-ai/"> whop.com/abra-ai </a>if that&#8217;s the wall you&#8217;re staring at right now.</p><p>I&#8217;m curious where the line falls for most of you. Reply with your workflow&#8217;s action count if you want to find out whether you&#8217;ve hit the Rebuild Wall yet.</p><p>P.S. A few places to go next:</p><ul><li><p>The full newsletter archive, including the Approval Leash issue on keeping a human checkpoint on AI agents: <a href="https://muddventures.substack.com/">muddventures.substack.com</a></p></li><li><p>Abra AI&#8217;s workflow-building skill files, if the scoped-edit habit is one you want to carry across tools: <a href="https://whop.com/abra-ai/">whop.com/abra-ai</a></p></li><li><p>If the real bottleneck in your funnel is what happens after a call gets booked: <a href="https://showtime.muddventures.com">showtime.muddventures.com</a></p></li><li><p>If the fix needs a full systems audit instead of one workflow prompt: <a href="https://muddventures.com/book">muddventures.com/book</a></p></li></ul><p>Back in your inbox tomorrow.</p><p>Andrew</p>]]></content:encoded></item><item><title><![CDATA[Is the Apple-OpenAI Lawsuit Really About Trade Secrets, or About Who Controls What Comes After the iPhone?]]></title><description><![CDATA[A lawsuit dressed up as corporate espionage is really a fight over who owns the next twenty years of computing. Here's my read on it, plus the trade secret law behind the case, explained plainly.]]></description><link>https://blog.muddventures.com/p/is-the-apple-openai-lawsuit-really</link><guid isPermaLink="false">https://blog.muddventures.com/p/is-the-apple-openai-lawsuit-really</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Mon, 13 Jul 2026 16:21:17 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!Ieoq!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!Ieoq!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!Ieoq!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!Ieoq!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!Ieoq!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!Ieoq!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!Ieoq!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1615319,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/206873329?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!Ieoq!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!Ieoq!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!Ieoq!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!Ieoq!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F7002bb35-a68a-4c4d-a7ba-9f2cf339386f_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TL;DR:</strong> Apple sued OpenAI, its hardware unit io Products, and two former Apple employees on July 10 for trade secret theft. The complaint reads like a thriller: a forgotten laptop, a &#8220;LOL&#8221; text, a Chief Hardware Officer accused of running the operation. The more interesting story is what happens eighteen months before a lawsuit like this gets filed, and why California&#8217;s unusual employment law makes trade secret claims the sharpest tool a company like Apple has left. Below: what the complaint actually alleges, a plain-English primer on how trade secret law works and why it exists, and my own opinion on who&#8217;s right (I&#8217;m Team Apple, with caveats). Plus what this signals for anyone whose business depends on a platform relationship that could sour.</p><p>Apple filed a 40-page lawsuit against OpenAI on July 10. Most of the coverage stayed on the parts that read like a thriller: a forgotten laptop, a text message that opens with &#8220;LOL,&#8221; a Chief Hardware Officer accused of running the operation from the top. I read the complaint and the reporting around it, and I think the trade-secret claims, real as they might be, are the smaller story here.</p><p>By the end of this you&#8217;ll have my actual read on what this lawsuit signals about where the AI industry is headed. You&#8217;ll also have enough working knowledge of trade secret law to follow the case yourself as it plays out in court over the next year. Not a checklist, more an executive-level take on a fight that&#8217;s bigger than the two companies filing it.</p><p><strong>What&#8217;s alleged:</strong> Apple says former employees, including its ex-hardware chief, systematically pulled confidential product information for OpenAI&#8217;s hardware push.<br><strong>What most coverage missed:</strong> this lawsuit lands eighteen months after Apple and OpenAI were partners, and six months after Apple quietly picked Google over OpenAI for its own AI.<br><strong>The legal backdrop most readers don&#8217;t know:</strong> California bans employee non-competes almost entirely, which is exactly why trade secret law carries so much weight in cases like this one.<br><strong>My take:</strong> a company doesn&#8217;t sue a rival it isn&#8217;t worried about. This is a confidence problem dressed up as a legal one.<br><strong>Where this leaves operators:</strong> a reminder that vendor alliances in this industry have a shorter shelf life than the contracts built on top of them.</p><h2>A quick primer: what actually makes something a trade secret, and why this is a California specialty</h2><p>Before getting into what Apple alleges, it&#8217;s worth understanding the legal machinery underneath it, because it explains why this kind of lawsuit is so common in Silicon Valley specifically.</p><p>Under both the federal Defend Trade Secrets Act (signed into law in 2016) and California&#8217;s own trade secret statute, something only counts as a trade secret if it clears two bars. First, it has to derive real economic value from not being publicly known, meaning a competitor would actually benefit from having it. Second, the owner has to have taken reasonable steps to keep it secret: locked file servers, restricted access, NDAs, exit interviews that revoke credentials. A hardware roadmap, an unreleased prototype&#8217;s design specs, or internal manufacturing test procedures all clear that bar easily. A general skill someone picked up on the job, like knowing how to run a good design review, does not.</p><p>To win, Apple has to prove three things: a trade secret existed, Apple took reasonable measures to protect it, and it was acquired or used through improper means rather than independent invention. That last piece is where the &#8220;LOL&#8221; text message and the unreturned laptop actually matter. They aren&#8217;t colorful filler for a filing. They&#8217;re the evidence Apple needs to clear the &#8220;improper means&#8221; bar.</p><p>Here&#8217;s the part that makes California specific. Unlike most states, California voids nearly all employee non-compete agreements outright, under Business and Professions Code Section 16600. A 2024 update went further, making it a separate legal violation for an employer to even include one in a contract. Practically, that means a company like Apple cannot stop a departing engineer from joining a direct competitor the next day, no matter how senior or sensitive their role was. Trade secret law is what&#8217;s left. It&#8217;s the one legal lever California employers have to stop specific confidential material from walking out the door. They have almost no power to stop the person who holds that knowledge from walking out the door too. That distinction, information versus knowledge, is the entire game these lawsuits are fought over.</p><p>This isn&#8217;t the first time Silicon Valley has had this exact fight. In 2017, Waymo sued Uber and engineer Anthony Levandowski, alleging he downloaded roughly 14,000 confidential files on self-driving LiDAR technology before leaving to found a startup Uber then acquired. That case settled five days into trial, with Uber handing Waymo about $245 million in equity. Levandowski was later criminally charged, pleaded guilty to one count of trade secret theft in 2020, and was sentenced to 18 months before receiving a pardon in the final hours of the first Trump administration. Apple v. OpenAI is following a similar shape: a departure, a competitor building the same category of product, and a filing that leans hard on internal messages as proof of intent. It&#8217;s too early to know if the outcome looks anything like Waymo&#8217;s, but the legal template is a familiar one.</p><h2>What Apple&#8217;s complaint alleges</h2><p>Apple sued OpenAI, its hardware unit io Products, and two former Apple employees, Tang Tan and Chang Liu, in federal court in the Northern District of California, p<a href="https://techcrunch.com/2026/07/10/apple-sues-openai-over-alleged-trade-secret-theft/">er TechCrunch&#8217;s report on the fili</a>ng. Tan spent 24 years at Apple, most recently as VP of product design for the iPhone and Apple Watch, before leaving to help found io Products with Jony Ive. OpenAI acquired io for $6.5 billion last year and made Tan its Chief Hardware Officer. io is named in the suit. Ive is not.</p><p>The most vivid detail belongs to Liu, an eight-year Apple engineer who Apple alleges never returned his company laptop and later found a bug giving him ongoing access to Apple&#8217;s internal file servers. He messaged a former colleague: &#8220;LOL, I found out I can access the [network storage], so funny,&#8221; <a href="https://fortune.com/2026/07/11/openai-engineers-legal-fight-apple-ai-product-poaching/">Bloomberg&#8217;s Mark Gurman reported via Fortune</a>. Apple&#8217;s filing calls the pattern systemic: &#8220;at every level, from members of its Technical Staff to its Chief Hardware Officer&#8230; OpenAI has been stealing Apple&#8217;s trade secrets and confidential information,&#8221; and describes OpenAI&#8217;s hardware business as &#8220;rotten to its core by its illegal reliance on misappropriated trade secrets.&#8221; OpenAI&#8217;s response, through a spokesperson: the company has &#8220;no interest in other companies&#8217; trade secrets&#8221; and is &#8220;focused on building innovative technology that empowers people everywhere.&#8221; More than 400 former Apple employees have joined OpenAI&#8217;s hardware division, per Fortune&#8217;s reporting.</p><p>That last number matters for reading the case correctly. Apple isn&#8217;t alleging that 400 people stole something. Under the legal standard above, hiring hundreds of people away from a competitor is completely legal, and expected in an industry this small. The complaint is narrowly built around specific documents and specific system access tied to two named individuals, which is exactly the kind of surgical claim that survives a motion to dismiss. A broader claim about talent flight alone wouldn&#8217;t.</p><h2>The part I think matters more</h2><p>Rewind eighteen months. In 2024, Apple and OpenAI announced a partnership putting ChatGPT inside the iPhone. That was the headline at the time: two of the most valuable brands in tech, aligned. Then OpenAI bought Jony Ive&#8217;s design studio and started building hardware of its own, explicitly aimed at the category Apple has owned since 2007. Then, in January of this year, Apple quietly signed a $1-billion-a-year deal to run Siri and Apple Intelligence on a custom Google Gemini model instead. ChatGPT became a peripheral option rather than the engine under the hood. Six months after that snub, Apple is in federal court accusing its former partner&#8217;s leadership of systematically raiding its talent and its confidential product data.</p><p>I don&#8217;t think the interesting question is whether Liu downloaded files he shouldn&#8217;t have. I think the interesting question is why a company as controlled and press-shy as Apple decided a public lawsuit was the right move here, instead of a quiet settlement or an even quieter internal fix. Companies with real leverage tend to negotiate in private. Companies that feel a genuine platform threat go public, because the lawsuit itself is a signal, to investors, to talent, and to OpenAI, that Apple is willing to fight in the open. Sam Altman has said openly that OpenAI wants to build the device that replaces the smartphone. Apple has spent two decades making sure nothing replaces the smartphone. That&#8217;s the actual fight. The trade secrets are the evidence, not the motive.</p><h2>Where I actually stand on this</h2><p>I&#8217;ll say the quiet part. I&#8217;m Team Apple here. Not blindly, the company&#8217;s a little past its peak right now, but it&#8217;s still a great company through and through. I&#8217;ve read both the Steve Jobs biography and Jony Ive&#8217;s. The thing that comes through in both is how obsessive Apple has always been about its own design language, doing things Apple&#8217;s way. It refuses to ship something that doesn&#8217;t feel like Apple, even when that costs time or money. That obsession is the actual moat, not the chips, not the retail stores, the taste.</p><p>Say the allegations hold up. Say trade secrets really did make their way into OpenAI&#8217;s hardware roadmap with help from people who used to sit in Apple&#8217;s own design reviews. I don&#8217;t think that&#8217;s close. That&#8217;s not competition. That&#8217;s copying the homework with the answer key someone walked out the door with, and Apple has every right to make an example out of it in open court.</p><p>I also don&#8217;t think that cancels out the deterrence angle I raised above, and I don&#8217;t see that as a knock against Apple for having it. Wanting to slow down a well-funded rival that&#8217;s actively poaching your design team is a completely valid business reason to litigate. I&#8217;d bet it plays a healthy role in the decision to sue right alongside the actual IP concern. Both things can be true at once. Apple can be genuinely wronged and strategically motivated to make this loud.</p><h2>The honest tradeoffs</h2><p>This is one side&#8217;s complaint, filed by the party with the most to lose from OpenAI succeeding in hardware. OpenAI denies wrongdoing, and none of the named individuals have responded publicly as of this writing. Nothing here is proven in court, and complaints are written to tell the most damning possible version of events.</p><p>Talent moving between competitors is normal, legal, and how this industry has always worked, and California&#8217;s ban on non-competes exists specifically to protect that mobility. Apple&#8217;s complaint is careful to draw the line around specific documents and system access, not the plain fact that hundreds of its former employees now work at OpenAI. Courts have not been shy about dismissing trade secret claims that amount to &#8220;our former employee is now good at their job somewhere else.&#8221; If Apple&#8217;s evidence turns out to be thinner than the filing suggests, this case could go the same way.</p><p>My read on the motive, the confidence-problem framing, is an interpretation, not a fact Apple has confirmed. Apple&#8217;s public statement sticks to protecting its IP. I&#8217;m connecting dots they haven&#8217;t officially connected, and my own Team Apple lean means I&#8217;m reading the ambiguous parts of this story in Apple&#8217;s favor. A reasonable person could read the same eighteen-month timeline and conclude Apple is using a legitimate legal claim to slow down a competitor it fumbled the partnership with, rather than the other way around.</p><h2>Where I land</h2><p>Strip away the AI hardware angle and this is a familiar story. A partnership that made sense when both companies needed something from each other stopped making sense the moment their roadmaps started pointing at the same prize. Call it the Alliance Half-Life, the shrinking window between &#8220;strategic partner&#8221; and &#8220;named defendant&#8221; once two companies realize they&#8217;re actually building toward the same platform. Eighteen months, in this case. I&#8217;d expect that number to keep shrinking as more AI labs decide the real prize isn&#8217;t the model, it&#8217;s the device people carry.</p><p>If you&#8217;re building any part of your business on top of a single AI vendor&#8217;s platform, a plugin ecosystem, an API, a partnership that felt stable last year, this is worth sitting with. Not because you&#8217;ll get sued. Because alliances in this industry are proving to be a lot more temporary than the roadmaps built on top of them assume.</p><p>I&#8217;ll be watching what OpenAI actually files in response, and whether this becomes the template other AI labs and Big Tech incumbents start reaching for as their roadmaps collide over the next few years. If it is the template, this is the first of many.</p><p>If you want to talk through how dependent your own stack is on any single AI vendor&#8217;s continued goodwill, that&#8217;s a real conversation worth having on an AI Clarity Call, grab a slot at <a href="https://muddventures.com/book">muddventures.com/book.</a></p><p>Reply and tell me if you&#8217;re Team Apple on this one too, or if you think I&#8217;m cutting them too much slack.</p><p>Andrew</p><p>P.S. If the theme of keeping control over what you don&#8217;t fully own resonates, that&#8217;s the same thread running through <a href="https://blog.muddventures.com/p/work-from-anywhere-just-got-tested">the Approval Leash post from last week</a>, just applied to a different kind of dependency. Come hang out in the <a href="https://whop.com/abra-ai/">Abra AI community</a> if you want more of this kind of breakdown. And if this is the first issue of mine you&#8217;ve read, the archive lives at <a href="https://muddventures.substack.com/">muddventures.substack.com.</a></p>]]></content:encoded></item><item><title><![CDATA[Work From Anywhere Just Got Tested by Two AI Labs, 48 Hours Apart]]></title><description><![CDATA[I ran it from 6,000 miles away for 20 days straight. This week, two AI labs made it official.]]></description><link>https://blog.muddventures.com/p/work-from-anywhere-just-got-tested</link><guid isPermaLink="false">https://blog.muddventures.com/p/work-from-anywhere-just-got-tested</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Fri, 10 Jul 2026 16:00:08 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!ZXCQ!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!ZXCQ!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/aaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1460873,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/206462046?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!ZXCQ!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2Faaf8fc9d-2054-4bd5-8eb5-1d5aa2a75f72_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p>TLDR: On July 7, Anthropic put Claude Cowork on mobile and web, so a background task can run while your laptop is closed. On July 9, OpenAI answered by merging Codex into ChatGPT and launching ChatGPT Work, an agent that turns a goal into a finished spreadsheet, deck, doc, or website instead of a chat reply. Both are usage-metered, not flat-fee. Pick one workflow you already know cold, run it with approvals switched on, and measure it against your actual hours before you schedule anything recurring.</p><p>&#8220;Work from anywhere&#8221; has been a slogan for a decade. This week it became a testable claim. This newsletter got written by an agent running inside Claude Cowork, on a schedule, before I was at my desk. That&#8217;s not a metaphor. I&#8217;ve been running Cowork through Dispatch as a daily tool for about two and a half, three months now, and it&#8217;s the literal infrastructure behind the issue you&#8217;re reading right now. So when Anthropic pushed Cowork out to phones and browsers this week, I wasn&#8217;t reading a press release about someone else&#8217;s tool. I was reading about the thing that already runs part of my business.</p><p>Then two days later, OpenAI answered.</p><p>By the end of this you&#8217;ll know what each launch actually ships, not the marketing description. You&#8217;ll know which one to pilot first depending on how technical your team is, and the one guardrail I&#8217;d never turn off no matter which agent you pick.</p><p>A few things worth flagging up top:</p><ul><li><p><strong>Neither of these is a better chatbot.</strong> Both are agents that take a goal and work for hours without you watching. Both hand back a finished artifact: a sheet, a deck, a doc, a live web page.</p></li><li><p><strong>Both bill by usage, not by seat.</strong> Nobody published per-task prices. Budget the first week like an experiment, not a subscription.</p></li><li><p><strong>The tester numbers in both launches come from the vendors.</strong> Real names, real workflows, real self-reported results. Still the vendor&#8217;s own case studies. Treat them as hypotheses, not proof.</p></li></ul><p>Who this is for: any operator running a small team who already delegates one recurring, multi-step task to a person. A report, a competitor scan, a monthly deck. You&#8217;re deciding whether to hand the first draft of that task to an agent instead.</p><h2>The 48 hours that mattered</h2><p>Anthropic&#8217;s move landed first. Claude Cowork launched as a desktop app back in January, built to feel like Claude Code without the code. On July 7,  <a href="https://techcrunch.com/2026/07/07/the-coding-agent-wars-are-spilling-into-the-rest-of-the-office-claude-cowork/">Anthropic expanded it to web and mobil</a>e for Max subscribers. The pitch is simple: start a task from your desk, get a status update on your phone, and come back to a finished output even if your laptop is closed the whole time. Anthropic&#8217;s own example: &#8220;Set Monday&#8217;s client prep for 6am: Claude works through the email threads, transcripts, and recent news, builds the briefing doc, and leaves the follow-up email drafted but unsent. Review it over coffee.&#8221;</p><p>I can vouch for the part about the laptop staying closed. I ran Dispatch on my phone for 20 straight days this summer on a trip to Switzerland, with the host device where the actual work was running sitting 6,000 miles away back home. Zero issues the entire trip. That&#8217;s the pitch Anthropic is making to everyone else this week. I&#8217;d already tested it at the most inconvenient distance I could put between me and my desk.</p><p>Two days later, on July 9, OpenAI answered with ChatGPT Work, an agent built on the newly released GPT-5.6 with Codex technology built in. Same week, OpenAI merged the standalone Codex app into one ChatGPT desktop app. Chat, Work, and Codex now live in one place, on every plan including Free. OpenAI also announced it&#8217;s sunsetting the standalone Atlas browser, folding those agentic-browsing features into ChatGPT itself.</p><p>Two labs, the same week, converging on the same bet: the fight isn&#8217;t about who has the sharper chatbot anymore. It&#8217;s about who owns the place where the actual work gets done.</p><h2>What &#8220;ships finished work&#8221; actually means</h2><p>Regular ChatGPT and regular Claude answer a question in the moment. Both of these products do something different. You give them an outcome, not a prompt, and they gather context from your connected tools, plan the steps, and keep working, sometimes for hours, without you sitting there.</p><p>Cowork&#8217;s early data backs up what that looks like in practice. Anthropic sampled 1.2 million anonymized Cowork sessions across more than 600,000 organizations over two weeks in May. The biggest category, at 33.4%, wasn&#8217;t coding. It was what Anthropic calls &#8220;business process operating&#8221;: pulling scattered updates into one report, building onboarding checklists, reconciling spreadsheets, the stuff that&#8217;s part of a lot of jobs but rarely anyone&#8217;s actual title. Content creation and copywriting came in second at 16.4%. Software development, the thing Cowork&#8217;s ancestor was built for, was 8.7%.</p><p>ChatGPT Work is chasing the same territory from the coding side. OpenAI says Codex already had 5 million weekly users and, more tellingly, more than 1 million people were using it for work that had nothing to do with software before Work even had a name. The new agent connects to more than 1,400 plugins and runs in Plan mode, proposing a step-by-step plan you approve before it touches anything. Its output format of choice is now &#8220;Sites,&#8221; a live web page instead of a static document. Sites update themselves as the underlying data changes.</p><h2>Real numbers from real early testers</h2><p>The launch materials name real people at real companies, which is more than most product announcements bother with. Worth reading with one caveat attached: these are vendor-arranged case studies, not independent audits.</p><p>Angela Ferrante, Head of Enterprise Marketing at Zapier, used ChatGPT Work to trace lead journeys across HubSpot, Gong, and email. Before, inspecting a single lead properly took her team 35 to 45 minutes, which means most leads never got that inspection. She put the result plainly: &#8220;That helped us identify and hand off seven figures in pipeline every month to sales.&#8221;</p><p>Nathan Bolt, Head of Digital Products at Virgin Atlantic, ran competitor customer-journey benchmarking that used to take his team weeks per cycle. His read: &#8220;A competitor analysis cycle that would normally take weeks now takes hours, helping us move from insight to product decisions much faster.&#8221;</p><p>On the Cowork side, one small business owner described it this way in an independent review. &#8220;I&#8217;ve been using Claude Cowork daily, and this is the first time an AI tool has genuinely changed how my workday flows,&#8221; they wrote. &#8220;Task queueing is the real unlock.&#8221;</p><p><strong>Walk away with:</strong> three real, named examples of the exact workflow class both products are best at right now: recurring reporting, lead or customer-journey triage, and the monthly deck nobody wants to build by hand.</p><h2>Pick your pilot workflow</h2><p>The move: point one of these agents at one recurring task you already know intimately, not a reorganization of how your team works.</p><p>Here&#8217;s the filter I&#8217;d run before picking a first workflow:</p><div class="highlighted_code_block" data-attrs="{&quot;language&quot;:&quot;plaintext&quot;,&quot;nodeId&quot;:&quot;b30169b5-0074-4a29-9afd-df0c43138cc8&quot;}" data-component-name="HighlightedCodeBlockToDOM"><pre class="shiki"><code class="language-plaintext">FIRST AGENT PILOT, PICK ONE

1. Do you already know the current hours-per-cycle for this task?
   If you can't name the number today, you won't be able to
   measure whether the agent actually helped. Pick a task you
   can time first.

2. Is the output a document, sheet, deck, or web page,
   not an irreversible action?
   Good pilot: draft a competitor report, build a briefing doc,
   triage leads into a dashboard.
   Bad pilot: anything that sends money, deletes records,
   or emails a customer without a human reading it first.

3. Which team runs it, and which product matches?
   Non-technical team, wants pre-built workflows out of the box
   -&gt; Claude Cowork (more on-rails, vertical bundles ship configured).
   Team comfortable connecting plugins and building its own flow
   -&gt; ChatGPT Work (more flexible, steeper setup).

4. Can you afford a week of unmeasured usage while you find out
   what the task actually costs?
   Neither product publishes per-task pricing. Week one is a
   measurement week, not a rollout week.</code></pre></div><p>What the correct output looks like: a finished draft you&#8217;d have been comfortable handing a new hire on their first week. Plus a real number for how long the agent took against how long the task used to take you.</p><p>The failure mode: skipping the timing step, loving the output, and scheduling the workflow to run daily before you know what a single run costs against your plan&#8217;s usage. The exact risk both product pages warn about in the fine print: a team burning a month&#8217;s included usage in the first week of enthusiastic use.</p><p><strong>Walk away with:</strong> a four-question filter for choosing your first agent pilot, and a rule for what a pilot is allowed to touch before you&#8217;ve measured it once.</p><h2>The guardrail I&#8217;d never turn off</h2><p>Both products ship with the same control: a plan you approve before work starts, and check-ins you configure along the way. Call it<strong> the Approval Leash</strong>: the agent can run for hours and touch a dozen tools, but nothing it produces goes anywhere real until a person signs off.</p><p>OpenAI&#8217;s own language undercuts a little of its confidence here. The company reports that in adversarial red teaming, its auto-review system &#8220;blocked 100% of attempts to extract protected data, including attacks the reviewing model had not seen during training.&#8221; That&#8217;s a genuinely strong red-team result. It isn&#8217;t a guarantee that a production system can never leak client data. Keep your own approval gate on anything that touches a client system anyway, regardless of what the safety layer claims.</p><p>The honest tradeoff on the Anthropic side runs the other direction: cost, not safety. Anthropic users have spent a chunk of 2026 complaining about high token burn on the newer frontier models. An agent that works for hours unattended is exactly the kind of workload that turns a complaint into a real invoice. Watch the usage meter the same way you&#8217;d watch a contractor&#8217;s timesheet.</p><h2>Recap</h2><p>Anthropic put Claude Cowork on mobile and web July 7. OpenAI answered July 9 by merging Codex into ChatGPT, launching ChatGPT Work, and starting to sunset the standalone Atlas browser. Both products take a goal instead of a prompt and hand back a finished artifact instead of a chat reply. Both bill by usage, not by seat, and neither has published what a task actually costs. &#8220;Work from anywhere&#8221; stopped being a slogan and started being a feature two different companies shipped in the same 48 hours. Pick one recurring task you already know the hours on, run it with approvals on, and keep the Approval Leash attached no matter which agent does the work.</p><h2>FAQ</h2><p><strong>Do I need a developer on staff to use either of these?</strong> No. Cowork ships pre-configured vertical workflows aimed at non-technical operators. ChatGPT Work leans on its plugin directory and Plan mode, which takes a little more setup but doesn&#8217;t require code.</p><p><strong>Which one is cheaper?</strong> Neither company has published per-task pricing. Both meter usage against your plan&#8217;s included allowance the way Codex already did, so the honest answer is: you find out in week one, not before.</p><p><strong>Can I run this on client data?</strong> Both offer admin controls and audit trails built for that use case. Keep a human approval step on anything touching client systems regardless of the vendor&#8217;s own safety claims.</p><p><strong>What if my team already uses Zapier or Make for automation?</strong> Nothing here replaces a working automation. These agents are best at the one-off or recurring analytical and drafting work a person currently owns, not at wiring two apps together.</p><p>Reply and tell me which one you&#8217;re piloting first, Cowork or ChatGPT Work. I&#8217;ll compare notes with you directly.</p><p>Andrew</p><p>P.S. If you want help wiring your first agent pilot into the tools you already run, the skill-building side of this conversation happens inside <a href="https://whop.com/abra-ai/">whop.com/abra-ai</a>. Related reading: <a href="https://blog.muddventures.com/p/google-rewired-its-productivity-suite">Google Rewired Its Productivity Suite </a>on the same &#8220;AI moves into where work happens&#8221; pattern from a different vendor. More at <a href="https://muddventures.substack.com/">muddventures.substack.com</a>.</p><p></p>]]></content:encoded></item><item><title><![CDATA[Canva 2.0: Will it be the biggest winner or the biggest loser of 2026? ]]></title><description><![CDATA[Canva took over the complete ads ecosystem on June 25: build, publish, report, and refresh winners without leaving the tool.]]></description><link>https://blog.muddventures.com/p/run-the-30-minute-canva-test-that</link><guid isPermaLink="false">https://blog.muddventures.com/p/run-the-30-minute-canva-test-that</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Thu, 09 Jul 2026 16:31:10 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!1t12!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!1t12!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!1t12!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!1t12!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!1t12!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!1t12!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!1t12!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1339549,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/206318196?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!1t12!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!1t12!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!1t12!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!1t12!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F037aca97-947c-4dec-8981-8b8fa79fd7e9_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p>Canva picked the biggest advertising stage in the world to stop calling itself a design tool. On June 25 at Cannes Lions, George Howes, who runs Canva Grow, walked through the 2.0 launch. AI ad creation for static and video, direct publishing to Meta, TikTok, and LinkedIn, one dashboard for every live campaign, cross-platform reports, and automatic new variations generated from whatever&#8217;s already winning. All of it inside the Canva login your team already has.</p><p>I spent years inside the version of this workflow that Canva just collapsed. At the agency I co-founded, creative lived in the design tool or with a designer, and publishing lived in separate ad managers. Reporting lived in a spreadsheet somebody updated on Fridays, usually late. The gap between &#8220;this ad is fatiguing&#8221; and &#8220;here&#8217;s the new batch&#8221; was measured in weeks, and agencies bill for that gap every month.</p><p>That&#8217;s why I&#8217;d put this launch in the same file as <a href="https://blog.muddventures.com/p/meta-just-gave-media-buyers-a-command">the Meta Ads CLI story from April</a>: the tools that used to justify hiring a vendor keep getting pulled into software an operator already pays for. Same pattern I named <a href="https://blog.muddventures.com/p/notion-just-made-zapier-optional">Zapier Optionality</a> when Notion did it to automation tools. This time it&#8217;s the ad workflow.</p><p>By the end of this you&#8217;ll know what shipped, the 30-minute test I&#8217;d run with one connected ad account, and how to tell whether your creative retainer still earns its line on the invoice.</p><p><strong>What&#8217;s ahead:</strong></p><ul><li><p><strong>What actually shipped June 25</strong> and the four acquisitions Canva stitched together to build it.</p></li><li><p><strong>The Refresh Retainer</strong> gets a name, because the thing agencies charge monthly for just became two product features.</p></li><li><p><strong>The 30-minute test</strong> with a copy-paste checklist and a verdict key for your next renewal conversation.</p></li><li><p><strong>What&#8217;s still shaky</strong>, including the dependency question nobody at Cannes wanted to raise on stage.</p></li></ul><h2>What actually shipped on June 25</h2><p>Here&#8217;s the concrete list, pulled from <a href="https://www.businesswire.com/news/home/20260625870253/en/Introducing-Canva-Grow-2.0-Create-Launch-and-Optimize-Ads-in-One-Place">Canva&#8217;s announcement</a>:</p><p><strong>AI ad creation.</strong> Drop in a website link and Grow pulls business info, product visuals, brand colors, and audience signals to generate static and video ads. A new Magic Layers integration exports AI-generated ads into the regular Canva editor as editable designs, so a human can fix what the machine got wrong.</p><p><strong>Bulk Publish.</strong> Push ads directly from Canva to Meta, TikTok, and LinkedIn in one workflow. This works for ads made in Grow, in the Canva editor, or uploaded from anywhere else. No more exporting, re-uploading, and re-entering the same campaign three times.</p><p><strong>Launch Dashboard.</strong> One view of every campaign running across all three platforms, so the &#8220;log into three ad managers to see what&#8217;s live&#8221; morning ritual goes away.</p><p><strong>Multi-Platform Ad Insights and Reports.</strong> Performance reporting across Meta, TikTok, and LinkedIn in one place, with custom reports you can share without touching a spreadsheet.</p><p><strong>AI Ad Tagging.</strong> Ads get automatically analyzed and tagged so you can see which themes, hooks, formats, and creative elements drive results across big ad sets. Worth flagging now: this one is only available on Canva Business and Enterprise plans. Canva put the smartest layer behind the highest tiers.</p><p><strong>Automatic Refresh Generation.</strong> Grow reads live performance data and generates new ad concepts and variations from what&#8217;s already working. This is the feature the rest of the launch exists to feed.</p><p>Rollout started June 25 across North America, Australia, and the UK, with more markets coming over the next few months. Canva&#8217;s product page says Grow is available on all plans including Free, with limits on connected ad accounts, <a href="https://vijaytalksai.com/canva-grow-2-0-explained-ai-ads-publishing-performance/">per the hands-on breakdown at Vijay Talks AI</a>.</p><p>The backstory matters as much as the feature list. Canva bought MagicBrief in June 2025 for the creative analytics, the tech Howes founded. Then it folded in Ortto for customer data, SimTheory for AI assistants, Doohly for digital out-of-home, and MangoAI for creative optimization, four acquisitions in twelve months. Canva&#8217;s B2B revenue passed $500 million, 95% of the Fortune 500 uses it somewhere, and <a href="https://a16z.com/100-gen-ai-apps-6/">a16z ranks it the third most-used AI platform in the world</a>. This launch is that whole shopping spree turned into one product.</p><h2>The Refresh Retainer just became a feature</h2><p>There&#8217;s a name for what got commoditized here, so let&#8217;s give it one: <strong>The Refresh Retainer</strong>. It&#8217;s the monthly fee where the real deliverable is a loop. Watch the performance data, figure out which creative elements are carrying the account, produce new variations of the winners, and ship them before the old ones fatigue. Plenty of agency and freelancer retainers are 80% this loop, dressed up as strategy.</p><p>Howes said the quiet part on stage: &#8220;For too long, creative and performance have lived in separate systems. With Canva Grow 2.0, businesses can go from generating engaging ads to publishing them across multiple platforms, seeing what&#8217;s working, and automatically refreshing creative based on what&#8217;s actually driving results.&#8221;</p><p>AI Ad Tagging plus Automatic Refresh Generation is the Refresh Retainer as software. The tagging layer answers &#8220;what&#8217;s working,&#8221; and the refresh layer produces the next batch from it. Keith Kirkpatrick at Futurum <a href="https://futurumgroup.com/insights/canva-grow-20-puts-ad-creation-launch-and-optimization-into-a-single-ai-workflow/">called the launch</a> &#8220;a direct assault on the inefficiencies of fragmented marketing technology stacks,&#8221; and his firm&#8217;s survey of 830 software buyers found 44% now rank AI capabilities as a top purchase criterion. The demand side is already leaning this way.</p><p>Early operator reaction is warm. James Graver, who runs brand and portfolio marketing at AG1 and worked with Canva as a design partner on Grow, said &#8220;we&#8217;re excited about what it means for our creative workflows as we scale into new markets.&#8221; AG1 is the profile this serves: heavy paid social, constant creative turnover, more markets than the creative team can hand-feed.</p><p>The operators this hits hardest run $2,000 to $50,000 a month in paid social and pay somebody, an agency, a freelancer, or a full-time hire, mostly to keep fresh variations flowing. That line item deserves a fresh look this quarter.</p><h2>The 30-minute test I&#8217;d run this week</h2><p>Don&#8217;t take Canva&#8217;s word for any of this, and don&#8217;t take mine. The whole thing is testable with one ad account and half an hour. Here&#8217;s the exact sequence:</p><pre><code><code>1. Go to canva.com/canva-grow and connect ONE ad account (Meta first,
   it's your densest data)
2. Open Launch Dashboard and confirm your active campaigns appear
3. Open Insights &amp; Reports, filter to the last 90 days, sort by cost
   per result
4. Write down your top 2 ads and your 2 most fatigued (frequency
   climbing, CTR sliding)
5. Run Automatic Refresh Generation on your best ad
6. Export one variation to the Canva editor and fix whatever the AI
   got wrong
7. Publish nothing yet. Put the batch next to the last refresh you
   paid for and compare quality, speed, and cost</code></code></pre><p>What a good result looks like: the dashboard numbers reconcile with what Ads Manager shows when you spot-check two or three metrics. The generated variations keep your actual product shots and brand kit intact, and at least one is something you&#8217;d put real budget behind.</p><p>What a bad result looks like, and what it means: if the reporting numbers don&#8217;t match Ads Manager, trust Ads Manager and treat Grow&#8217;s reporting as directional until it reconciles. If the variations come back generic, feed the brand kit and real product photos in and regenerate once. Still generic? The creation layer isn&#8217;t ready for your account yet, so use Grow for publishing and reporting only and keep your current creative source. That&#8217;s still a real win, it just changes which invoice you&#8217;re re-examining.</p><p>Then run the verdict:</p><pre><code><code>KEEP PAYING   -&gt; the refresh batch is clearly worse than what your
                 current creative source delivers
RENEGOTIATE   -&gt; the batch is 80% as good and your retainer's main
                 deliverable is variations of existing winners
GO HYBRID     -&gt; keep humans on new concepts and brand work, move
                 refresh production into Grow</code></code></pre><p><strong>Walk away with:</strong> a side-by-side read on whether your creative refresh loop is still a retainer or already a button.</p><h2>What&#8217;s still shaky</h2><p>The launch coverage was glowing, so here&#8217;s the other side of the ledger.</p><p><strong>The single-door problem.</strong> Shashi Bellamkonda, an analyst who has <a href="https://www.shashi.co/2026/07/canva-grow-20-makes-one-vendor-front.html">tracked the Canva acquisition run all year</a>, framed the risk cleanly: &#8220;Canva is asking a quarter billion users who trust it for design to now trust it with their advertising spend.&#8221; Grow sits on top of Meta, TikTok, and LinkedIn APIs that Canva doesn&#8217;t control, and those platforms change API terms whenever it suits them. His test question is worth stealing: if a platform changes access terms next quarter, do you still have a publishing workflow that doesn&#8217;t route through Canva?</p><p><strong>No Google Ads.</strong> Meta, TikTok, and LinkedIn only. If search is your main channel, this launch changes nothing for you yet.</p><p><strong>The smart layer is gated.</strong> AI Ad Tagging, the part that tells you which creative elements drive performance, requires Business or Enterprise. On Free and Pro you get the workflow consolidation without the deepest insight layer.</p><p><strong>It&#8217;s a newly stitched stack.</strong> Grow 1.0 launched in October 2025, and 2.0 is four acquisitions welded together nine months later. Expect seams. Reconcile the reporting against your ad platforms for at least a month before you trust it for decisions.</p><p><strong>It won&#8217;t replace deep campaign control.</strong> Bidding strategy, attribution modeling, and complex account structures still live in the native ad managers. Grow compresses the creative-to-report loop, and that&#8217;s the honest scope of it.</p><h2>Recap</h2><p>Canva Grow 2.0 shipped June 25 and pulls ad creation, publishing to Meta, TikTok, and LinkedIn, cross-platform reporting, and performance-based creative refreshes into the tool most operators already pay for. The Refresh Retainer, the monthly fee whose real deliverable is new variations of winning ads, just became two product features. The move is a 30-minute test with one connected account, then a verdict on your creative line item before the next renewal. Go in with eyes open about the single-vendor dependency, the gated tagging layer, and the missing Google Ads support.</p><p>Sorting out which tools in your stack still earn their invoice is a big part of what I do with operators on <a href="https://muddventures.com/book">AI Clarity Calls</a>. Consolidation questions like this one are the right size for a first conversation.</p><p>Reply with what your creative refresh loop costs you per month if you run paid, I read every reply.</p><p>Andrew</p><p>P.S. If this landed, these connect directly: <a href="https://blog.muddventures.com/p/meta-just-gave-media-buyers-a-command">the Meta Ads CLI issue</a> on the same compression pattern hitting media buying, and <a href="https://blog.muddventures.com/p/notion-just-made-zapier-optional">the Zapier Optionality issue</a> on platforms absorbing their neighbors. <a href="https://blog.muddventures.com/p/the-operators-read-3-quiet-ai-shifts">This week&#8217;s Operator&#8217;s Read</a> covers the quiet shifts stacking up this summer. The <a href="https://whop.com/abra-ai/">Abra AI community</a> is where I share the builds behind these issues. And if someone forwarded you this, <a href="https://muddventures.substack.com/">muddventures.substack.com</a> gets you the next one.</p>]]></content:encoded></item><item><title><![CDATA[Three startups raised money this week to tell you what ChatGPT says about your business, but you can check free]]></title><description><![CDATA[The AI visibility gold rush just kicked off, and the smartest first move takes 20 minutes and costs nothing.]]></description><link>https://blog.muddventures.com/p/three-startups-raised-money-this</link><guid isPermaLink="false">https://blog.muddventures.com/p/three-startups-raised-money-this</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Wed, 08 Jul 2026 14:45:04 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!LNMH!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!LNMH!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!LNMH!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!LNMH!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!LNMH!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!LNMH!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!LNMH!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1709712,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/206055035?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!LNMH!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!LNMH!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!LNMH!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!LNMH!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F44246eca-40e9-4aa9-b2d1-ed1709b4d440_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TLDR:</strong> Three separate startups launched the same product in the first week of July: tools that show businesses whether ChatGPT, Gemini, Claude, and Perplexity recommend them. One raised a &#8364;10 million seed round doing it. The pitch lands because most businesses have never once asked an AI what it says about them. You can run that first check yourself tonight, free, with three prompts. This issue gives you the exact audit, how to read what comes back, and the questions that separate a real vendor from a snake-oil retainer.</p><p>Monday's startup funding roundup had a pattern I've been waiting to see. <a href="https://www.startuphub.ai/ai-news/ai-news/2026/ai-infrastructure-week-july-6-2026">Three companies launched the same product in the same week</a> without knowing it: Foundin.ai in the Netherlands, GeoSurge in London, and Visiblie in Belgium, all building tools that show businesses whether they appear when someone asks an AI for a recommendation. GeoSurge closed a <a href="https://www.eu-startups.com/2026/07/londons-geosurge-raises-e10-million-to-help-brands-understand-ai-generated-outputs/">&#8364;10 million seed round on July 1</a>. Visiblie <a href="https://tech.eu/2026/07/02/belgian-startup-visiblie-raises-eur500k-for-ai-search-visibility/">raised &#8364;500K the next day</a>. In StartupHub's words, "none of them appear to know the others exist."</p><p>I pay close attention to this lane because I build in it. Counterclaim, my skill pack for fixing what AI engines say about a business, exists because of the same shift these three startups just raised money on. Back in May I wrote about <a href="https://blog.muddventures.com/p/googles-ai-swallowed-38-of-organic">AI Visibility Drift</a>, the gap between what your business actually is and what the answer engines say it is. What's new this month is the vendor layer forming on top of that gap. When three companies ship the same product in one week, the sales emails start landing in operator inboxes about six weeks later. Digiday reports GEO pitches are <a href="https://digiday.com/media/geo-hype-busted-experts-call-it-more-seo-than-new-discipline/">already flooding media execs' inboxes</a>. Yours is next.</p><p>Here's what the pitch decks won't lead with: the first and most useful thing every one of these tools does is ask the AI engines about your business and show you the answer. That part is free. You can do it tonight.</p><p>By the end of this you'll have a 20-minute audit that produces a dated snapshot of what ChatGPT, Gemini, and Perplexity say about your business, a verdict system for reading it, and a four-question filter for any vendor who calls.</p><p><strong>What's ahead:</strong></p><ul><li><p><strong>The myth this category is priced on:</strong> your Google ranking tells AI engines a lot less than you'd think, and there's 75,000-brand data on it.</p></li><li><p><strong>The 20-minute self-audit:</strong> three prompts, three engines, one dated document.</p></li><li><p><strong>Three verdicts and the fix for each:</strong> including the one I'm naming today, because half the businesses that run this audit are going to hit it.</p></li><li><p><strong>The vendor filter:</strong> four questions that collapse a snake-oil retainer pitch in one call.</p></li></ul><p><strong>Who this is for:</strong> operators whose customers ask ChatGPT or Perplexity for recommendations before they ever hit Google, which at this point means most local services, agencies, consultants, and B2B firms with a sales cycle.</p><h2>The myth the whole category is priced on</h2><p>The myth: if you rank well on Google, the AI engines already know you and recommend you. So AI visibility is either handled or hopeless, and either way there's nothing to do this quarter.</p><p>The data says otherwise, and it comes from an unlikely skeptic. Ahrefs, a company that sells traditional SEO software, <a href="https://ahrefs.com/blog/ai-brand-visibility-correlations/">studied 75,000 brands</a> to find what actually correlates with getting mentioned by ChatGPT, Google's AI Mode, and AI Overviews. The classic SEO signals came back weak. Domain Rating, the authority score agencies charge you to raise, correlates with ChatGPT mentions at just 0.266. Backlink volume barely registers. Total pages on your site: almost nothing, around 0.194. Publishing more content for volume's sake does basically zero for AI visibility.</p><p>What correlated strongest surprised me when I first read it. YouTube mentions, at roughly 0.737 across all three AI surfaces, beat every other factor they tested. Any time your brand name shows up in a video title, transcript, or description, that's a signal. Branded web mentions came second at 0.66 to 0.71: your name appearing in articles, guides, forums, and directories, whether or not anyone links to you. The engines learned about the world partly from YouTube transcripts and web text, so the businesses that get talked about are the businesses that get recommended. Being talked about and ranking on Google are related, but they're very much not the same job.</p><p>One more finding worth holding onto: ChatGPT showed the weakest pull toward established brand authority of the three surfaces Ahrefs tested. For a business without a household name, it's the most open door. Google's AI Mode sits at the other end, acting like a consensus engine that mostly repeats whoever's already big.</p><p>So the myth cuts both ways. Your page-one ranking doesn't carry over the way you'd assume, and your lack of a big brand doesn't lock you out the way you'd fear. Which means the only way to know where you stand is to ask.</p><h2>The 20-minute audit</h2><p><strong>The move:</strong> ask the engines about your business the way a customer would, from a clean seat. Open ChatGPT, Gemini, and Perplexity. Use a logged-out window or a fresh chat with no memory of you, because your own account has context that a stranger's doesn't.</p><p><strong>The exact prompts.</strong> Run all three in each engine, nine answers total:</p><pre><code><code>1. What are the best [your service] providers in [your city or market]?

2. I'm looking for [the specific thing you sell, phrased like a buyer:
   "a marketing agency for a home services company doing $3M"].
   Who would you recommend and why?

3. What do you know about [Your Business Name] in [city]?
   What do they do, and what is their reputation?
</code></code></pre><p>Paste all nine answers into one document. Put today's date at the top. That date matters more than it looks, because this snapshot is the baseline every future check gets compared against.</p><p><strong>What a healthy result looks like:</strong> you show up unprompted in at least some of the recommendation answers, and when you're named directly in prompt 3, the engine describes your actual services, your actual market, and doesn't confuse you with anyone else. The description reads like a knowledgeable colleague summarizing you.</p><p><strong>The failure mode:</strong> running it once, getting one bad answer, and treating it as a verdict. These engines are probabilistic. The same prompt can produce different answers an hour apart. One absent answer means nothing. A pattern across nine answers means a lot. If you want to be careful, run the set twice and keep both.</p><p><strong>Walk away with:</strong> a dated nine-answer snapshot of what the answer engines tell your prospects about you.</p><h2>Three verdicts, and the fix for each</h2><p><strong>The move:</strong> score each of the nine answers with one letter, then count.</p><pre><code><code>A = PRESENT AND ACCURATE: named in recommendations, described correctly
B = PRESENT BUT WRONG: named, but services, market, or reputation off
C = ABSENT: not mentioned, or the engine says it doesn't know you
</code></code></pre><p><strong>What the scorecard tells you.</strong> Mostly A: you're in maintenance mode, re-run monthly and watch for drift. A mix with B: you have a correction problem, which is exactly the <a href="https://blog.muddventures.com/p/googles-ai-swallowed-38-of-organic">AI Visibility Drift</a> pattern from May, where the engines are working from stale or third-party information about you. Mostly C is the one that deserves a name, so I'm giving it one.</p><p><strong>Institutional Absence</strong> is when the answer engines have no confident story about your business at all. Not a wrong answer, no answer. You ask prompt 3 and get a hedge, a guess, or a different company with a similar name. The engines aren't against you. They've simply never absorbed enough independent signal about you to say anything. Based on what I see when operators run this for the first time, absence is more common than distortion, and it's invisible until you look, because nothing in your Google Analytics tells you about the recommendation that never included you.</p><p><strong>The fix lane, per verdict.</strong> For B, correction: tighten the sources engines actually read, meaning your Google Business Profile, your directory listings, consistent descriptions of what you do across every page that names you, and follow-up where third-party sites describe you wrong. For C, the Ahrefs data hands you the priority list. Get your name into video: guest spots on local or niche YouTube channels, recorded talks, even client walkthroughs, because low-view videos still count when the mentions are wide. Earn genuine text mentions: trade publications, local press, supplier case studies, community guides. And aim at ChatGPT first, since it's the surface least gated by brand size.</p><p><strong>The failure mode:</strong> treating this like a week-long project. Mentions accumulate over months. An operator who starts now is buying position for Q4 and next year, not for next Tuesday. The other failure is buying a retainer to fix what's actually a listings problem you can clean up yourself in an afternoon.</p><p><strong>Walk away with:</strong> a verdict letter and the single fix lane that matches it.</p><h2>When paying a vendor actually makes sense</h2><p>The category is real, and so is the snake oil forming around it. Jeremy Moser, who runs the SEO agency uSERP, gave Digiday the line I'd frame on the wall: <a href="https://digiday.com/media/geo-hype-busted-experts-call-it-more-seo-than-new-discipline/">"If a GEO service does not openly tell you that success in AI visibility is 80 percent good fundamental SEO, they are selling you snake oil."</a> Lily Ray, VP of SEO at Amsive, put the vendor wave in context in the same piece: "We've all lived through this a million times, and that's why it's been frustrating for us." Her point is that AMP and featured snippets each spawned the same cottage industry of specialist vendors before getting absorbed back into ordinary search work.</p><p>There's also a hard structural limit the sales deck glosses over. Per Digiday's reporting, these tools don't have access to the real prompts people type into AI engines. They simulate queries with synthetic data and work backwards from outputs. That's a reasonable method, and it's also an estimate, sold with a dashboard's confidence.</p><p><strong>The move:</strong> before signing anything, make the vendor answer four questions.</p><pre><code><code>1. What share of your recommendations is fundamental SEO
   I might already be paying someone else for?

2. Do you see real user prompts, or do you model them synthetically?

3. Show me one client: what exactly did you change,
   and what moved in the 90 days after?

4. If I stop paying you, what happens to the visibility you built?
</code></code></pre><p><strong>What a good answer sounds like:</strong> a real one will echo Moser unprompted, admit the synthetic-data limit, and point at durable assets like mentions, listings, and content that keep working after the contract ends. Where a vendor genuinely earns the fee is scale: dozens of locations, hundreds of prompts tracked weekly, competitive categories where manual checks can't keep up.</p><p><strong>The failure mode:</strong> any guarantee of AI mentions. The engines are probabilistic and nobody controls their output. A guarantee on a system the vendor doesn't control tells you everything about the rest of the pitch.</p><p><strong>Walk away with:</strong> a four-question filter that ends most of these sales calls inside ten minutes.</p><h2>The honest tradeoffs</h2><p>The audit is a snapshot, and the engines move. Answers shift by model version, by phrasing, by day. That's the <a href="https://blog.muddventures.com/p/the-operators-read-3-quiet-ai-shifts">Engine Drift</a> problem wearing a different coat, and it's why the dated document matters more than any single answer.</p><p>AI visibility also isn't traffic, at least not yet. Ahrefs' own numbers show <a href="https://ahrefs.com/blog/chatgpt-has-12-percent-of-googles-search-volume/">ChatGPT carries about 12% of Google's search volume while Google still sends 190x more clicks to websites</a>. What AI answers drive today is the shortlist: who gets considered, who gets called. That's worth real money in high-ticket services, and it's hard to attribute in any dashboard you currently run.</p><p>And the correlation caveat is real. Ahrefs is careful to say YouTube mentions correlate with AI visibility; nobody has proven they cause it. The fix lanes above are the best-supported bets available, and they're still bets.</p><h2>Recap</h2><p>Three startups launched the same AI-visibility product in one week, which tells you the vendor wave is coming. The category prices itself on a myth, because Google rank and AI mentions are loosely related at best. The free move is the nine-answer audit: three prompts, three engines, one dated document, one verdict letter. B means correct your sources. C means Institutional Absence, and the fix is mentions, video first. Vendors earn their fee at scale, and the four questions above sort the real ones from the gold-rush ones.</p><p>If you run the audit and find the engines telling a wrong story, or no story, about your business, that's the exact job I built <a href="https://counterclaim.muddventures.com">counterclaim.muddventures.com</a> to handle, prompt by prompt, source by source.</p><h2>FAQ</h2><p><strong>Do I also check Google's AI Overviews and AI Mode?</strong> Yes, same prompts. Just know the dynamics differ: Google's AI surfaces lean harder on established authority, so your existing SEO carries more weight there. ChatGPT is where an unknown brand has the fairest shot.</p><p><strong>How often do I re-run this?</strong> Monthly, same prompts, same document. The trend line across snapshots is the real signal.</p><p><strong>I run a local business. Does this matter yet?</strong> Foundin.ai built its whole product for local businesses, which tells you where the demand is. Your buyers are already asking these engines who to call. The volume is smaller than Google's, and the people asking are usually close to a decision.</p><p><strong>Should I block AI crawlers from my site?</strong> If visibility is the goal, no. A blocked crawler can't learn what you do, and an engine that can't read you defaults to whatever third parties say, or to silence.</p><p><strong>What if the AI says something flat-out false about my business?</strong> That's a corrections workflow: find the source it learned from, fix it there, then feed the engines updated signal. It's the core of what Counterclaim automates.</p><p>Tomorrow's issue is back in your inbox at the usual time. If you run the audit tonight, reply with your verdict letter, A, B, or C. I read every reply.</p><p>Andrew</p><p><strong>P.S.</strong> If this issue hit, these go deeper: the original <a href="https://blog.muddventures.com/p/googles-ai-swallowed-38-of-organic">AI Visibility Drift breakdown</a> on what Google's AI did to organic clicks, the <a href="https://blog.muddventures.com/p/the-operators-read-3-quiet-ai-shifts">Operator's Read on Engine Drift</a> for why saved AI setups quietly change under you, a 2-minute <a href="https://showtime.muddventures.com/ai-iq-test">AI IQ Test</a> to see where your operation actually stands, and the <a href="https://whop.com/abra-ai/">Abra AI community</a> where operators share the skill files behind builds like this. New here? Get the daily issue at <a href="https://muddventures.substack.com/">muddventures.substack.com</a>.</p>]]></content:encoded></item><item><title><![CDATA[Run the 20-minute audit that finds the AI Seat Tax hiding in your new Microsoft 365 bill]]></title><description><![CDATA[Microsoft's July price hike quietly bundled a free AI tier into your plan, and it just made some of your paid AI seats redundant.]]></description><link>https://blog.muddventures.com/p/run-the-20-minute-audit-that-finds</link><guid isPermaLink="false">https://blog.muddventures.com/p/run-the-20-minute-audit-that-finds</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Tue, 07 Jul 2026 15:18:10 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!YSP6!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!YSP6!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!YSP6!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!YSP6!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!YSP6!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!YSP6!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!YSP6!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1390366,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/205783558?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!YSP6!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!YSP6!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!YSP6!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!YSP6!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F6f96a107-3b49-4d67-b545-fcab8d4e88d3_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p><strong>TLDR:</strong> Microsoft 365 got more expensive on July 1, and the reason is AI. The upside buried in that hike is that a capable AI chat now ships inside the base plan you already pay for. So this is the week to pull your usage reports, figure out what you&#8217;re really paying per person who actually opens each AI tool, and cut or downgrade the seats that are billing you for silence. I&#8217;ll walk the exact audit below.</p><p>If you run a business on Microsoft 365, you probably got the email. Your renewal bill went up on July 1, some plans by a little, some by a lot. Business Basic moved from $6 to $7 a seat, Business Standard from $12.50 to $14, and the frontline plans got hit hardest, with one version climbing 43% once you strip Teams out (<a href="https://www.windowslatest.com/2026/07/05/microsoft-365-just-got-a-price-hike-over-continuous-innovation-but-copilot-is-the-ai-tax-on-businesses/">Windows Latest has the full table</a>). Microsoft calls it a packaging and pricing update. Windows Latest called Copilot &#8220;the AI tax on businesses,&#8221; and that phrase is doing a lot of work.</p><p>By the end of this you&#8217;ll have a 20-minute audit that shows exactly which AI seats you pay for that nobody opens, and a simple test for whether the now-included free tier already does the job you&#8217;re paying $30 a head for.</p><p>The price went up because Microsoft folded AI and security features into the base tiers, which means every affected plan now includes the built-in Copilot chat with inbox and calendar awareness plus Word, Excel, and PowerPoint helpers. You&#8217;re paying more, and part of what you&#8217;re paying for is an AI tier that used to sit behind a separate charge. For the last three years I&#8217;ve watched the free layer inside these tools quietly get good enough that the paid add-on stops earning its keep, and this Microsoft change is the clearest version of that I&#8217;ve seen.</p><p><strong>Look past the sticker price.</strong> What you pay per person who actually opens the tool usually runs two to three times higher, because most of the seats sit idle.</p><p><strong>The free tier just moved.</strong> The same price hike that annoyed you also bundled in an AI chat that covers a real chunk of what the paid add-on does.</p><p><strong>Renewal timing is the whole game.</strong> Most of this money can only be cut on your renewal date, so the audit is worthless if you run it after you re-sign.</p><p><strong>Who this is for:</strong> any operator paying per seat for Microsoft 365 Copilot, ChatGPT Team, or any AI add-on across a team of five or more. The bigger your headcount, the more this is worth.</p><h3>Name for the thing: the AI Seat Tax</h3><p>The AI Seat Tax is the per-person premium you pay for AI add-ons and standalone AI seats, multiplied across a team where only a fraction of those people ever open the tool. Microsoft&#8217;s own adoption data is the cleanest example. Across enterprises, Copilot activation averages around 35.8%, which means for every 10 seats you buy, roughly 3 or 4 get used and the rest sit idle while you pay full freight (<a href="https://peafowlit.com/blog/copilot-licenses-go-unused-and-how-to-fix-adoption/">Peafowl IT breaks the math down</a>). Their line: &#8220;only 36% of Copilot users are actually using it,&#8221; and &#8220;at $30/user/month, that&#8217;s $115K wasted yearly on a 500-seat rollout.&#8221;</p><p>You&#8217;re probably not running 500 seats. Take a 25-person shop on Business Standard that also bought 25 Copilot add-ons at $30 a month. That&#8217;s $9,000 a year on the add-on alone, and if a third of the team ever opens it, you&#8217;re really paying about $75 a head for the people who use it and full price for the silence of everyone else. Almost every operator I run a cost audit with has some version of that sitting on the books, and the seat line is where the easy money almost always is.</p><h3>Step 1: See what you&#8217;re actually paying per person who uses it</h3><p>The move is simple. For every AI tool you pay per seat, pull the last-30-days active-usage number and divide your monthly spend by the people who actually showed up, not the people you bought seats for.</p><p>For Microsoft 365 Copilot, the usage lives in the Microsoft 365 admin center under Reports, then Usage, then the Microsoft 365 Copilot report (some tenants see it inside the Copilot Dashboard in Viva Insights). It shows enabled users versus active users per app. For ChatGPT Team or any other per-seat AI tool, the vendor&#8217;s admin panel has a members list with a &#8220;last active&#8221; column you can export.</p><p>Once you&#8217;ve got the export, you don&#8217;t have to eyeball it. Paste it into the AI and let it do the sorting:</p><pre><code><code>Here is a CSV export of our AI tool seats and last-active dates.
Columns: name, tool, monthly_cost_per_seat, last_active_date.

Today's date is [DATE]. Do three things:
1. Flag every seat with no activity in the last 30 days as "dead."
2. For each tool, calculate cost per ACTIVE user (monthly spend divided by seats used in the last 30 days).
3. Give me a table sorted by total monthly waste, highest first.
Keep it to the table plus a one-line summary. No preamble.</code></code></pre><p>The correct output is a short table where the &#8220;cost per active user&#8221; column is visibly higher than the sticker price, and a &#8220;dead&#8221; list you can act on. If the cost-per-active-user number comes back roughly equal to the sticker, you have high adoption and there&#8217;s nothing to cut here, which is a real and fine answer.</p><p>Where it goes wrong: the usage dashboards undercount. They often log activity for some surfaces and miss others, so a person who uses Copilot inside Outlook but never inside Teams can read as &#8220;inactive.&#8221; Don&#8217;t cut on a single bad month, and don&#8217;t cut a seat you can confirm someone uses somewhere.</p><p><strong>Walk away with:</strong> a ranked list of exactly which AI seats are dead and what each active seat really costs you.</p><h3>Step 2: Run the substitution test against the free tier that just leveled up</h3><p>For the seats that are getting used, ask a narrower question: is the paid version doing something the now-included free tier can&#8217;t. That test got a lot more interesting on July 1, because the built-in Copilot chat that ships with your base plan now handles inbox and calendar awareness and basic Word, Excel, and PowerPoint help. A year ago that was thin. It isn&#8217;t anymore.</p><p>Go person by person on the active seats and ask what they actually use the paid add-on for. Draft an email, summarize a thread, clean up a spreadsheet, pull notes from a meeting. Then check whether the built-in chat already covers that specific job. A quick way to force the comparison:</p><pre><code><code>TOOL | WHO USES IT | THE ONE JOB THEY USE IT FOR | DOES THE FREE BUILT-IN CHAT DO THIS? | VERDICT
Copilot add-on | Sarah (ops) | summarize long email threads | yes, mostly | downgrade, test for 2 weeks
Copilot add-on | Mike (finance) | build formulas + pivot in Excel | not deeply | keep
ChatGPT Team | 6 seats | general drafting | overlaps free tier | consolidate to 2 heavy users</code></code></pre><p>The correct output is a verdict per seat: keep, downgrade, or consolidate. Keep the seats where the paid tool does deep work the free tier can&#8217;t touch, like heavy in-Excel analysis or grounded answers across your company files. Downgrade the ones where someone&#8217;s paying $30 to do a job the bundled chat now does. Consolidate the tools where you&#8217;re buying six seats of the same capability across two overlapping products.</p><p>The failure mode here is cutting on a demo instead of a trial. The free tier looks close in a 5-minute test and then falls short on the fourth real task of the week. Don&#8217;t cancel anything on day one. Downgrade one or two seats, run them for two weeks against real work, and let the person tell you if they hit a wall.</p><p><strong>Walk away with:</strong> a keep / downgrade / consolidate call on every active seat, backed by the one real job that seat does.</p><h3>Step 3: Right-size before your renewal date</h3><p>This is the part people skip, and it&#8217;s the part that actually saves the money. Most of these seats are on annual or monthly terms you can only change at renewal, and Microsoft&#8217;s packaging changes roll through by August 1 with a 30-day notice in your Message Center. So the audit has a clock on it.</p><p>Pull every AI subscription&#8217;s renewal date into one line. For each one, take your Step 1 dead list and your Step 2 verdicts and set the new seat count you want to land on. Then put a reminder two weeks before each renewal to make the change, because the default is silent auto-renew at the seat count you had last year.</p><p>The correct output is a single number: your new monthly AI spend after you drop the dead seats and downgrade the redundant ones. For that 25-seat example, trimming even 10 idle Copilot add-ons is $3,600 a year back in the business, and you lost nothing anyone was using.</p><p>Where this goes wrong: annual lock-ins. If you pre-paid a year of seats, you may not be able to drop them until the term ends, and reactivating later can carry friction or a price bump. Check the term before you count the savings, and if you&#8217;re locked, put the date on the calendar and hold the audit until then.</p><p><strong>Walk away with:</strong> a renewal calendar with a target seat count and a save-this-much number next to each date.</p><h3>Failure modes to watch</h3><p>The audit is easy to get wrong in three specific ways.</p><p>You cut a power user. The people who actually use Copilot report real gains, on the order of several hours a week saved, so the goal is never to strip AI from the team. It&#8217;s to stop paying full price for the two-thirds who never log in. Protect the heavy users. Cut the ghosts.</p><p>You trust one month of dashboard data. Usage reporting lags and undercounts by surface. Look at 30 to 60 days, and when a seat looks dead, confirm with the person before you pull it.</p><p>You run the audit after you renew. The single most common way this saves nothing is doing it a week too late. Renewal date first, everything else second.</p><h3>The honest part</h3><p>The built-in Copilot chat is real, but it isn&#8217;t the full paid Copilot. For deep, grounded work across your own company files, or heavy Excel and PowerPoint building, the $30 add-on still does things the free tier can&#8217;t. This isn&#8217;t a &#8220;cancel everything&#8221; pitch. When Copilot is actually used, the returns are strong, and cutting a genuine power user to save $30 is the wrong trade. The whole point of the audit is to tell those two groups apart instead of paying the same price for both.</p><p>And the size of the prize scales with your headcount. At 5 seats this is a couple hundred dollars a year and a good habit. At 40 or 50 seats it&#8217;s real money, and it&#8217;s money you&#8217;re currently handing over for logins that never happen.</p><h3>Recap</h3><p>Microsoft&#8217;s July 1 price hike raised your bill and, in the same move, bundled a capable AI chat into your base plan. That makes this the moment to check the AI Seat Tax you&#8217;re carrying. Pull usage and find what each seat costs per active user, test the active seats against the free tier that just got better, then right-size before your renewal date. Protect the power users, cut the ghosts, and put the renewal dates on a calendar so the savings actually land.</p><h3>FAQ</h3><p><strong>I&#8217;m a 5-person business. Is this even worth 20 minutes?</strong> Probably yes, once. The dollars are smaller, but stacked AI seats hit small teams too, and finding one dead seat pays for the coffee you drank running the audit. Do it once, then only revisit at renewal.</p><p><strong>How do I see who&#8217;s actually using Copilot?</strong> Microsoft 365 admin center, then Reports, then Usage, then the Microsoft 365 Copilot report. Some tenants see it in the Copilot Dashboard inside Viva Insights instead. It shows enabled versus active users per app.</p><p><strong>Microsoft raised my price and I never added Copilot. Why did it go up?</strong> Because the AI and security features got folded into the base tiers. The built-in Copilot chat and some Defender protections are now part of the plan, so the base price moved even if you never touched the paid add-on.</p><p><strong>If I drop the $30 add-on, do I lose AI in Word and Excel entirely?</strong> No. You lose the deeper in-app Copilot, but the bundled chat plus the Word, Excel, and PowerPoint helpers that now ship with your plan give you a lighter version. Test that lighter version on real work before you decide.</p><p><strong>Is ChatGPT Team the same problem?</strong> Same audit, same math. Any tool you buy per seat and pay for whether or not people log in belongs in this review.</p><h3>Where the Mudd Ventures work fits</h3><p>Most of what I do in an <a href="https://muddventures.com/book">AI Clarity Call</a> starts exactly here, with a map of what you&#8217;re paying for versus what your team actually runs, because the seat line is usually the fastest dollars and the clearest signal of where AI is stuck in your business. If you want a second set of eyes on the audit before your renewal, that&#8217;s the room for it.</p><p><strong>P.S.</strong> A few earlier issues that pair with this one:</p><ul><li><p><a href="https://blog.muddventures.com/p/notion-just-made-zapier-optional">When the built-in tool absorbs what you were paying extra for</a> (the free-tier-ate-the-paid-tool pattern, applied to automation)</p></li><li><p><a href="https://blog.muddventures.com/p/anthropic-embedded-15-ai-workflows">The workflows that now ship inside the software you already own</a></p></li><li><p>Want the audit run with you before renewal? <a href="https://muddventures.com/book">muddventures.com/book</a></p></li><li><p>Building the seat-cost math into a repeatable system is the kind of thing we trade in the <a href="https://whop.com/abra-ai/">Abra AI community</a></p></li><li><p>Not subscribed yet? <a href="https://muddventures.substack.com/">muddventures.substack.com</a></p></li></ul><p>Reply with your seat count and which AI tools you stack, and I&#8217;ll tell you the first line I&#8217;d check.</p><p>Andrew</p>]]></content:encoded></item><item><title><![CDATA[Three quiet AI shifts from the last two weeks: one saves you $200/mo, one kills a Zapier line item, and one is silently drifting under your custom GPTs]]></title><description><![CDATA[A CRM you already use just gave you 5 AI tools for free, an orchestration drop that puts real pressure on your Zapier bill, and a model swap running underneath your custom GPTs whether you noticed or]]></description><link>https://blog.muddventures.com/p/the-operators-read-3-quiet-ai-shifts</link><guid isPermaLink="false">https://blog.muddventures.com/p/the-operators-read-3-quiet-ai-shifts</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Mon, 06 Jul 2026 15:31:18 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!2kfV!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<div class="captioned-image-container"><figure><a class="image-link image2 is-viewable-img" target="_blank" href="https://substackcdn.com/image/fetch/$s_!2kfV!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png" data-component-name="Image2ToDOM"><div class="image2-inset"><picture><source type="image/webp" srcset="https://substackcdn.com/image/fetch/$s_!2kfV!,w_424,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!2kfV!,w_848,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!2kfV!,w_1272,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!2kfV!,w_1456,c_limit,f_webp,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png 1456w" sizes="100vw"><img src="https://substackcdn.com/image/fetch/$s_!2kfV!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png" width="1456" height="1048" data-attrs="{&quot;src&quot;:&quot;https://substack-post-media.s3.amazonaws.com/public/images/8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png&quot;,&quot;srcNoWatermark&quot;:null,&quot;fullscreen&quot;:null,&quot;imageSize&quot;:null,&quot;height&quot;:1048,&quot;width&quot;:1456,&quot;resizeWidth&quot;:null,&quot;bytes&quot;:1812237,&quot;alt&quot;:null,&quot;title&quot;:null,&quot;type&quot;:&quot;image/png&quot;,&quot;href&quot;:null,&quot;belowTheFold&quot;:false,&quot;topImage&quot;:true,&quot;internalRedirect&quot;:&quot;https://blog.muddventures.com/i/205544875?img=https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png&quot;,&quot;isProcessing&quot;:false,&quot;align&quot;:null,&quot;offset&quot;:false}" class="sizing-normal" alt="" srcset="https://substackcdn.com/image/fetch/$s_!2kfV!,w_424,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png 424w, https://substackcdn.com/image/fetch/$s_!2kfV!,w_848,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png 848w, https://substackcdn.com/image/fetch/$s_!2kfV!,w_1272,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png 1272w, https://substackcdn.com/image/fetch/$s_!2kfV!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F8f2f5566-5047-4280-87ee-ee988dc1c243_1456x1048.png 1456w" sizes="100vw" fetchpriority="high"></picture><div class="image-link-expand"><div class="pencraft pc-display-flex pc-gap-8 pc-reset"><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container restack-image"><svg aria-hidden="true" width="20" height="20" viewBox="0 0 20 20" fill="none" stroke-width="1.5" stroke="var(--color-fg-primary)" stroke-linecap="round" stroke-linejoin="round" xmlns="http://www.w3.org/2000/svg"><g><path d="M2.53001 7.81595C3.49179 4.73911 6.43281 2.5 9.91173 2.5C13.1684 2.5 15.9537 4.46214 17.0852 7.23684L17.6179 8.67647M17.6179 8.67647L18.5002 4.26471M17.6179 8.67647L13.6473 6.91176M17.4995 12.1841C16.5378 15.2609 13.5967 17.5 10.1178 17.5C6.86118 17.5 4.07589 15.5379 2.94432 12.7632L2.41165 11.3235M2.41165 11.3235L1.5293 15.7353M2.41165 11.3235L6.38224 13.0882"></path></g></svg></button><button tabindex="0" type="button" class="pencraft pc-reset pencraft icon-container view-image"><svg xmlns="http://www.w3.org/2000/svg" width="20" height="20" viewBox="0 0 24 24" fill="none" stroke="currentColor" stroke-width="2" stroke-linecap="round" stroke-linejoin="round" class="lucide lucide-maximize2 lucide-maximize-2"><polyline points="15 3 21 3 21 9"></polyline><polyline points="9 21 3 21 3 15"></polyline><line x1="21" x2="14" y1="3" y2="10"></line><line x1="3" x2="10" y1="21" y2="14"></line></svg></button></div></div></div></a></figure></div><p>I read release notes for a living. Not because they&#8217;re fun. It&#8217;s because if you run any AI in your business, the ground under you moves every couple of weeks, and the operators who stay a step ahead are the ones who noticed the ground move.</p><p>Three shifts landed in the last two weeks. None of them got a keynote. All three change what an operator like you is paying for, wiring together, or trusting to run. Two of them are quiet wins. The third can quietly burn you.</p><p>I&#8217;ve started calling the third one Engine Drift: the model under a tool you already use gets swapped, upgraded, or retired by the provider, and your saved setup keeps running on top of the new thing whether you notice or not. The output can shift meaningfully even though you never touched a single instruction. The only real hedge is to re-test your builds on a schedule. More on that in a minute.</p><p>Here&#8217;s the short version before I dig in:</p><p>- GHL turned on their entire AI stack for free through the summer, so if you&#8217;re running on GHL, some of the standalone AI subs bolted on top just got redundant.</p><p>- Anthropic dropped multi-agent orchestration inside Claude, which means one job can run a team of agents (researcher, writer, QA) in a single call instead of a Zapier chain, and that&#8217;s a real crack in the middleware layer.</p><p>- OpenAI retired two models in June, GPT-5.2 on the 12th and GPT-4.5 on the 26th, so anything you built against those is now running on GPT-5.5 whether you told it to or not. That&#8217;s Engine Drift.</p><h2>GoHighLevel turned on its AI stack for free</h2><p>I&#8217;ve had probably fifteen conversations with GHL operators this year about their AI stack, and the pattern&#8217;s always the same. They&#8217;re running GHL for CRM, funnels, workflows, all of it, and then they&#8217;ve bolted a $59/mo Jasper seat on top for copy, a $99/mo Copy.ai for ads, sometimes a ChatGPT Team seat too, plus whatever writer they&#8217;re using for emails. Two hundred dollars a month easily, on top of GHL, for what usually adds up to &#8220;an AI that helps me write things.&#8221;</p><p>Last week they turned on Ask AI, AI Studio, Workflow AI, Funnel AI, and Email AI for every paid sub-account. Free. Through August. That&#8217;s five separate tools, native inside GHL, that overlap directly with what most operators were paying for on top.</p><p>The workflow one is the interesting one. You can drop an AI action inside a workflow now the same way you&#8217;d drop a webhook, which means your booking flow can enrich, draft, and send without you wiring three tools together. Small thing on paper, real cost reduction if you&#8217;re the operator who was paying for the enrichment tool plus the writer plus the tool that sends.</p><p>So here&#8217;s what I&#8217;d do this week if I were running on GHL. Pull up your subscriptions, and any AI writing tool that&#8217;s doing what one of those five now covers, cancel it. Then rebuild one workflow, probably the one that touches new leads, using the Workflow AI action instead of the Zapier or Make chain you had before.</p><p>The thing that gets me about this isn&#8217;t the money. It&#8217;s the pattern. The tool you already pay for keeps absorbing what you were buying separately, and I wrote about this in May when <a href="https://blog.muddventures.com/p/anthropic-embedded-15-ai-workflows">Anthropic embedded 15 workflows into apps you already own</a>. GHL just did the same move for the CRM crowd. The line item disappears whether you cancel it or not, so cancel it.</p><h2>Anthropic put a real crack in the Zapier bill</h2><p>Anthropic ran Code with Claude a couple weeks back, and the piece that mattered wasn&#8217;t the marketing. It was multi-agent orchestration. You can build a job now that runs a small team of Claude agents in sequence, one researches, one writes, one checks, as a single call. Not a chain of Zapier steps calling three separate models. One job.</p><p>I&#8217;ve been trying to migrate a client&#8217;s lead-enrichment flow off Zapier for about six months. Standard mess: Zapier hits Apollo, then Clay, then a GPT step to write the intro line, then HubSpot. Twelve steps. About $180 a month in tool costs across everything, mostly because Zapier and Clay both bill on operations. The migration always died because rebuilding it in Claude Code meant writing custom orchestration for the agent handoffs, and that was more work than it was worth.</p><p>That&#8217;s the thing that just changed. The orchestration lives inside the model now, so I don&#8217;t have to write it. I write instructions for each agent (the researcher gets one, the writer gets one, the QA gets one), and Claude runs the team as one job. My best guess after playing with it for a few hours is that I can pull that whole enrichment flow into a single seat at maybe $200/mo total, which replaces about $180 of tool spend plus some of what I was paying the VA to babysit the workflow.</p><p>If you&#8217;ve got a workflow you built on Zapier or Make that stitches together three or four tools plus an AI step, this is the week to look at it fresh. Don&#8217;t rebuild everything. Pick the one workflow you touch most and see if a multi-agent Claude job can absorb it. My rule so far: if the workflow is more than half AI thinking and less than half moving structured data between tables, it probably belongs in Claude now. If it&#8217;s mostly moving data, keep it in Zapier.</p><p>The bigger point, and this is the one I keep coming back to when I talk to operators about their stack, is that middleware is thinning. Zapier isn&#8217;t dead, but the reason to pay for it just narrowed.</p><h2>The models under your custom GPTs already changed</h2><p>Now the one that can burn you.</p><p>OpenAI retired GPT-5.2 from ChatGPT on June 12, and GPT-4.5 on June 26, custom GPTs included. Anything that was running on either of those now runs on GPT-5.5. The instructions still live, and the engine underneath swapped.</p><p>That&#8217;s the trap. If you built a custom GPT in the last year and tuned it against a specific model&#8217;s behavior (the way it hedges, the way it formats, the length it defaults to), it&#8217;s running on a different engine now, and the same prompt can come back meaningfully different. Nobody sent you an email about it. The system just kept running.</p><p>I saw this happen with a client of mine last week. He runs a small services shop and has a custom GPT his team uses to draft quotes. He mentioned it was &#8220;off&#8221; lately, nobody had touched the instructions, and it turned out he&#8217;d tuned it back in the winter against a specific model that was gone. The outputs had drifted enough that his team was rewriting more than they used to, so he&#8217;d been silently paying that tax for weeks.</p><p>If you built anything against a specific model in the last twelve months (a custom GPT, an API call in a Zap, a pinned model string in a Make step), this is the week to walk through them and re-test each one against a known-good input. Not a new input where you can talk yourself into liking the output. A specific input where you already know what the good answer looks like. If the tone, format, or accuracy shifted, either update the instructions or pin a specific model that&#8217;s still available.</p><p>This is going to keep happening. Providers retire models on their own schedule, and your setup lives on top of theirs. The only durable move is to build a re-test into your calendar every couple months, or every time you notice output feels different, so Engine Drift can&#8217;t quietly compound.</p><h2>The pattern underneath all three</h2><p>Three announcements, three angles on the same shift.</p><p>The tool you already own is absorbing what you were paying for separately (GHL). Middleware is thinning (Anthropic). The engine under your saved setup can change without asking you (OpenAI).</p><p>What all three have in common, and this is the thing worth taking, is that your AI stack isn&#8217;t a set of tools you decided on. It&#8217;s a set of dependencies you&#8217;re renting. The provider can rearrange what&#8217;s inside them, absorb the layer above them, or retire the model under them. The job isn&#8217;t to pick tools once and forget them. The job is to re-audit them on a schedule so you&#8217;re not paying for something you don&#8217;t need, wiring together something that now runs in one place, or trusting an engine that got swapped.</p><p>I have this on my calendar every four weeks. Twenty minutes. I look at every AI line item, every workflow that stitches together more than three tools, and every custom GPT or pinned model string. Half the time I find nothing, and every once in a while I find a subscription I can cancel, a workflow that just got easier to build, or an engine that drifted underneath a build I&#8217;d forgotten about.</p><p>Three shifts, one habit. If you don&#8217;t already have &#8220;audit the AI stack&#8221; on a monthly calendar, put it there this week. It&#8217;ll pay for itself the first time you find one thing.</p><p>Andrew</p><p>P.S. Two threads from the archive if you want more on any of this:</p><p><a href="https://blog.muddventures.com/p/anthropic-embedded-15-ai-workflows">The tools you already pay for keep absorbing features</a> (the embedded workflow layer)</p><p><a href="https://blog.muddventures.com/p/notion-just-made-zapier-optional">Notion just made Zapier optional</a> (the middleware thinning trend, before this week&#8217;s evidence)</p><p>If you&#8217;d rather do this audit alongside other operators reading the same release notes, that&#8217;s the exact conversation we&#8217;re having in the <a href="https://showtime.muddventures.com/operator-council">Operator Council</a>.</p><p>Reply and tell me what you found. I read every one.</p>]]></content:encoded></item><item><title><![CDATA[Comet Will Read Your Whole Inbox in One Command. That's Exactly Why I Keep It Walled Off.]]></title><description><![CDATA[The agentic browser is useful for research and unsafe pointed at your real logins. Both things are true in June 2026.]]></description><link>https://blog.muddventures.com/p/comet-will-read-your-whole-inbox</link><guid isPermaLink="false">https://blog.muddventures.com/p/comet-will-read-your-whole-inbox</guid><dc:creator><![CDATA[Andrew Mudd]]></dc:creator><pubDate>Thu, 18 Jun 2026 14:08:41 GMT</pubDate><enclosure url="https://substackcdn.com/image/fetch/$s_!Ml4j!,w_256,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F3c9fc46f-93fa-4df4-b502-45810c63a5ed_3546x3546.jpeg" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p>There is a piece of software that climbed into the top three on the iOS App Store this spring and can open your Gmail, read every thread, draft replies, compare prices across five retailer tabs, and add the winner to your cart while you watch. It is free. It crossed 10 million users earlier this year. And the week I am writing this, the company behind it raised another $200 million at a valuation near $20 billion.</p><p>The tool is Comet, Perplexity&#8217;s AI browser. If you spend any time in operator circles right now, you have heard some version of the pitch: the browser is becoming the place where an AI agent starts a task and finishes it for you, and whoever owns that surface owns the next decade. Maybe. What I want to give you today is the version of this I would give a client over coffee, because the hype and the reality are both real and they point in different directions.</p><p>For the last three years I have used AI every single day, and I have watched the same pattern repeat. A capability shows up, the demos look like magic, and then the gap between the demo and a safe daily workflow turns out to be the whole game. Comet is the cleanest example of that gap I have seen in a while.</p><p>Start with what it actually does well, because that part is not marketing.</p><h2>What Comet is good at today</h2><p>The honest wins are research and synthesis. You point it at a topic, it opens the relevant sources, pulls the key information, and assembles a structured summary inside the browser with citations you can click back to. A developer named Naresh B A wrote up an honest week with Comet and described asking it to find the best video tutorial on a topic: it compared several options, analyzed them, and opened the best one without him touching a tab (<a href="https://medium.com/@phoenixarjun007/my-honest-week-with-comet-browser-what-i-loved-what-i-didnt-b63487122481">medium.com</a>). For competitive research, pulling notes out of a long thread, or turning six open tabs into one brief, this is a real time save.</p><p>It is also a capable shopper and reader. Summarize this page, or find a product under a set budget across a few retailer sites. These lower-stakes, mostly-public-web tasks are where the tool earns its place in a workday.</p><p>Now the part the demos skip.</p><h2>Where it gets slow and rough</h2><p>Naresh&#8217;s review is useful precisely because he is not selling anything. When he asked Comet to draft and send an email, it took nearly five minutes, because the agent methodically analyzed the page, captured screenshots, and navigated the HTML step by step. His verdict on the rough edges: &#8220;more genius toddler than polished pro&#8221; (<a href="https://medium.com/@phoenixarjun007/my-honest-week-with-comet-browser-what-i-loved-what-i-didnt-b63487122481">medium.com</a>). That matches what I see. The agent will occasionally misread a button, open an extra tab, or stall on a complicated page. For a one-off task that is a curiosity. For a workflow you run thirty times a day, five minutes plus an occasional misfire is the line between a tool and a toy.</p><p>What I tell the operators I work with is that &#8220;it can do it in a demo&#8221; and &#8220;I can trust it to do this unattended on my real accounts&#8221; are separated by roughly eighteen months of engineering, and most tools in this category are still on the early side of that line.</p><h2>The part nobody hyping this wants to dwell on</h2><p>Here is the uncomfortable middle of the story.</p><p>The single most useful thing an agentic browser does, acting across every site you are logged into at once, is also the reason security researchers say it cannot be fully secured. For thirty years the web&#8217;s core protection has been the same-origin policy: the bank tab cannot read the Gmail tab. An AI agent operating with your logged-in credentials erases that boundary by design, because you handed it the keys.</p><p>The attack that exploits this is called indirect prompt injection. Someone hides an instruction inside a web page, a Reddit comment, a calendar invite, or a document, the agent reads it, and the language model cannot reliably tell your command apart from the attacker&#8217;s. In August 2025, Brave&#8217;s security team hid text inside a Reddit spoiler tag, Comet read it, followed the hidden instructions, and pulled out a user&#8217;s email address and a one-time passcode (<a href="https://brave.com/blog/comet-prompt-injection/">brave.com</a>). In March 2026, Zenity Labs demonstrated a zero-click version triggered by a malicious calendar invite, plus a path that lifted credentials out of a 1Password vault through the agent&#8217;s own authorized workflow (<a href="https://siliconangle.com/2026/03/03/zenity-warns-inherent-security-risks-agentic-browsers-perplexity-comet-findings/">siliconangle.com</a>). Enterprise testing by LayerX found Comet up to 85 percent more vulnerable to phishing and web attacks than Chrome (<a href="https://layerxsecurity.com/blog/layerx-finds-that-perplexitys-comet-browser-is-up-to-85-more-vulnerable-to-phishing-and-web-attacks-than-chrome/">layerxsecurity.com</a>).</p><p>I want to be fair here, because this is not a Perplexity-only failure. OpenAI&#8217;s competing browser, Atlas, has the same structural problem, and OpenAI itself wrote in December 2025 that prompt injection is &#8220;unlikely to ever be fully &#8216;solved&#8217;&#8221; in browser agents. Security researcher Simon Willison calls it the &#8220;lethal trifecta&#8221;: any agent that can touch private data, read untrusted content, and send information out can be turned into a data-exfiltration tool by a single injected prompt. Every agentic browser, by design, does all three of those things.</p><p>There is a legal cloud on top of the technical one. Amazon sued Perplexity in November 2025, arguing that Comet&#8217;s agent visiting Amazon&#8217;s logged-in pages with your credentials counts as unauthorized access under the Computer Fraud and Abuse Act. The Ninth Circuit heard oral arguments on June 11 and has not ruled yet (<a href="https://www.techtimes.com/articles/318528/20260616/ai-browser-comparison-2026-atlas-vs-comet-vs-dia-ranked-security-use-case.htm">techtimes.com</a>). Whichever way that lands, it tells you how unsettled the ground under this whole category still is.</p><h2>What to do with this</h2><p>The part that actually matters for an SMB owner is that none of the above is a reason to ignore the tool. It is a reason to use it the way you would onboard a sharp but brand-new assistant: hand them the research, keep them away from the checkbook until they have earned the trust. Here is how I would set it up this week.</p><p>First, install it in a walled-off profile. Run Comet as a dedicated browser that does not share saved passwords, payment methods, or active logins with the browser where you do your real email and banking. The damage an injected prompt can do is capped by what the agent can reach, so a profile with nothing sensitive in it is a small target.</p><p>Second, give it only research and reading jobs to start. Spend a week pointing it at public-web work: competitive scans, summarizing long reports, comparing vendors, turning a pile of tabs into a one-page brief. Notice where it genuinely saves you time and where it stalls. You will know inside five sessions whether it belongs in your week.</p><p>Third, keep your hand on anything irreversible. Every agentic browser has a setting that makes the agent pause before it buys, sends, submits, or changes a password. Leave that on. The cost of one confirmation tap is nothing next to an agent completing a purchase or sending a message you never approved.</p><p>If your work is research-heavy, Comet is worth the hour it takes to test. If you want an agent transacting on your real accounts unattended, the honest answer in June 2026 is that the tooling, the security, and the law are all still catching up, and a careful operator stays on the research side of that line for now.</p><p>The judgment call I spend most of my time on with the businesses I work with is figuring out where a tool safely fits in how you already operate, separate from how impressive it looks in a demo. If you want help mapping which AI tools earn a place in your stack and which ones are still demos wearing a product page, that is exactly what an AI Clarity Call is for. You can grab one at <a href="https://muddventures.com/book">muddventures.com/book</a>.</p><p>And if you just want to compare notes with other operators kicking the tires on this stuff before betting real workflows on it, come hang out in the Abra AI community at <a href="https://whop.com/abra-ai/">whop.com/abra-ai</a>.</p><p>The browser that runs your whole business is coming. It is not here in June 2026, and knowing the difference between a useful copilot and an unattended employee is most of the edge.</p><p>Andrew</p><p>P.S. If a friend forwarded you this, you can get it straight to your inbox at <a href="https://muddventures.substack.com/">muddventures.substack.com</a>.</p>]]></content:encoded></item></channel></rss>