Markie Wagner Co-Writes
Jun 10, 2026
161
11
14
Share
Welcome to the 192 newly Not Boring people who have joined us since Monday (https://www.notboring.co/p/expanding-the-radius-of-daily-life) ! Join 269,285 smart, curious folks by subscribing here:
Subscribe
Hi friends đ,
Happy Wednesday and welcome back!
A couple of months ago, my friends Adam and Ben at Genius Ventures asked if they could introduce me to one of their favorite founders, Markie Wagner (https://x.com/markiewagner) .
The Markie Wagner? The Choose Good Quests (https://foundersfund.com/2023/06/choose-good-quests/) Markie Wagner? The drop an all-timer then go quiet for years, cooking up something spoken of in hushed tones Markie Wagner? The grew up inside of a computer and dreamed as a young girl in Southern California, seriously, of making computers do the work that humans shouldnât have to Markie Wagner?
Of course I wanted to meet Markie Wagner.
So we met a month ago at Soho Diner and I ordered a milkshake and she asked them to cut up a bowl of fruit. She asked for my lore, which was boring, and I asked for hers, which she weaved non-stop for the next hour, landing so naturally on why sheâs building what sheâs building that it seemed almost pre-destined.
She also told me, before everyone else came to the same conclusion, that tokenmaxxing was bullshit, because behind closed doors, the Fortune 500 CEOs she works with were all saying some version of âWe committed to all this token spend and I have no idea what weâre getting out of it.â
She was right, I think sheâs going to be right again, sheâs backed by Founders Fund, Kleiner Perkins, Genius Ventures, and OpenAI to go prove it, and now sheâs explaining her logic publicly in her first written piece since Good Quests .
So this is where we are heading, according to Markie Wagner.
Letâs get to it.
Todayâs Not Boring is brought to you by⌠Deel (https://www.deel.com/resources/a-guide-to-eor-for-startups/?utm_medium=sponsored-newsletter&utm_source=notboring&utm_campaign=ww_engage_download_notboring_sponnewsletter_smb-eorforstartups-may26_eor_smb&utm_content=engage_eor_sponnewsletter_eorforstartups-sponnews180-smb_en)
https://www.deel.com/resources/a-guide-to-eor-for-startups/?utm_medium=sponsored-newsletter&utm_source=notboring&utm_campaign=ww_engage_download_notboring_sponnewsletter_smb-eorforstartups-may26_eor_smb&utm_content=engage_eor_sponnewsletter_eorforstartups-sponnews180-smb_en
New to global hiring? Start here.
Hiring internationally is complex. Learn what an Employer of Record is and how startups use EORs to hire global talent compliantly.
Get the Guide (https://www.deel.com/resources/a-guide-to-eor-for-startups/?utm_medium=sponsored-newsletter&utm_source=notboring&utm_campaign=ww_engage_download_notboring_sponnewsletter_smb-eorforstartups-may26_eor_smb&utm_content=engage_eor_sponnewsletter_eorforstartups-sponnews180-smb_en)
# Return on Tokens (ROT)
Co-Written with Markie Wagner
https://substackcdn.com/image/fetch/$s_!WRho!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F511d1dba-22f1-422e-b540-f0fd671ad658_900x450.png
The promise of AI is that it will turn businesses into software so that they can evolve over millions of tiny iterations. Beautiful, ideal, complex things can only emerge as the result of tremendous trial and error over time. You cannot build perfection, only discover it.
Capitalism is organizational evolution. Millions of businesses compete in the marketplace with offerings that they think customers will want. Some thrive and grow. Others die. Each company evolves, too. People come and go. An experiment becomes a process, a process becomes a web of tacit knowledge. Products are introduced, and products are retired.
This constant evolution is why we enjoy the standard of living we enjoy today, and why ours will look primitive to future generations. Accelerating it is my Good Quest (https://foundersfund.com/2023/06/choose-good-quests/) , because if every business can evolve to its ideal form, it will create trillions of dollars of value and unblock all of the other Good Quests.
I dropped out of the research world because it felt like the wrong hill to climb, and I went out into America to just do work, so that I could figure out how to make computers do the work, so that humans could direct the computers in evolving the work.
What I suspected before and learned in my travels is that the way that the market has implemented AI thus far is the wrong way. Itâs not endgame. It is too wasteful, too forgetful, and too imprecise. Iâve been in the fucking Sahara Desert out here fighting demons to learn this wisdom.
Tokenmaxxing Clearly Isnât It
Tokenmaxxing - literally maximizing the amount of tokens you or your organization spends, tracked in leaderboards and rewarded with trinkets - was a mass delusion, something like a commercial form of AI psychosis.
Tokenmaxxing was a lab-grown supermeme that worked better than the labs could have hoped.
Picture this. Anthropic and OpenAI release a product, Agents, in the form of Claude Code/Cowork and Codex, respectively, that are basically lab employees working inside of customersâ companies and are given company credit cards with no spending limit (tokenmaxxing suggests the more they spend the better theyâre doing) to spend on behalf of their real employer, the lab. Anthropic ships a bunch of Agents into, say, KPMG, which commits to a certain spend in exchange for discounts ( token commits ), KPMGâs employees are encouraged to use Agents to do everything they can possibly think of (lots of dashboards), and then these Agents, which again you can think of as digital Anthropic employees with no-limit KPMG credit cards that they can use to spend on Anthropic, run up token bills to their heartâs content. Employees who direct their Agents to use the most tokens are recognized as AI Innovators.
Certainly, some people recognized that it was a delusion. They would ask questions like, âBut are the Agents doing anything useful? Arenât they just building dashboards? Please can someone show me something useful theyâve built with an Agent?â but those sane few were met with the killer retort: â Skill Issue .â
Some people, they were told, were building immensely valuable things with Agents, the same way that some people had a super hot girlfriend at summer camp but youâve never met her. If you couldnât figure out how to do the same, well, welcome to the Permanent Underclass.
Everyone fell for it, for a while. The market incentivized companies to spend tokens, so boards incentivized leaders to spend tokens, so leaders incentivized managers to spend tokens, so managers incentivized employees to spend tokens. Nobody had an incentive to say that the tokens arenât doing useful stuff.
https://substackcdn.com/image/fetch/$s_!6Y9o!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F76eecab3-d7a9-4c57-a11f-c94ea7bf5551_908x356.png OpenAI âTokens of Appreciationâ for Customers That Spend the Most Tokens
I talk to these people all the time and every company has some version of the same conversation. Someone whoâs running the AI team goes, âWeâve made a ton of progress this quarter. We spent $50 million on tokens.â And everyone nods and claps. âUsage is up. Weâve built 3,000 Agents. We shipped 10 million lines of code.â And youâre like⌠what? And then Iâd ask, âHey did you measure accuracy for the fraud Agent?â And theyâd go, âYeah⌠itâs about 50%.â People are at 99%!
But the models improve and everyoneâs tokenmaxxing and you donât want to be the luddite, so you keep lazily throwing Agents at everything and hoping they learn.
All of this happened, by the way, right as the labs switched from subscription-based to consumption-based revenue models, so companies had no time to prepare.
It is no wonder token usage, and therefore lab revenue, went parabolic.
It took Uber, that last eraâs poster child of VC-subsidized demand, to break the spell. Its CTO said that the company had burned through its 2026 Claude Code token budget by April. In May, its COO said that the company was having a harder time justifying its AI spend, because the link between AI consumption and shipped features âis not there yet.â
What followed was like that scene in Mean Girls where Tina Fey asks the students to raise their hands if they feel personally victimized by Regina George, and one hand goes up, then all of the rest of the hands go up.
There was the consultant saying his client accidentally burned half a billion dollars on Claude Code. Amazon shut down its AI leaderboard (https://www.businessinsider.com/amazon-ai-leaderboard-tokenmaxxing-2026-5) . Legora CTO Jacob Lauritzen told Harry Stebbings (https://www.businessinsider.com/legora-cto-tokenmaxxing-jacob-lauritzen-encourage-ai-usage-2026-6) that token leaderboards âlead to tokenmaxxing, which is people just burn tokens just to look good. Thatâs a really stupid way to do anything.â Rampâs Veeral Patel called it the Token Casino (https://x.com/vral/status/2064067264991654236) : âuseful software wrapped in mechanics that make spend feel like progress. It starts with the oldest trick in the book: abstract the money.â Palantir CEO Alex Karp told the TBPN boys (https://x.com/jawwwn_/status/2062631893900607504?s=46) that tokenmaxxing is like âa porn addiction.â
Even Sam Altman, a prominent token vendor himself, admitted on CNBC that âYou hear companies saying, âI am spending a ton of money on AI, and I know some great stuff is happening, but I know thereâs a ton of waste, and you know, when⌠how long do I have to wait for it to really show up in revenue, and how long do I have to wait to really get the costs under control?ââ It had become, he admitted, a âhuge issue.â
The issue is the companies have focused on maximizing tokens, assuming that tokens = value.
Every cycle has its dumb metric. In the mid-nineteenth century, the market wanted miles of railroad track as a proxy for future monopoly and the benefits thereof, and so railroads raced to lay miles, often along the same routes as competitors. At the turn of the 21st century, the market wanted eyeballs, and so dot coms attracted eyeballs and served them up on a platter. In the 2010s, the market wanted top-line gross revenue, and so companies like WeWork delivered top line gross revenue.
This cycle has tokenmaxxing.
Which is not to say that tokens canât be valuable. Cornelius Vanderbiltâs New York Central ended up becoming very valuable, as did the Pennsylvania Railroad. Google and Facebook have converted eyeballs to cashflow better than anyone has ever converted anything to cashflow. Uber ended up turning top line growth into market dominance and turning that into $10 billion in 2025 free cash flow.
The question is always: can the thing generate returns?
For tokens, the question is: what is your Return on Tokens (ROT) ?
Return on Tokens
When you invest in a new machine, you expect it to generate a return. When you hire an employee, you expect them to generate a return. Business is the process of making investments big and small and expecting them to create more value than they cost.
Tokens need to be held to the same standard.
Return on Tokens = (Value of Output - Cost of Tokens) / Cost of Tokens x 100
There are two ways, then, to increase your ROT. You can create more valuable things with them, or you can spend less on them. Ideally, you spend less to create more value.
The first thing that companies are focused on, because it is easier to measure than output value, is spending less.
Now that the spell has been broken, cooler heads are proudly discussing âroutingâ as a means to lowering the cost. Use Anthropic and OpenAIâs best models for the really big brain stuff, but do most of the work with cheap Chinese open source models. Coinbase CEO Brian Armstrongâs recent tweet is a good example of this logic:
Brian Armstrong @brian_armstrong
Good take
My guess is
- demand for intelligence is near infinite
- but 80% of workloads will be running on 99% cheaper models within 12-18 months
- 20% of workloads will still run on latest gen models where IQ maxing is important (scientific breakthroughs, higher level
Tommy @Shaughnessy119
The most basic way AI could blow up imo. I'm not saying it does but this is the most obvious way I can see it happening
- Per seat subscriptions are massively subsidized. The flat fee was priced way below what heavy usage actually costs
- For real business use you have to move
12:38 AM ¡ Jun 8, 2026 ¡ 2.67M Views
467 Replies ¡ 602 Reposts ¡ 6.39K Likes (https://x.com/brian_armstrong/status/2063782620815876515?s=20) You can see this in the OpenRouter AI Model Rankings (https://openrouter.ai/rankings) . The move to Chinese models actually showed up in lockstep with consumption-based pricing, although this is a self-selecting group of users that were already thinking about routing tasks to the right models. The rest are scrambling to do the same now.
https://substackcdn.com/image/fetch/$s_!06ly!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F14b4aaa0-152b-4ca2-8cdb-0de87106fc54_908x829.png
Itâs a good start, but Agents spending tokens, American or Chinese, to figure everything out from scratch is not endgame, either.
Because you know whatâs cheaper than Chinese models? Code.
Code, good old fashioned deterministic code, is not only cheap, it is a better fit for most economically valuable work. We have learned this lesson.
In the past, companies hired humans to do all manner of repetitive tasks. Before âcomputersâ were digital, they were humans.
https://substackcdn.com/image/fetch/$s_!rq77!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F58cc50d5-70a0-4fb1-8109-a39a6ce361c5_908x528.png When Computers Were Human, NASA (https://www.nasa.gov/centers-and-facilities/jpl/when-computers-were-human/)
About half a century ago, we began the process of taking repetitive tasks that humans did, like calculating the trajectory of a missile or the profits of a business, and handing them over to software. Code ran more reliably than even the most reliable human computer. It made no mistakes. It answered instantly. Enter the same numbers and same formulae in the same cells in an Excel spreadsheet anywhere in the world, at any time, and it spit out the same number.
Then we got Agents, and we forgot the lesson. We decided that we needed to throw these pseudo-humans at everything, because everyone else was. Agents are great at some things, but theyâre not the right shape for a lot of others.
Itâs no wonder companies arenât getting a positive ROT on their tokens. All the dashboards have been dashboarded, and now theyâre sending Agents to do softwareâs job.
There is an argument to be made that a lot of companies arenât getting a positive Return on Tokens because they donât know how to use them yet, or because they havenât re-architected their companies to be AI-native yet. This is one of the reasons, perhaps, that both Anthropic and OpenAI have launched consulting subsidiaries to help companies better deploy tokens. And there are certainly examples of startups built during the AI era that seem to be using tokens to great effect, which shows up in their own supersonic revenue growth. If the Old Economy canât generate a ROT, well, this is creative destruction baby.
And while itâs certainly true that not all companies are deploying AI equally well, we believe that there are more fundamental structural reasons for the negative ROT: Agents are the wrong architecture for most work.
Agents Arenât It (For Most Work) Either
There are three structural reasons for Agentsâ negative ROT:
The Agentic architecture canât do long-running work at the nines of quality that real economic work requires.
Agents improvise. Theyâre spawned fresh onto repetitive tasks like every day is their first day on the job, which hurts consistent accuracy. For new features, prototypes, or dashboards, 80% accuracy is fine. For the real repetitive work on which the economy runs, like fraud detection or underwriting decisions, 80% accuracy is 0% usable.
Engineers donât know what to build because they donât do the work.
Most of the process-driven work weâre describing exists as a combination of written rules, which Agents can ingest, and then like 3,000 tacit rules and sub-rules that live in peopleâs heads, in offices around the country, far away from the engineersâ San Francisco desks. AI can only evolve what it can touch, which is why itâs been great at coding but has largely failed to do useful things in the enterprise.
The original sin is that there are no goals.
If people have no goals then the Agent has no goals, and then the thing achieves no end. Without a goal to hill-climb against, code (whether written by humans or generated by Agents) decays into slop in the limit because thereâs no purifying force to evaluate whatâs good and bad.
One of the beautiful things about Agents, from a laziness perspective, particularly when you are being encouraged to spend a lot of tokens, is that you can just set them loose without knowing exactly what it is that youâre solving for. They can go spin on a vague instruction for a while, bring something back thatâs decent but not perfect, and then go out and spin some more.
This process drives more token spend without delivering any value, which is a fast track to negative ROT.
People are searching for new things for Agents to do assuming that AI will do for everything else what itâs done for code. But it doesnât have to.
There is a surprising amount of work that is best done with plain old code. The challenge has been that, until recently, there were not enough engineers to turn everything every business does into code, and then update it as things changed. There are now. AI makes writing code trivial, so if we can get the knowledge out of peopleâs heads, we can turn businesses into code.
AI is a Compiler, Not a Runtime
Basically, software works in two steps, thinking and doing.
First, thinking : you take the goals and requirements for what a piece of software should do and compile it into code that a computer can run.
Then, doing (and doing and doing and doing): every time it needs to do the thing, the code runs cheaply and predictably (or deterministically).
While computer science has a precise definition of a compiler, you can also think of a software company or a software engineer as a compiler. They take the goals and requirements and turn it into code. Then customers buy the code they built, and run it over and over again. This is the beauty of the zero marginal cost software business, and itâs why companies can sell software that took millions of dollars to develop for $20 per month and still generate mouthwatering margins.
The way that most people think about (and use) Agents today is that they replace both the software company and the software. That is the wrong way to think about it.
Agents should replace the software company; they should take goals and requirements in English (or whatever language), and turn it into code that runs over and over and over, deterministically.
Thinking is expensive but happens rarely. Doing is cheap and happens forever.
Agents should do the thinking, code should do the doing.
For most economic work, you want to use humans to figure out the rules, use AI to turn the rules into code, and then run that code forever at near-zero token cost, only bringing the AI back in when the rules change.
Why would you use a prompt to add two numbers? Just write a line of Python, dawg.
The current Thinking-Doing Ratio (TDR) in AI implementations is roughly 1000:1, which is not surprising. San Francisco is a Thinking town. Anthropicâs hats say, simply, âthinking.â
https://substackcdn.com/image/fetch/$s_!eZl-!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F90e85292-a643-4475-9d0c-c7b0a83aeac4_708x533.png Anthropic thinking cap + OpenAI Token Appreciation Award, Business Insider (https://www.businessinsider.com/ai-status-symbols-openai-anthropic-cursor-2025-10)
Silicon Valley built AI assuming work is mostly thinking, but work is mostly doing .
Chat is the rare exception where you genuinely donât know what comes next. So maybe customer support chat continues to churn through thinking tokens (although even customer support Agents kick complex problems to humans). Almost nothing else in a business looks like constant improvisation.
So we use Agents for Doing, but Agents are the BlackBerry of doing. They are not where most work will get done inside of companies in five yearsâ time. It will get done in the deterministic code that they write.
The Agentsâ role is in compiling into code, not into running and doing work day to day. Which means that itâs more like CapEx than OpEx. Everything you think is gonna be AI running is just going to be code running.
Everyone thinks the thing that is going to change in the world is that AI is going to become a person, but the real change is that a business is going to become a piece of software.
Thatâs the world weâre building at Poetic (https://poetic.com/) .
Turning Businesses Into Software That Evolves
We are building the antidote to tokenmaxxing: software that tokenminns itself.
Poetic (https://poetic.com/) is a new class of software: adaptive like AI, reliable like code .
We use AI as the compiler. We learn everything that a business does by taking in all of the processes that are written down, then going on-site in Nebraska or Providence or wherever the work is done, sitting on peopleâs shoulders, and asking âWhat did you just do?â âWhy did you do that?â hundreds of times to learn the thousands of hidden tacit rules on which every company runs. Then we turn it into code.
The code is the runtime. When the world stays the same, this code runs the exact same steps every time. When the world changes, it learns, regenerates, tests itself against the objective, and then runs the new code until the world changes again.
The result is 100x less token usage and nines of accuracy on complex tasks. Put differently, each token you spend does 100x more, and it does it right.
The value of the output increases, because Poetic does something that your business actually needs to do, over and over. And the cost of tokens is lower, because Poetic only uses tokens when the world changes. Combined, Poetic generates a clear, measurable Return on Tokens.
We are doing it today, for companies like AIG, SoFi, and Chime. AIG CEO Peter Zaffino said that Poetic has already âachieved 99%+ quality outcomes on multi-hour processes - delivering real enterprise value.â
These companiesâ leaders believe what we believe: that every business will have to be re-founded as a software business. The story of the next decade is the beginning of those new businesses. Some will be truly new, built from scratch. Others will be businesses that have existed for hundreds of years, brave enough to reinvent themselves.
Everyone talks about the fact that it took reconfiguring factories around electricity to benefit from electricity, and follows that with the AI equivalent of âso the new businesses that are built to throw a ton of electricity at the problem will win.â What you really need to do is refactor the businesses into code.
Doing that takes a ground game , going deep into the guts of companies, wherever they are, to understand how they work and migrate their logic into programs. We need people to get out there into Minnesota to be like, what the hell do you guys do all day?
Most of our team are engineers, a lot of them ex-Palantir, who spend weeks at a time on-site with customers, learning from them, getting into the nittiest of gritty details.
The term gets a bad rap, but relative to engineers who spend all day at a desk prompting Claude, they are the most Social Engineers . Engineers who understand people, business, and AI will rule the world. If that sounds like you, come join us at Poetic.
Itâs hard work, but the biggest mistake of the AI era so far has been believing that anything worth doing could be easy.
This is worth doing. Itâs what Iâve wanted to do for as long as I can remember, because all of the institutions that run our world, every business, every government, does so much stuff that doesnât make sense, operates way more slowly than it should, and gums up the works.
Our goal is to discover the perfect process for every business - the plan, the set of steps that is ideal for achieving your goals. This will evolve as the world evolves.
The role of AI is not âAgentic,â improvising as it goes. It is an evolutionary force : changing, testing, evaluating what plan is most successful and sticking to it until a better one emerges.
In the end, this is what a business is. A piece of living, evolving software reconfiguring, testing, evaluating itself, hurtling towards ideal form at the fastest clip imaginable. Humans exist to define what good looks like, not how to get there. Shaping behavior, directing behavior.
Through running in production, the system gains a record of what happened. What happened, step by step, for every dispute, underwriting case, insurance claim. Thus, every process change becomes testable. The answer to âwhat ifâ is known after minutes of backtesting. Run both scenarios in shadow, compare outcomes, decide which is better.
When impact is entirely known, there is little risk - you know exactly what would have happened. Change simply becomes a choice. Then the choice becomes: which outcome is ideal? The process lead, the person responsible for making sure the process achieves the goals above it, simply makes choices.
To make a change, you have to know what the impact of the change is. Itâs easy to generate code, hard to know what happens if you run it. After months of running, youâll be able to ask questions like âWhat if we approved every dispute under $25?â and know in minutes.
The hill-climb towards the most beautiful process will then begin. Experts experiment, asking what-if questions. Now that humans are no longer bottlenecks, they can begin searching.
This is endgame.
Every business is not just a piece of software; itâs a piece of software constantly editing, testing, evaluating changes. Evolving at the highest frame-rate possible, climbing towards its most correct form. All energy is spent evolving, figuring out the ideal form of the rules.
We donât use tokens to run the business. We use tokens to turn the business into code and evolve it. We tokenminn to ROTmaxx.
Over billions of years, we have evolved from ocean slime, through trial and error, into fish, lizards, voles, monkeys, and humans.
We donât want to have to wait billions of years for businesses to evolve into their diverse and ideal forms, and Agents wonât build them. You cannot build the butterfly.
Beautiful, ideal, complex things can only emerge through evolution. I want to speed it up and see how far we can go.
Thanks to Markie for dropping her knowledge and to Adam and Ben for introducing us!
Thatâs all for today! Weâll be back in your inbox on Friday with a Weekly Dose.
Thanks for reading,
Packy
161
11
14
Share
