OpenAI cut Luna 80%. The plumbing underneath just got funded. ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
The Agent Stack mascot
The Agent Stack _
Daily B2B AI automation brief · Friday, July 31, 2026 · Issue #49

Hey there 👋

OpenAI posted a new price sheet yesterday, and I read it three times this morning to be sure I had the math right. GPT-5.6 Luna, their cheapest capable tier, dropped 80%. Three weeks after it launched. That is not how a confident market leader behaves.

So here is my bias, up front. I think the 2026 agent race quietly stopped being about which model is smartest and turned into a fight over price and rails, not raw IQ. Today's issue is that fight: one model that got cheap overnight, and the money, identity, and testing plumbing that got funded and shipped on the very same day.


The Big Thing

OpenAI made its cheapest good model 80% cheaper.

Starting July 30, GPT-5.6 Luna costs $0.20 per million input tokens and $1.20 per million output, down from roughly $1 and $6 at launch. Terra fell 20%, to $2 and $12. Sol, the heavy reasoning tier, did not budge: still $5 and $30. Every dollar of the cut landed on the tiers you run in bulk, and none of it touched the model you reach for when a task is genuinely hard.

The timing is the whole story. Luna shipped July 9. Price cuts on a fresh frontier model usually arrive months later, once a successor is close. This one came in three weeks. Axios and CNBC both trace it to cheaper Chinese open-weight models, DeepSeek among them, chewing on the low end. Read it plainly: the frontier labs are losing pricing power on commodity inference, and they clearly know it.

OpenAI's own explanation earns a raised eyebrow. The company says Sol rewrote its production inference kernels by itself, trimming end-to-end serving cost about 20%, and that it is handing that saving to customers. Believe the mechanism if you want, but that 20% is OpenAI's figure measured on OpenAI's own hardware, so hold it loosely. Same for the testimonials in the post. Ramp, Dust, and Blitzy all say Luna is their new default, and those names were hand-picked by OpenAI, so read the percentages as marketing copy and wait for an outside benchmark before you cite them.

The move for your own stack is small and it pays this week. Take your high-volume, low-stakes agent steps, the classification, the test-writing, the background automation nobody watches, and route them to Luna now. Keep Sol for the reasoning that earns the premium. Everything in between is a routing decision you can make in an afternoon.

Ship it? Deploy. Pin your model IDs like dependencies, split traffic by cost per step, and let Luna carry the boring 90% of calls. The cheapest capable tier just got five times cheaper. Leaving it on the shelf is a line item you chose to keep paying.

CNBC · VentureBeat


Tour de Headlines

💸 Natural raised $30M to own the money rail.

While the timeline argues about model IQ, Natural closed a $30M Series A led by Forerunner's Kirsten Green to build the payments stack agents will run on. Six products are already live: FDIC-insured wallets for agents, one-way vaults, and send, request, and transfer tools, with its own ledger and multi-bank settlement underneath. That "FDIC-insured" label is Natural's own description, and a named bank partner would firm it up. TechCrunch covered the raise back on July 20; the formal close landed July 30. The bet is blunt: agents will move real money this decade, and someone owns the pipe. Natural wants to be Stripe before Stripe gets there.

PR Newswire · Natural

🔑 Hush Security wants every agent on a leash.

Pair the money rail with an identity rail. Hush Security raised $30M with Akamai joining as a strategic backer, all to kill standing credentials for machines. Instead of handing each agent a permanent password, Hush grants runtime just-in-time scopes, logs every action, keeps a central registry, and holds one kill switch for the whole fleet. The framing number, an average Fortune 500 running 150,000-plus agents by 2028 up from under 15 last year, comes per the company, citing Gartner, so weigh it as a vendor slide rather than a checked forecast. The problem underneath is real enough: every agent you deploy is one more credential someone can steal.

PR Newswire · FinTech Global

🔍 Stairwell hunts malware from the files up.

Stairwell launched Backstory, an agentic investigation tool that flips the SOC workflow. Rather than triaging an endless alert queue, it starts from the files sitting on your endpoints, traces related variants, names the affected systems, and maps an incident's blast radius in seconds. The pitch leans on Stairwell's own Hidden Malware Report, which across 1,085 public threat reports found an average of 2.4 related variants for every published sample. That research is the vendor's, and so is the claim to be the first agentic investigation platform of its kind, so grade both accordingly. Still, for an IR team buried in triage this week, the shape is right: begin where the damage already lives.

Help Net Security · SecurityBrief

🤖 500 small businesses hired a robot before lunch.

The fastest real adoption of agents is happening down-market, far from the frontier labs. Vendasta pushed two autonomous "AI Employees," a Social Media Manager and an AI Blogger, to general availability and says more than 500 of them went into production in the first 24 hours. Nobody outside Vendasta has counted those installs, so treat the number as a press release, not a census. The agents generate, schedule, publish, and analyze location-specific social and blog posts for local businesses. While the enterprise debates governance in a conference room, a plumber in Ohio just handed his Instagram to a bot. Demand lives in the boring long tail.

Vendasta newsroom


Sponsor

Read the human signal in every conversation.

Every agent you run gets scored on the calls it handles. RapportScore does that for the humans, reading the communication signals in live sales and support conversations so your team can see what lands and adjust before the next one. Built by the crew behind this newsletter.

See your team's score →

Tool of the Day

⚙️ BrowserStack Test Companion

What it is for: an agent that writes, runs, and maintains your tests without leaving your IDE.

Coding agents happily generate test code. Keeping those tests alive across a real CI stack is a different job, and that is the one Test Companion is built for. It installs straight from the marketplace into VS Code, JetBrains, Cursor, and Antigravity, points at your existing framework, and asks for no setup scaffolding. BrowserStack says more than 1,000 teams are already authoring and debugging tests with it up to 4x faster, though both numbers are the company's own, so run it against your own suite before you trust the multiplier. One detail to clock: it names Cursor and Antigravity as targets, which tells you which agent-IDEs have real traction. The shape here is augmentation riding inside the tools engineers already open, the kind that respects your stack instead of promising to replace your QA team.

Ship it? Try it on one flaky test file. If it fixes the thing you have been avoiding, widen from there.

Read the BrowserStack launch


Worth a Click

  • The price war, laid out. VentureBeat on the 80% Luna cut and where the pressure is coming from. VentureBeat
  • Why it came so fast. CNBC on enterprise cost sensitivity forcing a three-week price drop. CNBC
  • Natural takes on Stripe. The company's own read on agent-native payments and whether the category is real. Natural

Three weeks ago the moat was governance. Last week it was your proprietary data. This week the model itself got cheap enough to treat like electricity, and on the same day, the money, identity, and testing layers under agents got funded and shipped. The frontier is commoditizing in public.

If I had to bet, the durable position is not the model you rent. It is the rails the agent rides: who moves its money, who issues its credentials, who catches it when a test breaks. Pin your model IDs like dependencies so a price cut is a config change. Route every step by its cost. Treat payments, identity, and your eval harness as first-class infrastructure you own or pick on purpose. Rent the brain cheap. Own the plumbing.

Ron

You’re receiving this because you subscribed to The Agent Stack. · Unsubscribe