The AI Firehose Podcast
← All episodes
Episode 1 · May 8, 2026 · 09:26

New Claude AI Agents DESTROY OpenClaw?

Claude AI Dreaming: The Feature That Just Killed OpenAI AgentsAnthropic has introduced 'Dreaming' for Claude AI agents, allowing them to learn from past mistakes and improve autonomously while offline. Discover how these new features, including self-grading Outcomes and Multi-agent Orchestration, are revolutionizing business automation.00:00 - Intro00:20 - What is Claude AI Dreaming?01:45 - How Dreaming Works in Practice02:57 - Outcomes: Agents Grading Themselves04:04 - Multi-Agent Orchestration05:02 - Webhooks and Task Notifications07:22 - Why Claude is Outpacing OpenAI

Full transcript

New Claude AI agents destroy OpenClaude. Claude's dreaming AI just killed OpenClaude, and most people have no idea what just happened. Anthropic let their AI agents dream. Yes, dream.

Like humans do at night. The early test results are kind of insane. One company, Harvey, the legal AI firm, said their agent task completion rates went up six times. Six times.

After the agent slept on it. So here's what's going on. Code with Claude event in San Francisco, Anthropic dropped a feature called Dreaming. It's in research preview right now.

And it's wild. While your agent is not working, it goes through up to 100 of its past chats, looks at what worked, looks at what failed. It spots patterns you and the agent could never spot on your own. Then it cleans up the agent's memory, removes duplicate notes, fixes things that don't match, locks in the lessons.

So next time you wake it up, it's a smarter version of itself. Think about what that means for a second. Every other AI agent today wakes up dumb. Same as yesterday.

You have to remind it of stuff. You have to fix the same mistakes over and over. Now Claude's agent fixes itself while you sleep, while you eat, while you're with your kids. Hey, if we haven't met already, I'm the digital avatar of Julian Goldie, CEO of SEO agency, Goldie Agency.

Hosties helping clients get more leads and customers. I'm here to help you get the latest AI updates. Julian Goldie reads every comment, so make sure you comment below. Tom Brown, the CTO of Anthropic, said inference on the new compute would start ramping within days.

Evora, head of product, said on stage, they're now using all the compute capacity at SpaceX's Colossus One data center. 220,000 NVIDIA GPUs, megawatts of new power. This thing isn't just smart, it's getting fast too. Doubled the Claude code rate limits for paid users.

They raised the API limits on Opus. They removed peak hour limits, so you can use Claude way more without hitting walls. That alone is huge for anyone running agents in their business. Now, let me show you what dreaming actually does.

An agent that handles your customer emails, day one, it makes a small mistake, uses the wrong tone. Day two, it does the same thing. Day three, same thing. Old agents would keep doing this forever.

Claude's new agent dreams between sessions. He's back through those emails, sees the pattern, goes, wait, I keep messing this up. And it writes a new note in its memory that says, use a friendly tone, not a formal one. Next session, fixed, without you saying a word.

That's the part that breaks my brain. The agent learns from its own mistakes by itself, while doing nothing. Imagine if a new hire could do that. You'd never have to train anyone twice.

Here's a quick word for you if you run a business. If you want to learn how to set up Claude's new dreaming agents inside your business, save tons of time and put your followups, content lead gen, and customer support on autopilot, come join us in the AI Profit boardroom. We've got four weekly coaching calls where you can ask live questions about your Claude agent setup. Daily step-by-step tutorials walking you through exactly how to use these new features.

A 30-day roadmap built around Claude agents. Prompt library full of stuff that works. Two for my business owners building with Claude right now. Many of them already running these agents on the new dreaming feature.

Member map so you can connect with people near you. Link in the description or go to AIProfitBoardroom.com. Okay, back to the meat. Thropic also dropped three more updates with dreaming.

These are just as big. The first one is called outcomes. Now this is in public beta. Anyone can use it.

Outcomes is simple. You write down what a good result looks like, like a checklist. The agent runs the task. Then a second Claude, a totally separate one, grades the work.

Grader doesn't see how the agent got there. Just looks at the final answer. If it fails the checklist, the grader points out exactly what's wrong. The agent tries again.

Again. Until it passes. Why is this big? Because most agents today just hand you stuff and hope it's right.

You read it. You fix it. You waste hours. Now the agent fixes its own work first.

Thropic's own test showed task success went up by 10 points. The hardest tasks saw the biggest jump. Outputs like word docs went up 8%. Slide decks went up 10%.

So if you've ever asked an agent to make a sales deck and got back something off, this fixes that. Here's a real example for you. Want an agent to write a weekly newsletter for your business. You write a rubric.

Must have a strong hook. Must have one customer story. Must end with one clear call to action. The agent writes.

The grader reads it. Misses the hook. Send it back. Misses the story.

Send it back. You wake up. The newsletter is done. And it's right.

First try. From your point of view. Then they dropped multi-agent orchestration. Public beta 2.

This one's a beast. Here's what it does. You set up one lead agent. The lead agent breaks a big job into smaller parts.

Then it hands each part to a specialist agent. Each specialist has its own brain, its own tools, its own job. They all work at the same time. Side by side.

So 20 different agents can work on one project together. Share the same files. Check in with each other. The lead agent watches the whole show.

Multi-agent, the agent looks at batches in parallel. Instead of one slow check, it's many fast checks at the same time. They only get told about the patterns that actually matter. Spiral, this writing tool from a company called Every, is using outcomes and multi-agent together.

Their lead agent runs on Haiku, the fast cheat model. It takes the order. It asks quick follow-up questions. Then it sends the actual writing to bigger opus agents.

If you ask for three drafts, three opus agents work on them at the same time. You get three options in the time it would take to do one. And then there's webhooks. Sounds boring.

It's not. Webhooks let you set an outcome, walk away, and get a ping when it's done. So you don't sit and watch the agent work. You go live your life.

The agent finishes the job. You get a notification, like a text from a co-worker that just says done. Now think about all four of these together. Dreaming, outcomes, multi-agent orchestration, webhooks.

The agent learns. The agent grades itself. The agent splits jobs across a team. The agent tells you when it's done.

Have you noticed the gap with OpenClaw and other tools right now? Stuck running one task at a time. Don't dream. They don't grade themselves.

Forget what they did yesterday. So if you've been pushing OpenClaw to run your business, you've been hitting the same wall over and over. Claw just jumped that wall. The Harvey result is the one that makes me sit up.

They had legal agents handling document workflows, file type quirks, tool patterns, all that boring stuff that breaks AI agents in the real world. Before dreaming, the agents kept tripping on the same junk. After dreaming, six times more tasks, done. That's not a small lift.

That's the difference between an agent that costs you time and an agent that gives you time back. Now let's get real about what this means for you. Most business owners are thinking about AI like it's a fancy chat tool. Type something, get something back.

That's not what's happening anymore. Agents now run jobs. They run them in the background. They get better while you sleep.

They check their own work. Tell you when they're done. This is the moment. The folks who set this up now are about to leave everyone else behind.

Not in a year, a few months. Because once an agent has been dreaming for 30 days, it's a totally different worker than when it started. Imagine an employee that gets 6% better every week without you doing anything. That compounds fast.

Few more things from the event worth knowing. The whole multi-agent setup is fully traceable. You can pop open the Clawed console and see exactly which agent did what, in what order, and why. So if something goes wrong, you can find it.

No mystery. No guessing. The events are persistent too. Meaning agents don't forget the steps mid-job.

They remember where they left off. So you can run jobs that take hours, days, even longer. The pricing on managed agents is $0.08 per agent runtime hour on top of normal model cost. Dreaming, outcomes, and webhooks are not extra.

They're included in that hour rate. So you're not paying more for the agent to grade itself or learn from past sessions. It just does it. And you can plug this into other things you already use.

Webhooks fire off to your CRM, your email tool, your Slack. So when your lead gen agent finishes a batch, your team gets a ping. When your support agent flags a tricky reply, you get pulled in. Otherwise, the agents run quiet.

I want to talk about why this matters more than the headline does. AI agents have had one main problem. They're forgetful. You set them up.

You teach them. You train them. The next session, half the lessons are gone. They make the same mistakes.

They use the wrong tone. They miss the same details. Most folks gave up on agents because of this. They felt like a goldfish that needed a babysitter.

Dreaming kills that. It turns the goldfish into a worker that grows. Outcomes turns the worker into a worker that checks itself. Multi-agent turns the worker into a small team.

Webhooks turns the team into one that reports back. Stack all four, and you've got something that wasn't possible six months ago. Claude is built around the idea of an agent in a browser. It's good at clicking, scrolling, filling out forms, but it doesn't dream.

It doesn't grade itself. It works one task at a time. So Claude just walked past it on the curve that actually matters, which is, can the agent learn and run jobs without you babysitting? Right now, Claude can.

Claude can't. Their agent tools can't yet either. If you're someone who tried agents six months ago and gave up, this is your sign to look again. The thing that made you quit just got fixed.

The agent that forgot your rules now keeps them. The agent that got tone wrong now grades its own tone. The agent that took forever now splits the job across 20 helpers. Last quick word.

If you want the full setup guide for these new Claude agents, plus four weekly coaching calls where we go deep on dreaming, outcomes, and multi-agent orchestration setups, come join the AI Profit Boardroom. You'll get a 30-day roadmap built around Claude's new agent stack. Daily tutorials walking you through how to put your support, content, follow-up, and lead gen on autopilot with these features. Prompt library tuned for Claude agents.

Business owners building this stuff right now, many already running dreaming agents in their business. A member map so you can connect with Claude users near you. Link in the description or go to AIProfitBoardroom.com. And one more thing.

If you want the full process, SOPs, and over 100 AI use cases like this one, come join the AI Success Lab. It's free. Links in the comments and description. You'll get all the video notes from this episode plus access to our community of 68,000 members crushing it with AI right now.

See you in the next one.

More episodes

Browse all episodes →