back to top
HomeTechClaude Just Doubled Its Usage Limits. The Real Story Is the SpaceX...

Claude Just Doubled Its Usage Limits. The Real Story Is the SpaceX Deal Behind It

- Advertisement -

Claude users have spent months playing a very specific game, how much work can you squeeze out of Opus before the rate limits slam shut?

Anthropic is finally loosening things up. The company says it’s doubling Claude Code limits, removing peak-hour reductions for paid users, and significantly raising Opus API caps. The reason is also there in the same announcement. Anthropic now has access to all the compute capacity at SpaceX’s Colossus 1 data center. That’s over 220,000 NVIDIA GPUs.

That’s the kind of announcement that makes you realize AI companies aren’t just shipping models anymore. They’re building power infrastructure.

This Isn’t really about Chatbot Anymore

The easiest way to understand Anthropic’s announcement is to look at what people are actually using Claude for now.

A normal chatbot session doesn’t burn through compute like this. You ask a few questions, maybe upload a file, then leave. That’s not the workload forcing companies to secure hundreds of thousands of GPUs.

Claude Code is different. People are leaving it running for hours inside terminals. They’re asking it to refactor projects, debug broken dependencies, analyze huge codebases, write tests, retry failed tasks, call tools, and keep context alive across long sessions. One coding agent can easily consume far more tokens than a casual chat user ever would.

That’s probably why the double limits part of Anthropic’s announcement matters more than it first appears.

For months, Claude had this weird split reputation among developers. The model itself was excellent, especially for coding, but heavy users kept running into walls. Long sessions would suddenly slow down. Opus users learned to ration prompts. Some developers even changed workflows entirely around avoiding rate limits.

Now Anthropic is signaling something pretty clearly, they expect usage to keep getting heavier And honestly, that tracks with where AI tooling is heading. The industry keeps talking about chatbots, but the real compute monster may end up being autonomous agents quietly running in the background for hours at a time.

TierMaximum Input Tokens per MinuteMaximum Output Tokens per Minute
130,000 -> 500,0008,000 -> 80,000
2450,000 -> 2,000,00090,000 -> 200,000
3800,000 -> 5,000,000160,000 -> 400,000
42,000,000 -> 10,000,000400,000 -> 800,000

Those new limits aren’t small bumps either. Some Opus tiers are seeing output caps increase by 10x. Which also explains why Anthropic immediately followed the announcement, a compute partnership with SpaceX.

You May Like Best AI Coding Models for Consumer Hardware

The real flex wasn’t the higher limits

The company says its SpaceX partnership gives it access to all of the compute capacity at Colossus 1, adding more than 300 megawatts of capacity and over 220,000 NVIDIA GPUs within the month.

That’s an absurd amount of compute for what most people still casually describe as “a chatbot.”

And Anthropic isn’t talking like this in isolation anymore. Over the past year, major AI companies have increasingly started announcing infrastructure deals the way they used to announce model launches. Amazon, Google, Microsoft, NVIDIA, xAI everybody suddenly sounds half software company, half utility provider.

The shift makes sense once you look at where AI usage is heading. Reasoning models are expensive to run. Coding agents stay active for long stretches. Enterprise workloads don’t disappear after a few prompts. The more capable these systems become, the more compute they consume in the background.

A year ago, companies competed on benchmark screenshots. Now they’re competing on who can secure enough GPUs to keep the systems online at scale.

This feels like the next phase of the AI race

For a while, AI companies competed mostly on model quality. Better benchmarks, reasoning, coding. But now the conversation is shifting towards, who can actually keep these systems running at scale without constantly slamming users into limits.

Anthropic partnering with SpaceX for compute sounds unusual today. It probably won’t a year from now. Because at this point, the bottleneck isn’t really ideas anymore. It’s power, GPUs, and who can secure enough infrastructure before everyone else does.

Don’t miss any Tech Story

Subscribe To Firethering NewsLetter

You Can Unsubscribe Anytime! Read more in our privacy policy

LEAVE A REPLY

Please enter your comment!
Please enter your name here

YOU MAY ALSO LIKE

The Biggest AI Companies Are All Building Their Own Chips. That’s Not a Coincidence.

0
Anthropic confirmed this week it's hiring a custom silicon team to design chips for running Claude. The announcement was quiet a job listing, a spokesperson confirmation, no big launch event. Easy to file under "interesting but expected" and move on. But zoom out for a second. OpenAI shipped its first custom inference chip in June. Google has been running models on its own TPUs for years. Meta has designed and deployed its own silicon. Mistral is reportedly exploring the same path. And now Anthropic. Five of the most important AI labs in the world, all arriving at the same decision, within roughly the same window. None of them are copying each other. All of them looked at the same competitive landscape and reached the same conclusion independently. That kind of convergence doesn't happen by accident. It happens when an entire industry agrees that the thing everyone assumed was someone else's problem is actually the problem and that whoever solves it first has an advantage that's very hard to close later.
AI Was Supposed to Stop Cheating. Instead, 58,000 Students Must Retake Their Exams

AI Was Supposed to Stop Cheating. Instead, 58,000 Students Must Retake Their Exams.

0
UNAM runs the largest university in Mexico. Every year, hundreds of thousands of students take an entrance exam that determines whether they get in. This year, for the first time, the whole thing went remote. They deployed a lockdown browser, AI webcam monitoring, and one human supervisor per 150 applicants. The kind of setup that sounds serious on paper. Then the scores came in. Students hitting 100 or above jumped from 3.5 percent in previous years to 16.3 percent this year. At the very top end, scores of 110 or higher went from 0.9 percent to 5.5 percent. Not a small shift. Not noise. A roughly fivefold increase in top scores, in one year, under one new format. An expert commission investigated. Their conclusion: administer the entire exam again, in person, to around 58,000 people. The rector apologized to students who hadn't cheated. They now have to prepare for and sit another exam anyway.
Claude Chats Ended Up on Google Search. Here's How It Happened

Claude Chats Ended Up on Google Search. Here’s How It Happened.

0
A single line typed into Google was all it took. Type "site:claude.ai/share" into the search bar, and over the weekend, it surfaced a long list of conversations people had shared through Claude, Anthropic's AI chatbot. Not conversations they'd shared with the world on purpose. Conversations they'd shared with one person, or thought they had.Some of what turned up reads like exactly the kind of thing you'd never want indexed anywhere. Medical records. Children's names and phone numbers. Internal company documents marked for employees only. This wasn't a hack, no one broke into anything. It was a feature working exactly as built, surfacing exactly what people had typed into it, in ways most of them almost certainly never intended.