AI Tech News
By H.O.

Why Claude's Free Tier Uses a Rolling 5-Hour Window Instead of a Daily Cap: How the Billing Mechanic Actually Works

The Window Is Not a Day. Here's Why That Matters.

Claude's usage limit runs on a rolling 5-hour window that starts from the moment you send your first message. This is the single most misunderstood fact about Claude's free tier, and it explains half the frustration people post on Reddit about getting cut off mid-afternoon.

The key difference: Claude's usage limit is not a daily message cap and does not reset at midnight. The five-hour window starts rolling when you send your first message. If you burn through your allocation at 10 AM, you won't see new capacity until 3 PM—not at 12:01 AM the next day. Heavy morning usage will cut you off by the afternoon.

This is a deliberate architectural choice by Anthropic, not a bug. To understand why, you need to see the problem it solves: capacity management and cost control in an era of exploding AI infrastructure spending.

The Scale Problem Behind the Design

Claude's rate limits are primarily about compute capacity, not money. Running Claude Opus 4.8 on complex tasks requires enormous GPU resources. A daily reset would be vulnerable to thundering herd effects—everyone worldwide could hammer the service at 12:01 AM their local time, creating massive simultaneous load spikes.

A rolling window flattens that problem. If your window started at 10 AM and someone else's started at 11 AM and a third user's at 9:30 AM, usage spreads across the hour rather than spiking at a fixed moment. Anthropic governs Claude Code through a dual-layer usage system: a 5-hour rolling window for short-term activity, and a weekly cap on active compute hours. The same bucket is shared across Claude Code, Claude.ai chat, and Cowork — burn tokens in one, you lose capacity in the others.

This change likely stems from rising AI infrastructure costs and capacity management, while also nudging power users toward paid tiers. But the first reason is real engineering: a rolling window is measurably better at keeping GPU utilization stable than a fixed daily reset.

What the Numbers Actually Are (and Why They're Fuzzy)

Anthropic publishes no fixed message count for any Claude tier. Every number you see online, including here, is reverse-engineered from users counting their own conversations. Reports cluster around 15 to 40 messages per rolling 5-hour window, with longer messages and attachments burning through it faster.

The fuzziness is intentional. Usage is a rolling budget per five-hour session, consumed by conversation length, file size, and complexity. A one-line question costs less than a 10,000-token research document with file attachments. Anthropic reduced 5-hour limits during weekday peak hours (5–11 AM PT) starting March 2026, and acknowledged on March 31 that users are hitting limits faster than expected. That's not a bug—it's the system doing exactly what it was designed to do.

Anthropic meters you with two things at once: a rolling five-hour session window that starts at your first message, and a weekly cap layered on top for paid plans. What's published is relative — Pro is at least 5x Free per session, Max 5x is about five times Pro, Max 20x about twenty.

Plan Rolling 5-Hour Window (Per-Session) Weekly Cap (Paid Only) Model Access
Free Approximately 15–40 messages None (rolling window only) Sonnet 4.6, Haiku 4.5
Pro ($20/month USD) Approximately 44,000 tokens per 5-hour period Yes (weekly allocation) Opus 4.8 + Sonnet/Haiku
Max 5x ($100/month USD) ~5x Pro window Higher weekly allocation All models, Claude Code
Max 20x ($200/month USD) ~20x Pro window Highest weekly allocation All models, Claude Code, priority

The Paid Plans Add a Second Meter on Top

Every plan has usage limits that reset on a rolling five-hour session window, and paid plans add weekly limits on top. Your activity across Claude on web, desktop, mobile, and Claude Code all draws from the same pool. This is crucial: you don't get a fresh budget on Pro after 5 hours. You get *both* the rolling window *and* a separate, slower weekly cap.

Chat sessions run on a five-hour clock. Weekly caps reset at a fixed hour assigned to your account. So hitting your 5-hour session limit doesn't mean you're done for the day—you can come back in 5 hours and continue. But Pro users also have a weekly ceiling, so heavy usage Monday through Thursday can lock you out by Friday regardless of the rolling window.

Why Rolling Windows Instead of Buckets

This rolling-window model is common across LLM providers — OpenAI and Google use similar quota structures. The elegance is that it meters both short bursts and sustained usage: Time, and nothing else. The allowance refills on its own when the current window ends, and the interface names that time when it stops you.

Unlike a bucket system (where you accumulate credits over time up to a maximum), a rolling window is stateless and fair. Your allocation five hours from now doesn't depend on how much you used yesterday. It depends only on what you used in the previous five hours.

From infrastructure perspective, this is elegant: This rate-limiting approach resembles a centralized quota manager: efficient for fairness, but rigid for flexibility. Anthropic knows at any moment how much total compute capacity the fleet has, and can apportion it across active windows. A fixed daily reset would require predicting tomorrow's load and either under-provisioning or over-spending on GPUs.

The Real Constraint: Capacity, Not Cost Recovery

One thing worth clarifying: Anthropic isn't using rolling windows to trick free users into upgrading. Since July 1, 2026, the free plan runs Claude Sonnet 5 as its default model, alongside web search, file uploads, Projects, and Artifacts, all of which used to sit behind the paywall. A genuinely crippled experience would push people to paid plans faster. Instead, Anthropic is trying to make the free tier useful—good enough for daily assistance—while preventing individual users from monopolizing GPU capacity during peak hours.

Our weekly tracking of frontier model releases shows how this economics scales: since July 1, 2026, the free plan runs Claude Sonnet 5 as its default model , reflecting Anthropic's bet that a strong free tier with smart capacity management beats a weak free tier and aggressive upselling.

The honest answer to "how many times can I use Claude for free" is: enough for short daily questions, and not enough for a working session. The free plan is metered by usage inside a rolling window rather than by a fixed lifetime quota, so it refills on its own and you can come back later the same day.

Practical Takeaway: When the 5-Hour Window Helps You

The rolling window is a feature if you use it right. Heavy usage in the morning affects your availability in the afternoon. But if you spread work across time—a code review at 9 AM, a research prompt at 2 PM, another task at 6 PM—you can sustain useful work across a full day despite a tight per-session limit. The window resets independently for each session, so your 3 PM prompt count is unrelated to your 9 AM usage.

For teams considering Pro or higher: the economics tilt toward a subscription once you need sustained, daily access. API billing only beats Pro if you're below roughly 50 sessions per month. For everything above that, a $20 USD monthly subscription gives you a much larger rolling window, a weekly cap to rely on, and Claude Code access. For US and UK readers using Claude as part of daily development or writing work, the Pro tier typically pays for itself in productivity within a single week.

Our tracked data

AI Intelligence Index (Top 3 Frontier Models)

01632476305-1706-0106-0807-0607-1307-2007-2708-0308-1008-1708-2408-3109-07Claude Opus 4.7 (Adaptive Reasoning, Max Effort) — Anthropic: 57 (2026-05-17)Claude Opus 4.8 (Adaptive Reasoning, Max Effort) — Anthropic: 61 (2026-06-01)Claude Opus 4.8 (Adaptive Reasoning, Max Effort) — Anthropic: 61 (2026-06-08)Claude Opus 4.8 (Adaptive Reasoning, Max Effort) — Anthropic: 56 (2026-07-06)Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) — Anthropic: 60 (2026-07-13)Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) — Anthropic: 59.9 (2026-07-20)Claude Opus 5 (Adaptive Reasoning, Max Effort) — Anthropic: 61 (2026-07-27)Claude Opus 5 (Adaptive Reasoning, Max Effort) — Anthropic: 61 (2026-08-03)Claude Opus 5 (Adaptive Reasoning, Max Effort) — Anthropic: 63 (2026-08-10)Claude Opus 5 (Adaptive Reasoning, Max Effort) — Anthropic: 63 (2026-08-17)Claude Opus 5 (Adaptive Reasoning, Max Effort) — Anthropic: 63 (2026-08-24)Claude Opus 5 (Adaptive Reasoning, Max Effort) — Anthropic: 63 (2026-08-31)Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) — Anthropic: 57 (2026-09-07)57GPT-5.5 (xhigh) — OpenAI: 60 (2026-05-17)GPT-5.5 (xhigh) — OpenAI: 60 (2026-06-01)GPT-5.5 (xhigh) — OpenAI: 60 (2026-06-08)GPT-5.5 (xhigh) — OpenAI: 55 (2026-07-06)GPT-5.6 Sol (max) — OpenAI: 59 (2026-07-13)GPT-5.6 Sol (max) — OpenAI: 58.9 (2026-07-20)GPT-5.6 Sol (max) — OpenAI: 59 (2026-07-27)GPT-5.6 Sol (max) — OpenAI: 59 (2026-08-03)GPT-5.6 Sol (max) — OpenAI: 61 (2026-08-10)GPT-5.6 Sol (max) — OpenAI: 61 (2026-08-17)GPT-5.6 Sol (max) — OpenAI: 61 (2026-08-24)GPT-5.6 Sol (max) — OpenAI: 61 (2026-08-31)GPT-6 Astra (max) — OpenAI: 55 (2026-09-07)55Gemini 3.1 Pro Preview — Google DeepMind: 57 (2026-05-17)Gemini 3.1 Pro Preview — Google DeepMind: 57 (2026-06-01)Gemini 3.1 Pro Preview — Google DeepMind: 57 (2026-06-08)Gemini 3.1 Pro Preview — Google DeepMind: 46 (2026-07-06)Gemini 3.5 Flash (high) — Google DeepMind: 55 (2026-07-13)Gemini 3.1 Pro Preview — Google DeepMind: 46 (2026-07-20)Gemini 3.6 Flash (high) — Google DeepMind: 50 (2026-07-27)Gemini 3.6 Flash (high) — Google DeepMind: 50 (2026-08-03)Gemini 3.6 Flash (high) — Google DeepMind: 52 (2026-08-10)Gemini 3.7 Flash (high) — Google DeepMind: 56 (2026-08-17)Gemini 3.7 Flash (high) — Google DeepMind: 56 (2026-08-24)Gemini 3.7 Flash (high) — Google DeepMind: 56 (2026-08-31)Gemini 3.8 Flash (high) — Google DeepMind: 59 (2026-09-07)59
  • Anthropic
  • OpenAI
  • Google DeepMind

Intelligence Index — Trend

Hover over each point to see the specific model version at that date.

Last updated: 2026-09-07 · 13 data points · artificialanalysis.ai

Collected weekly by our editorial team from primary sources.

See the full dataset