The Great AI Bait-and-Switch: How Usage Limits Quietly Tighten After Adoption

INVESTIGATION SUMMARY

► Pattern Identified: Industry-Wide Bait-and-Switch

Every major AI service follows the same playbook: subsidize generous access to build dependency, then quietly tighten limits once users are locked in.

This pattern is now documented across OpenAI, Anthropic, Cursor, GitHub Copilot, Google Gemini, Perplexity, and Replit—with specific dates, limit reductions, and community backlash tracked in detail.

The practice is enabled by deliberate opacity: not a single consumer-facing AI chat product publishes exact usage calculation formulas, making it nearly impossible for subscribers to verify they're getting what they pay for.

Approximately 200 individual complaints sit in the FTC's Consumer Sentinel database as of September 2025, though no formal enforcement actions have targeted this pattern yet. The underlying economics make further tightening inevitable: AI companies collectively spend hundreds of billions more on inference than they earn, and the VC subsidies funding current pricing are unsustainable.

01

Anthropic's Claude: From Concrete Numbers to Deliberate Vagueness

The Evolution of Opacity: When Claude Pro launched in late 2023, Anthropic stated a concrete limit: approximately 100 messages per 8 hours for $20/month subscribers. This was one of the last times Anthropic gave a specific number.

Late 2023

Claude Pro Launch — Approximately 100 messages per 8 hours for $20/month. Concrete, verifiable numbers.

March 2024

Claude 3 Launch — Window quietly shifted to 5 hours rolling. Community testing pegged Pro limits at roughly 45 messages per 5-hour window—less than half the original promise. Anthropic began removing specific numbers from documentation, replacing them with relative language like "5x more usage than free."

The Meaningless Multiplier: Without publishing what "1x" equals, this multiplier is mathematically meaningless. As CheckThat.ai noted: "The stated 'usage multiplier' quantifies relative allowances, but without published baselines, the absolute amounts remain unknown."

Token-Based Reality: The system underneath is token-based, not message-based, though Anthropic has never formally confirmed this. Each message's cost includes the entire conversation history (up to 200K tokens reprocessed each turn), meaning a message at turn 50 costs dramatically more than one at turn 1. File uploads, extended thinking, web search, and tool calls all consume additional tokens. Community reverse-engineering suggests Pro allocations of roughly 44,000 tokens per 5-hour window—but Anthropic provides no token counter in the consumer UI.

The January 2026 Controversy

60% Reduction Revealed: Anonymous Anthropic customer provided The Register with analysis showing "roughly a 60 percent reduction in token usage limits, based on a token-level analysis of Claude Code logs."

Anthropic doubled limits as a "holiday gift" from December 25–31, 2025, using idle enterprise compute. When the bonus expired on January 1, users experienced what felt like a sharp reduction. Anthropic dismissed this as users "simply reacting to the withdrawal of bonus usage"—but GitHub Issue #9424 documented users across Pro, Max 5x, and Max 20x reporting "weekly usage limits being depleted in 1-2 days of normal use, making paid subscriptions effectively unusable for 5-6 days per week."

Perhaps Most Damning: Anthropic stores detailed token usage data locally in undocumented JSONL files (~/.claude/projects/), with exact input_tokens and output_tokens fields per message—but deliberately does not surface this data to users. The AI Engineering Report's exposé called this out directly: "Anthropic hides your cost data, but here's how to track tokens." The data exists. The company simply chooses not to show it.

02

Cursor's June 2025 Pricing Shock and the Death of Predictability

Cursor's trajectory is the most dramatic single-event case of the bait-and-switch pattern. For over a year, Pro subscribers ($20/month) received a straightforward deal: 500 fast requests per month with unlimited slower requests as a safety net. Users understood the system, could plan around it, and generally found it generous enough for serious coding work. Cursor exploded to $500M+ ARR faster than OpenAI.

June 16, 2025

The Rug Pull — Weeks after raising $900M at a $9.9B valuation, Cursor replaced the 500-request system with a $20 credit pool pegged to API pricing. Under the new math, that $20 bought approximately 225 Claude Sonnet 4 requests, not 500. The previously unlimited slow requests were eliminated. Users received no advance email notification. Existing subscribers were silently migrated.

July 3, 2025

Old Option Removed — The old 500-request fallback option was quietly removed.

Immediate Backlash: One Reddit user reported going from roughly $100/month to "$20-30/day." A long-time Pro user found $10 in unexpected usage charges within the first week, with the request counter having vanished from the UI. The r/cursor subreddit erupted with "rug pull" language.

Hacker News Reaction: "Cursor raised $900M, are losing market share to Claude Code... AND they're decreasing the value of their product? Huge red flag."

CEO Apology (July 4, 2025): Michael Truell issued a public apology: "We recognize that we didn't handle this pricing rollout well, and we're sorry." The company offered full refunds for unexpected charges and temporarily allowed users to revert to the old model.

Developer Oliver Holl (Medium): "Had they been open and transparent, I'd have defended them. Instead, they quietly changed things, repeatedly, and rightfully lost user trust."

The Hidden Pro+ Tier

A hidden tier called "Pro+" ($60/month) was never listed on the pricing page—users only discovered it when hitting rate limits and seeing an upsell prompt. Truell admitted the company was "still figuring out how to include the tier without introducing too much complexity." The pattern is clear: reduce the base tier's value, then monetize the frustration with premium upsells.

03

The Industry-Wide Pattern: OpenAI, Copilot, Perplexity, and Gemini

The bait-and-switch is not limited to Claude and Cursor. Every major AI service has followed some version of this playbook.

OpenAI / ChatGPT

🎁 API Free Credits Evolution

  • Initial: $18 in free credits for new signups
  • Mid-2023: Dropped to $5
  • Mid-2025: Eliminated entirely

💬 ChatGPT Plus Message Caps

  • March 2023: 25 messages/3 hours
  • July 2023: Briefly doubled to 50
  • Late 2023: ~24 messages (despite official cap staying "40")
  • Users hit limits at 21 messages when official number said 40

Corporate Vagueness: OpenAI's explanation—"we're currently dynamically adjusting usage caps as we learn more about demand"—became a meme for corporate vagueness.

Unprofitable Economics: Sam Altman publicly admitted the company loses money on $200/month ChatGPT Pro subscriptions, having spent $8.67 billion on inference in the first nine months of 2025 alone.

GitHub Copilot

Most Dramatic Structural Change: From truly unlimited completions to a capped "premium requests" system. The shift began in December 2024 with the free tier (50 premium requests/month) and expanded to all paid plans by mid-2025.

The Multiplier System: Pro users received 300 premium requests/month—sounds reasonable until you encounter the multiplier system. A single GPT-4.5 interaction costs 50 premium requests (16.7% of the monthly allowance). Claude Opus 4 costs 10x.

Silent Downgrades: Visual Studio Magazine ran a piece titled "Beware Project-Wrecking GitHub Copilot Premium SKU Quotas" after a journalist's project was derailed when auto-downgrade to GPT-4.1 triggered silently mid-work.

Student Impact: Student Pack users who previously had unrestricted access were especially vocal: "It now enforces 300 premium request limits, where previously there was no such restriction."

Perplexity

Most Aggressive Reduction: MakeUseOf documented that Deep Research queries were slashed from approximately 500-600 per day to just 20 per month—a 99%+ reduction.

  • File uploads went from unlimited to 50/week
  • Pro Search shifted to "vague weekly limits that Perplexity's own support staff admitted they can't disclose exact numbers for"
  • Annual subscribers who paid $200 upfront found "the product looks nothing like the one you signed up for"
  • No email notification sent to existing subscribers
  • No grandfathering offered

Google Gemini

API Free Tier Cuts (December 2025): One of the industry's most generous free tiers—up to 10,000 requests/day for Gemini 2.5 Pro. Free-tier rates were cut 50-80% across the board, with Gemini 2.5 Pro removed from the free tier entirely for many users.

The Accidental Generosity: Google's Lead PM Logan Kilpatrick admitted on Reddit that the generous 2.5 Pro free tier was "originally only supposed to be available for a single weekend" but accidentally lingered for seven months.

Credit Where Due: To Google's credit, they eventually published specific daily prompt limits per tier—making them the most transparent consumer AI service on this dimension.

04

Free Credits to Dependency: The Economics of an Unsustainable Model

Timing Pattern: The pattern correlates tightly with adoption milestones and funding rounds. Cursor's pricing change came within weeks of hitting $500M ARR. Replit's shift to "effort-based pricing" (which saw some users' costs jump from $180/month to $1,000/week) coincided with a $250M funding round.

The Burn Rate Reality

💸 Company Economics

  • Anthropic: Burns 70 cents of every revenue dollar
  • OpenAI: Spent nearly double its revenue on inference
  • Max Plan Users: Consuming over $1,000 worth of API calls daily for $200/month subscription

📈 Future Projections

  • Uptech Studio: Current API pricing may need to increase 3-10x to reach sustainable economics
  • WEKA Analysis: "Surge pricing for tokens will be a 2026 wake-up call"
  • VC Warfel: "A meaningful share of AI demand may be inflated by venture-capital subsidies"

The Cost-Reduction Treadmill

Impressive Efficiency Gains: Stanford's AI Index found that inference costs for GPT-3.5-level performance dropped 280-fold between November 2022 and October 2024. Wing VC calculated GPT-4-equivalent intelligence costs fell 240x in 18 months.

But Usage Grows Faster: Reasoning models use 5-10x more tokens per request. Agentic coding tools run continuously. The treadmill accelerates.

Anthropic CEO Dario Amodei (DealBook Summit, December 2025): "There are some players who are YOLO, and I'm very concerned."

Dr. Christina Inge (Harvard): "The idea is to get people to try your product by offering it for an unsustainably low fee, and then when you get traction and users relying on it, increase the price to the point that you can be profitable. It's the same pricing strategy that has been used by companies like Uber, DoorDash, Canva."

05

The Regulatory Gap: 200 Complaints But No Enforcement

Despite the pattern's prevalence, regulatory action has been minimal. The FTC's Consumer Sentinel database contained approximately 200 complaints against major AI companies (OpenAI, Anthropic, xAI) as of September 2025, obtained through FOIA by FedScoop.

Actual User Complaints (FTC Database)

📋 Complaint Excerpts

  • "I upgraded to the Claude Pro Max plan ($100/month) based on clear advertising that promised 10x higher usage limits. Despite paying for this tier, I am still unable to initiate basic new chats."
  • "I signed up for an AI program called Claude...paid for a full year subscription to the Pro account, which at the time said unlimited usage...I barely got started...when I was told I had to quit for the day. So, not very unlimited."

No Formal Investigation: However, the FTC has not launched any formal investigation specifically targeting the progressive limit reduction pattern. Its AI enforcement (Operation AI Comply, September 2024) focused on deceptive capability claims—DoNotPay ($193K settlement for claiming AI could replace lawyers), not on subscription pricing practices.

Legal Framework Exists But Remains Unapplied

🇺🇸 US Framework

  • FTC Warning (Feb 2024): "A business that collects user data based on one set of privacy commitments cannot then unilaterally renege on those commitments"
  • Instacart Settlement: $60M for deceptive "free delivery" claims establishes relevant precedent
  • State Level: Transparency Coalition tracks AI pricing bills across 27 US states
  • Kentucky: Price Fairness Act prohibiting surveillance pricing using AI

🇪🇺 EU Framework

  • AI Act: Implementation through 2027
  • Multiple algorithmic pricing investigations underway
  • None specifically target AI service limit reductions yet

The Gap: No class action lawsuits have been filed against any major AI provider specifically for the free-to-paid limit reduction pattern. The gap between documented consumer harm and regulatory action remains wide.

06

Who Publishes Formulas and Who Hides Them

The transparency gap between API products and consumer products is stark. A forensic assessment of major services reveals a clear hierarchy:

✅ Transparent

  • Google Gemini: Publishes exact daily prompt limits per tier and feature
  • OpenAI API: Documents RPM, TPM, and RPD per model per tier with real-time dashboard
  • Windsurf: Publishes credit multipliers per model, rated 10/10 for pricing metric clarity
  • GitHub Copilot: Publishes specific premium request counts and model multipliers

🔄 Partially Transparent

  • Cursor: Publishes dollar-based credit pools but shifted systems without adequate explanation
  • Perplexity: Publishes some limits but support staff "can't disclose exact numbers" for others
  • Amazon Q Developer: Lists some limits in AWS service quotas but lacks clarity on general interaction caps

❌ Deliberately Opaque

  • Anthropic Claude (Consumer): Publishes ZERO specific numbers—no message counts, token thresholds, or formulas for any tier
  • ChatGPT (Consumer): No usage dashboard, no remaining-message indicator, uses "dynamic adjustment" language
  • Replit: Credit system described as "buried or missing from the pricing page"

🛠️ Community Response

  • At least 6 Claude usage trackers exist (lugia19 extension leads with crowdsourced data)
  • Over 8 Cursor trackers built (Cursor Stats most popular)
  • Multiple ChatGPT trackers (Chatterclock ~5,000 active users)
  • All built by unpaid volunteers — indictment of providers' refusal to surface data they already collect

The "Compute Unit" Abstraction Trend

The trend accelerates the opacity by making cross-service comparison impossible:

  • GitHub: "Premium requests" with multipliers up to 50x
  • Cursor: Dollar-denominated token pools
  • Windsurf: "Prompt credits"
  • Replit: "Usage credits"
  • Cognition's Devin: "Agent Compute Units"

The Point: When every service uses a different abstract currency, comparison shopping becomes impossible—which is precisely the point. As one industry analysis noted: changing credit exchange rates is less visible than raising prices, and abstract units create "perception of abundance" while enabling "dynamic cost shifting."

07

What the Tightening Cycle Means Going Forward

The evidence points to an industry approaching an inflection. The VC-subsidized era of generous AI access is ending, but the opacity with which limits are being tightened—rather than prices honestly raised—is eroding the trust that companies spent billions building.

Three Dynamics Will Intensify

🧠 Reasoning & Agentic Models

Consume 5-100x more tokens than simple chat, meaning the same "message" now costs providers dramatically more, accelerating limit pressure.

🔧 Proprietary Development

Cursor's October 2025 "Composer" model, Google's Gemini 3 family may eventually reduce third-party API costs and allow more generous limits—but only for providers who succeed in building competitive models.

Regulatory Pressure Building Slowly:

  • 27 US states have active AI pricing bills
  • EU's algorithmic pricing investigations expanding
  • 200 FTC complaints represent a growing paper trail

Most Practical Takeaway: The services most likely to maintain stable pricing are those with the deepest pockets and most diversified revenue—Google, Microsoft, and Amazon—rather than the venture-backed pure-play AI companies where, as Dr. Inge warned, they're "most likely to struggle to find a profitable pricing model and pull the rug."

The pattern isn't a bug. It's the business model.

📊 Related Research

🚫
ChatGPT Evolving Restrictions

How progressive restrictions reshaped the AI landscape, driving professionals to multi-tool strategies.

READ REPORT
📉
ChatGPT Reality Gap

Disparities between claimed and actual capabilities across coding, math, and creative domains.

READ REPORT
🗺️
The Real Power Map

Behind the AI industry: who controls what, how consolidation works, and the actual infrastructure.

COMING SOON