AI Limit Watcher
See which AI tools quietly tighten their usage limits, quotas and pricing — and how stable each vendor has been. Every change dated and sourced.
Across 108 tracked tightenings on 30 AI tools, 9% arrived unannounced — a figure we compute from our sourced record and publish nowhere else.
01 / RECENT CHANGES
- Aug 2026DeepSeek—↓[Pricing]DeepSeek V4 Flash input price $0.0882 → $0.14 per 1M tokens (+59%)
- Aug 2026DeepSeek—↓[Pricing]DeepSeek V4 Flash output price $0.18 → $0.28 per 1M tokens (+59%)
- Aug 2026DeepSeek—↑[Pricing]DeepSeek V4 Flash input price $0.14 → $0.0882 per 1M tokens (-37%)
- Aug 2026DeepSeek—↑[Pricing]DeepSeek V4 Flash output price $0.28 → $0.18 per 1M tokens (-37%)
- Jul 2026Cursor—→New product announcement; the blog and changelog listings disclose no change to limits, pricing, or model access.
- Jul 2026GitHub Copilot—↑Adds a new model option for agentic coding workflows.
02 / MOST TIGHTENING — LAST 90 DAYS
6 tightening · 7 shifting · 17 steady
↓ DeepSeek V4 Flash input price $0.0882 → $0.14 per 1M tokens (+59%) (Aug 2026)
↓ Shorter default fallback window for the backup LLM cascade in ConvAI agents; remains user-configurable within the 2–15s range. (Jul 2026)
↑ Restores GPT model access for Agent requests on Enterprise SaaS deployments after a period of intermittent failures. (Jul 2026)
↑ Adds a new model option for agentic coding workflows. (Jul 2026)
↑ Adds a new model capability for embodied navigation. (Jul 2026)
↓ Users lose unlimited slow-lane drafting; every generation now costs credits (though Max gives ~4x more credits than Unlimited's 2,250) (Jun 2026)
↓ Changes the metric for session sizing to focus exclusively on ACU consumption, potentially increasing cost visibility or impact for high-message-volume users. (Jul 2026)
↑ New features added including live avatar APIs and cloud rendering capabilities. (Jul 2026)
→ New product announcement; the blog and changelog listings disclose no change to limits, pricing, or model access. (Jul 2026)
→ Default-model swap for any app that does not pin a model — output behaviour can shift with no action by the builder (announced in the changelog). (Jul 2026)
↑ Faster model performance and new capabilities like website publishing and memory improve utility. (Jul 2026)
↑ New flagship Opus model at identical per-token pricing; Opus 4.8 still listed, so no forced migration or cost increase yet. (Jul 2026)
→ New model variants introduced, expanding available model options. (Jul 2026)
→ New $100 Pro mid-tier (Apr 2026)
↓ Free-tier creators effectively blocked from iterative daily use (Nov 2025)
→ Adds world-building context for Characters; beta with staged early access. (Jul 2026)
↓ Pro/Premier subscribers locked out of the flagship model at launch — must upgrade to Ultra (~$180/mo) for day-one access (Feb 2026)
↑ Annual subscribers benefit; monthly subscribers see no relief — pressure to commit annually (Sep 2025)
↑ Lower entry price for light users; the cheap tier is daily-capped vs the monthly-pooled higher tiers (Nov 2024)
→ Adds a new feature for reusable context and workflow management, reducing setup friction. (Jul 2026)
→ Roadmap risk of shifting toward enterprise/Canva integration over individual creators (Jul 2024)
↑ Adds a new advanced stem separation feature restricted to Premier subscribers. (Jun 2026)
↑ Adds new workflow automation capabilities for users. (Jun 2026)
↑ Self-hosters can drop API dependency/billing lock-in entirely (Jun 2026)
→ Reduces friction by enabling direct workflow integration between Claude and Replit. (Jun 2026)
↑ New model version released; default-model status not stated in the announcement. (Jul 2026)
↑ Allows users to manage network access policies directly within the chat interface, reducing the need to switch contexts to the web app. (Jul 2026)
↑ Adds granular control over audio processing, reducing potential for missed speech in noisy environments. (Jul 2026)
↑ Increases the maximum execution time for workflow steps to 30 minutes, allowing for longer-running background tasks. (Jul 2026)
→ Introduces visibility into usage metrics for Notion Workers, preparing users for future billing transitions post-beta. (Jul 2026)
03 / WHAT USERS REPORT
- “Blew past my five-hour limit after a single Claude Code prompt.” Claude source
- “Limits feel inconsistent and change without notice.” ChatGPT
- “Quota changed without prior notice and is very confusing.” Cursor
- “Free Gemini Pro access disappeared overnight.” Gemini
- “You have exceeded your premium request allowance.” GitHub Copilot
- “Bought an annual Pro subscription and the limits got quietly gutted.” Perplexity
- “The credit-to-quota switch changed how much I can actually do per day.” Windsurf
- “My credits now expire after six months — they didn't before.” Replit
- “The $300 Heavy tier is steep for what you get.” Grok
- “Cheap per-token, but the model lineup churns — V3.1, then V3.2-Exp, now V4, with deprecation dates each time.” DeepSeek
- “Video eats fast hours fast — about 8x an image job, so a Basic plan drains in no time.” Midjourney
- “Pinned eleven_multilingual_v1 in production — then it got a December removal date and I had to re-validate every voice.” ElevenLabs
- “I spent over $35 in a single day doing what previously cost me $20 for the entire month.” v0 source
- “ACUs cost $2.25 on the $20 plan, a hike from the $2 they cost on the $500-per-month subscription.” Devin source
- “Certain bots demand huge point quotas per message, making it impractical without a subscription.” Poe source
- “I don't know how many credits an operation will use before I run it, and credits are spent even on failed outputs.” Lovable source
- “Each failed fix attempt drains tokens without resolving the issue — easy to blow through a monthly allocation on one broken component.” Bolt.new source
- “Runway charges credits for failed or erroneous generations that need regenerating — you're charged for both attempts.” Runway source
- “Free songs can't be used commercially, and free users don't get the latest model.” Suno source
- “Credits shouldn't expire — at Kling they don't. If you keep your membership they should stack.” Luma Dream Machine source
- “Long-time users worry the Canva acquisition could shift focus from creators toward business customers.” Leonardo.ai source
- “Sora lead Bill Peebles: 'the economics are currently completely unsustainable.'” Sora source
- “The 'unlimited' messaging comes with a soft cap of ~150 messages/day under fair use.” Mistral Le Chat source
- “Topics that worked fine six months ago now trigger instant blocks, with unclear boundaries.” Character.AI source
- “200 Creator credits vanished in a single Avatar IV session — every render iteration counts, not just the final video.” HeyGen source
- “Re-renders consume additional minutes from the monthly allowance; production costs exceed advertised pricing.” Synthesia source
- “Teams that previously ran Plus + AI for $18/user now must jump to Business at $20/user for AI.” Notion AI source
- “Tabnine discontinued their Dev plan and pivoted to enterprise-only pricing.” Tabnine source
- “Free tier provides only 10 slow credits/week (~40 images), with no priority-queue access.” Ideogram source
We track whether these changes are real.
04 / EXPLORE
See exactly what changed in each AI tool →