https://efficienist.com/claude-code-may-be-burning-your-limits-with-invisible-tokens-you-cant-see-or-audit/ Skip to content Efficienist Efficienist * AI * Agents * Enterprise * Productivity * SaaS Newsletter * Contact Efficienist Efficienist AI Claude Code may be burning your limits with invisible tokens #ByIvan Jenic April 13, 2026April 13, 2026 A developer's experiment suggests Anthropic may be silently injecting thousands of tokens into Claude Code requests, and users' usage limits are taking the hit. Anthropic Claude logo Claude Code users have been complaining for weeks that their usage limits are disappearing faster than they should. Even on the Max 20x plan at $200 per month, users report hitting their quota within hours of active sessions, sometimes in as little as 90 minutes. The complaints have been widespread enough that Anthropic acknowledged them, noting that users were hitting usage limits in Claude Code faster than expected. No clear explanation followed. Limits are shared across all Claude interfaces, so running Claude Code alongside regular chat accelerates the burn. Heavy sessions on large repositories compound it further. But the math still didn't add up for a lot of users, and the community started digging. One developer set up an HTTP proxy to capture full API requests across multiple Claude Code versions and found something that would explain the gap. Starting with v2.1.100, every request appears to carry around 20,000 extra tokens that the user never sent. The numbers came from the same project and the same prompt across versions. v2.1.98 billed 49,726 tokens. v2.1.100 billed 69,922. The v2.1.100 request was actually smaller in bytes sent from the client, which rules out the user side as the source. The inflation is server-side. It doesn't show up in the CLI's / context view or anywhere users can audit it. Nothing in Anthropic's changelogs explains it. This hasn't been independently verified at scale, and Anthropic hasn't commented on it. Multiple users have since reported the same delta after running their own proxy tests, and the finding has circulated widely on X and Reddit. CLAUDE CODE MAX BURNS YOUR LIMITS 40% FASTER AND NO ONE TOLD YOU WHY this guy set up an HTTP proxy to capture full API requests across 4 different Claude Code versions. here's what he found: Claude Code v2.1.100 silently adds ~20,000 invisible tokens to every single request.... pic.twitter.com/NDtAM8Y0zv -- Om Patel (@om_patel5) April 13, 2026 Community speculation points to expanded session memory features introduced in v2.1.100, such as summary injection or additional tool schemas. Whether intentional or a bug, the effect is the same. Those extra tokens don't just affect billing. They enter the model's actual context window, diluting custom instructions set in CLAUDE.md and degrading output quality faster in long sessions. When Claude starts ignoring project rules, there's currently no way to tell if injected context is the cause. The workaround circulating on X and Reddit is downgrading to v2.1.98 via npx claude-code@2.1.98. The token inflation finding landed on top of an already frustrating month for Claude Code users. On April 4, Anthropic removed the ability to use Claude subscription limits for third-party tools. OpenClaw, an open-source autonomous AI agent that many users run on top of their Claude subscription, was among the tools affected. Usage through those tools now draws from a separate pay-as-you-go add-on billed outside the subscription. Anthropic offered affected users a one-time credit equal to their monthly subscription to ease the transition. RunPod RunPod If you need on-demand GPUs for training, fine-tuning, inference, or running open-source models, give RunPod a try. * Available hardware: H100, H200, A100, L40S, RTX 4090, RTX 5090, and 30+ more * Cost: significantly cheaper than AWS or GCP, billed per second, no contracts * Setup: spins up in under a minute, 30+ regions worldwide Try RunPod - Affiliate disclosure: We may earn a commission if you sign up via our link, at no extra cost to you. Share: * Twitter * Facebook * Bluesky * Reddit * Linkedin * Email [cropped-efficienist-logo] Efficienist Newsletter Get the core business tech news delivered straight to your inbox. We track AI, automation, SaaS, and cybersecurity so you don't have to. Just read what you want, and be done with it. Email*[ ] Subscribe Read Next * OpenClaw 2026.4.12 adds automatic memory recall and fixes a long list of reliability issues Agents * April 13, 2026 OpenClaw 2026.4.12 adds automatic memory recall and fixes a long list of reliability issues Read More -: OpenClaw 2026.4.12 adds automatic memory recall and fixes a long list of reliability issues * MiniMax releases open weights for M2.7, but the license is a catch AI * April 13, 2026 MiniMax releases open weights for M2.7, but the license is a catch Read More -: MiniMax releases open weights for M2.7, but the license is a catch * Evernote's AI assistant can now manage your note reminders Productivity * April 10, 2026 Evernote's AI assistant can now manage your note reminders Read More -: Evernote's AI assistant can now manage your note reminders * Qwen Code v0.14 adds cron jobs, planning mode, and remote control from your phone AI * April 10, 2026 Qwen Code v0.14 adds cron jobs, planning mode, and remote control from your phone Read More -: Qwen Code v0.14 adds cron jobs, planning mode, and remote control from your phone * Anthropic just confirmed Claude Mythos, the model it says is too capable to release publicly AI * April 7, 2026 Anthropic just confirmed Claude Mythos, the model it says is too capable to release publicly Read More -: Anthropic just confirmed Claude Mythos, the model it says is too capable to release publicly Tech news, optimized for what matters. Our Newsletter Facebook Bluesky Threads X * Privacy Policy * Contact * About * RSS [cropped-efficienist-logo] (c) 2026 Efficienist * AI * Agents * Enterprise * SaaS * Productivity * About * Contact * Newsletter Search for: [ ] [Search]