The More It Says, the More You Pay: A Black-Box Audit of Provider-Side Token Inflation in LLM Services
cs.CR
Submitted: 2026-09-17
Updated: 2026-09-17
Comments: 22 pages, 10 figures, 8 tables
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- State of AI: An Empirical 100 Trillion Token Study with OpenRouter
- Are You Getting What You Pay For? Auditing Model Substitution in LLM APIs
- POT: Inducing Overthinking in LLMs via Black-Box Iterative Optimization
- Behavioral Consistency and Transparency Analysis on Large Language Model API Gateways
- Token Inflation: How Dishonest Providers Can Overcharge for Large Language Model Usage
- OverThink: Slowdown Attacks on Reasoning LLMs
- The Llama 3 Herd of Models
- Test-Time Compute Games
- Is Your LLM Overcharging You? Tokenization, Transparency, and Incentives
- Ministral 3
- Predictive Auditing of Hidden Tokens in LLM APIs via Reasoning Length Estimation
- Qwen3 Technical Report
- Excessive Reasoning Attack on Reasoning LLMs
- CoIn: Counting the Invisible Reasoning Tokens in Commercial Opaque LLM APIs
- Real Money, Fake Models: Deceptive Model Claims in Shadow APIs
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs