Beyond Leaderboards: Tokenomics of Agentic Small Language Model Ensembles
cs.CL
Submitted: 2026-10-01
Updated: 2026-10-01
Terminology
Sources
- Small Language Models are the Future of Agentic AI
- FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance
- Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
- SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
- AgentBench: Evaluating LLMs as Agents
- GAIA: a benchmark for General AI Assistants
- Evaluating LLM Metrics Through Real-World Capabilities
- Small Language Models (SLMs) Can Still Pack a Punch: A survey (updated 2026)
- Instruction-Following Evaluation for Large Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering