When Compression Helps and When It Hurts: Condition-Aware Analysis of Chain-of-Thought Distillation
cs.CL
Submitted: 2026-06-19
Updated: 2026-09-01
Terminology
Sources
- Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models
- What Makes Effective Supervision in Latent Chain-of-Thought: An Information-Theoretic Analysis
- Reasoning Beyond Language: A Comprehensive Survey on Latent Chain-of-Thought Reasoning
- Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
- Reconsidering Overthinking: Penalizing Internal and External Redundancy in CoT Reasoning
- Training Verifiers to Solve Math Word Problems
- O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?
- Stepwise Perplexity-Guided Refinement for Efficient Chain-of-Thought Reasoning in Large Language Models
- EPiC: Towards Lossless Speedup for Reasoning Training through Edge-Preserving CoT Condensation
- FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
- DRP: Distilled Reasoning Pruning with Skill-aware Step Decomposition for Efficient Large Reasoning Models
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- From LLMs to LRMs: Rethinking Pruning for Reasoning-Centric Models
- C3oT: Generating Shorter Chain-of-Thought without Compromising Effectiveness
- Break the Chain: Large Language Models Can be Shortcut Reasoners
- Making Slow Thinking Faster: Compressing LLM Chain-of-Thought via Step Entropy
- The Llama 3 Herd of Models
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
- Token-Budget-Aware LLM Reasoning
- A*-Thought: Efficient Reasoning via Bidirectional Compression for Low-Resource Settings
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering