In Their Own Words: Reasoning Traces Tailored for Small Models Make Them Better Reasoners
cs.CL
Submitted: 2025-09-26
Updated: 2026-09-26
Code: https://github.com/jaeh8nkim/equigranular
Terminology
Sources
- Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
- SplitReason: Learning To Offload Reasoning
- Role of obstacle softness in the diffusive behavior of active Particles
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Gemma 3 Technical Report
- The First Few Tokens Are All You Need: An Efficient and Effective Unsupervised Prefix Fine-Tuning Method for Reasoning Models
- Stationary flows for viscous heat-conductive fluid in a perturbed half-space
- Stabilizing Knowledge, Promoting Reasoning: Dual-Token Constraints for RLVR
- Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
- Effect of Dielectric Wakefields in a Capillary Discharge for Plasma Wakefield Acceleration
- Phi-4-Mini-Reasoning: Exploring the Limits of Small Reasoning Language Models in Math
- Qwen2.5 Technical Report
- Speculative Thinking: Enhancing Small-Model Reasoning with Large Model Guidance at Inference Time
- Scaling Relationship on Learning Mathematical Reasoning with Large Language Models
- Transformers Utilization in Chart Understanding: A Review of Recent Advances & Future Trends
- Short-time Fourier Transform-based Signal Recovery for Modulo Analog-to-Digital Converters
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering