LIMIT: Less Is More for Instruction Tuning in Text-to-SQL
cs.AI
Submitted: 2026-09-21
Updated: 2026-09-21
License: http://creativecommons.org/licenses/by/4.0/
The gist: Large language models have achieved remarkable progress on Text-to-SQL through reasoning-enhanced fine-tuning, yet existing approaches predominantly rely on massive instruction corpora under the
Terminology
Abstract
Large language models have achieved remarkable progress on Text-to-SQL through reasoning-enhanced fine-tuning, yet existing approaches predominantly rely on massive instruction corpora under the assumption that scale drives performance. We challenge this paradigm by investigating a fundamental question: what is the minimal data requirement for effective Text-to-SQL instruction tuning? We propose LIMIT(Less Is More for Instruction Tuning in Text-to-SQL), a data-centric framework that demonstrates strong database reasoning can emerge from an extremely compact training set when examples are strategically selected. LIMIT operates through four stages: difficulty-aware filtering that identifies samples within the model's learning frontier, chain-of-thought synthesis with consistency-based selection, multi-dimensional quality scoring via LLM-as-judge, and genetic algorithm optimization that jointly maximizes schema coverage and sample quality. On the BIRD and Spider benchmark, LIMIT selects only 796 and 863 samples while achieving 100% table coverage, enabling Qwen3-8B to reach 69.1% and 88.9% execution accuracy.This result surpasses methods trained on 20 times more data and establishes a new state-of-the-art among open-source approaches. Our findings suggest that careful data curation, rather than scale, is the key to efficient Text-to-SQL learning.
Sources
- A Preview of XiYan-SQL: A Multi-Generator Ensemble Framework for Text-to-SQL
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- SQL-Trail: Multi-Turn Reinforcement Learning with Interleaved Feedback for Text-to-SQL
- GPT-4o System Card
- The Dawn of Natural Language to SQL: Are We Fully Ready?
- OmniSQL: Synthesizing High-quality Text-to-SQL Data at Scale
- RSL-SQL: Robust Schema Linking in Text-to-SQL Generation
- CodeS: Towards Building Open-source Language Models for Text-to-SQL
- Can LLM Already Serve as A Database Interface? A BIg Bench for Large-Scale Database Grounded Text-to-SQLs
- C3: Zero-shot Text-to-SQL with ChatGPT
- PET-SQL: A Prompt-Enhanced Two-Round Refinement of Text-to-SQL with Cross-consistency
- Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
- DeepSeek-V3 Technical Report
- ROUTE: Robust Multitask Tuning and Collaboration for Text-to-SQL
- EPI-SQL: Enhancing Text-to-SQL Translation with Error-Prevention Instructions
- SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQL
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- DB-Explore: Automated Database Exploration and Instruction Synthesis for Text-to-SQL
- HybridFlow: A Flexible and Efficient RLHF Framework
- SQL-R1: Training Natural Language to SQL Reasoning Model By Reinforcement Learning
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection