CoCurve: Cross-Module Co-Pruning Curvature for Structured LLM Pruning
cs.LG, cs.AI
Submitted: 2026-07-20
Updated: 2026-09-25
Terminology
Sources
- Program Synthesis with Large Language Models
- Evaluating Large Language Models Trained on Code
- Optimal Brain Connection: Towards Efficient Structural Pruning
- Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
- Training Verifiers to Solve Math Word Problems
- Everybody Prune Now: Structured Pruning of LLMs with only Forward Passes
- The Llama 3 Herd of Models
- LLM Compression by Block Removal with Constrained Binary Optimization
- Mistral 7B
- Shortened LLaMA: Depth Pruning for Large Language Models with Comparison of Retraining Methods
- T'yr-the-Pruner: Structural Pruning LLMs via Global Sparsity Distribution Optimization
- ShortGPT: Layers in Large Language Models are More Redundant Than You Expect
- Efficient LLMs with AMP: Attention Heads and MLP Pruning
- 2SSP: A Two-Stage Framework for Structured Pruning of LLMs
- SlimLLM: Accurate Structured Pruning for Large Language Models
- DarwinLM: Evolutionary Structured Pruning of Large Language Models
- Faster gaze prediction with dense networks and Fisher pruning
- CFSP: An Efficient Structured Pruning Framework for LLMs with Coarse-to-Fine Activation Information
- Qwen2.5 Technical Report
- LaCo: Large Language Model Pruning via Layer Collapse
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks