Task-Aware Spectral Pruning: A Mixture-of-Masks Framework for Efficient LLM Inference

arXiv:2609.29499 · cs.LG · Submitted 2026-08-24 · Read on arXiv

cs.LG

Submitted: 2026-08-24

Updated: 2026-08-24

Code: https://github.com/tatsu-lab/alpaca_eval

Terminology

Sources

Related papers