ResMerge: Residual-based Spectral Merging of Large Language Models
cs.CL
Submitted: 2026-06-01
Updated: 2026-08-26
Comments: 19 pages including appendix. Accepted to the EMNLP 2026 Main Conference
Code: https://github.com/sunyd0303-cpu/ResMerge-release
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
- LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
- AdaRank: Adaptive Rank Pruning for Enhanced Model Merging
- Program Synthesis with Large Language Models
- When Shared Knowledge Hurts: Spectral Over-Accumulation in Model Merging
- SALSA: Soup-based Alignment Learning for Stronger Adaptation in RLHF
- Let's Verify Step by Step
- Evaluating Large Language Models Trained on Code
- Task Singular Vectors: Reducing Task Interference in Model Merging
- General-Reasoner: Advancing LLM Reasoning Across All Domains
- Qwen2.5 Technical Report
- No Task Left Behind: Isotropic Model Merging with Common and Task-Specific Subspaces
- Language Models are Super Mario: Absorbing Abilities from Homologous Models as a Free Lunch
- Behavior Knowledge Merge in Reinforced Agentic Models
- SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
- WARP: On the Benefits of Weight Averaged Rewarded Policies
- GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering