Past, Future, All at Once: Mitigating Stability-Plasticity Dilemma via Post-hoc JANUS Rectification
cs.LG, cs.AI
Submitted: 2026-09-17
Updated: 2026-09-17
Code: https://github.com/EleutherAI/lm-evaluation-harness
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models
- WizardCoder: Empowering Code Large Language Models with Evol-Instruct
- KeepLoRA: Continual Learning with Residual Gradient Adaptation
- Gradient Projection Memory for Continual Learning
- SC-LoRA: Balancing Efficient Fine-tuning and Knowledge Preservation via Subspace-Constrained LoRA
- Orthogonal Low-rank Adaptation in Lie Groups for Continual Learning of Large Language Models
- AdaLoRA: Adaptive Budget Allocation for Parameter-Efficient Fine-Tuning
- Data Shapley in One Training Run
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- Training Verifiers to Solve Math Word Problems
- Evaluating Large Language Models Trained on Code
- Program Synthesis with Large Language Models
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks