Dimensionless Controls of Plasticity Under Alternating Tasks: From Evolutionary Biology to Continual Learning
math.OC, cs.LG
Submitted: 2026-08-24
Updated: 2026-08-24
Terminology
Sources
- Disentangling Linear Mode-Connectivity
- The Loss Surfaces of Multilayer Networks
- Gradient Descent on Neural Networks Typically Occurs at the Edge of Stability
- Deep Ensembles: A Loss Landscape Perspective
- Out-of-distribution forgetting: vulnerability of continual learning to intra-class distribution shift
- The Platonic Representation Hypothesis
- Deep linear neural networks with arbitrary loss: All local minima are global
- Linear Mode Connectivity in Multitask and Continual Learning
- On the curvature of the loss landscape
- Experience Replay for Continual Learning
- There Will Be a Scientific Theory of Deep Learning
- When Representations Align: Universality in Representation Learning Dynamics
- Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
- Gradient Surgery for Multi-Task Learning
- Why flatness does and does not correlate with generalization for deep neural networks
- Non-equilibrium physics: from spin glasses to machine and neural learning
Related papers
- Lions and Muons: Optimization via Stochastic Frank-Wolfe under Heavy-Tailed Noise
- Adam-HNAG: A Convergent Reformulation of Adam with Accelerated Rate
- Incremental Learning in Mirror Flows
- Online Control via Counterfactual Tracking
- Asynchronous Replanning in Two Population Linear Quadratic Mean Field Games: Information Requirements and Stability
- Petrov-Galerkin operator inference with application to stability-encouraging identification