AirLLM: Diffusion Policy-based Adaptive LoRA for Remote Fine-Tuning of LLM over the Air
cs.LG, cs.AI, cs.CL
Submitted: 2025-07-15
Updated: 2026-08-27
Comments: 11 pages, 8 figures
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- GPT-4 Technical Report
- DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
- A Survey of Large Language Models
- Scaling Down to Scale Up: A Guide to Parameter-Efficient Fine-Tuning
- Adaptive Sampling and Joint Semantic-Channel Coding under Dynamic Channel Environment
- Proximal Policy Optimization Algorithms
- PEFT-U: Parameter-Efficient Fine-Tuning for User Personalization
- Off-Policy Reinforcement Learning with High Dimensional Reward
- OPT: Open Pre-trained Transformer Language Models
- Wider and Deeper LLM Networks are Fairer LLM Evaluators
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks