Learning to Remember: Attentive Reinforcement Learning for Edge Serverless Autoscaling
cs.DC, cs.LG
Submitted: 2026-03-21
Updated: 2026-09-23
Code: https://github.com/Azure/AzurePublicDataset
Project page: https://ai-ran.org
Terminology
Sources
- Auto-scaling Approaches for Microservice Applications: A Survey and Taxonomy
- An SLO Driven and Cost-Aware Autoscaling Framework for Kubernetes
- A Hybrid Reactive-Proactive Auto-scaling Algorithm for SLA-Constrained Edge Computing
- Proximal Policy Optimization Algorithms
Related papers
- iScheduler: Reinforcement Learning-Driven Continual Optimization for Large-Scale Resource Investment Problems
- SAMM: Sharded Automated Market Maker
- InferScale: GPU-Native KV Injection for Personalized LLM Serving
- Vigil: Accountable Liveness against Selective Silence
- Steelhead: Interleaving Partially Synchronous and Asynchronous Commit Rules on a Shared DAG
- Pushing CPU Speech Synthesis to the Wall: Extreme Inference Tuning under Serverless Architecture and Billing