Para-Pipe: Exploiting Hierarchical Operator Parallelism of ML Computational Graphs on SoCs
cs.DC, cs.LG, cs.PF
Submitted: 2026-09-03
Updated: 2026-09-03
Code: https://github.com/ssvb/tinymembench
Terminology
Sources
- Model Parallelism on Distributed Infrastructure: A Literature Review from Theory to LLM Case-Studies
Related papers
- iScheduler: Reinforcement Learning-Driven Continual Optimization for Large-Scale Resource Investment Problems
- SAMM: Sharded Automated Market Maker
- InferScale: GPU-Native KV Injection for Personalized LLM Serving
- Vigil: Accountable Liveness against Selective Silence
- Steelhead: Interleaving Partially Synchronous and Asynchronous Commit Rules on a Shared DAG
- Pushing CPU Speech Synthesis to the Wall: Extreme Inference Tuning under Serverless Architecture and Billing