Dora: QoE-Aware Hybrid Parallelism for Distributed Edge AI
Hybrid pipeline parallelism for LLM training across heterogeneous edge devices: a DP-based Pareto-optimal model partitioner under latency and energy constraints, plus a network-aware wavefront scheduler with online re-planning — 1.1×–6.3× end-to-end improvement over the state of the art.