Abstract
Accelerators in mission-critical applications consume substantial power even during low workloads. Due to strict availability requirements, they cannot be stopped or restarted. This leads to unnecessary power consumption during low-load periods. We propose an accelerator pool architecture that aggregates accelerators beyond traditional server constraints. This approach enables dynamic resource allocation through middleware-level management while keeping applications running. It allows online resource scaling rather than relying on application start/stop scheduling. Our contributions include: (1) a novel pooling architecture for resource-intensive applications (virtual Radio Access Network (vRAN), generative Artificial Intelligence (AI), automated driving) that reduces power consumption while maintaining performance, (2) a comprehensive evaluation of the architecture’s practicality for vRAN use cases with identified implementation challenges and mitigation strategies, (3) a mathematical model formulating the accelerator allocation problem as a bin-packing optimization, and (4) empirical validation through comparative simulations. Preliminary simulations using vRAN demonstrate up to 64.3% reduction in operating accelerators compared to conventional architectures. They also show a 29.0% reduction compared to existing server-internal sharing approaches. While this architectural concept offers limited benefits for applications with consistently high accelerator utilization, it shows significant potential in scenarios with geographical load imbalances or temporal variations.
Author supplied keywords
Cite
CITATION STYLE
Saito, S., Natori, K., Otani, I., & Fujimoto, K. (2025). Accelerator pool: accelerator sharing architecture for energy efficiency. Discover Applied Sciences, 7(7). https://doi.org/10.1007/s42452-025-07388-1
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.