Abstract
Any successful solution to using multicore processors to scale general-purpose program performance will have to contend with rising intercore communication costs while exposing coarse-grained parallelism. Recently proposed pipelined multithreading (PMT) techniques have been demonstrated to have general-purpose applicability and are also able to effectively tolerate inter-core latencies through pipelined interthread communication. These desirable properties make PMT techniques strong candidates for program parallelization on current and future multicore processors and understanding their performance characteristics is critical to their deployment. To that end, this paper evaluates the performance scalability of a general-purpose PMT technique called decoupled software pipelining (DSWP) and presents a thorough analysis of the communication bottlenecks that must be overcome for optimal DSWP scalability. © 2008 ACM.
Author supplied keywords
Cite
CITATION STYLE
Rangan, R., Vachharajani, N., Ottoni, G., & August, D. I. (2008). Performance scalability of decoupled software pipelining. Transactions on Architecture and Code Optimization, 5(2). https://doi.org/10.1145/1400112.1400113
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.