Abstract
The research in Connected Component Labeling, although old, is still very active and several efficient algorithms for CPUs and GPUs have emerged during the last years and are always improving the performance. This article introduces a new SIMD run-based algorithm for CCL. We show how RLE compression can be SIMDized and used to accelerate scalar run-based CCL algorithms. A benchmark done on Intel, AMD and ARM processors shows that this new algorithm outperforms the State-of-the-Art by an average factor of ×1.7 on AVX2 machines and ×1.9 on Intel Xeon Skylake with AVX512.
Cite
CITATION STYLE
Lemaitre, F., Hennequin, A., & Lacassagne, L. (2020). How to speed connected component labeling up with SIMD RLE algorithms. In WPMVP 2020 - Proceedings of the 2020 6th Workshop on Programming Models for SIMD/Vector Processing, co-located with PPoPP 2020. Association for Computing Machinery. https://doi.org/10.1145/3380479.3380481
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.