copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

PLSSVM: A (multi-)GPGPU-accelerated Least Squares Support Vector Machine

A. Van Craen, M. Breyer, and D. Pflüger. 2022 IEEE International Parallel and Distributed Processing Symposium Workshops (IPDPSW), (2022)
DOI: 10.1109/IPDPSW55747.2022.00138.

Abstract

Machine learning algorithms must be able to efficiently cope with massive data sets. Therefore, they have to scale well on any modern system and be able to exploit the computing power of accelerators independent of their vendor. In the field of supervised learning, Support Vector Machines (SVMs) are widely used. However, even modern and optimized implementations such as LIBSVM or ThunderSVM do not scale well for large non-trivial dense data sets on cutting-edge hardware: Most SVM implementations are based on Sequential Minimal Optimization, an optimized though inherent sequential algorithm. Hence, they are not well-suited for highly parallel GPUs. Furthermore, we are not aware of a performance portable implementation that supports CPUs and GPUs from different vendors. We have developed the PLSSVM library to solve both issues. First, we resort to the formulation of the SVM as a least squares problem. Training an SVM then boils down to solving a system of linear equations for which highly parallel algorithms are known. Second, we provide a hardware independent yet efficient implementation: PLSSVM uses different interchangeable backends--OpenMP, CUDA, OpenCL, SYCL--supporting modern hardware from various vendors like NVIDIA, AMD, or Intel on multiple GPUs. PLSSVM can be used as a drop-in replacement for LIBSVM. We observe a speedup on CPUs of up to 10 compared to LIBSVM and on GPUs of up to 14 compared to ThunderSVM. Our implementation scales on many-core CPUs with a parallel speedup of 74.7 on up to 256 CPU threads and on multiple GPUs with a parallel speedup of 3.71 on four GPUs. The code, utility scripts, and documentation are all available on GitHub: https://github.com/SC-SGS/PLSSVM.

Description

PLSSVM: A (multi-)GPGPU-accelerated Least Squares Support Vector Machine

Links and resources

BibTeX key: vancraen2022plssvm
entry type: article
year: 2022
journal: 2022 IEEE International Parallel and Distributed Processing Symposium Workshops (IPDPSW)
pages: 818-827
preprinturl: https://arxiv.org/abs/2202.12674
language: english
DOI: 10.1109/IPDPSW55747.2022.00138.

@ipvs-sc's tags highlighted

Cite this publication

%0 Journal Article %1 vancraen2022plssvm %A Van Craen, Alexander %A Breyer, Marcel %A Pflüger, Dirk %D 2022 %J 2022 IEEE International Parallel and Distributed Processing Symposium Workshops (IPDPSW) %K %P 818-827 %R 10.1109/IPDPSW55747.2022.00138. %T PLSSVM: A (multi-)GPGPU-accelerated Least Squares Support Vector Machine %X Machine learning algorithms must be able to efficiently cope with massive data sets. Therefore, they have to scale well on any modern system and be able to exploit the computing power of accelerators independent of their vendor. In the field of supervised learning, Support Vector Machines (SVMs) are widely used. However, even modern and optimized implementations such as LIBSVM or ThunderSVM do not scale well for large non-trivial dense data sets on cutting-edge hardware: Most SVM implementations are based on Sequential Minimal Optimization, an optimized though inherent sequential algorithm. Hence, they are not well-suited for highly parallel GPUs. Furthermore, we are not aware of a performance portable implementation that supports CPUs and GPUs from different vendors. We have developed the PLSSVM library to solve both issues. First, we resort to the formulation of the SVM as a least squares problem. Training an SVM then boils down to solving a system of linear equations for which highly parallel algorithms are known. Second, we provide a hardware independent yet efficient implementation: PLSSVM uses different interchangeable backends--OpenMP, CUDA, OpenCL, SYCL--supporting modern hardware from various vendors like NVIDIA, AMD, or Intel on multiple GPUs. PLSSVM can be used as a drop-in replacement for LIBSVM. We observe a speedup on CPUs of up to 10 compared to LIBSVM and on GPUs of up to 14 compared to ThunderSVM. Our implementation scales on many-core CPUs with a parallel speedup of 74.7 on up to 256 CPU threads and on multiple GPUs with a parallel speedup of 3.71 on four GPUs. The code, utility scripts, and documentation are all available on GitHub: https://github.com/SC-SGS/PLSSVM.

@article{vancraen2022plssvm, abstract = {Machine learning algorithms must be able to efficiently cope with massive data sets. Therefore, they have to scale well on any modern system and be able to exploit the computing power of accelerators independent of their vendor. In the field of supervised learning, Support Vector Machines (SVMs) are widely used. However, even modern and optimized implementations such as LIBSVM or ThunderSVM do not scale well for large non-trivial dense data sets on cutting-edge hardware: Most SVM implementations are based on Sequential Minimal Optimization, an optimized though inherent sequential algorithm. Hence, they are not well-suited for highly parallel GPUs. Furthermore, we are not aware of a performance portable implementation that supports CPUs and GPUs from different vendors. We have developed the PLSSVM library to solve both issues. First, we resort to the formulation of the SVM as a least squares problem. Training an SVM then boils down to solving a system of linear equations for which highly parallel algorithms are known. Second, we provide a hardware independent yet efficient implementation: PLSSVM uses different interchangeable backends--OpenMP, CUDA, OpenCL, SYCL--supporting modern hardware from various vendors like NVIDIA, AMD, or Intel on multiple GPUs. PLSSVM can be used as a drop-in replacement for LIBSVM. We observe a speedup on CPUs of up to 10 compared to LIBSVM and on GPUs of up to 14 compared to ThunderSVM. Our implementation scales on many-core CPUs with a parallel speedup of 74.7 on up to 256 CPU threads and on multiple GPUs with a parallel speedup of 3.71 on four GPUs. The code, utility scripts, and documentation are all available on GitHub: https://github.com/SC-SGS/PLSSVM.}, added-at = {2023-11-02T11:23:18.000+0100}, author = {Van Craen, Alexander and Breyer, Marcel and Pflüger, Dirk}, biburl = {https://puma.ub.uni-stuttgart.de/bibtex/2f6796639496f7f2f152b61d2a4f6c7da/ipvs-sc}, description = {PLSSVM: A (multi-)GPGPU-accelerated Least Squares Support Vector Machine}, doi = {10.1109/IPDPSW55747.2022.00138.}, interhash = {56e21001d5ae2bbf746c88b06325d16c}, intrahash = {f6796639496f7f2f152b61d2a4f6c7da}, journal = {2022 IEEE International Parallel and Distributed Processing Symposium Workshops (IPDPSW)}, keywords = {}, language = {english}, pages = {818-827}, preprinturl = {https://arxiv.org/abs/2202.12674}, timestamp = {2023-11-02T11:23:18.000+0100}, title = {PLSSVM: A (multi-)GPGPU-accelerated Least Squares Support Vector Machine}, year = 2022 }

PUMA

copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

PLSSVM: A (multi-)GPGPU-accelerated Least Squares Support Vector Machine

Abstract

Description

Links and resources

Tags

community

Cite this publication

More citation styles

search on

Meta data

Comments and Reviews
(0)

PUMA

copydeleteadd this publication to your clipboardcommunity posthistory of this postURLDOIBibTeXEndNoteAPAChicagoDIN 1505HarvardMSOffice XML PLSSVM: A (multi-)GPGPU-accelerated Least Squares Support Vector Machine

Abstract

Description

Links and resources

Tags

community

Cite this publication

More citation styles

search on

Meta data

Comments and Reviews (0)

copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

PLSSVM: A (multi-)GPGPU-accelerated Least Squares Support Vector Machine

Comments and Reviews
(0)