copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

A Scalable Sparse Matrix-vector Multiplication Kernel for Energy-efficient Sparse-blas on FPGAs

R. Dorrance, F. Ren, and D. Marković. Proceedings of the 2014 ACM/SIGDA International Symposium on Field-programmable Gate Arrays, page 161--170. New York, NY, USA, ACM, (2014)
DOI: 10.1145/2554688.2554785

Abstract

Sparse Matrix-Vector Multiplication (SpMxV) is a widely used mathematical operation in many high-performance scientific and engineering applications. In recent years, tuned software libraries for multi-core microprocessors (CPUs) and graphics processing units (GPUs) have become the status quo for computing SpMxV. However, the computational throughput of these libraries for sparse matrices tends to be significantly lower than that of dense matrices, mostly due to the fact that the compression formats required to efficiently store sparse matrices mismatches traditional computing architectures. This paper describes an FPGA-based SpMxV kernel that is scalable to efficiently utilize the available memory bandwidth and computing resources. Benchmarking on a Virtex-5 SX95T FPGA demonstrates an average computational efficiency of 91.85%. The kernel achieves a peak computational efficiency of 99.8%, a >50x improvement over two Intel Core i7 processors (i7-2600 and i7-4770) and showing a >300x improvement over two NVIDA GPUs (GTX 660 and GTX Titan), when running the MKL and cuSPARSE sparse-BLAS libraries, respectively. In addition, the SpMxV FPGA kernel is able to achieve higher performance than its CPU and GPU counterparts, while using only 64 single-precision processing elements, with an overall 38-50x improvement in energy efficiency.

Description

A scalable sparse matrix-vector multiplication kernel for energy-efficient sparse-blas on FPGAs

Links and resources

BibTeX key: Dorrance:2014:SSM:2554688.2554785
entry type: inproceedings
address: New York, NY, USA
booktitle: Proceedings of the 2014 ACM/SIGDA International Symposium on Field-programmable Gate Arrays
year: 2014
pages: 161--170
publisher: ACM
series: FPGA '14
acmid: 2554785
isbn: 978-1-4503-2671-1
numpages: 10
location: Monterey, California, USA
DOI: 10.1145/2554688.2554785
url: http://doi.acm.org/10.1145/2554688.2554785

@loroch's tags highlighted

Cite this publication

@inproceedings{Dorrance:2014:SSM:2554688.2554785, abstract = {Sparse Matrix-Vector Multiplication (SpMxV) is a widely used mathematical operation in many high-performance scientific and engineering applications. In recent years, tuned software libraries for multi-core microprocessors (CPUs) and graphics processing units (GPUs) have become the status quo for computing SpMxV. However, the computational throughput of these libraries for sparse matrices tends to be significantly lower than that of dense matrices, mostly due to the fact that the compression formats required to efficiently store sparse matrices mismatches traditional computing architectures. This paper describes an FPGA-based SpMxV kernel that is scalable to efficiently utilize the available memory bandwidth and computing resources. Benchmarking on a Virtex-5 SX95T FPGA demonstrates an average computational efficiency of 91.85%. The kernel achieves a peak computational efficiency of 99.8%, a >50x improvement over two Intel Core i7 processors (i7-2600 and i7-4770) and showing a >300x improvement over two NVIDA GPUs (GTX 660 and GTX Titan), when running the MKL and cuSPARSE sparse-BLAS libraries, respectively. In addition, the SpMxV FPGA kernel is able to achieve higher performance than its CPU and GPU counterparts, while using only 64 single-precision processing elements, with an overall 38-50x improvement in energy efficiency.}, acmid = {2554785}, added-at = {2018-06-12T16:13:52.000+0200}, address = {New York, NY, USA}, author = {Dorrance, Richard and Ren, Fengbo and Markovi\'{c}, Dejan}, biburl = {https://www.bibsonomy.org/bibtex/268e7bc2be40e4513ba2e0ec5b28a42a2/loroch}, booktitle = {Proceedings of the 2014 ACM/SIGDA International Symposium on Field-programmable Gate Arrays}, description = {A scalable sparse matrix-vector multiplication kernel for energy-efficient sparse-blas on FPGAs}, doi = {10.1145/2554688.2554785}, interhash = {2509afbb8ba26d981768a345a2f68102}, intrahash = {68e7bc2be40e4513ba2e0ec5b28a42a2}, isbn = {978-1-4503-2671-1}, keywords = {FPGA comparison matrix_vector_mult performance sparse sparsity}, location = {Monterey, California, USA}, numpages = {10}, pages = {161--170}, publisher = {ACM}, series = {FPGA '14}, timestamp = {2018-06-16T10:46:03.000+0200}, title = {A Scalable Sparse Matrix-vector Multiplication Kernel for Energy-efficient Sparse-blas on FPGAs}, url = {http://doi.acm.org/10.1145/2554688.2554785}, year = 2014 }

BibSonomy

copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

A Scalable Sparse Matrix-vector Multiplication Kernel for Energy-efficient Sparse-blas on FPGAs

Abstract

Description

Links and resources

Tags

community

Cite this publication

More citation styles

search on

Meta data

Comments and Reviews
(0)

BibSonomy

copydeleteadd this publication to your clipboardcommunity posthistory of this postURLDOIBibTeXEndNoteAPAChicagoDIN 1505HarvardMSOffice XML A Scalable Sparse Matrix-vector Multiplication Kernel for Energy-efficient Sparse-blas on FPGAs

Abstract

Description

Links and resources

Tags

community

Cite this publication

More citation styles

search on

Meta data

Comments and Reviews (0)

copy delete add this publication to your clipboard
community post
history of this post
URL
DOI
BibTeX
EndNote
APA
Chicago
DIN 1505
Harvard
MSOffice XML

A Scalable Sparse Matrix-vector Multiplication Kernel for Energy-efficient Sparse-blas on FPGAs

Comments and Reviews
(0)