Accelerating GPU kernels for dense linear algebra
From MaRDI portal
Recommendations
- Accelerating numerical dense linear algebra calculations with GPUs
- KBLAS: an optimized library for dense matrix-vector multiplication on GPU accelerators
- Redesigning triangular dense matrix computations on GPUs
- Towards dense linear algebra for hybrid GPU accelerated manycore systems
- Batched triangular dense linear algebra kernels for very small matrix sizes on GPUs
Cited in
(12)- KBLAS: an optimized library for dense matrix-vector multiplication on GPU accelerators
- Accelerating numerical dense linear algebra calculations with GPUs
- Efficient 6D Vlasov simulation using the dynamical low-rank framework \texttt{Ensign}
- Randomized GPU Algorithms for the Construction of Hierarchical Matrices from Matrix-Vector Operations
- CUBLAS-aided long vector algorithms
- Implementing Multifrontal Sparse Solvers for Multicore Architectures with Sequential Task Flow Runtime Systems
- Performance and energy consumption of accurate and mixed-precision linear algebra kernels on GPUs
- Changes in dense linear algebra kernels: decades-long perspective
- Redesigning triangular dense matrix computations on GPUs
- Towards dense linear algebra for hybrid GPU accelerated manycore systems
- nuGPR: GPU-accelerated Gaussian process regression with iterative algorithms and low-rank approximations
- Batched triangular dense linear algebra kernels for very small matrix sizes on GPUs
This page was built for publication: Accelerating GPU kernels for dense linear algebra
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q3081346)