| Publication | Date of Publication | Type |
|---|
Algorithm 1039: automatic generators for a family of matrix multiplication routines with Apache TVM ACM Transactions on Mathematical Software | 2024-09-12 | Paper |
Efficient update of determinants for many-electron wave function overlaps Computer Physics Communications | 2023-05-25 | Paper |
<scp>Ginkgo</scp> : A Modern Linear Operator Algebra Framework for High Performance Computing ACM Transactions on Mathematical Software | 2022-03-29 | Paper |
DMRlib: Easy-Coding and Efficient Resource Management for Job Malleability IEEE Transactions on Computers | 2022-03-23 | Paper |
A data-parallel ILUPACK for sparse general and symmetric indefinite linear systems Euro-Par 2016: Parallel Processing Workshops | 2022-03-09 | Paper |
Adaptive Precision Block-Jacobi for High Performance Preconditioning in the Ginkgo Linear Algebra Software ACM Transactions on Mathematical Software | 2022-02-01 | Paper |
Analysis of Threading Libraries for High Performance Computing IEEE Transactions on Computers | 2020-10-02 | Paper |
Cholesky and Gram-Schmidt orthogonalization for tall-and-skinny QR factorizations on graphics processors Lecture Notes in Computer Science | 2020-07-20 | Paper |
FloatX ACM Transactions on Mathematical Software | 2020-04-24 | Paper |
Reproducibility strategies for parallel preconditioned conjugate gradient Journal of Computational and Applied Mathematics | 2020-02-18 | Paper |
Solving matrix equations on multi-core and many-core architectures Algorithms | 2019-03-26 | Paper |
Look-ahead in the two-sided reduction to compact band forms for symmetric eigenvalue problems and the SVD Numerical Algorithms | 2019-02-07 | Paper |
Look-ahead in the two-sided reduction to compact band forms for symmetric eigenvalue problems and the SVD Numerical Algorithms | 2019-02-07 | Paper |
| Exploiting task-parallelism in message-passing sparse linear system solvers using OmpSs | 2018-01-11 | Paper |
Parallel solution of hierarchical symmetric positive definite linear systems Applied Mathematics and Nonlinear Sciences | 2017-12-14 | Paper |
Analytical modeling is enough for high-performance BLIS ACM Transactions on Mathematical Software | 2017-06-30 | Paper |
Programming matrix algorithms-by-blocks for thread-level parallelism ACM Transactions on Mathematical Software | 2017-05-19 | Paper |
A Runtime System for Programming Out-of-Core Matrix Algorithms-by-Tiles on Multithreaded Architectures ACM Transactions on Mathematical Software | 2017-05-19 | Paper |
A fast band-Krylov eigensolver for macromolecular functional motion simulation on multicore architectures and graphics processors Journal of Computational Physics | 2016-12-20 | Paper |
Exploring large macromolecular functional motions on clusters of multicore processors Journal of Computational Physics | 2016-12-05 | Paper |
Deriving dense linear algebra libraries Formal Aspects of Computing | 2014-11-10 | Paper |
Improved accuracy and parallelism for MRRR-based eigensolvers -- a mixed precision approach SIAM Journal on Scientific Computing | 2014-08-13 | Paper |
Improved accuracy and parallelism for MRRR-based eigensolvers -- a mixed precision approach SIAM Journal on Scientific Computing | 2014-08-13 | Paper |
A factored variant of the Newton iteration for the solution of algebraic Riccati equations via the matrix sign function Numerical Algorithms | 2014-07-03 | Paper |
Solving dense generalized eigenproblems on multi-threaded architectures Applied Mathematics and Computation | 2013-12-23 | Paper |
Parallel computation of 3-D soil-structure interaction in time domain with a coupled FEM/SBFEM approach Journal of Scientific Computing | 2013-01-11 | Paper |
Increasing data locality and introducing level-3 BLAS in the neville elimination Applied Mathematics and Computation | 2012-06-11 | Paper |
A mixed-precision algorithm for the solution of Lyapunov equations on hybrid CPU-GPU platforms Parallel Computing | 2011-11-10 | Paper |
Large-scale linear system solver using secondary storage: self-energy in hybrid nanostructures Computer Physics Communications | 2011-05-31 | Paper |
Large scale simulation of wave propagation in soils interacting with structures using FEM and SBFEM Journal of Computational Acoustics | 2011-05-18 | Paper |
Exploiting thread-level parallelism in the iterative solution of sparse linear systems Parallel Computing | 2011-05-10 | Paper |
Extending OpenMP to survive the heterogeneous multi-core era International Journal of Parallel Programming | 2010-11-03 | Paper |
Out-of-core solution of linear systems on graphics processors International Journal of Parallel, Emergent and Distributed Systems | 2010-05-21 | Paper |
Parallel Implementation of LQG Balanced Truncation for Large-Scale Systems Large-Scale Scientific Computing | 2009-03-26 | Paper |
Solving linear-quadratic optimal control problems on parallel computers Optimization Methods & Software | 2009-02-23 | Paper |
| Parallelization of multilevel preconditioners constructed from inverse-based ILUs on shared-memory multiprocessors | 2009-02-09 | Paper |
| Strategies for parallelizing the solution of rational matrix equations | 2009-02-09 | Paper |
Design, Tuning and Evaluation of Parallel Multilevel ILU Preconditioners High Performance Computing for Computational Science - VECPAR 2008 | 2009-02-03 | Paper |
An Algorithm-by-Blocks for SuperMatrix Band Cholesky Factorization High Performance Computing for Computational Science - VECPAR 2008 | 2009-02-03 | Paper |
Accumulating Householder transformations, revisited ACM Transactions on Mathematical Software | 2008-12-21 | Paper |
Blocked algorithms for the reduction to Hessenberg-triangular form revisited BIT | 2008-12-16 | Paper |
Efficient algorithms for generalized algebraic Bernoulli equations based on the matrix sign function Numerical Algorithms | 2008-02-18 | Paper |
Solution of Band Linear Systems in Model Reduction for VSLI Circuits Scientific Computing in Electrical Engineering | 2008-01-02 | Paper |
Solving stable Sylvester equations via rational iterative schemes Journal of Scientific Computing | 2006-09-14 | Paper |
High Performance Computing for Computational Science - VECPAR 2004 Lecture Notes in Computer Science | 2005-11-10 | Paper |
| scientific article; zbMATH DE number 2222851 (Why is no real title available?) | 2005-11-04 | Paper |
Representing linear algebra algorithms in code: the FLAME application program interfaces ACM Transactions on Mathematical Software | 2005-07-22 | Paper |
The science of deriving dense linear algebra algorithms ACM Transactions on Mathematical Software | 2005-07-22 | Paper |
Formal derivation of algorithms ACM Transactions on Mathematical Software | 2005-07-21 | Paper |
Spectral division methods for block generalized Schur decompositions Mathematics of Computation | 2004-08-13 | Paper |
| scientific article; zbMATH DE number 2090653 (Why is no real title available?) | 2004-08-12 | Paper |
| scientific article; zbMATH DE number 2089176 (Why is no real title available?) | 2004-08-12 | Paper |
| scientific article; zbMATH DE number 2080151 (Why is no real title available?) | 2004-08-04 | Paper |
Parallel algorithms for model reduction of discrete-time systems International Journal of Systems Science. Principles and Applications of Systems and Integration | 2004-03-17 | Paper |
Efficient algorithms for the block Hessenberg form The Journal of Supercomputing | 2003-12-14 | Paper |
| scientific article; zbMATH DE number 1953305 (Why is no real title available?) | 2003-07-27 | Paper |
NUMERICAL SOLUTION OF DISCRETE STABLE LINEAR MATRIX EQUATIONS ON MULTICOMPUTERS Parallel Algorithms and Applications | 2003-02-06 | Paper |
Parallel solvers for discrete‐time algebric Riccati equations Concurrency and Computation: Practice and Experience | 2003-02-04 | Paper |
The Generalized Newton Iteration forthe Matrix Sign Function SIAM Journal on Scientific Computing | 2003-01-05 | Paper |
Specialized parallel algorithms for solving Lyapunov and Stein equations Journal of Parallel and Distributed Computing | 2002-07-31 | Paper |
Parallel algorithms for LQ optimal control of discrete-time periodic linear systems Journal of Parallel and Distributed Computing | 2002-07-04 | Paper |
| scientific article; zbMATH DE number 1740441 (Why is no real title available?) | 2002-06-16 | Paper |
Balanced Truncation Model Reduction of Large-Scale Dense Systems on Parallel Computers Mathematical and Computer Modelling of Dynamical Systems | 2002-01-13 | Paper |
Parallel spectral division using the matrix sign function for the generalized eigenproblem International Journal of High Speed Computing | 2001-12-03 | Paper |
Parallel partial stabilizing algorithms for large linear control systems The Journal of Supercomputing | 2000-09-19 | Paper |
Solving algebraic Riccati equations on parallel computers using Newton's method with exact line search Parallel Computing | 2000-08-21 | Paper |
Parallel codes for computing the numerical rank Linear Algebra and its Applications | 1999-12-19 | Paper |
Solving stable generalized Lyapunov equations with the matrix sign function Numerical Algorithms | 1999-06-29 | Paper |
Efficient Solution Of The Rank-Deficient Linear Least Squares Problem SIAM Journal on Scientific Computing | 1999-06-24 | Paper |
Parallel solution of Riccati matrix equations with the matrix sign function Automatica | 1999-01-05 | Paper |