All Stories

  1. Announcing the Shaw Prize in Computer Science
  2. Durable Engines of Discovery
  3. Generalizing Random Butterfly Transforms to Arbitrary Matrix Sizes
  4. Task-Based Polar Decomposition Using SLATE on Massively Parallel Systems with Hardware Accelerators
  5. GPU-based LU Factorization and Solve on Batches of Matrices with Band Structure
  6. Using Additive Modifications in LU Factorization Instead of Pivoting
  7. A Set of Batched Basic Linear Algebra Subprograms and LAPACK Routines
  8. Using Advanced Vector Extensions AVX-512 for MPI Reductions
  9. Extreme-Scale Task-Based Cholesky Factorization Toward Climate and Weather Prediction Applications
  10. Load-balancing Sparse Matrix Vector Product Kernels on GPUs
  11. Guest editors’ note: Special issue on clusters, clouds, and data for scientific computing
  12. Massively Parallel Automated Software Tuning
  13. PLASMA
  14. Big data and extreme-scale computing
  15. A look back on 30 years of the Gordon Bell Prize
  16. GPU-accelerated co-design of induced dimension reduction
  17. Exascale computing and big data