NAS Parallel Benchmarks
The NAS Parallel Benchmarks (NPB) are a small set of programs designed to help evaluate the performance of parallel supercomputers. The benchmarks are derived from computational fluid dynamics (CFD) applications and consist of five kernels and three pseudo-applications in the original "pencil-and-paper" specification (NPB 1). The benchmark suite has been extended to include new benchmarks for unstructured adaptive meshes, parallel I/O, multi-zone applications, and computational grids. Problem sizes in NPB are predefined and indicated as different classes. Reference implementations of NPB are available in commonly-used programming models like MPI and OpenMP (NPB 2 and NPB 3).
- Design and implementation of an agent home scheme strategy for prefetch-based DSM systems
- Using cost to control instrumentation overhead
- Elkhound
- HPF/JA
- MPI-CHECK
- Paje
- QUAFF
- OpenSHMEM
- Parallel benchmarks of turbulence in complex geometries
- A parallel finite element method for the analysis of crystalline solids
- A parallelized ENO procedure for direct numerical simulation of compressible turbulence
- ParoC++
- CableS
- BSP2OMP
- FLEXSIM
- VXDL
- EARTH--MANNA
- OVERFLOW-MLP
- Algorithms for the parallel alternating direction access machine
- MPICH
- TPVM
- BSPlib
- HPCC
- CoArray
- PVM
- An object-oriented parallel programming language for distributed-memory parallel computing platforms
- eSkel
- Code modernization strategies to 3-D stencil-based applications on intel Xeon Phi: KNC and KNL
- ARMCI
- LAM-MPI
- KAAPI
- Topology-aware strategy for MPI-IO operations in clusters
- Unstructured adaptive meshes: Bad for your memory?
- PAPI
- CPMD
- MPI/MPICH
- MPI
- Reducing division latency with reciprocal caches
- ParaProf
- Online malleable job scheduling for m 3
- SKaMPI
- Parallelization of a multiblock flow code: An engineering implementation
- Ibis
- HPCTOOLKIT
- KOJAK
- VAMPIR
- MARMOT
- mpiP
- Jumpshot
- SPLASH-2
- ReSHAPE
- Nimrod/G
- SIMGRID
- NetLogger
- Scalasca
- Self-similarity of parallel machines
- SAC -- a functional array language for efficient multi-threaded execution
- A proposal for error handling in OpenMP
- Direct and inverse problems of high-viscosity fluid dynamics
- High-scalability parallelization of a molecular modeling application: Performance and productivity comparison between OpenMP and MPI implementations
- Sweep3d
- Intel MPI Benchmarks
- hwloc
- LogGOPSim
- MOCCA
- Performance characteristics of the multi-zone NAS parallel benchmarks
- Parallel 3D mortar element method for adaptive nonconforming meshes
- Sphinx
- UJMP
- Sequoia Benchmark
- Sisal
- PARAVER
- ParADE
- Omega
- ickp
- Data optimizations for constraint automata
- MPJ Express
- PCG
- scientific article; zbMATH DE number 991438 (Why is no real title available?)
- OProfile
- JCuda
- SICOSYS
- Bsp2omp: A Compiler For Translating Bsp Programs To Openmp
- VXDL: virtual resources and interconnection networks description language
- Kendo
- MCT
- iperf
- Performance modeling of hybrid MPI/OpenMP scientific applications on large-scale multicore supercomputers
- ADAPT
- Comments on PVPs, MPPs, NOWS, and future computer architectures
- Comments on PVPs, MPPs, NOWS, and future computer architectures
- scientific article; zbMATH DE number 1287887 (Why is no real title available?)
- Copperhead
- Chapel
- Algorithm-system scalability of heterogeneous computing
- A two-stage hardware scheduler combining greedy and optimal scheduling
- Implementation and evaluation of a communication intensive application on the EARTH multithreaded system
- scientific article; zbMATH DE number 2087789 (Why is no real title available?)
- Efficient communication using message prediction for clusters of multiprocessors
- HPF/JA: extensions of High Performance Fortran for accelerating real‐world applications
This page was built for software: NAS Parallel Benchmarks