International Journal of High Performance Computing Applications

Papers
(The median citation count of International Journal of High Performance Computing Applications is 2. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Visualization at exascale: Making it all work with VTK-m89
Dynamic spawning of MPI processes applied to malleability82
chipStar : Making HIP/CUDA applications cross-vendor portable by building on open standards78
HPC I/O innovations in the exascale era46
Accelerating atmospheric physics parameterizations using graphics processing units28
Compressed basis GMRES on high-performance graphics processing units24
Automatizing the creation of specialized high-performance computing containers19
Refining HPCToolkit for application performance analysis at exascale19
HPL-MxP benchmark: Mixed-precision algorithms, iterative refinement, and scalable data generation17
Julia versus C++ Kokkos for performance portable Cartesian CFD solvers on heterogeneous architectures16
Direct numerical simulations for hybrid rocket boundary layers: Performance modeling and scaling16
Running ahead of evolution—AI-based simulation for predicting future high-risk SARS-CoV-2 variants16
Orchestration of materials science workflows for heterogeneous resources at large scale16
Scalable multilevel Monte Carlo methods exploiting parallel redistribution on coarse levels15
Modeling, evaluating, and orchestrating heterogeneous environmental leverages for large-scale data center management15
A tale of two codes: CUDA vs OpenACC for mass-zero constrained dynamics13
HDF5 in the exascale era: Delivering efficient and scalable parallel I/O for exascale applications13
Ginkgo - A math library designed to accelerate Exascale Computing Project science applications12
General framework for re-assuring numerical reliability in parallel Krylov solvers: A case of bi-conjugate gradient stabilized methods12
Preparing MPICH for exascale11
Architecture specific generation of large scale lattice Boltzmann methods for sparse complex geometries10
GPU-based molecular dynamics of fluid flows: Reaching for turbulence10
Data-driven analysis to understand GPU hardware resource usage of optimizations10
Hypergraph-based locality-enhancing methods for graph operations in Big Data applications10
Retraction Notice10
Performance of explicit and IMEX MRI multirate methods on complex reactive flow problems within modern parallel adaptive structured grid frameworks10
Massively parallel nodal discontinous Galerkin finite element method simulator for room acoustics9
Integrating ytopt and libEnsemble to autotune OpenMC9
Special issue introduction9
Technology trends in computing hardware and their impacts on high-performance scientific computing Part II: Memory systems, interconnects, and system integration8
Cache blocking of distributed-memory parallel matrix power kernels8
PeleC: An adaptive mesh refinement solver for compressible reacting flows8
Heterogeneous programming using OpenMP and CUDA/HIP for hybrid CPU-GPU scientific applications8
A study on the performance of distributed training of data-driven CFD simulations8
Accelerated dynamic data reduction using spatial and temporal properties8
Technology trends in computing hardware and their impacts on high-performance scientific computing Part I: General-purpose processors and hardware accelerators8
Special issue: Introduction7
Preparing the TAU performance system for exascale and beyond6
TransGRU-X – A fusion Seq2Seq network enhanced with multiresolution analysis and gating for forecasting of AI/ML workloads in cloud environments6
FNPF-SEM: A parallel spectral element model in Firedrake for fully nonlinear water wave simulations6
Fair-sharing simulator: Toward fair scheduling in batch computing systems6
Fast truncated SVD of sparse and dense matrices on graphics processors6
Data-driven scalable pipeline using national agent-based models for real-time pandemic response and decision support6
Enable : A CPU-GPU framework for dynamic data-driven agent-based population health simulations6
Experiences with nested parallelism in task-parallel applications using malleable BLAS on multicore processors6
HOPPS: A performance portable spectral difference solver for high-fidelity computational fluid dynamics6
Accelerating cluster dynamics simulation of fission gas behavior in nuclear fuel on deep computing unit–based heterogeneous architecture supercomputer5
Abisko: Deep codesign of an architecture for spiking neural networks using novel neuromorphic materials5
HPC-AI coupling methodology for scientific applications5
Democratizing responsible artificial intelligence for innovation and impact5
Understanding power and energy utilization in large scale production physics simulation codes5
Semi-Lagrangian 4d, 5d, and 6d kinetic plasma simulation on large-scale GPU-equipped supercomputers5
NUMA-aware parallel sparse LU factorization for SPICE-based circuit simulators on ARM multi-core processors5
Clacc: OpenACC for C/C++ in Clang5
A randomized point-block Schwarz preconditioner for multiphysics problems on fully unstructured meshes on GPUs5
Asynchronous-many-task systems: Challenges and opportunities - Scaling an AMR astrophysics code on exascale machines using Kokkos and HPX5
Sequence length scaling in vision transformers for scientific images on frontier5
Bricks: A high-performance portability layer for computations on block-structured grids4
Cache-optimized and low-overhead implementations of additive Schwarz methods for high-order FEM multigrid computations4
Advances in ArborX to support exascale applications4
Feynman and computation: From Los Alamos to quantum computers4
UMap: An application-oriented user level memory mapping library4
P4IRS: An intermediate representation and compiler for parallel and performance-portable particle simulations4
PoCL-R: An open standard based heterogeneous offloading layer with server side scalability4
MAGMA: Enabling exascale performance with accelerated BLAS and LAPACK for diverse GPU architectures4
PaRSEC: Scalability, flexibility, and hybrid architecture support for task-based applications in ECP3
Guest editors note: Special issue on clusters, clouds, and data for scientific computing3
Scalable cosmic AI inference using cloud serverless computing3
Guest editor’s note: Special issue on system-level innovations for performance and fairness at scale: From interconnects to schedulers3
An integrated three-dimensional aeromechanical analysis for the prediction of stresses on modern coaxial rotors3
The ECP ALPINE project: In situ and post hoc visualization infrastructure and analysis capabilities for exascale3
Exploiting mesh structure to improve multigrid performance for saddle-point problems3
A GPU-based compressible combustion solver for applications exhibiting disparate space and time scales3
ECP libraries and tools: An overview3
IO-aware Job-Scheduling: Exploiting the Impacts of Workload Characterizations to select the Mapping Strategy3
Black-box statistical prediction of lossy compression ratios for scientific data3
Towards exascale simulations of granular fluidization in offshore wind turbine foundations3
Fault-tolerant numerical iterative algorithms at scale3
#COVIDisAirborne: AI-enabled multiscale computational microscopy of delta SARS-CoV-2 in a respiratory aerosol2
Evolution of the SLATE linear algebra library2
End-to-end GPU acceleration of low-order-refined preconditioning for high-order finite element discretizations2
Efficient solution of batched band linear systems on GPUs2
Simulation-based machine learning for real-time assessment of side-branch hemodynamics in coronary bifurcation lesions2
High-performance conjugate gradient benchmark: A comprehensive survey2
Fixed-work versus fixed-time checkpointing on large-scale failure-prone platforms2
Mixed precision LU factorization on GPU tensor cores: reducing data movement and memory footprint2
Role-shifting threads: Increasing OpenMP malleability to address load imbalance at MPI and OpenMP2
Performance evaluation of mixed-precision Runge–Kutta methods for the solution of partial differential equations2
An HPC benchmark survey and taxonomy for characterization2
Deep learning foundation and pattern models: Challenges in hydrological time series2
Detecting interference between applications and improving the scheduling using malleable application clones2
SWARM: Reimagining scientific workflow management systems in a distributed world2
An implicit barotropic mode solver for MPAS-ocean using a modern Fortran solver interface2
A GPU-accelerated simulation of rapid intensification of a tropical cyclone with observed heating2
Numerics-driven uplifting, automatic parallelization, and performance optimizations with deep kernel fusion for ocean models on heterogeneous architectures2
Corrigendum to large-scale direct numerical simulations of turbulence using GPUs and modern Fortran2
0.040080070495605