International Journal of Parallel Programming

Papers
(The median citation count of International Journal of Parallel Programming is 1. The table below lists those papers that are above that threshold based on CrossRef citation counts [max. 250 papers]. The publications cover those that have been published in the past four years, i.e., from 2022-08-01 to 2026-08-01.)
ArticleCitations
Efficient High-Performance Computing Strategies for the Legendre Pairs Search19
Parallelization as a Source of Randomness in Iterative Tomography: A Case Study17
Scalable and Accurate Parallel Timing Simulations with parti-gem513
SkePU-Streaming: Distributed Pipelining of Portable Data-Parallel Skeleton Computations for the Heterogeneous Edge-Cloud Continuum11
Accelerating OCaml Programs on FPGA10
Calculation of Distributed-Order Fractional Derivative on Tensor Cores-Enabled GPU8
Special Issue on SAMOS 20227
High-level Programming of Vulkan-based GPUs Through OpenMP7
Meerkat: A Framework for Dynamic Graph Algorithms on GPUs7
Declarative Data Flow in a Graph-Based Distributed Memory Runtime System6
Erasure-Coded Hybrid Writes Based on Data Delta6
SGgraph: A Scalable GPU-Based Edge-Centric Graph Processing Framework6
Investigating Methods for ASPmT-Based Design Space Exploration in Evolutionary Product Design6
A Practical Approach for Employing Tensor Train Decomposition in Edge Devices5
Scaling the Maximum Flow Computation on GPUs5
ControlPULP: A RISC-V On-Chip Parallel Power Controller for Many-Core HPC Processors with FPGA-Based Hardware-In-The-Loop Power and Thermal Emulation5
Using Machine Learning Hardware to Solve Linear Partial Differential Equations with Finite Difference Methods4
Portable C++ Code that can Look and Feel Like Fortran Code with Yet Another Kernel Launcher (YAKL)4
Design and Performance Evaluation of a Novel High-Speed Hardware Architecture for Keccak Crypto Coprocessor4
K*-Means: An Efficient Clustering Algorithm with Adaptive Decision Boundaries4
Advancing Interactive Parallelization: iCetus4
Optimizing Three-Dimensional Stencil-Operations on Heterogeneous Computing Environments3
Automatic Heterogeneous Runtime Using Signal Processing Domain-Specific and Parallel Patterns3
RMOWOA: A Revamped Multi-Objective Whale Optimization Algorithm for Maximizing the Lifetime of a Network in Wireless Sensor Networks3
Efficient Implementation of AI Algorithms on an FPGA-Based System for Enhancing Blood Vessel Segmentation3
Generic Exact Combinatorial Search at HPC Scale3
Larger-Than-Memory Stateful Stream Processing with WindFlow2
Self-Adaptive Micro-Batching for Low-Latency GPU-Accelerated Stream Processing2
Generating Sparse Matrices for Large-Scale Spectral Clustering on a Single GPU2
SymTensor: Symbolic and Adaptive Tensor Partitioning by Unified Parallelism for Deep Learning2
Accelerating the Conjugate Gradient Method on Distributed-Memory Computers2
Programming Parallelism on FPGAs with Eclat2
Enabling Pinning Strategies for Stream Processing Applications on Multicores2
Acknowledgement 20251
Thread and Data Mapping in Software Transactional Memory: an Overview1
East of Eden: Parallel Functional Programming in Idris1
A High-Level API for End-to-End Data Compression in Multi-GPU Cluster Applications1
Giraph-Based Distributed Algorithms for Coloring Large-Scale Graphs1
SMSG: Profiling-Free Parallelism Modeling for Distributed Training of DNN1
NLock: A Scalable Lock for NUMA Architectures1
Retraction Note: QoS and QoE Enhanced Resource Allocation for Wireless Video Sensor Networks Using Hybrid Optimization Algorithm1
Yet Another Lock-Free Atom Table Design for Scalable Symbol Management in Prolog1
CAPIO-CL: The CAPIO Coordination Language1
A Fault-Model-Relevant Classification of Consensus Mechanisms for MPI and HPC1
High-Level Programming of FPGA-Accelerated Systems with Parallel Patterns1
RISC-V Instruction Fetch Architecture Optimized for Harsh Environments1
MICPAT: Micro-Architecture Independent Characteristics Profiling Analysis Tool for GPU Programs1
0.083353996276855