Results 41 to 50 of about 7,328 (165)
HARD: A performance portable radiation hydrodynamics code based on FleCSI framework
Hydrodynamics And Radiation Diffusion (HARD) is an open-source application for high-performance simulations of compressible hydrodynamics with radiation-diffusion coupling.
Julien Loiseau +9 more
doaj +1 more source
DBEFT: A Dependency-Ratio Bundling Earliest Finish Time Algorithm for Heterogeneous Computing
Performance effective task scheduling algorithms are essential for taking advantage of the heterogeneous multi-processor in heterogeneous computing environments.
Tao Li +6 more
doaj +1 more source
The Dynamic Partial Reconfiguration function of reconfigurable devices permits tasks to be performed simultaneously on a single device. Nevertheless, task placement and resource management problems emerge with the parallelism of reconfigurable devices ...
Tingyu Zhou +4 more
doaj +1 more source
Enhancement of GPU-accelerated smoothed particle hydrodynamics (SPH) method with dynamic parallelism
An innovative GPU programming architecture leveraging CUDA Dynamic Parallelism (CDP) is introduced in this study, aiming to enhance the computational efficiency of Smoothed Particle Hydrodynamics (SPH) simulations.
Liwen Xue, Shenglong Gu, Songdong Shao
doaj +1 more source
Response errors explain the failure of independent-channels models of perception of temporal order
Independent-channels models of perception of temporal order (also referred to as threshold models or perceptual latency models) have been ruled out because two formal properties of these models (monotonicity and parallelism) are not borne out by data ...
Miguel A García-Pérez +1 more
doaj +1 more source
Towards Task-Parallel Reductions in OpenMP
Reductions represent a common algorithmic pattern in many scientific applications. OpenMP* has always supported them on parallel and worksharing constructs. OpenMP 3.0’s tasking constructs enable new parallelization opportunities through the annotation of irregular algorithms.
Jan Ciesko +10 more
openaire +2 more sources
Task-Level Checkpointing System for Task-Based Parallel Workflows
Scientific applications are large and complex; task-based programming models are a popular approach to developing these applications due to their ease of programming and ability to handle complex workflows and distribute their workload across large infrastructures.
Pere Vergés +3 more
openaire +2 more sources
CPU has insufficient resources to satisfy the efficient computation of the convolution neural network (CNN), especially for embedded applications. Therefore, heterogeneous computing platforms are widely used to accelerate CNN tasks, such as GPU, FPGA ...
Li Luo +9 more
doaj +1 more source
Distributed Memory Implementation of Bron-Kerbosch Algorithm
This paper proposes a parallel implementation of the Bron-Kerbosch algorithm, which finds all maximal cliques in large, complex graphs using CPU and thread-level parallelism and distributed memory with multiple cores. With the growing size and complexity
Tejas Ravindra Rote +4 more
doaj +1 more source
Scalable deep text comprehension for Cancer surveillance on high-performance computing
Background Deep Learning (DL) has advanced the state-of-the-art capabilities in bioinformatics applications which has resulted in trends of increasingly sophisticated and computationally demanding models trained by larger and larger data sets.
John X. Qiu +8 more
doaj +1 more source

