Results 81 to 90 of about 926 (201)
Heterogeneous nodes composed of a multicore CPU and accelerators are today’s norm in high-performance computing (HPC) platforms due to their superior performance and energy efficiency.
Simon Farrelly +2 more
doaj +1 more source
Design of Neko—A Scalable High‐Fidelity Simulation Framework With Extensive Accelerator Support
ABSTRACT Recent trends and advancements in including more diverse and heterogeneous hardware in High‐Performance Computing (HPC) are challenging scientific software developers in their pursuit of efficient numerical methods with sustained performance across a diverse set of platforms.
Niclas Jansson +4 more
wiley +1 more source
From OpenACC to OpenMP 4 [PDF]
For the past few years, OpenACC has been the primary directive-based API for programming accelerator devices like GPUs. OpenMP 4.0 is now a competitor in this space, with support from different vendors. In this paper, we describe an algorithm to convert (a subset of) OpenACC to OpenMP 4; we implemented this algorithm in a prototype tool and evaluated ...
Nawrin Sultana +3 more
openaire +1 more source
Optimization strategies for multi‐block structured CFD simulation based on Sunway TaihuLight
We proposed several optimization strategies to reducing the consuming time for multi‐block structured CFD simulation based on Sunway TaihuLight super computing system. Numerical cases were performed and show the high load balance and efficiency of our parallel implementation that achieve a speedup of 8× + for one loop step, and a speed up of 2× + for ...
Xiaojing Lv +5 more
wiley +1 more source
Implementing OpenMP 4.0 for the NVIDIA PTX architecture in GCC compiler
The paper describes the approach used in implementing OpenMP offloading to NVIDIA accelerators in GCC. Offloading refers to a new capability in OpenMP 4.0 specification update that allows the programmer to specify regions of code that should be executed ...
A. V. Monakov, V. A. Ivanishin
doaj +1 more source
Predictive Performance Tuning of OpenACC Accelerated Applications [PDF]
Graphics Processing Units (GPUs) are gradually becoming mainstream in supercomputing as their capabilities to significantly accelerate a large spectrum of scientific applications have been clearly identified and proven. Moreover, with the introduction of
Feki, Saber, Siddiqui, Shahzeb
core
In this paper, we developed an adaptive signal control system with the help of deep reinforcement learning and autonomous vehicles. The system is designed to improve intersection safety, mobility and energy efficiency simultaneously. The system performance evaluation verifies the effectiveness of the proposed system compared with conventional fixed ...
Amir Hossein Karbasi, Hao Yang
wiley +1 more source
GPU-Accelerated High-Resolution Geostationary-Satellite-Based Atmospheric Motion Vector Retrieval
Satellite-derived atmospheric motion vectors (AMVs) provide essential wind field data crucial for numerical weather prediction (NWP) and nowcasting applications. However, current operational AMV products typically offer relatively low spatial resolution,
Rundong Zhou +8 more
doaj +1 more source
Gap setting control strategy for connected and automated vehicles in freeway lane‐drop bottlenecks
Here, two strategies are proposed that control the gap setting of CAVs equipped with ACC systems, and the longitudinal behaviours of different gap settings were calibrated from the state‐of‐the‐art commercial AVs’ trajectory dataset. Simulation experiments were conducted in both hypothetical and real‐world networks using VISSIM to validate the proposed
Sungyong Chung +3 more
wiley +1 more source
Accelerator Benchmark Suite Using OpenACC Directives [PDF]
In recent years, GPU computing has been very popular for scientific applications, especially after the release of programming languages like CUDA, OpenCL, and OpenACC.
Chitral, Pooja 1986-
core

