Results 101 to 110 of about 926 (201)
Application Experiences on a GPU-Accelerated Arm-based HPC Testbed. [PDF]
Elwasif W +33 more
europepmc +1 more source
Historic Learning Approach for Auto-tuning OpenACC Accelerated Scientific Applications [PDF]
The performance optimization of scientific applications usually requires an in-depth knowledge of the hardware and software. A performance tuning mechanism is suggested to automatically tune OpenACC parameters to adapt to the execution environment on a ...
Feki, Saber +2 more
core +1 more source
Performance of Portable Sparse Matrix-Vector Product Implemented Using OpenACC [PDF]
Kinga Stec, Przemyslaw Stpiczynski
doaj +1 more source
Accelerating the A2D Ocean Model With Standard Language Parallelism
Parallel computing is crucial for enhancing the computational efficiency of ocean numerical models. Traditionally, parallelization in such models relied primarily on MPI. Subsequent developments introduced GPU-oriented computing tools, including CUDA and
Bingrui Chen +3 more
doaj +1 more source
Programming GPU-Accelerated OpenPOWER Systems with OpenACC [PDF]
This 2h tutorial interactively teaches how to handle the massive computing performance offered by POWER systems with NVLink- attached GPUs – the technology also powering Sierra and Summit, two of the fastest supercomputers in the US.
Herten, Andreas
core
Programming GPU-Accelerated POWER Systems with OpenACC [PDF]
In the tutorial, attendees will learn how to leverage the full performance offered by GPU-equipped IBM POWER systems – the architecture which powers the world’s fastest supercomputers Summit and Sierra.
Herten, Andreas
core
Performance portability in reverse time migration and seismic modelling via OpenACC [PDF]
Heterogeneity among the computational resources within a single machine has significantly increased in high performance computing to exploit the tremendous potential of graphics processing units (GPUs).
Ahmad Qawasmeh +3 more
core +1 more source
Multiple GPUs using MPI and OpenACC [PDF]
A scalable parallelization algorithm to port an explicit marching-on-in-time (MOT)-based time domain volume integral equation (TDVIE) solver onto multi-GPUs is described. The algorithm makes use of MPI and OpenACC for efficient implementation.
Feki, Saber +2 more
core
GPU-I-TASSER: a GPU accelerated I-TASSER protein structure prediction tool. [PDF]
MacCarthy EA, Zhang C, Zhang Y, Kc DB.
europepmc +1 more source
Evaluating the OpenACC API for Parallelization of CFD Applications [PDF]
Directive-based programming of graphics processing units (GPUs) has recently appeared as a viable alternative to using specialized low-level languages such as CUDA C and OpenCL for general-purpose GPU programming. This technique, which uses directive or
Pickering, Brent Phillip
core

