Results 31 to 40 of about 1,012 (185)

Parallelization and Performance Optimization on Face Detection Algorithm with OpenCL: A Case Study

open access: yesTsinghua Science and Technology, 2012
Face detect application has a real time need in nature. Although Viola-Jones algorithm can handle it elegantly, today’s bigger and bigger high quality images and videos still bring in the new challenge of real time needs.
Weiyan Wang   +4 more
doaj   +1 more source

Machine learning guided thermal management of Open Computing Language applications on CPU‐GPU based embedded platforms

open access: yesIET Computers & Digital Techniques, 2023
As embedded devices start supporting heterogeneous processing cores (Central Processing Unit [CPU]–Graphical Processing Unit [GPU] based cores), performance aware task allocation becomes a major issue. Use of Open Computing Language (OpenCL) applications
Rakesh Kumar, Bibhas Ghoshal
doaj   +1 more source

Coupling UV Camera‐Derived SO2 Emission Rate and Thermal Output at Mount Etna: A Decadal Analysis Across Eruptive Phases

open access: yesGeochemistry, Geophysics, Geosystems, Volume 27, Issue 7, July 2026.
Abstract At persistently degassing open‐vent volcanoes, SO2 emissions are especially useful to capture the transition from quiescence to eruption, though accurate quantification is challenged by complex volcanic plume dynamics and limited resolution of observations. Here, we present an exceptionally long (∼10–year) record of SO2 emission rates from the
Giovanni Lo Bue Trisciuzzi   +11 more
wiley   +1 more source

Optimizing OpenCL Code for Performance on FPGA: k-Means Case Study With Integer Data Sets

open access: yesIEEE Access, 2020
High Level Synthesis (HLS) tools targeting Field Programmable Gate Arrays (FPGAs) aim to provide a method for programming these devices via high-level abstractions. Initially, HLS support for FPGAs focused on compiling C/C++ to hardware circuits.
Nuno Paulino   +2 more
doaj   +1 more source

Rethinking Per‐Thread Computation for Machine Learning Design Exploration: A Work‐Efficient GPU Strategy for K‐Means and XGBoost

open access: yesConcurrency and Computation: Practice and Experience, Volume 38, Issue 11, June 2026.
ABSTRACT This paper presents novel GPU implementation strategies that effectively exploit the available parallelism, based on the “more work per thread” approach, for three machine learning design exploration tasks: Multiple K‐means evaluation, dimensionality reduction through parallel K‐means encoding for XGBoost trees, and XGBoost tree pruning ...
Olavo Barros   +8 more
wiley   +1 more source

Beyond TG‑43: A PRISMA‐based systematic review on model‐based dose‐calculation algorithms in brachytherapy

open access: yesJournal of Applied Clinical Medical Physics, Volume 27, Issue 5, May 2026.
Abstract Background and purpose The AAPM TG‐43 formalism has long served as the clinical standard for brachytherapy dose calculation but assumes a homogeneous water equivalent medium, overlooking limited scattering conditions and tissue heterogeneities.
Dat Tran   +4 more
wiley   +1 more source

Optimized Parallel Reduction for Regular and Irregular Segments on GPU

open access: yesConcurrency and Computation: Practice and Experience, Volume 38, Issue 8, April 2026.
ABSTRACT Reduction is an operation that combines all the elements of a collection by applying a binary operation, such as sum, maximum, or minimum, to all the elements to obtain a single resulting value. This paper investigates implementation strategies for both segmented and non‐segmented reduction on GPUs.
Michel B. Cordeiro, Wagner M. Nunan Zola
wiley   +1 more source

OpenCL realization of some many-body potentials [PDF]

open access: yesКомпьютерные исследования и моделирование, 2015
Modeling of carbon nanostructures by means of classical molecular dynamics requires a lot of computations. One of the ways to improve the performance of basic algorithms is to transform them for running on SIMD-type computing systems such as systems with
A. S. Minkin   +2 more
doaj   +1 more source

Performance and Cost Evaluation of StarPU on AWS: Case Studies With Dense Linear Algebra Kernels and N‐Body Simulations

open access: yesConcurrency and Computation: Practice and Experience, Volume 38, Issue 3, February 2026.
ABSTRACT Task‐based programming interfaces introduce a paradigm in which computations are decomposed into fine‐grained units of work known as “tasks”. StarPU is a runtime system originally developed to support task‐based parallelism on on‐premise heterogeneous architectures by abstracting low‐level hardware details and efficiently managing resource ...
Vanderlei Munhoz   +5 more
wiley   +1 more source

Performance Comparison of Different OpenCL Implementations of LBM Simulation on Commodity Computer Hardware

open access: yesAdvances in Electrical and Computer Engineering, 2022
Parallel programming is increasingly used to improve the performance of solving numerical methods used for scientific purposes. Numerical methods in the field of fluid dynamics require the calculation of a large number of operations per second.
TEKIC, J., TEKIC, P., RACKOVIC, M.
doaj   +1 more source

Home - About - Disclaimer - Privacy