Results 51 to 60 of about 34,039 (292)
Part of the report Parallelizing PageRank for a Volta GPU. -- Comparing various launch configs for CUDA based vector element sum. A floating-point vector x, with number of elements from 1E+6 to 1E+9 was summed up using CUDA (Σx).
Subhajit Sahu
core +1 more source
Gap‐Free Information Transfer in 4D‐STEM via Fusion of Complementary Scattering Channels
Fused Full‐Field STEM (FF‐STEM) is introduced as a 4D‐STEM imaging modality that combines direct ptychography with tilt‐corrected dark‐field reconstruction in a single acquisition. Fourier‐space fusion using Wiener‐type spectral weighting closes the low‐frequency contrast gap inherent to bright‐field methods, delivering gap‐free, dose‐efficient, near ...
Shengbo You +15 more
wiley +1 more source
CU2CL: A CUDA-to-OpenCL Translator for Multi- and Many-core Architectures [PDF]
The use of graphics processing units (GPUs) in high-performance parallel computing continues to become more prevalent, often as part of a heterogeneous system.
Martinez, Gabriel +2 more
core +1 more source
Analysis and development tools for efficient programs on parallel architectures
The article proposes methods for supporting development of efficient programs for modern parallel architectures, including hybrid systems. Specialized profiling methods designed for programmers tasked with parallelizing existing code are proposed.
Alexander Monakov +3 more
doaj +1 more source
Machine Learning of Temperature‐Dependent Chemical Kinetics Using Parallel Droplet Microreactors
An integrated droplet microfluidics and machine learning framework enables high‐throughput characterization of temperature‐dependent reaction kinetics. Time‐resolved measurements from thousands of droplets train Neural ODE models that accurately predict nonlinear reaction dynamics across diverse thermal environments, bridging large‐scale ...
Mamoru Saita, Yutaka Hori
wiley +1 more source
Acceleration computing process in wavelength scanning interferometry [PDF]
The optical interferometry has been widely explored for surface measurement due to the advantages of non-contact and high accuracy interrogation. Eventually, some interferometers are used to measure both rough and smooth surfaces such as white light ...
Gao, F. +2 more
core +3 more sources
Аlgorithm of fast computation of local image histograms on video card1
An algorithm of parallel computation of image histograms of different types, including brightness and oriented gradient ones, on video cards of various types is presented.
Ph. S. Trotski, B. A. Zalesky
doaj
An Expertise Transfer Framework For Autonomous Surgical Assistance
This study introduces an expertise transfer framework for procedure‐spanning autonomous surgical assistance. By emulating expert logic through hierarchical perception, attention modeling, and knowledge graph‐based decision‐making, the system provides near‐expert surgical view assistance.
Yuan Gao +12 more
wiley +1 more source
Analysis and design of massively parallel channel estimation algorithms on graphic cards [PDF]
The necessity of accurate channel estimation for coherent multiuser detectors is well known. Indeed they are based on the assumption that signals are perfectly estimated, and this is never completely achieved in practice.
Da Lio, Davide
core
Machine learning interatomic potentials bridge quantum accuracy and computational efficiency for materials discovery. Architectures from Gaussian process regression to equivariant graph neural networks, training strategies including active learning and foundation models, and applications in solid‐state electrolytes, batteries, electrocatalysts ...
In Kee Park +19 more
wiley +1 more source

