Event-Based Vision at the Edge: A Review. [PDF]
Middleton M +9 more
europepmc +1 more source
The three‐step test in international copyright—a global framework for generative AI training
Abstract The training of generative artificial intelligence (genAI) models involves the processing of large datasets, which often include copyrighted materials, thereby raising questions of copyright infringement. These questions tend to center on whether such use constitutes fair use or falls under text and data mining (TDM) exceptions.
Nicola Lucchi +2 more
wiley +1 more source
Multimodal Navigation System for Visually Impaired Users Using Environmental Perception and Vision-Language Models. [PDF]
Lin HY, Fan YH, Chang CC.
europepmc +1 more source
Cache‐Cache: Dans les souffles du Grand Fleuve…
ABSTRACT Between ventriloquism and divination, devotion and consecration, an old and sublime marine game unfolds each day, with living breaths shared in the waters of the great river. In Tadoussac, on the north shore of the St. Lawrence Estuary, at the mouth of the Saguenay, the rumor of whales has long resonated.
David Jaclin +3 more
wiley +1 more source
Exploring vision transformers for deep feature extraction and classification in video genre recognition for digital media. [PDF]
Alarfaj FK, Naz A.
europepmc +1 more source
Democratising Multi‐Projector Displays
Spatially augmented reality (SAR) transforms large, surround, collaborative experiences out of VR/AR headsets to the real world by merging content from projectors with the physical environment. This detailed state‐of‐the‐art survey reports on the advancements in multi‐projector aggregation and hardware technologies used to achieve SAR and build ...
Aditi Majumder, Muhammad Twaha Ibrahim
wiley +1 more source
Modular Framework for Responsive and Explainable Robotic Assistance with Intention Prediction Using Human-Centric Digital Twins. [PDF]
Asad U +4 more
europepmc +1 more source
A Survey of Methods for Constructing 3D Urban Models From Point Clouds
The survey outlines a general process for constructing 3D models from point clouds and categorizes the methods based on their use of templates, basic surface primitives, hybrid approaches, or linear primitives. Additionally, the survey reviews the datasets and benchmarks that are essential for testing and training the developed methods.
Chiara Romanengo +5 more
wiley +1 more source
A Multi-Module Fusion Framework for Restoring Human and Machine Vision Quality in Compressed Video. [PDF]
He K, Xiang K, Gao Y, Yu Y, Zhou J.
europepmc +1 more source
Story2Board: A Training‐Free Approach for Expressive Visual Storytelling
Abstract We present Story2Board, a training‐free framework for expressive storyboard generation from natural language. Existing methods narrowly focus on subject identity, overlooking key aspects of visual storytelling such as spatial composition, background evolution, and narrative pacing.
D. Dinkevich +4 more
wiley +1 more source

