Results 61 to 70 of about 14,258 (262)

Virtual Experience Toolkit: An End-to-End Automated 3D Scene Virtualization Framework Implementing Computer Vision Techniques

open access: yesSensors
Virtualization plays a critical role in enriching the user experience in Virtual Reality (VR) by offering heightened realism, increased immersion, safer navigation, and newly achievable levels of interaction and personalization, specifically in indoor ...
Pau Mora   +4 more
doaj   +1 more source

Improving the Robustness of Visual Teach‐and‐Repeat Navigation Using Drift Error Correction and Event‐Based Vision for Low‐Light Environments

open access: yesAdvanced Robotics Research, EarlyView.
Visual teach‐and‐repeat (VTR) navigation allows robots to learn and follow routes without building a full metric map. We show that navigation accuracy for VTR can be improved by integrating a topological map with error‐drift correction based on stereo vision.
Fuhai Ling, Ze Huang, Tony J. Prescott
wiley   +1 more source

An Approach to Global Illumination Calculation Based on Hybrid Cone Tracing

open access: yesIEEE Access, 2020
For 3D geographic information systems (GIS) or video game systems, global illumination (GI) effect can greatly improve the understanding of the volumetric structure and the spatial relationships of objects in the scene.
Tao Liu, Jin Gao, Zhengling Lei
doaj   +1 more source

DC-Scene: Data-Centric Learning for 3D Scene Understanding

open access: yesCoRR
3D scene understanding plays a fundamental role in vision applications such as robotics, autonomous driving, and augmented reality. However, advancing learning-based 3D scene understanding remains challenging due to two key limitations: (1) the large scale and complexity of 3D scenes lead to higher computational costs and slower training compared to 2D
Ting Huang   +3 more
openaire   +2 more sources

Using Ignorance in 3D Scene Understanding [PDF]

open access: yesMathematical Problems in Engineering, 2014
Awareness of its own limitations is a fundamental feature of the human sight, which has been almost completely omitted in computer vision systems. In this paper we present a method of explicitly using information about perceptual limitations of a 3D vision system, such as occluded areas, limited field of view, loss of precision along with distance ...
Bogdan Harasymowicz-Boggio   +1 more
openaire   +1 more source

A Self‐Healing Permanent Magnet Putty for Soft Robot Skins With Force Sensing and Functional Recovery

open access: yesAdvanced Robotics Research, EarlyView.
Permanent magnet putty (PMP) integrates high‐coercivity NdFeB particles with a dynamic polyborosiloxane–Ecoflex matrix, achieving rapid self‐healing (90% mechanical recovery in 10 s) and magnetic recovery within 20 min. With twice the sensitivity of commercial putties, PMP enables precise 5–30 N force detection and discrimination between pressing and ...
Ruotong Zhao   +5 more
wiley   +1 more source

Open-Vocabulary 3-D Segmentation With 3-D-Aware Mask Alignment and Identity Expansion

open access: yesIEEE Access
Open-Vocabulary 3D segmentation assigns semantic labels to 3D scenes from free-form text prompts. Many approaches fuse multi-view 2D embeddings into a 3D representation, but performance depends on strong cross-view consistency, which degrades under large
Weijie Lin   +4 more
doaj   +1 more source

Articulate3D: Holistic Understanding of 3D Scenes as Universal Scene Description

open access: yes2025 IEEE/CVF International Conference on Computer Vision (ICCV)
3D scene understanding is a long-standing challenge in computer vision and a key component in enabling mixed reality, wearable computing, and embodied AI. Providing a solution to these applications requires a multifaceted approach that covers scene-centric, object-centric, as well as interaction-centric capabilities. While there exist numerous datasets
Anna-Maria Halacheva   +5 more
openaire   +2 more sources

Multimodal Human–Robot Interaction Using Human Pose Estimation and Local Large Language Models

open access: yesAdvanced Robotics Research, EarlyView.
A multimodal human–robot interaction framework integrates human pose estimation (HPE) and a large language model (LLM) for gesture‐ and voice‐based robot control. Speech‐to‐text (STT) enables voice command interpretation, while a safety‐aware arbitration mechanism prioritizes gesture input for rapid intervention.
Nasiru Aboki   +2 more
wiley   +1 more source

Home - About - Disclaimer - Privacy