Results 81 to 90 of about 11,177,407 (275)
Generating Visual Representations for Zero-Shot Classification
International audienceThis paper addresses the task of learning an image clas-sifier when some categories are defined by semantic descriptions only (e.g. visual attributes) while the others are defined by exemplar images as well.
Frederic Jurie +5 more
core +2 more sources
The Future of Research in Cognitive Robotics: Foundation Models or Developmental Cognitive Models?
Research in cognitive robotics founded on principles of developmental psychology and enactive cognitive science would yield what we seek in autonomous robots: the ability to perceive its environment, learn from experience, anticipate the outcome of events, act to pursue goals, and adapt to changing circumstances without resorting to training with ...
David Vernon
wiley +1 more source
Zero-Shot Image Classification Method Based on Attention Mechanism and Semantic Information Fusion. [PDF]
Wang Y, Feng L, Song X, Xu D, Zhai Y.
europepmc +1 more source
Grounding Large Language Models for Robot Task Planning Using Closed‐Loop State Feedback
BrainBody‐Large Language Model (LLM) introduces a hierarchical, feedback‐driven planning framework where two LLMs coordinate high‐level reasoning and low‐level control for robotic tasks. By grounding decisions in real‐time state feedback, it reduces hallucinations and improves task reliability.
Vineet Bhat +4 more
wiley +1 more source
Robots can learn manipulation tasks from human demonstrations. This work proposes a versatile method to identify the physical interactions that occur in a demonstration, such as sequences of different contacts and interactions with mechanical constraints.
Alex Harm Gert‐Jan Overbeek +3 more
wiley +1 more source
Automated poultry processing lines still rely on humans to lift slippery, easily bruised carcasses onto a shackle conveyor. Deformability, anatomical variance, and hygiene rules make conventional suction and scripted motions unreliable. We present ChicGrasp, an end‐to‐end hardware‐software co‐designed imitation learning framework, to offer a ...
Amirreza Davar +8 more
wiley +1 more source
Multimodal Human–Robot Interaction Using Human Pose Estimation and Local Large Language Models
A multimodal human–robot interaction framework integrates human pose estimation (HPE) and a large language model (LLM) for gesture‐ and voice‐based robot control. Speech‐to‐text (STT) enables voice command interpretation, while a safety‐aware arbitration mechanism prioritizes gesture input for rapid intervention.
Nasiru Aboki +2 more
wiley +1 more source
Zero-shot image privacy classification with Vision-Language Models
While specialized learning-based models have historically dominated image privacy prediction, the current literature increasingly favours adopting large Vision-Language Models (VLMs) designed for generic tasks. This trend risks overlooking the performance ceiling set by purpose-built models due to a lack of systematic evaluation.
Alina Elena Baia +2 more
openaire +3 more sources
This review maps the methods to monitor robots’ health by fusing vibration, sound, control signals, vision, force, and oil information with artificial intelligence. It identifies deep learning, transfer learning, digital twins, and physics‐informed models as key methodological pathways enabling earlier diagnosis, safer human–robot collaboration, and ...
Yuting Qiao +6 more
wiley +1 more source
Zero-Shot Image Classification Algorithm Based on Particle Swarm Optimization Fusion Feature
Aiming at the limitation of describing the semantic attributes of target classification, this paper proposes an adaptive weighted fusion feature based zero-sampling image classification algorithm.
doaj +1 more source

