-
We present a working implementation of a dynamics based architecture for visual sensing. This architecture provides field rate estimates of the positions and velocities of two independent falling balls in the face of repeated visual occlusions and departures f…
openalex
A.A. Rizzi, Daniel E. Koditschek
1996-01-01
置信度 0.72
SoundnessEstimatorComputer scienceArchitectureController (irrigation)
-
This paper focuses on the problem of manipulating the orientation of a polygonal object in hand. We propose a control technique which integrates the use of palm and fingers to pick up a given object on the table, to drop it on a specific spot on the palm, and …
openalex
Yunfei Bai, C. Karen Liu
2014-05-01
置信度 0.72
GRASPPalmComputer scienceDissipationOrientation (vector space)
-
We use reinforcement learning (RL) to learn dexterous in-hand manipulation policies which can perform vision-based object reorientation on a physical Shadow Dexterous Hand. The training is performed in a simulated environment in which we randomize many of the …
openalex
OpenAI, Andrychowicz, Marcin, Baker, Bowen, Chociej, Maciek 等
2018-08-01
置信度 0.72
Reinforcement learningObject (grammar)Computer scienceShadow (psychology)Artificial intelligence
-
Object-level control of a dexterous robot hand provides an intuitive high-level interface to solve fine manipulation tasks. In the past, many algorithms were proposed based on a weighted pseudoinverse of the grasp map. In a different approach, Stramigioli intr…
openalex
Thomas Wimböck, Christian Ott, Alin Albu‐Schäffer, Gerd Hirzinger
2011-09-12
置信度 0.72
GRASPObject (grammar)Controller (irrigation)Control theory (sociology)Inertia
-
Optimizing behaviors for dexterous manipulation has been a longstanding challenge in robotics, with a variety of methods from model-based control to model-free reinforcement learning having been previously explored in literature. Such prior work often require …
openalex
Sridhar Pandian Arunachalam, Sneha Silwal, Ben Evans, Lerrel Pinto
2023-05-29
置信度 0.72
ImitationComputer scienceHuman–computer interactionArtificial intelligencePsychology
-
Dexterous multi-fingered hands can provide robots with the ability to flexibly perform a wide range of manipulation skills. However, many of the more complex behaviors are also notoriously difficult to control: Performing in-hand object manipulation, executing…
openalex
Anusha Nagabandi, Kurt Konoglie, Sergey Levine, Vikash Kumar
2019-09-25
置信度 0.72
Computer scienceTask (project management)Object (grammar)Artificial intelligenceControl (management)
-
This paper presents a minimalist, four-finger hand comprised of two pairs of tendon-driven, underactuated fingers decoupled by an independent, central, rotating axis. This mechanical configuration allows for finger-gaiting while also retaining the passive adap…
openalex
R. Raymond, Aaron M. Dollar
2014-12-01
置信度 0.72
UnderactuationActuatorComputer scienceRobotic handSet (abstract data type)
-
openalex
Dafni Antotsiou, Guillermo Garcia-Hernando, Tae‐Kyun Kim
2019-01-01
置信度 0.72
RetargetingComputer scienceInverse kinematicsArtificial intelligenceGRASP
-
In this work, a cartesian impedance controller purposely designed for dexterous manipulation is described. Based on the main features of the DLR Hand II, concerning kinematic structure and sensory equipment of fingers, this control strategy allows to overcome …
openalex
Luigi Biagiotti, HaiGe LIU, G. Hirzinger, Claudio Melchiorri
2004-03-30
置信度 0.72
Cartesian coordinate systemImpedance controlKinematicsControl theory (sociology)Electrical impedance
-
Abstract This paper introduces a new technique to synthesize dexterous manipulation of cloth. Given a simple description of the desired cloth motion, our algorithm computes appropriate joint torques for physically simulated hands, such that, via contact forces…
openalex
Yunfei Bai, Wenhao Yu, C. Karen Liu
2016-05-01
置信度 0.72
Computer scienceLaundryKinematicsTorqueSet (abstract data type)
-
Graphical Abstract
openalex
Ryuta Ozawa, Kenji Tahara
2017-08-29
置信度 0.72
GRASPRobotic handComputer sciencePoint (geometry)Control (management)
-
We propose to perform imitation learning for dexterous manipulation with multi-finger robot hand from human demonstrations, and transfer the policy to the real robot hand. We introduce a novel single-camera teleoperation system to collect the 3D demonstrations…
openalex
Yuzhe Qin, Hao Su, Xiaolong Wang
2022-08-22
置信度 0.72
TeleoperationImitationComputer scienceArtificial intelligenceRobot
-
In this paper, we propose a new method for the motion planning problem of rigid object dexterous manipulation with a robotic multi-fingered hand, under quasi-static movement assumption. This method computes both object and finger trajectories as well as the fi…
openalex
Jean-Philippe Saut, Anis Sahbani, Sahar El-Khoury, Véronique Perdereau
2007-10-01
置信度 0.72
GRASPComputer scienceProbabilistic logicLinear subspaceProbabilistic roadmap
-
We present an algorithm called finger tracking for in-hand manipulation of three-dimensional objects with independent robot fingers. We describe and analyze the differential control for finger tracking and extend it to on-line continuous control for a set of c…
openalex
Daniela Rus
1999-04-01
置信度 0.72
PiecewiseTracking (education)Computer scienceSet (abstract data type)Artificial intelligence
-
We propose the Dexterous Manipulation Graph as a tool to address in-hand manipulation and reposition an object inside a robot's end-effector. This graph is used to plan a sequence of manipulation primitives so to bring the object to the desired end pose. This …
openalex
Silvia Cruciani, Christian Smith, Danica Kragić, Kaiyu Hang
2018-10-01
置信度 0.72
GrippersComputer scienceRobotRobot end effectorObject (grammar)
-
Most objects that we manipulate have curved surfaces. We have analyzed how subjects during a prototypical manipulatory task use visual and tactile sensory information for adapting fingertip actions to changes in object curvature. Subjects grasped an elongated …
openalex
Per Jenmalm, Seth Dahlstedt, Roland S. Johansson
2000-12-01
置信度 0.72
GRASPCurvatureComputer visionKinematicsArtificial intelligence
-
The authors formulate the dextrous manipulation problem for a robot hand. First, dextrous manipulation is decomposed into coordinated manipulation, rolling motion, sliding motion, and finger relocation. Then the authors develop motion constraints for each of t…
openalex
Zexiang Li, John Canny, S. Shankar Sastry
2003-01-07
置信度 0.72
Motion (physics)HolonomicMotion planningComputer scienceLift (data mining)
-
The quality of robotic dexterous manipulation, in real or in virtual environments, relies on a fine control of the fingertips to perform stable grasps and inside-hand manipulation. In practice, teleoperating a robotic hand requires to capture the human hand co…
openalex
C. Mizera, Thibault Delrieu, Vincent Weistroffer, Claude Andriot 等
2019-10-15
置信度 0.72
TeleoperationMotion captureComputer visionAvatarArtificial intelligence
-
openalex
Roland S. Johansson
2002-01-01
置信度 0.72
Sensory systemObject (grammar)Computer scienceTactile sensorAfferent
-
Quadrupedal robots are skillful at locomotion tasks while lacking manipulation skills, not to mention dexterous manipulation abilities. Inspired by the animal behavior and the duality between multi-legged locomotion and multi-fingered manipulation, we showcase…
openalex
Fan Shi, Timon Homberger, Joonho Lee, Takahiro Miki 等
2021-05-30
置信度 0.72
QuadrupedalismRobotComputer scienceReinforcement learningArtificial intelligence
-
openalex
Chen Wang, Haochen Shi, Weizhuo Wang, Ruohan Zhang 等
2024-07-15
置信度 0.72
Computer scienceScalabilityHuman–computer interactionData collectionComputer graphics (images)
-
In this letter, we show that soft robotic hands provide a robust means of performing basic primitives of in-hand manipulation in the presence of uncertainty. We first discuss the design of a prototype hand with dexterous soft fingers capable of moving objects …
openalex
Sylvain Abondance, Clark B. Teeple, Robert J. Wood
2020-07-07
置信度 0.72
Computer scienceArtificial intelligenceComputer visionObject (grammar)Robotic hand
-
When robotic hands or arms are capable of enveloping the workpieces that they manipulate, envelopment of the workpiece ensures grasp maintenance even if the object experiences significant external forces in directions unknown prior to grasp synthesis. The enve…
openalex
Jeff Trinkle, Ranganathan Charath Ram, A.O. Farahat, Peter F. Stiller
2002-12-30
置信度 0.72
GRASPPlan (archaeology)Task (project management)Computer scienceStability (learning theory)
-
This paper presents a model based approach to autonomous dexterous manipulation, developed as part of the DARPA Autonomous Robotic Manipulation (ARM) program. The developed autonomy system uses robot, object, and environment models to identify and localize obj…
openalex
Nicolas Hudson, Thomas M. Howard, Jeremy Ma, Abhinandan Jain 等
2012-05-01
置信度 0.72
Computer scienceTask (project management)Human–computer interactionRobotPlan (archaeology)
-
openalex
Pedro J. Sanz, Pere Ridao, Gabriel Oliver, Claudio Melchiorri 等
2010-01-01
置信度 0.72
TridentIntervention (counseling)UnderwaterEuropean commissionEngineering
-
Dexterous manipulation of objects in virtual environments with our bare hands, by using only a depth sensor and a state-of-the-art 3D hand pose estimator (HPE), is challenging. While virtual environments are ruled by physics, e.g. object weights and surface fr…
openalex
Guillermo Garcia-Hernando, Edward Johns, Tae‐Kyun Kim
2020-10-24
置信度 0.72
Reinforcement learningComputer scienceTask (project management)Artificial intelligenceHuman–computer interaction
-
In this paper, a method for dexterous manipulation of 3D soft objects for real-time deformation control is presented, relying on Finite Element modelling. The goal is to generate proper forces on the fingertips of an anthropomorphic device during in-hand manip…
openalex
Fanny Ficuciello, Anna Migliozzi, Eulalie Coevoet, A. Petit 等
2018-10-01
置信度 0.72
Finite element methodComputer scienceDeformation (meteorology)Control (management)Computer vision
-
In this work, we present a new software environment for the comparative evaluation of algorithms for grasping and dexterous manipulation. The key aspect in its development is to provide a tool that allows the reproduction of well-defined experiments in real-li…
openalex
Stefan Ulbrich, Daniel Kappler, Tamim Asfour, Nikolaus Vahrenkamp 等
2011-09-01
置信度 0.72
BenchmarkingSuiteComputer scienceArtificial intelligenceHuman–computer interaction
-
We explore learning-based approaches for feedback control of a dexterous five-finger hand performing non-prehensile manipulation. First, we learn local controllers that are able to perform the task starting at a predefined initial state. These controllers are …
openalex
Vikash Kumar, Abhishek Gupta, Emanuel Todorov, Sergey Levine
2016-11-15
置信度 0.72
Computer scienceTeleoperationTrajectoryArtificial intelligenceTask (project management)
-
To enable general-purpose robots, we will require the robot to operate daily articulated objects as humans do. Current robot manipulation has heavily relied on using a parallel gripper, which restricts the robot to a limited set of objects. On the other hand, …
openalex
Chen Bao, Helin Xu, Yuzhe Qin, Xiaolong Wang
2023-06-01
置信度 0.72
Computer scienceArtificial intelligenceRobotBenchmarkingBenchmark (surveying)
-
The ability to grasp a wider range of objects in size and shape directly relates to the performance of robotic grippers. Adapting to complex geometries of objects requires large degrees of freedom to allow complex configurations. However, complexity in control…
openalex
Sohee John Yoon, Minsik Choi, Bomin Jeong, Yong‐Lae Park
2022-02-07
置信度 0.72
UnderactuationRevolute jointGrippersGRASPTactile sensor
-
Humans can handle and manipulate objects with ease; however, human dexterity has yet to be matched by artificial systems. Receptors in our fingers and hands provide essential tactile information to the motor control system during dexterous manipulation such th…
openalex
Wei Chen, Heba Khamis, Ingvars Birznieks, Nathan F. Lepora 等
2018-09-03
置信度 0.72
GRASPSlip (aerodynamics)Tactile sensorComputer scienceVibration
-
We propose a framework for simulation and control of the human musculoskeletal system, capable of reproducing realistic animations of dexterous activities with high-level coordination. We present the first controllable system in this class that incorporates vo…
openalex
Seung-Hwan Lee, Ri Yu, Jungnam Park, Mridul Aanjaneya 等
2018-07-30
置信度 0.72
Computer scienceController (irrigation)ActuatorTrajectoryCurse of dimensionality
-
In-hand manipulation of objects is an important capability to enable robots to carry-out tasks which demand high levels of dexterity. This work presents a robot systems approach to learning dexterous manipulation tasks involving moving objects to arbitrary 6-D…
openalex
Arthur Allshire, Mayank MittaI, Varun Lodaya, Viktor Makoviychuk 等
2022-10-23
置信度 0.72
Computer scienceReinforcement learningWorkspaceCodebaseRobot
-
openalex
Zhanat Kappassov, Juan Antonio Corrales Ramón, Véronique Perdereau
2015-07-28
置信度 0.72
GRASPTactile sensorComputer scienceRobotExploit
-
This paper presents a method for recognizing dexterous manipulation actions that we usually perform using our hands. Our method is based on model representation using spatio-temporal vector fields and a spotting algorithm that gives segmentation-free and frame…
openalex
Kazutoshi Takahashi, Shogo Seki, E. Kojima, R. Oka
2002-12-17
置信度 0.72
SpottingArtificial intelligenceComputer scienceComputer visionFrame (networking)
-
This paper investigates the equivalent mechanism structure of origami cartons and for the first time proposes a quantitative model of cartons and the interactive configuration space for folding origami cartons. With an analysis of the equivalent mechanism, gus…
openalex
Wei Yao, Jian S. Dai
2007-12-27
置信度 0.72
CartonFolding (DSP implementation)KinematicsKinematic chainConfiguration space
-
In this paper, a high-speed multi-fingered reconfigurable gripper is presented. The aim is to create a robotic end effector that is capable of handling parts of different geometries and weight. It consists of three fingers, accounting for a total of eight Degr…
openalex
Jason Spiliotopoulos, George Michalos, Sotiris Makris
2018-01-10
置信度 0.72
GRASPControl reconfigurationDegrees of freedom (physics and chemistry)GrippersSimplicity
-
Dexterous manipulation is one of the primary goals in robotics. Robots with this capability could sort and package objects, chop vegetables, and fold clothes. As robots come to work side by side with humans, they must also become human-aware. Over the past dec…
openalex
Aude Billard, Danica Kragić
2019-06-20
置信度 0.72
GrippersRobotRoboticsArtificial intelligenceComputer science
-
This paper presents Contact Mode Guided Manipulation Planning (CMGMP) for 3D quasistatic and quasi-dynamic rigid body motion planning in dexterous manipulation. The CMGMP algorithm generates hybrid motion plans including both continuous state transitions and d…
openalex
Xianyi Cheng, Eric Huang, Yifan Hou, Matthew T. Mason
2022-05-23
置信度 0.72
Motion (physics)Computer scienceKey (lock)Object (grammar)Plan (archaeology)
-
An electronically controlled acoustic tweezer was used to demonstrate two acoustic manipulation phenomena: superposition of Bessel functions to allow independent manipulation of multiple particles and the use of higher-order Bessel functions to trap particles …
openalex
C. R. P. Courtney, Christine Démoré, Hongxiao Wu, A. Grinenko 等
2014-04-14
置信度 0.72
Optical tweezersTweezersSuperposition principleSPHERESBessel function
-
In this paper, we present a novel method for achieving dexterous manipulation of complex objects, while simultaneously securing the object without the use of passive support surfaces.We posit that a key difficulty for training such policies in a Reinforcement …
openalex
Gagan Khandate, Siqi Shang, Eric T. Chang, Tristan L. Saidi 等
2023-07-10
置信度 0.72
Reinforcement learningComputer scienceSampling (signal processing)Artificial intelligenceReinforcement
-
Integrating vision-language models (VLMs) into clinical radiology workflows requires exporting two-dimensional images that preserve diagnostic viewing context, including imaging plane, slice position, window/level, slab parameters, and overlays. Existing appro…
pubmed
Albera M, Colarieti A, Albera I, Carriero A
2026 Jul 1
置信度 0.82
-
The problem of Human Action Recognition (HAR) continues to be difficult with the intricate temporal interactions, superfluous frames, and minor visual variations that usually define similar actions. A large number of current approaches are based on either tran…
pubmed
Majid A, Wang Y, Aqsa, Xing Z 等
2026 Jun 30
置信度 0.82
-
Vision-and-language navigation (VLN) traditionally relies on explicit reasoning chains, which, despite being interpretable, impose severe constraints on inference efficiency and scalability in long-range environments. Existing multimodal large language models …
pubmed
Zhu R, Li S, Yang M
2026 Jun 15
置信度 0.82
-
Vision-and-Language Navigation in continuous environments (VLN-CE) requires embodied agents to ground natural language instructions into reliable long-horizon motion decisions under partial observability. Despite their strong semantic understanding and reasoni…
pubmed
Liu T, Qi X, Wang L, Li J 等
2026 Jun 10
置信度 0.82
-
Vision-language models (VLMs) have shown promising performance in surgical visual question answering (VQA). However, existing surgical VQA datasets often contain linguistic shortcuts, where question phrasing implicitly constrains the answer space. In safety-cr…
pubmed
Shin J, Kim KY, Cho E, Kim ST 等
2026 Jun 18
置信度 0.82
-
Recent Vision-Language-Action (VLA) models have rapidly emerged as general-purpose robotic policies that integrate language understanding, visual perception, and robot control. However, prior studies and surveys have primarily emphasized backbone architectures…
pubmed
Ko BC
2026 Jun 3
置信度 0.82
-
Reasoning capability has significantly advanced complex logical inference and robotic decision-making in general domains. However, its potential in the Artificial Intelligence (AI) copilot robot—particularly implemented based on the Vision-Language-Acti…
pubmed
Wang G, Bai L, Ren H
2026 Jun 11
置信度 0.82
-
Effective physical exposure assessment for manual materials handling (MMH) is essential for identifying activities that increase the risk of work-related musculoskeletal disorders and for guiding ergonomic interventions. However, existing methods are labor-int…
pubmed
Rajabi MS, Ojelade A, Kim S, Nussbaum MA
2026 Jun 5
置信度 0.82
-
Embodied navigation and manipulation are fundamental capabilities for embodied agents operating in physical environments. A key challenge in this process is understanding the spatial context and the affordances of the environment, which involves recognizing ho…
pubmed
Hao X, Tang Y, Zhang L, Chen L 等
2026
置信度 0.82
-
This study proposes a psychology informed framework for exploring the application of large language models (LLMs) in dance education. To address the limited personalization of traditional dance instruction and its insufficient adaption to learners' cognitive a…
pubmed
Zhao L, Huang Q, Zhou K, Teng L 等
2026
置信度 0.82
-
As animal brain size increases, cognitive performance generally increases. However, other key brain functions, such as the regulation of somatic processes, processing of sensory input, and planning and initiation of motor actions, are also closely tied to brai…
pubmed
Song Z, Heldstab SA, Kay RF, Kirk EC 等
2026 Jun
置信度 0.82
-
Vision-Language-Action (VLA) policies promise flexible long-horizon manipulation, but deployment under domain shift requires both reliable uncertainty estimates and a workable runtime-assurance policy. We study a model-agnostic uncertainty-calibrated safety-ga…
pubmed
Ghaleb AM, Allahloh AS, Mejjaouli S, Ali MAH 等
2026 May 15
置信度 0.82
-
Neurology and psychiatry have operated as separate disciplines for over a century, yet this division reflects historical and institutional developments rather than the underlying biology of the brain. Contemporary neuroscience shows that brain and mental healt…
pubmed
Bègue I, Mohr P, Bassetti CLA, Fiorillo A 等
2026 May 25
置信度 0.82
-
The 2023 iteration of the Global Burden of Diseases, Injuries, and Risk Factors Study (GBD) estimated prevalence, incidence, and health burden for 375 diseases and injuries, including 12 mental disorders. We assess past, current, and emerging trends in the pre…
pubmed
GBD 2023 Mental Disorder Collaborators
2026 May 23
置信度 0.82
-
Industrial processes must operate robustly in unpredictable environments, where errors are costly and difficult to detect. AI-based control systems offer a path forward but typically rely on large, labeled datasets, limiting their generalization to variable, d…
pubmed
Margadji C, Pattinson SW
2026 May 18
置信度 0.82
-
Artificial intelligence (AI) for surgical workflow analysis often fails to generalize because surgical actions lack a standardized, fine-grained representation. Gesture-level "tokenization" of surgery, capturing instrument-tissue interactions as the smallest i…
pubmed
Morais MC, Godbole AA, Iqbal E, Ballo M 等
2026 Jun
置信度 0.82
-
Forecasting how human hands move in egocentric views is critical for applications like augmented reality, human-robot policy transfer, and service/assistive technologies. Recently, several hand trajectory prediction (HTP) methods have been developed to generat…
pubmed
Ma J, Bao W, Xu J, Sun G 等
2026 May 13
置信度 0.82
-
In this era, the success of large language models and text-to-image models can be attributed to the driving force of large-scale datasets. However, in the realm of 3D vision, while significant progress has been achieved in object-centric tasks through large-sc…
pubmed
Li C, Liao H, Zhi Y, Yang X 等
2026 May 8
置信度 0.82
-
Recent advances in vision-language cross-modal learning have substantially improved the performance of video temporal grounding. However, most existing methods directly associate global video features with sentence-level features, overlooking the fact that tex…
pubmed
Tian Y, Guo X, Wang J, Zhao Y 等
2026 Apr 1
置信度 0.82
-
Vision-Language Action (VLA) models have enabled language-driven robotic manipulation by integrating language instructions, visual perception, and action generation. However, existing VLA approaches heavily rely on large-scale human demonstration datasets, whi…
pubmed
Yi T, Yang Q, Chen E
2026
置信度 0.82
-
Embodied AI is widely recognized as a cornerstone of artificial general intelligence (AGI) because it involves controlling embodied agents to perform tasks in the physical world. Building on the success of large language models (LLMs) and vision-language model…
pubmed
Ma Y, Song Z, Zhuang Y, Hao J 等
2026 Jul
置信度 0.82
-
"Population Health Management (PHM), Fit Lifecycles in Analytics" examines the policy and practice of AI-driven methodologies to enhance public health and patient safety in the context of the Human Phenotype Ontology (HPO). It aims for personalized healthcare …
pubmed
Henry JA
2025
置信度 0.82
-
Action Quality Assessment (AQA) has gained significant attention due to its potential real-world applications, which require a fine-grained understanding of action sequences. Recent works have attempted to utilize multimodal video features and address some exi…
pubmed
Gedamu K, Ji Y, Zuo W, Bentahar J 等
2026
置信度 0.82
-
Foundation models for embodied artificial intelligence (Embodied AI) increasingly adopt diffusion modules as the action generation core of vision-language-action (VLA) policies, but the diffusion module's iterative denoising imposes prohibitive inference laten…
pubmed
Shi X, Hu Y, Jin J
2026
置信度 0.82
-
Autonomous vehicles, such as Unmanned Aerial Vehicles (UAVs), have the potential to completely reshape various industries such as parcel delivery, agriculture, surveillance, monitoring, and search-and-rescue missions. Consequently, the demand for safe, cost-ef…
pubmed
Sarker GC, Azad A, Rahman S, Hasan MM
2026
置信度 0.82
-
The accelerating antimicrobial resistance (AMR) crisis continues to render more and more conventional antibiotics ineffective. Antimicrobial peptides (AMPs) are promising alternatives to traditional antibiotics due to their broad-spectrum activity, diverse mec…
pubmed
Ibisanmi TA, Jiang X, Willcox M, Kumar N
2026
置信度 0.82
-
Understanding animal actions and interactions is essential for behavior analysis and ecological monitoring. Although large-scale in-the-wild datasets have advanced animal action recognition, existing methods still struggle with fine-grained motion, spatial rel…
pubmed
Yang Y, Nakagawa R, Shinoda R, Santo H 等
2026 Mar 21
置信度 0.82
-
Reliable fault detection along transmission corridors is essential for preventing small defects from developing into long outages and costly emergency operations. This study aims to improve the field reliability of an open vocabulary vision language backbone w…
pubmed
Yu R, Mai L, Weng Y, Cui Q 等
2026 Feb 28
置信度 0.82
-
Among the many excellent presentations and posters of this conference, I was tasked with making sense of two of the most esoteric. This is my meat and potatoes. I hasten to add that they are not esoteric because they are peripheral; au contraire , they are cen…
pubmed
Killeen PR
2026 Mar
置信度 0.82
-
While recent research suggests Large Language Models match human creative performance in divergent thinking tasks, visual creativity remains underexplored. This study compared image generation in human participants (Visual Artists and Non-Artists) and using an…
pubmed
Rondini S, Alvarez-Martin C, Angermair-Barkai P, Penacchio O 等
2026 May
置信度 0.82
-
Unconstrained fall detection is essential for real-world applications. However, it remains underexplored due to the scarcity of real-world fall data and the limited generalization ability of existing methods. To address these challenges, we first introduce HUS…
pubmed
Wu S, Chen T, Zha Z, Wu B 等
2026 Mar 23
置信度 0.82
-
Trajectory prediction is a fundamental problem in computer vision, vision-language-action models, world models, and autonomous systems, with broad impact on applications including autonomous driving, robotics, and surveillance. Most existing approaches assume …
pubmed
Zhang H, Xu Y, Fu Y
2026 Aug
置信度 0.82
-
Botulinum toxin (BoNT) is a highly specific molecular enzyme whose therapeutic action is based on the proteolytic cleavage of SNARE proteins, most notably SNAP-25. Despite the deterministic nature of this molecular mechanism, the clinical effects of BoNT exhib…
pubmed
Armenti AF, Armenti F
2026 Mar 9
置信度 0.82
-
Video temporal grounding, including moment retrieval and highlight detection, is an emerging topic aiming to identify specific clips within videos. In addition to pre-trained video models, contemporary methods utilize pre-trained vision-language models (VLMs) …
pubmed
Wang Y, Jiang X, Cheng D, Li D 等
2026
置信度 0.82
-
Generalization to unseen environments remains a fundamental challenge in Vision-Language Navigation. To tackle this issue, we propose a novel framework that leverages world knowledge embedded within Multimodal Large Language Models. We introduce Collaborative …
pubmed
Zhu R, Li S, Zhu Z, Jia J 等
2026 Feb 14
置信度 0.82
-
The paper presents a visio-verbal teleimpedance interface for commanding 3D stiffness ellipsoids to the remote robot with a combination of the operator's gaze and verbal interaction. The gaze is detected by an eye-tracker, allowing the system to understand the…
pubmed
Jekel HHA, Díaz Rosales A, Peternel L
2026
置信度 0.82
-
Vision-language pre-training (VLP) offers unique advantages for surgery by aligning language with surgical videos, enabling workflow understanding and transfer across tasks without relying on expert-labeled datasets. However, progress in surgical VLP remains c…
pubmed
Perez A, Nwoye C, Raji Kermani R, Mohareri O 等
2026 May
置信度 0.82
-
Recent advances in surgical robotics and computer vision have greatly improved intelligent systems' autonomy and perception in the operating room (OR), especially in endoscopic and minimally invasive surgeries. However, for open surgery, which is still the pre…
pubmed
Xu B, Wu J, Liang J, Sun Z 等
2026
置信度 0.82
-
The preschool years (ages 3-5) represent a critical window for promoting development and lifelong health. However, in many low-resource settings, developmental delays, sensory impairments and emerging health risks often go undetected. Although early, integrate…
pubmed
Smith R, Jordaan EM, Russell DC, de Milander M 等
2026 Mar
置信度 0.82
-
Foundation Models (FMs) are profoundly transforming the clinical management pathway for liver cancer. Their core value lies in enhancing diagnostic accuracy, enabling personalized therapeutic decision-making, and optimizing clinical efficiency through the inte…
pubmed
Wang J, Xue S, Xia H, Cui P 等
2026 Jan 29
置信度 0.82
-
Developing a unified algorithm that can learn from and generate across modalities such as text, images and video has been a fundamental challenge in artificial intelligence. Although next-token prediction has driven major advances in large language models 1 , …
pubmed
Wang X, Cui Y, Wang J, Zhang F 等
2026 Feb
置信度 0.82
-
Autonomous AI-to-AI creative systems promise new frontiers in machine creativity, yet we show that they systematically converge toward generic outputs. We built iterative feedback loops between Stable Diffusion XL (SDXL; image generation) and Large Language an…
pubmed
Hintze A, Proschinger Åström F, Schossau J
2026 Jan 9
置信度 0.82
-
In the era of data-driven research and artificial intelligence, proper dataset licensing and attribution practices are crucial for legal compliance and ethical data usage. Open-access datasets are often shared under various licensing schemes, such as Creative …
pubmed
Ayyamperumal V, Aswath S, Vignesh S, Thamaraimanalan T
2026 Jan 22
置信度 0.82
-
The growing demand for therapeutic support increasingly exceeds the capacity of available professionals. A virtual agent capable of performing motivational interviewing (MI) offers a promising solution to assist patients in reaching their goal of behavior chan…
pubmed
Galland L, Younsi N, Baudonne C, Chaby L 等
2025 Dec 23
置信度 0.82
-
Traditional natural disaster response involves significant coordinated teamwork, where speed and efficiency are key. Nonetheless, human limitations can delay critical actions and inadvertently increase human and economic losses. Agentic Large Vision Language M…
pubmed
Chen Z, Asadi Shamsabadi E, Jiang S, Shen L 等
2026 Jan 10
置信度 0.82
-
For a humanoid robot, it is difficult to predict a motion trajectory through end-to-end imitation learning when performing complex operations and multi-step processes, leading to jittering in the robot arm. To alleviate this problem and reduce the computationa…
pubmed
Ren B, Shi D
2025 Dec 26
置信度 0.82
-
Clinical documentation demands are increasingly eroding clinician time and morale. Large language models (LLMs) are emerging as practical allies, drafting notes in real-time and laying the groundwork for decision support. This narrative review examines both re…
pubmed
Elechi U, Orobator ET, Udoh K, Ngozi EO 等
2025
置信度 0.82
-
More knowledge and resources are required to strengthen 'leadership and governance' (L+G) as a central building block to further develop emergency care (EC) systems in low-income and middle-income countries (LMICs).
pubmed
Phillips G, Sharma D, O'Reilly G, Romero L 等
2025 Dec 19
置信度 0.82
-
Transfer learning from image to video has become a widely adopted strategy in action recognition. Existing mainstream approaches typically fine-tune the entire network after initialising with pre-trained parameters, which inevitably leads to substantial traini…
pubmed
Wu C, Xu T, Feng Z, Wu XJ 等
2026 Apr
置信度 0.82
-
Procedure planning in instructional videos entails predicting an action sequence that transitions a given start state to a desired goal state. This task is particularly challenging due to two key sources of uncertainty: limited visual observations and an enorm…
pubmed
Fang F, Yang M, Wu M, Yang Y 等
2026 Apr
置信度 0.82
-
Despite impressive performance in various tasks, large language models (LLMs) are subject to the symbol grounding problem, so from the cognitive science perspective, one can argue that they are merely statistics-driven distributional models without a deeper un…
pubmed
Farkaš I, Vavrečka M, Wermter S
2025
置信度 0.82
-
Hearing loss is a sensory damage, a hidden disability that affects numerous adults globally. However, gesture recognition plays a significant role in overcoming several problems and challenges in human life, particularly for persons with hearing impairments an…
pubmed
Maashi M, Aljohani N, Alsahafi YA, Rizwanullah M
2025 Dec 2
置信度 0.82
-
Reproducibility in biological research and manufacturing remains constrained by the complexity of multi-step protocols, fragmented data-analysis pipelines, and the intrinsic variability of experimental execution. Here, we present Agentic Lab, an agentic-physic…
pubmed
Wang W, Swain S, Lee J, Lin Z 等
2025 Nov 13
置信度 0.82
-
Understanding human intentions and actions through egocentric videos is important on the path to embodied artificial intelligence. As a branch of egocentric vision techniques, hand trajectory prediction plays a vital role in comprehending human motion patterns…
pubmed
Ma J, Chen X, Bao W, Xu J 等
2026 Mar
置信度 0.82
-
Multimodal contrastive learning has achieved significant performance advantages in self-supervised skeleton-based action recognition. Previous methods are limited by modality imbalance, which reduces alignment accuracy and makes it difficult to combine importa…
pubmed
Xin W, Teng Y, Zhang J, Liu Y 等
2025 Oct 23
置信度 0.82
-
Automatic furniture layout is long desired for convenient interior design. Leveraging the remarkable visual reasoning capabilities of multimodal large language models (MLLMs), recent methods address layout generation in a static manner, lacking the feedback-dr…
pubmed
Wang C, Zhong H, Chai M, He M 等
2026 Feb
置信度 0.82
-
Computer vision offers a promising approach to automating the observation of animal behavior, thereby contributing to improved animal welfare and precision livestock management. However, the absence of standardized behavioral definitions limits the accuracy an…
pubmed
Zhou S, Li W, Zhou M, Dilger RN 等
2025 Oct 19
置信度 0.82
-
Autonomous driving in complex real-world environments requires robust perception, reasoning, and physically feasible planning, which remain challenging for current end-to-end approaches. This paper introduces VLA-MP, a unified vision-language-action framework …
pubmed
Ge M, Ohtani K, Niu Y, Zhang Y 等
2025 Oct 5
置信度 0.82