-
Timely and comprehensive analyses of causes of death stratified by age, sex, and location are essential for shaping effective health policies aimed at reducing global mortality. The Global Burden of Diseases, Injuries, and Risk Factors Study (GBD) 2023 provide…
pubmed
GBD 2023 Causes of Death Collaborators
2025 Oct 18
置信度 0.82
-
Human-robot interaction (HRI) via voice command has significantly advanced in recent years, with large Vision-Language-Action (VLA) models demonstrating particular promise in human-robot voice interaction. However, these systems still struggle with environment…
pubmed
Li M, Xu W, Zeng C, Wang N
2025 Sep 9
置信度 0.82
-
Reliance on internal predictive models of the world is central to many theories of human cognition. Yet it is unknown whether humans acquire multiple separate internal models, each evolved for a specific domain, or maintain a globally unified representation. U…
pubmed
Yazin F, Majumdar G, Bramley N, Hoffman P
2025 Sep 25
置信度 0.82
-
Fine-grained action recognition (FGAR) aims to identify subtle and distinctive differences among fine-grained action categories. However, current recognition methods often capture coarse-grained motion patterns but struggle to identify subtle details in local …
pubmed
Sun B, Wang Y, Ma X, Wang Z 等
2026 Jan
置信度 0.82
-
Robotic automation is instrumental in the valorization of construction and demolition waste (CDW), facilitating scalable and efficient material recovery in response to rising waste volumes from accelerated urban development. AI-driven computer vision (CV) has …
pubmed
Prasad V, Arashpour M
2025 Oct
置信度 0.82
-
Considering how to make the model accurately understand and follow natural language instructions and perform actions consistent with world knowledge is a key challenge in robot manipulation. This mainly includes human fuzzy instruction reasoning and the follow…
pubmed
Ren P, Zhang K, Zheng H, Li Z 等
2025 Dec
置信度 0.82
-
Navigation means getting from here to there. Unfortunately, for biological navigation, there is no agreed definition of what we might mean by 'here' or 'there'. Computer vision ('Simultaneous Localisation and Mapping', SLAM) uses a 3D world-based coordinate fr…
pubmed
Glennerster A
2025 Dec 15
置信度 0.82
-
Current vision-language models (VLMs) are well-adapted for general visual understanding tasks. However, they perform inadequately when handling complex visual tasks related to human poses and actions due to the lack of specialized vision-language instruction-f…
pubmed
Zhang D, Hussain T, An W, Shouno H
2025 Aug 21
置信度 0.82
-
Background/Objectives : The rapid development of artificial intelligence is transforming the face of medicine. Due to the large number of imaging studies (pre-, intra-, and postoperative) combined with histopathological and molecular findings, its impact may b…
pubmed
Szmyd B, Podstawka M, Wiśniewski K, Zaczkowski K 等
2025 Aug 11
置信度 0.82
-
Food product labels serve as a critical source of information, providing details about nutritional content, ingredients, and health implications. These labels enable Food and Drug Authorities (FDA) to ensure compliance and take necessary health-related and log…
pubmed
Assiri FY, Alahmadi MD, Almuashi MA, Almansour AM
2025 Aug 13
置信度 0.82
-
Video summarization aims to generate a compact summary of the original video by selecting and combining the most representative parts. Most existing approaches only focus on recognizing key video segments to generate the summary, which lacks holistic considera…
pubmed
Ye C, Chen W, Hu B, Zhang L 等
2025
置信度 0.82
-
There is contradictory evidence on the effect that visual experience has on haptic abilities. Indeed, some studies have documented that a lack of vision (blindness) results in decreased haptic perception, whereas other studies report an enhanced haptic ability…
pubmed
Coelho LA, Ramirez DEA, Basta S, Guarischi M 等
2025 Aug 19
置信度 0.82
-
With the significant development of large models in recent years, large vision-language models (LVLMs) have demonstrated remarkable capabilities across a wide range of multimodal understanding and reasoning tasks. Compared with traditional large language model…
pubmed
Liu D, Yang M, Qu X, Zhou P 等
2025 Nov
置信度 0.82
-
Spiking Neural Networks (SNNs), which simulate the spiking behavior of biological neurons, have attracted significant attention in recent years due to their distinctive low-power characteristics. Meanwhile, Transformer models, known for their powerful self-att…
pubmed
Zhang H, Sboev A, Rybka R, Yu Q
2025 Nov
置信度 0.82
-
Behavior analysis across species represents a fundamental challenge in neuroscience, psychology, and ethology, typically requiring extensive expert knowledge and labor-intensive processes that limit research scalability and accessibility. We introduce BehaveAg…
pubmed
Aljović A, Lin Z, Wang W, Zhang X 等
2025 May 20
置信度 0.82
-
Parkinson's Disease (PD) is a neurodegenerative disorder that affects motor and non-motor functions. Speech impairments, such as reduced variability in pitch (F0SD) and intensity (IntSD), are commonly observed. Early identification of these changes through voi…
pubmed
Karimi A, Moein N, D'Alessandro E, Bearss KA 等
2025 May 27
置信度 0.82
-
Three-dimensional human pose estimation (3D HPE) from monocular RGB cameras is a fundamental yet challenging task in computer vision, forming the basis of a wide range of applications such as action recognition, metaverse, self-driving, and healthcare. Recent …
pubmed
Guo Y, Gao T, Dong A, Jiang X 等
2025 Apr 10
置信度 0.82
-
Vision-and-Language Navigation (VLN), as a crucial research problem of Embodied AI, requires an embodied agent to navigate through complex 3D environments following natural language instructions. Recent research has highlighted the promising capacity of large …
pubmed
Lin B, Nie Y, Wei Z, Chen J 等
2025 Jul
置信度 0.82
-
Sign language (SL) is an effective mode of communication, which uses visual-physical methods like hand signals, expressions, and body actions to communicate between the difficulty of hearing and the deaf community, produce opinions, and carry significant conve…
pubmed
Alabduallah B, Al Dayil R, Alkharashi A, Alneil AA
2025 Mar 18
置信度 0.82
-
Human-centric perception tasks, e.g., pedestrian detection, skeleton-based action recognition, and pose estimation, have wide industrial applications, such as metaverse and sports analysis. There is a recent surge to develop human-centric foundation models tha…
pubmed
Wang Y, Wu Y, He W, Guo X 等
2025 Jul
置信度 0.82
-
Who, What and Where (3W)are the three core elements of storytelling, and accurately identifying the 3W semantics is critical to understanding the story in a video. This paper studies the 3W composite-semantics video Instance Search (INS) problem, which aims to…
pubmed
Guo J, Lu A, Wu Z, Wang Z 等
2025
置信度 0.82
-
Challenging behaviors in children with autism is a serious clinical condition, oftentimes leading to aggression or self-injurious actions. The Revised Family Observation Schedule 3 rd Edition (FOS-R-III) is an intensive and fine-grained scale used to observe a…
pubmed
Zhao Z, Chung E, Chung KM, Park CH
2025 Sep
置信度 0.82
-
In modern telehealth and healthcare information systems medical image analysis is essential to understand the context of the images and its complex structure from large, inconsistent-quality, and distributed datasets. Achieving desired results faces a few chal…
pubmed
Al-Hammuri K, Gebali F, Kanan A
2025 Aug
置信度 0.82
-
Traditional action recognition methods predominantly rely on a single modality, such as vision or motion, which presents significant limitations when dealing with fine-grained action recognition. These methods struggle particularly with video data containing c…
pubmed
Li Y, Yang X, Chen C
2024
置信度 0.82
-
Humans excel at applying learned behavior to unlearned situations. A crucial component of this generalization behavior is our ability to compose/decompose a whole into reusable parts, an attribute known as compositionality. One of the fundamental questions in …
pubmed
Vijayaraghavan P, Queißer JF, Flores SV, Tani J
2025 Jan 22
置信度 0.82
-
Vision-language navigation (VLN) is a challenging task that requires agents to capture the correlation between different modalities from redundant information according to instructions, and then make sequential decisions on visual scenes and text instructions …
pubmed
Zhou D, Deng J, Pang Z, Li W
2025 Apr
置信度 0.82
-
Odor source localization (OSL) technology allows autonomous agents like mobile robots to localize a target odor source in an unknown environment. This is achieved by an OSL navigation algorithm that processes an agent's sensor readings to calculate action comm…
pubmed
Hassan S, Wang L, Mahmud KR
2024 Dec 10
置信度 0.82
-
In the World Health Organization (WHO) African Region, many cases of serious and preventable diseases remain unmanaged because appropriate and good quality diagnostic support is not available at all levels within countries' health systems. Diagnostic and labor…
pubmed
Coulibaly SO, Amri M, Vuguziga C, Seydi AB 等
2024 Nov 28
置信度 0.82
-
In human interactions, gaze may be used to acquire information for goal-directed actions, to acquire information related to the interacting partner's actions, and in the context of multimodal communication. At present, there are no models of gaze behavior in t…
pubmed
Hessels RS, Li P, Balali S, Teunisse MK 等
2024 Nov
置信度 0.82
-
To align with the 2030 vision of the World Health Organization (WHO) to ensure 90% of girls receive the HPV vaccine before turning 15, Bangladesh has recently started the (HPV) vaccine campaign nationwide. Therefore, our study aimed to assess the level of its …
pubmed
Hawlader MDH, Eva FN, Khan MAS, Islam T 等
2024
置信度 0.82
-
Currently, the application of robotics technology in sports training and competitions is rapidly increasing. Traditional methods mainly rely on image or video data, neglecting the effective utilization of textual information. To address this issue, we propose:…
pubmed
Ma L, Tong Y
2024
置信度 0.82
-
To utilize artificial intelligence (AI) platforms to generate medical illustrations for refractive surgeries, aiding patients in visualizing and comprehending procedures like laser-assisted in situ keratomileusis (LASIK), photorefractive keratectomy (PRK), and…
pubmed
Petroff DJ, Nasir AA, Moin KA, Loveless BA 等
2024 Aug
置信度 0.82
-
crossref
Moo Kim, Chelsea Finn, Percy Liang
2025-09-08T21:47:42Z
置信度 0.70
-
The intelligent transformation of the power industry requires highly reliable robotic autonomous operations, yet general Vision-Language-Action (VLA) models often underperform in specialized power grid scenarios. Current mainstream methods—whether vision-langu…
europepmc
zhisong zhang、, guozheng peng, Peng Zhang, jin wang 等
2026
置信度 0.80
-
The convergence of vision, language, and action modeling has catalyzed a paradigm shift in robotic manipulation, enabling robots to interpret natural language commands and execute complex tasks through learned sensorimotor policies. This comprehensive review s…
europepmc
Md Selim Sarowar, Sungho Kim
2026
置信度 0.80
-
Vision-Language Action (VLA) models have enabled language-driven robotic manipulation by integrating language instructions, visual perception, and action generation. However, existing VLA approaches heavily rely on large-scale human demonstration datasets, whi…
europepmc
Tinghao Yi, Quantao Yang, Enhong Chen
2026
置信度 0.80
-
crossref
Jason Baldridge
2021-06-03T15:56:42Z
置信度 0.70
-
crossref
Mücahit Emre Kabaoğlu, Mustafa Hikmet Bilgehan Uçar
2026-06-03T19:38:22Z
置信度 0.70
-
crossref
Arshia Eslami, Mahsa Ardakani, Amin Roostaee, Hasti Zanganeh 等
2026-06-19T10:46:08Z
置信度 0.70
-
This work analyzes the generalization capabilities of modern Vision-Language- Action models for robotic agents. The work compares several current approaches and discusses their ability to connect visual perception, language instructions, and action generation.…
crossref
Sydor Bohdan
2026-05-31T11:00:03Z
置信度 0.70
-
crossref
Gilad Sharir, Tinne Tuytelaars
2014-07-28T20:51:07Z
置信度 0.70
-
crossref
Zihao Yuan, Fangfang Xie, Tingwei Ji
2024-11-26T18:45:20Z
置信度 0.70
-
Apprentissage par Imitation Indépendant de l’Incarnation et de l’Environnement pour les Robots : intégration de la Reconnaissance d’Actions Basée sur la Pose Humaine avec des Modèles de Langage et Vision L’apprentissage automatique, notamment l’apprentissage p…
crossref
Kevin Riou
2026-04-09T08:27:05Z
置信度 0.70
-
crossref
Grigorii Guz, Giuseppe Carenini, Mathias Lécuyer, Michiel van de Panne 等
2026-05-14T05:03:54Z
置信度 0.70
-
crossref
2020-12-22T06:24:09Z
置信度 0.70
-
crossref
2010-02-16T14:03:11Z
置信度 0.70
-
Vision-Language Models (VLMs) have demonstrated impressive performance across various multimodal tasks. However, deploying large teacher models in real-world applications is often infeasible due to their high computational cost. To address this, knowledge dist…
crossref
Jessica Liang, Jianbo Shi
2025-12-08T17:09:34Z
置信度 0.70
-
crossref
Yuji Sato, Yasunori Ishii, Takayoshi Yamashita
2025-09-26T17:35:13Z
置信度 0.70
-
Deploying dual-arm robots in human-centric environments demands not only dexterous task execution but also strict adherence to common sense safety constraints. While recent advancements in Vision-Language-Action (VLA) models enable complex policy reasoning fro…
crossref
Jiajun Gu, Weihao Cheng, Longsen Gao
2026-04-01T16:45:15Z
置信度 0.70
-
crossref
Thomas Schenk, Volker Franz, Nicola Bruno
2011-02-09T04:33:06Z
置信度 0.70
-
crossref
2019-10-30T17:03:00Z
置信度 0.70
-
crossref
Katrin Renz, Long Chen, Elahe Arani, Oleg Sinavski
2025-08-13T17:26:42Z
置信度 0.70
-
crossref
Zewei Zhou, Tianhui Cai, Seth Zhao, Yun Zhang 等
2026-08-06T14:44:29Z
置信度 0.70
-
Vision-Language-Action (VLA) models unify visual perception, natural-language understanding, and action generation within a single foundation model, allowing a robot to follow instructions such as “fold the towel” or “fly to the red building” directly from cam…
europepmc
Inkyu Sa, Chanoh Park, Ho Seok Ahn
2026
置信度 0.80
-
crossref
Bingxin Xu, Yuzhang Shang, Binghui Wang, Emilio Ferrara
2026-07-01T12:25:50Z
置信度 0.70
-
crossref
Michael A. Arbib, JinYong Lee
2007-09-20T06:44:59Z
置信度 0.70
-
crossref
Cuixin Yang, Rongkang Dong, Kin-Man Lam
2025-09-13T01:39:27Z
置信度 0.70
-
crossref
2020-12-22T06:24:09Z
置信度 0.70
-
Embodied navigation is a fundamental capability for intelligent agents, yet remains challenging in partially observable environments where navigation instructions can be difficult to interpret. However, existing tasks only provide unimodal instructions, which …
crossref
Jugang Fan, Peihao Chen, Changhao Li, Qing Du 等
2026-03-18T01:04:19Z
置信度 0.70
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
Recent Vision–Language–Action (VLA) models have rapidly emerged as general-purpose robotic policies that integrate language understanding, visual perception, and robot control. However, prior studies and surveys have primarily emphasized backbone architectures…
europepmc
Byoung Chul Ko
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
Artificial intelligence (AI) is rapidly reshaping orthopaedic surgery, supported by advances in data science, computational power, and perioperative digitalization. Within this evolving landscape, five "AI companions" structure the surgeon's workflow. The "AI …
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
Foundation models for embodied artificial intelligence (Embodied AI) increasingly adopt diffusion modules as the action generation core of vision-language-action (VLA) policies, but the diffusion module's iterative denoising imposes prohibitive inference laten…
pubmed
Shi X, Hu Y, Jin J
2026
置信度 0.82
-
The ability to autonomously navigate and explore complex 3D environments in a purposeful manner, while integrating visual perception with natural language interaction in a human-like way, represents a longstanding research objective in Artificial Intelligence …
pubmed
Li Z, Meng X, He X, Zhang Y 等
2026
置信度 0.82
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for embodied intelligence, enabling robots to perform complex actions over multimodal observations with one end-to-end policy. With the increasing deployment of VLA models in rea…
europepmc
Wei Yuan, Fengwen Liu, Ruize Wei, Zongwei Wang 等
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
The operating room remains a paradox: it is one of the most sensor-rich environments in the hospital, yet it produces largely underutilized data. While surgical artificial intelligence (AI) has achieved remarkable progress in recent years, the day-to-day pract…
pubmed
Oh N, Jung KH, Choi GS
2026
置信度 0.82
-
Anomaly detection is crucial in maintaining the safety, reliability, and optimal performance of complex systems across diverse domains, such as industrial manufacturing, cybersecurity, and autonomous systems. While conventional methods typically handle single …
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
Zero-shot goal navigation requires an agent to locate targets in unseen environments based on object categories, reference images, or text descriptions, placing high demands on scene understanding and reasoning. Existing methods mainly rely on online observati…
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
Vision–language–action (VLA) models often suffer from limited robustness in long-horizon manipulation tasks due to their inability to explicitly exploit structural symmetries and to react adaptively when such symmetries are violated by environmental uncertaint…
europepmc
Yina Jian, Tian Di, Zhen-Yuan Wei, Chen-Wei Liang 等
2026
置信度 0.80
-
The rapid advancement of artificial intelligence (AI) and the continuous reform of English education in China have greatly reshaped the professional requirements for college English teachers. In contrast to the predominance of general teacher perspectives in a…
pubmed
Zhou Q, Hu L, Zou X, Zhang X
2026
置信度 0.82
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
europepmc
2026
置信度 0.80
-
Accurately monitoring the screen exposure of young children is important for research related to screen use, such as childhood obesity, physical activity, and social interaction. Most existing studies rely upon self-report or manual measures from bulky wearabl…
europepmc
2026
置信度 0.80