-
Artificial intelligence (AI) is increasingly integrated into public health practice, offering new tools for surveillance, behavioral interventions, disease prevention, risk communication, and population health management. Although AI demonstrates promise for e…
europepmc
2026
置信度 0.80
-
Abstract Introduction : As artificial intelligence (AI) tools gain popularity in maternal healthcare, understanding community and health system perspectives is essential for inclusive and ethical implementation. This study explores stakeholder views on AI-supp…
europepmc
2026
置信度 0.80
-
Integrating medical data across hospitals has become a critical challenge in medical informatics, largely due to the heterogeneity of electronic medical record (EMR) systems. This study aims to address this issue by developing a diagnosis classification model …
europepmc
2026
置信度 0.80
-
Generative AI can help scale climate services to meet growing demand. This approach is illustrated here with two prototype AI systems designed to enhance access to climate data, support decision-making, and improve efficiency. Key concerns are identified aroun…
europepmc
2026
置信度 0.80
-
Optical experiments are essential across science and technology, yet their design, assembly, and alignment remain predominantly manual, limiting throughput, reproducibility, and scalability. Automating such experiments is challenging because of stringent preci…
europepmc
2026
置信度 0.80
-
Introduction Objective Structured Clinical Examinations (OSCEs) are widely used to assess clinical competence, but face challenges related to examiner workload, scoring variability, delayed feedback, and resource demands. Although AI may address these constrai…
europepmc
2026
置信度 0.80
-
ABSTRACT Individuals increasingly use conversational AI systems for symptom guidance. Whether multi-turn interactions improve clinical triage standard alignment remains uncertain. We conducted a retrospective, cross-sectional evaluation of 255 cases from three…
europepmc
2026
置信度 0.80
-
Large language models (LLMs), initially developed for generative AI, are now evolving into agentic AI systems, which make decisions in complex, real-world contexts. Unfortunately, while their generative capabilities are well-documented, their decision-making p…
europepmc
2026
置信度 0.80
-
Social interactions are foundational to learning, yet scalable one-on-one instructor-student interaction remains challenging in online video learning. We examine whether a brief, structured pre-lecture instructor-student interaction, led by either a human or l…
europepmc
2026
置信度 0.80
-
Artificial Intelligence (AI) has emerged as a transformative force in medical education, enabling customized, feedback-rich, and immersive learning experiences. Anatomy education, as a discipline that is simultaneously visual, spatial, and clinically foundatio…
europepmc
2026
置信度 0.80
-
Background Millions of people now use leading generative artificial intelligence (AI) tools (chatbots) for psychological support. Despite the promise related to availability and scale, the single most pressing question in AI for mental health is whether these …
europepmc
2026
置信度 0.80
-
Lower back pain (LBP) is a leading cause of disability globally, characterized by multifactorial origins that complicate accurate diagnosis and effective treatment planning. Artificial intelligence (AI), including machine learning (ML), deep learning (DL), and…
europepmc
2026
置信度 0.80
-
Scientific and systematic data collection and analysis have long been a crucial foundation in psychological assessment systems. It is only through this process that psychology professionals can effectively measure and interpret individuals' mental states, beha…
europepmc
2026
置信度 0.80
-
Abstract The integration of generative AI into educational assessment has transformed feedback practices, yet the extent to which AI-generated feedback supports compassionate teaching—integrating emotional attunement with precise cognitive scaffolding— remains…
europepmc
2026
置信度 0.80
-
Nuclear medicine (NM) is rapidly expanding with new radiopharmaceuticals, imaging equipment, and theranostic possibilities, necessitating a capable and competent workforce expansion. Consequently, NM education must undergo a transformation, powered in part by …
europepmc
2026
置信度 0.80
-
Artificial intelligence (AI) is rapidly reshaping healthcare through advances in prediction and personalization, yet these developments also raise questions that are not merely technical but also are moral and religious. This article offers a normative framewo…
europepmc
2026
置信度 0.80
-
Introduction This descriptive study aimed to longitudinally evaluate the performance of contemporary large language models - ChatGPT-5, Gemini 2.5 Flash, and Grok-3 - on orthopaedic clinical multiple-choice tasks, benchmarked against pooled clinician consensus…
europepmc
2026
置信度 0.80
-
Abstract Background Stillbirth remains a major global health problem, with most occurring in low- and middle-income countries (LMICs). Novel antenatal technologies, including Doppler ultrasound, maternal haemodynamic monitoring and artificial intelligence (AI)…
europepmc
2026
置信度 0.80
-
Background Appendicitis severity underpins contemporary management guidelines, where laparoscopic appendectomy remains gold standard. Preoperative measures poorly predict actual disease severity or complication risk, while operative grading systems such as the…
europepmc
2026
置信度 0.80
-
Introduction Artificial intelligence (AI) is being rapidly integrated into systematic review workflows, yet its impact on methodological rigor, transparency, and reporting quality remains poorly understood. This work examines the current use of AI assistance i…
europepmc
2026
置信度 0.80
-
The 2025 AMCP Foundation Symposium: Health Optimization 2.0 convened a select group of just over 70 managed care stakeholders in Dallas, Texas, on December 2-3, 2025. Via presentations and panel sessions, participants examined how precision medicine, data anal…
europepmc
2026
置信度 0.80
-
Background Artificial intelligence (AI) is increasingly used in cardiovascular care to support diagnosis, monitoring and clinical decision-making. However, its dynamic and adaptive nature challenges conventional health technology assessment (HTA) frameworks, w…
europepmc
2026
置信度 0.80
-
AI-assisted legal concept alignment, defined as identifying legal concepts that best correspond to a given fact description based on computational relevance, addresses the growing need for complex legal analysis that requires both semantic understanding and in…
europepmc
2026
置信度 0.80
-
Abstract Background: Access to obstetric ultrasound remains limited in many low- and middle-income countries, contributing to delays in diagnosis and care. Although AI-enabled portable ultrasound has the potential to expand access, successful implementation de…
europepmc
2026
置信度 0.80
-
Objective To compare answers to clinical questions between five publicly available large language model (LLM) chatbots and information scientists. Methods LLMs were prompted to provide 45 PICO (patient, intervention, comparison, outcome) questions addressing t…
europepmc
2026
置信度 0.80
-
Abstract Recent advances in agentic AI have accelerated a shift from monolithic language models toward modular, multi-agent systems enriched with memory, orchestration, and affect-aware interaction. Large Language Models now serve as reasoning and communicatio…
europepmc
2026
置信度 0.80
-
Abstract Artificial intelligence (AI) maturity is increasingly viewed as a source of advantage for technology startups, yet its value may depend on the planning assumptions guiding its use. Drawing on socio-technical systems theory, the attention-based view of…
europepmc
2026
置信度 0.80
-
Objectives This study aimed to compare the agreement and consistency of classifications generated by ChatGPT-4.5 and Meta AI (Llama 4 Scout) with those of the UpToDate Lexidrug reference standard for identifying potential drug interactions involving antihypert…
europepmc
2026
置信度 0.80
-
Abstract Background Conversational artificial intelligence (AI) is increasingly used for psychological support in everyday help-seeking. In real-world settings, many individuals use both AI tools and human counseling; however, evidence quantifying within-perso…
europepmc
2026
置信度 0.80
-
Abstract Artificial intelligence increasingly mediates how young people enter the workforce, with recommender systems guiding students’ educational and career decisions at the start of their working lives. Such systems are evaluated largely on predictive accur…
europepmc
2026
置信度 0.80
-
Background The process of aligning sequencing reads to a reference genome is a foundational step in genomic analysis, underpinning tasks from variant detection to pathogen surveillance. In viral genomics, however, this problem becomes substantially more challe…
europepmc
2026
置信度 0.80
-
Background: Artificial intelligence (AI) is increasingly used in nephrology, including rehabilitation planning for patients with chronic kidney disease (CKD). However, most AI systems are predominantly trained on English-language data, which may influence the …
europepmc
2026
置信度 0.80
-
Existing research on AI cultural biases predominantly focuses on Western models, overlooking critical gaps in non-Western models. We conduct a comparative analysis of AI models-ChatGPT (US developed) and ErnieBot (China developed)-from different cultures to in…
europepmc
2026
置信度 0.80
-
Background and objective Artificial intelligence has improved adenoma detection during colonoscopy, but most models remain black boxes, limiting clinical trust, especially for sessile serrated lesions (SSLs). Concept Bottleneck Models (CBMs) route predictions …
europepmc
2026
置信度 0.80
-
This study examines the global impacts of Artificial General Intelligence (AGI) and Artificial Superintelligence (ASI) through a conceptual review of the related literature. AGI refers to AI systems that may match human cognitive ability across multiple domain…
europepmc
2026
置信度 0.80
-
Generative AI has entered nursing education rapidly and pervasively. While considerable scholarly attention has been directed toward student use of these tools, including concerns about academic integrity, citation fabrication, and the design of AI-resistant a…
europepmc
2026
置信度 0.80
-
Background/Objectives: Artificial intelligence (AI) is increasingly introduced into clinical trial operations, but operational usefulness does not imply regulatory or site-level readiness. This review maps AI applications across trial operations and proposes a…
europepmc
2026
置信度 0.80
-
Objectives This study aimed to evaluate and compare the accuracy of multiple artificial intelligence (AI) models (ChatGPT 5.2 Pro, Gemini 3 Fast, Claude 4.5 Sonnet, and Microsoft Copilot) in detecting orthodontic malocclusion features in standardized multiview…
europepmc
2026
置信度 0.80
-
The growing integration of artificial intelligence tools, particularly conversational agents, is transforming how consumers access information, including content related to critical areas such as public health and personal healthcare. A portion of the American…
europepmc
2026
置信度 0.80
-
Background Artificial intelligence (AI) is transforming higher education and healthcare, yet nursing faculty lack practical guidance for determining appropriate AI use in specific academic tasks. Institutional AI policies establish boundaries but rarely addres…
europepmc
2026
置信度 0.80
-
Generative artificial intelligence (AI) can convert clinical information into patient-facing instructions, including discharge summaries, medication explanations, portal messages and plain-language educational materials. These outputs may improve accessibility…
europepmc
2026
置信度 0.80
-
Ensuring fairness and reliabilityin examination scoring remains a persistent challenge in educational assessment, particularly in the presence of subjective grading inconsistencies across evaluators. While automated essay scoring systems improve scalability, l…
europepmc
2026
置信度 0.80
-
Background Effective glaucoma management relies on accurate optical coherence tomography (OCT) quantification of the circumpapillary retinal nerve fiber layer (cpRNFL) to track disease progression. However, the clinical translation of artificial intelligence (…
europepmc
2026
置信度 0.80
-
In any field, unquestioningly accepting artificial intelligence (AI) results should be considered bad practise. Here, we devised a comparative modelling-based strategy for validating protein structures that exploits the well-known observation that protein fold…
europepmc
2026
置信度 0.80
-
Agentic AI systems plan over time, call tools, write and retrieve memory, and coordinate across modules and services. Some of the most consequential failures in such systems are structural before they are behavioral: hidden coordination, trace-mediated lock-in…
europepmc
2026
置信度 0.80
-
Conversational AI is being deployed into medical decision support, mental-health triage, and social companionship, where reinforcement of a user’s false or delusional belief can cause direct harm. Most deployed safety techniques are evaluated for factual accur…
europepmc
2026
置信度 0.80
-
Background Trustworthy artificial intelligence (AI) in health care requires assurance frameworks that translate ethical principles into measurable governance and evaluation practices. While a growing number of AI assurance frameworks have been proposed, they d…
europepmc
2026
置信度 0.80
-
Background AI is transforming health care, creating an imperative to integrate AI into medical education. While student perspectives are well-studied, faculty views, particularly in non-Western contexts, remain underexplored. Objective This qualitative study e…
europepmc
2026
置信度 0.80
-
Extreme heat, driven by climate change, poses significant health risks, particularly for vulnerable populations such as older adults, infants, and individuals with chronic illnesses. Generative AI may help tailor and scale climate adaptation messages, includin…
europepmc
2026
置信度 0.80
-
Abstract Purpose This study evaluated the relationship between Fixed flexion deformity (FFD), hip range of motion (ROM), and dynamic spinopelvic parameters and identified predictors of early postoperative sacral slope (SS) mobility in patients with advanced hi…
europepmc
2026
置信度 0.80
-
Cross-domain analysis of spatial transcriptomics is challenging because tissues from different organs, diseases and experimental platforms exhibit distinct cellular compositions, spatial organisations and technical biases, making direct comparison of tissue st…
europepmc
2026
置信度 0.80
-
Background Complex hip-spine syndrome combined with hip ankylosis is clinically challenging to manage. Traditional surgeries are associated with insufficient precision and significant trauma, which easily lead to poor lumbar-pelvic-hip alignment and affect pro…
europepmc
2026
置信度 0.80
-
Marketing in aesthetic plastic surgery has shifted from traditional print and television toward digital-first strategies, with social media and search engine optimization (SEO) now shaping how patients engage with practices. More than 90% of cosmetic patients …
europepmc
2026
置信度 0.80
-
Abstract Responsible reporting on suicide is an important public health strategy but monitoring whether news articles follow reporting recommendations is difficult at scale. While the Tool for Evaluating Media Portrayals of Suicide (TEMPOS) provides a structur…
europepmc
2026
置信度 0.80
-
Abstract Generative artificial intelligence is increasingly being explored as a pedagogical support tool in teacher education, yet limited empirical evidence exists on how preservice teachers interact with customized AI systems to create culturally responsive …
europepmc
2026
置信度 0.80
-
As artificial intelligence (AI) continues to advance, understanding public perceptions-including biases, risks, and benefits-is essential for guiding research priorities and AI alignment, shaping public discourse, and informing policy. This exploratory study i…
europepmc
2026
置信度 0.80
-
People increasingly turn to conversational AI for companionship, emotional support, and well-being, using both purpose-built companion apps such as Replika and Character.ai and general-purpose assistants such as ChatGPT and Claude. While some evidence suggests…
europepmc
2026
置信度 0.80
-
Implant dentistry has undergone a significant transformation with the integration of artificial intelligence (AI) into digital workflows. AI, a branch of computer science that enables machines to simulate human cognitive processes, has demonstrated the potenti…
europepmc
2026
置信度 0.80
-
Artificial intelligence (AI) is playing an increasingly central role in drug discovery and the pharmaceutical industry more broadly. However, despite some high-profile success stories, many technically successful pilots do not translate into sustained business…
europepmc
2026
置信度 0.80
-
Modern agentic systems that plan, reason, invoke tools, and self-reflect offer a promising path to autonomous cyber defense and SOC analyst augmentation. Yet this same autonomy introduces novel failure modes and attack surfaces absent from traditional security…
europepmc
2026
置信度 0.80
-
Generative artificial intelligence (AI), especially large language models, is playing an expanding role in shaping learning within health sciences education. Current discussions often focus on efficiency or academic integrity, with less attention to how learne…
europepmc
2026
置信度 0.80
-
Generative artificial intelligence (GenAI) has created new possibilities for teaching and learning, yet its pedagogical role in physical education and health (PEH) remains underdeveloped. Current AI-related research in PEH has largely focused on performance-or…
europepmc
2026
置信度 0.80
-
Statement of problem Artificial intelligence (AI)-based applications have increasingly been developed and integrated in different digital data acquisition technologies, including intraoral scanners (IOSs). A comprehensive evaluation of their accuracy and curre…
europepmc
2026
置信度 0.80
-
Background Charles Friedman's Fundamental Theorem of Biomedical Informatics holds that a person working in partnership with an information resource outperforms that same person unassisted. Since its publication, advances in artificial intelligence (AI), adapti…
europepmc
2026
置信度 0.80
-
Background Effective patient education requires accurate communication aligned with patients' emotional and semantical needs. Text-based large language models (LLMs) lack access to non-verbal cues, which may contribute to misaligned responses. Methods We evalu…
europepmc
2026
置信度 0.80
-
Background Correction of adult spinal deformity (ASD) using the Roussouly classification, SRS-Schwab classification, or the global alignment and proportion (GAP) score is associated with reduced mechanical complications. However, the benefit of combining these…
europepmc
2026
置信度 0.80
-
Conversational AI systems combine AI-based solutions with the flexibility of conversational interfaces. However, most existing testing solutions do not straightforwardly adapt to the characteristics of conversational interaction or to the behavior of AI compon…
arxiv
Elena Masserini
2026-02-03T09:38:59Z
置信度 0.78
cs.SE
-
In explainable AI, Concept Activation Vectors (CAVs) are typically obtained by training linear classifier probes to detect human-understandable concepts as directions in the activation space of deep neural networks. It is widely assumed that a high probe accur…
arxiv
Jacob Lysnæs-Larsen, Marte Eggen, Inga Strümke
2025-11-06T12:34:05Z
置信度 0.78
cs.AI
-
How much does a user's skill with AI shape what AI actually delivers for them? This question is critical for users, AI product builders, and society at large, but it remains underexplored. Using a richly annotated sample of 27K transcripts from WildChat-4.8M, …
arxiv
Christopher Potts, Moritz Sudhof
2026-04-28T17:51:13Z
置信度 0.78
cs.CL
-
The rapid advancement of artificial intelligence (AI) has brought about sophisticated models capable of various tasks ranging from image recognition to natural language processing. As these models continue to grow in complexity, ensuring their trustworthiness …
arxiv
Nishant Jagannath, Christopher Wong, Braden Mcgrath, Md Farhad Hossain 等
2025-04-07T07:38:29Z
置信度 0.78
cs.CRcs.DC
-
Increasing evidence suggests that many deployed AI systems do not sufficiently support end-user interaction and information needs. Engaging end-users in the design of these systems can reveal user needs and expectations, yet effective ways of engaging end-user…
arxiv
Christine P Lee, Min Kyung Lee, Bilge Mutlu
2024-05-26T22:18:38Z
置信度 0.78
cs.HCcs.AI
-
There is no consensus on what constitutes human-centeredness in AI, and existing frameworks lack empirical validation. This study addresses this gap by developing a hierarchical framework of 26 attributes of human-centeredness, validated through practitioner i…
arxiv
Aung Pyae
2025-02-05T15:51:03Z
置信度 0.78
cs.HC
-
Higher education remains largely reactive in its approach to student success. Institutions frequently identify academic problems only after students have failed courses, fallen behind in degree progression, accumulated excessive debt, or departed without a cre…
arxiv
Kaushik Dutta
2026-08-06T17:36:56Z
置信度 0.78
cs.CY
-
Agentic AI systems capable of autonomous goal setting and proactive intervention introduce new challenges for regulating moral-emotional processes in learning environments. Existing frameworks typically treat emotion as reactive feedback or engagement optimiza…
arxiv
Ji Yeon Kim
2026-04-07T03:38:27Z
置信度 0.78
cs.CYcs.AI
-
As generative AI becomes increasingly integrated into higher education, its frequent errors and hallucinations, often seen as limitations, offer a unique pedagogical opportunity. By framing AI as a ``learning companion'' whose imperfect outputs prompt analysis…
arxiv
Hadi Hosseini
2026-05-06T21:50:46Z
置信度 0.78
cs.CYcs.AI
-
Efficient orchestration of AI services in 6G AI-RAN requires well-structured, ready-to-deploy AI service repositories combined with orchestration methods adaptive to diverse runtime contexts across radio access, edge, and cloud layers. Current literature lacks…
arxiv
Yun Tang, Mengbang Zou, Udhaya Chandhar Srinivasan, Obumneme Umealor 等
2025-04-13T16:40:58Z
置信度 0.78
cs.AIcs.SE
-
Deep learning has become widely used in complex AI applications. Yet, training a deep neural network (DNNs) model requires a considerable amount of calculations, long running time, and much energy. Nowadays, many-core AI accelerators (e.g., GPUs and TPUs) are …
arxiv
Yuxin Wang, Qiang Wang, Shaohuai Shi, Xin He 等
2019-09-15T17:30:05Z
置信度 0.78
cs.DCcs.LG
-
The ideal AI safety moderation system would be both structurally interpretable (so its decisions can be reliably explained) and steerable (to align to safety standards and reflect a community's values), which current systems fall short on. To address this gap,…
arxiv
Jing-Jing Li, Valentina Pyatkin, Max Kleiman-Weiner, Liwei Jiang 等
2024-10-22T03:38:37Z
置信度 0.78
cs.CLcs.CY
-
Determining whether AI systems process information similarly to humans is central to cognitive science and trustworthy AI. While modern AI models can match human accuracy on standard tasks, such parity does not guarantee that their underlying decision-making s…
arxiv
Binxia Xu, Xiaoliang Luo, Luke Dickens, Robert M. Mok
2026-03-08T04:51:39Z
置信度 0.78
cs.AI
-
Aligning Large Language Models (LLMs) with human values and preferences is essential for making them helpful and safe. However, building efficient tools to perform alignment can be challenging, especially for the largest and most competent LLMs which often con…
arxiv
Gerald Shen, Zhilin Wang, Olivier Delalleau, Jiaqi Zeng 等
2024-05-02T17:13:40Z
置信度 0.78
cs.CLcs.AIcs.LG
-
Generative AI is likely to transform terminology work by creating new opportunities for automation. At the same time, it raises concerns about the future of terminologists and terminological resources, as efficiency pressures may encourage excessive automation…
arxiv
Antonio San Martin
2025-12-21T19:16:40Z
置信度 0.78
cs.CL
-
According to the recent European legislation, high-risk AI systems will have to adapt in order to comply with requirements related to specific areas, like risk management, data quality and governance, logging and traceability, technical documentation, transpar…
arxiv
Anna Gatzioura, Vrettos Moulos, Nina Baranowska
2026-07-14T10:00:27Z
置信度 0.78
cs.AI
-
One strategy in response to pluralistic values in a user population is to personalize an AI system: if the AI can adapt to the specific values of each individual, then we can potentially avoid many of the challenges of pluralism. Unfortunately, this approach c…
arxiv
Nandhini Swaminathan, David Danks
2024-04-30T04:41:47Z
置信度 0.78
cs.AIcs.CYcs.GTcs.HCcs.LG
-
This study presents a novel, AI-driven framework for assessing Environmental, Social, and Governance (ESG) performance in European small and medium-sized enterprises (SMEs). An initial phase established expert-validated ESG baseline scores from a subset of the…
arxiv
Viet Trinh, Tan Nguyen, Minh-Huyen Phan, Quan Luu
2026-04-05T20:44:14Z
置信度 0.78
cs.AIecon.GN
-
The growing presence of Artificial Intelligence (AI) in various sectors necessitates systems that accurately reflect societal diversity. This study seeks to envision the operationalization of the ethical imperatives of diversity and inclusion (D&I) within …
arxiv
Muneera Bano, Didar Zowghi, Vincenzo Gervasi
2023-12-11T02:44:39Z
置信度 0.78
cs.AIcs.SE
-
Large Language Models (LLMs) are increasingly used in decision-making scenarios that involve risk assessment, yet their alignment with human economic rationality remains unclear. In this study, we investigate whether LLMs exhibit risk preferences consistent wi…
arxiv
Jiaxin Liu, Yixuan Tang, Yi Yang, Kar Yan Tam
2025-03-09T14:47:31Z
置信度 0.78
econ.GNcs.CL
-
AI-empowered tools have emerged as a transformative force, fundamentally reshaping the software development industry and promising far-reaching impacts across diverse sectors. This study investigates the adoption, impact, and security considerations of AI-empo…
arxiv
Shidong Pan, Litian Wang, Tianyi Zhang, Zhenchang Xing 等
2024-09-20T09:17:10Z
置信度 0.78
cs.SEcs.CR
-
I study Full Justified Representation (FJR) in approval-based multiwinner elections under both the Hare and Droop quota conventions. I introduce a descending-budget algorithm in which voters distribute their remaining budgets across their current representatio…
arxiv
Yizhou Ai
2026-08-05T21:20:42Z
置信度 0.78
cs.GT
-
We propose a novel framework for combining datasets via alignment of their intrinsic geometry. This alignment can be used to fuse data originating from disparate modalities, or to correct batch effects while preserving intrinsic data structure. Importantly, we…
arxiv
Jay S. Stanley, Scott Gigante, Guy Wolf, Smita Krishnaswamy
2018-09-30T14:23:10Z
置信度 0.78
cs.LGstat.ML
-
We study how humans learn from AI, leveraging an introduction of an AI-powered Go program (APG) that unexpectedly outperformed the best professional player. We compare the move quality of professional players to APG's superior solutions around its public relea…
arxiv
Sukwoong Choi, Hyo Kang, Namil Kim, Junsik Kim
2023-10-12T20:28:05Z
置信度 0.78
econ.GN
-
Next generation of embedded Information and Communication Technology (ICT) systems are collaborative systems able to perform autonomous tasks. The remarkable expansion of the embedded ICT market, together with the rise and breakthroughs of Artificial Intellige…
arxiv
Miguel de Prado, Jing Su, Rabia Saeed, Lorenzo Keller 等
2019-01-15T21:27:28Z
置信度 0.78
cs.LGcs.DCcs.SDeess.ASstat.ML
-
Evaluating the value alignment of large language models (LLMs) has traditionally relied on single-sentence adversarial prompts, which directly probe models with ethically sensitive or controversial questions. However, with the rapid advancements in AI safety t…
arxiv
Yazhou Zhang, Qimeng Liu, Qiuchi Li, Peng Zhang 等
2025-03-28T03:31:37Z
置信度 0.78
cs.CLcs.AIcs.CY
-
The more AI agents are deployed in scenarios with possibly unexpected situations, the more they need to be flexible, adaptive, and creative in achieving the goal we have given them. Thus, a certain level of freedom to choose the best path to the goal is inhere…
arxiv
Francesca Rossi, Nicholas Mattei
2018-12-10T18:58:05Z
置信度 0.78
cs.AIcs.CYcs.LG
-
Automating AI research holds immense potential for accelerating scientific progress, yet current AI agents struggle with the complexities of rigorous, end-to-end experimentation. We introduce EXP-Bench, a novel benchmark designed to systematically evaluate AI …
arxiv
Patrick Tser Jern Kon, Jiachen Liu, Xinyi Zhu, Qiuyi Ding 等
2025-05-30T16:46:29Z
置信度 0.78
cs.AI
-
In AI research and practice, rigor remains largely understood in terms of methodological rigor -- such as whether mathematical, statistical, or computational methods are correctly applied. We argue that this narrow conception of rigor has contributed to the co…
arxiv
Alexandra Olteanu, Su Lin Blodgett, Agathe Balayn, Angelina Wang 等
2025-06-17T15:44:41Z
置信度 0.78
cs.CYcs.AIcs.LG
-
As AI attracts vast investment and attention, there are competing concerns about the technology's opportunities and uncertainties that blend technical and social questions. The public debate, dominated by a few powerful voices, tends to highlight extreme promi…
arxiv
Cian O'Donovan, Sarp Gurakan, Ananya Karanam, Xiaomeng Wu 等
2026-03-06T12:34:33Z
置信度 0.78
cs.CY
-
AI research agents now support large-scale AI-assisted scientific discovery. We examine whether AI-generated ideas broaden scientific exploration or primarily reinforce existing work. Using five agent frameworks and five large language models, we generate 219,…
arxiv
Yixuan Tang, Yi Yang
2026-05-27T03:26:43Z
置信度 0.78
cs.CL
-
The rapid adoption of Large Language Models (LLMs) has spurred interest in automated peer review; however, progress is currently stifled by benchmarks that treat reviewing primarily as a rating prediction task. We argue that the utility of a review lies in its…
arxiv
Bowen Li, Haochen Ma, Yuxin Wang, Jie Yang 等
2026-04-21T14:21:15Z
置信度 0.78
cs.CL
-
Public policies are being developed around the world to address privacy, economic, intellectual property, energy, and other risks that AI technologies pose. Involvement from the general public is essential to governance as an accountability and alignment mecha…
arxiv
Carter Buckner, Jennifer Mickel, Nandhini Swaminathan, William Agnew 等
2026-06-18T05:43:49Z
置信度 0.78
cs.CY
-
A framework is proposed that seeks to identify and establish a set of robust autonomous levels articulating the realm of Artificial Intelligence and Legal Reasoning (AILR). Doing so provides a sound and parsimonious basis for being able to assess progress in t…
arxiv
Lance Eliot
2020-08-04T16:12:30Z
置信度 0.78
cs.CYcs.AI