Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
Formal Analysis of AGI Decision-Theoretic Models and the Confrontation Question
arXiv:2601.04234v1 Announce Type: new Abstract: Artificial General Intelligence (AGI) may face a confrontation question: under what conditions would a rationally self-interested AGI choose to seize power or eliminate human control (a confrontation) rather than remain cooperative? We formalize this in a Markov decision process with a stochastic human-initiated shutdown event. Building on results on convergent instrumental incentives, we show that for almost all reward functions a misaligned agen
-
cs.AI, q-bio.NC updates on arXiv.org
-
Systems Explaining Systems: A Framework for Intelligence and Consciousness
arXiv:2601.04269v1 Announce Type: new Abstract: This paper proposes a conceptual framework in which intelligence and consciousness emerge from relational structure rather than from prediction or domain-specific mechanisms. Intelligence is defined as the capacity to form and integrate causal connections between signals, actions, and internal states. Through context enrichment, systems interpret incoming information using learned relational structure that provides essential context in an efficien
Systems Explaining Systems: A Framework for Intelligence and Consciousness
-
cs.AI, q-bio.NC updates on arXiv.org
-
An ASP-based Solution to the Medical Appointment Scheduling Problem
arXiv:2601.04274v1 Announce Type: new Abstract: This paper presents an Answer Set Programming (ASP)-based framework for medical appointment scheduling, aimed at improving efficiency, reducing administrative overhead, and enhancing patient-centered care. The framework personalizes scheduling for vulnerable populations by integrating Blueprint Personas. It ensures real-time availability updates, conflict-free assignments, and seamless interoperability with existing healthcare platforms by central
An ASP-based Solution to the Medical Appointment Scheduling Problem
-
cs.AI, q-bio.NC updates on arXiv.org
-
Sci-Reasoning: A Dataset Decoding AI Innovation Patterns
arXiv:2601.04577v1 Announce Type: new Abstract: While AI innovation accelerates rapidly, the intellectual process behind breakthroughs -- how researchers identify gaps, synthesize prior work, and generate insights -- remains poorly understood. The lack of structured data on scientific reasoning hinders systematic analysis and development of AI research agents. We introduce Sci-Reasoning, the first dataset capturing the intellectual synthesis behind high-quality AI research. Using community-vali
Sci-Reasoning: A Dataset Decoding AI Innovation Patterns
-
cs.AI, q-bio.NC updates on arXiv.org
-
Autonomous Agents on Blockchains: Standards, Execution Models, and Trust Boundaries
arXiv:2601.04583v1 Announce Type: new Abstract: Advances in large language models have enabled agentic AI systems that can reason, plan, and interact with external tools to execute multi-step workflows, while public blockchains have evolved into a programmable substrate for value transfer, access control, and verifiable state transitions. Their convergence introduces a high-stakes systems challenge: designing standard, interoperable, and secure interfaces that allow agents to observe on-chain s
Autonomous Agents on Blockchains: Standards, Execution Models, and Trust Boundaries
-
cs.AI, q-bio.NC updates on arXiv.org
-
ResMAS: Resilience Optimization in LLM-based Multi-agent Systems
arXiv:2601.04694v1 Announce Type: new Abstract: Large Language Model-based Multi-Agent Systems (LLM-based MAS), where multiple LLM agents collaborate to solve complex tasks, have shown impressive performance in many areas. However, MAS are typically distributed across different devices or environments, making them vulnerable to perturbations such as agent failures. While existing works have studied the adversarial attacks and corresponding defense strategies, they mainly focus on reactively det
ResMAS: Resilience Optimization in LLM-based Multi-agent Systems
-
cs.AI, q-bio.NC updates on arXiv.org
-
Beyond Monolithic Architectures: A Multi-Agent Search and Knowledge Optimization Framework for Agentic Search
arXiv:2601.04703v1 Announce Type: new Abstract: Agentic search has emerged as a promising paradigm for complex information seeking by enabling Large Language Models (LLMs) to interleave reasoning with tool use. However, prevailing systems rely on monolithic agents that suffer from structural bottlenecks, including unconstrained reasoning outputs that inflate trajectories, sparse outcome-level rewards that complicate credit assignment, and stochastic search noise that destabilizes learning. To a
Beyond Monolithic Architectures: A Multi-Agent Search and Knowledge Optimization Framework for Agentic Search
-
cs.AI, q-bio.NC updates on arXiv.org
-
Balancing Usability and Compliance in AI Smart Devices: A Privacy-by-Design Audit of Google Home, Alexa, and Siri
arXiv:2601.04403v1 Announce Type: cross Abstract: This paper investigates the privacy and usability of AI-enabled smart devices commonly used by youth, focusing on Google Home Mini, Amazon Alexa, and Apple Siri. While these devices provide convenience and efficiency, they also raise privacy and transparency concerns due to their always-listening design and complex data management processes. The study proposes and applies a combined framework of Heuristic Evaluation, Personal Information Protect
Balancing Usability and Compliance in AI Smart Devices: A Privacy-by-Design Audit of Google Home, Alexa, and Siri
-
cs.AI, q-bio.NC updates on arXiv.org
-
Surface-based Molecular Design with Multi-modal Flow Matching
arXiv:2601.04506v1 Announce Type: cross Abstract: Therapeutic peptides show promise in targeting previously undruggable binding sites, with recent advancements in deep generative models enabling full-atom peptide co-design for specific protein receptors. However, the critical role of molecular surfaces in protein-protein interactions (PPIs) has been underexplored. To bridge this gap, we propose an omni-design peptides generation paradigm, called SurfFlow, a novel surface-based generative algori
Surface-based Molecular Design with Multi-modal Flow Matching
-
cs.AI, q-bio.NC updates on arXiv.org
-
Self-MedRAG: a Self-Reflective Hybrid Retrieval-Augmented Generation Framework for Reliable Medical Question Answering
arXiv:2601.04531v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated significant potential in medical Question Answering (QA), yet they remain prone to hallucinations and ungrounded reasoning, limiting their reliability in high-stakes clinical scenarios. While Retrieval-Augmented Generation (RAG) mitigates these issues by incorporating external knowledge, conventional single-shot retrieval often fails to resolve complex biomedical queries requiring multi-step inferen
Self-MedRAG: a Self-Reflective Hybrid Retrieval-Augmented Generation Framework for Reliable Medical Question Answering
-
cs.AI, q-bio.NC updates on arXiv.org
-
A Vision for Multisensory Intelligence: Sensing, Synergy, and Science
arXiv:2601.04563v1 Announce Type: cross Abstract: Our experience of the world is multisensory, spanning a synthesis of language, sight, sound, touch, taste, and smell. Yet, artificial intelligence has primarily advanced in digital modalities like text, vision, and audio. This paper outlines a research vision for multisensory artificial intelligence over the next decade. This new set of technologies can change how humans and AI experience and interact with one another, by connecting AI to the hu
A Vision for Multisensory Intelligence: Sensing, Synergy, and Science
-
cs.AI, q-bio.NC updates on arXiv.org
-
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
arXiv:2601.04758v1 Announce Type: cross Abstract: The Patent Trial and Appeal Board (PTAB) of the USPTO adjudicates thousands of ex parte appeals each year, requiring the integration of technical understanding and legal reasoning. While large language models (LLMs) are increasingly applied in patent and legal practice, their use has remained limited to lightweight tasks, with no established means of systematically evaluating their capacity for structured legal reasoning in the patent domain. In
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
-
cs.AI, q-bio.NC updates on arXiv.org
-
Atlas 2 -- Foundation models for clinical deployment
arXiv:2601.05148v1 Announce Type: cross Abstract: Pathology foundation models substantially advanced the possibilities in computational pathology -- yet tradeoffs in terms of performance, robustness, and computational requirements remained, which limited their clinical deployment. In this report, we present Atlas 2, Atlas 2-B, and Atlas 2-S, three pathology vision foundation models which bridge these shortcomings by showing state-of-the-art performance in prediction performance, robustness, and
Atlas 2 -- Foundation models for clinical deployment
-
cs.AI, q-bio.NC updates on arXiv.org
-
PsychEval: A Multi-Session and Multi-Therapy Benchmark for High-Realism AI Psychological Counselor
arXiv:2601.01802v3 Announce Type: replace Abstract: To develop a reliable AI for psychological assessment, we introduce \texttt{PsychEval}, a multi-session, multi-therapy, and highly realistic benchmark designed to address three key challenges: \textbf{1) Can we train a highly realistic AI counselor?} Realistic counseling is a longitudinal task requiring sustained memory and dynamic goal tracking. We propose a multi-session benchmark (spanning 6-10 sessions across three distinct stages) that de
PsychEval: A Multi-Session and Multi-Therapy Benchmark for High-Realism AI Psychological Counselor
-
cs.AI, q-bio.NC updates on arXiv.org
-
A Framework for Responsible AI Systems: Building Societal Trust through Domain Definition, Trustworthy AI Design, Auditability, Accountability, and Governance
arXiv:2503.04739v2 Announce Type: replace-cross Abstract: Responsible Artificial Intelligence (RAI) addresses the ethical and regulatory challenges of deploying AI systems in high-risk scenarios. This paper proposes a comprehensive framework for the design of an RAI system (RAIS) that integrates five key dimensions: domain definition, trustworthy AI design, auditability, accountability, and governance. Unlike prior work that treats these components in isolation, our proposal emphasizes their in
A Framework for Responsible AI Systems: Building Societal Trust through Domain Definition, Trustworthy AI Design, Auditability, Accountability, and Governance
-
cs.AI, q-bio.NC updates on arXiv.org
-
SciClaims: An End-to-End Generative System for Biomedical Claim Analysis
arXiv:2503.18526v2 Announce Type: replace-cross Abstract: We present SciClaims, an interactive web-based system for end-to-end scientific claim analysis in the biomedical domain. Designed for high-stakes use cases such as systematic literature reviews and patent validation, SciClaims extracts claims from text, retrieves relevant evidence from PubMed, and verifies their veracity. The system features a user-friendly interface where users can input scientific text and view extracted claims, predic
SciClaims: An End-to-End Generative System for Biomedical Claim Analysis
-
cs.AI, q-bio.NC updates on arXiv.org
-
Multi-Modal AI for Remote Patient Monitoring in Cancer Care
arXiv:2512.00949v2 Announce Type: replace-cross Abstract: For patients undergoing systemic cancer therapy, the time between clinic visits is full of uncertainties and risks of unmonitored side effects. To bridge this gap in care, we developed and prospectively trialed a multi-modal AI framework for remote patient monitoring (RPM). This system integrates multi-modal data from the HALO-X platform, such as demographics, wearable sensors, daily surveys, and clinical events. Our observational trial
Multi-Modal AI for Remote Patient Monitoring in Cancer Care
-
Journal of Medical Internet Research
-
Developing an AI-Assisted Tool That Identifies Patients With Multimorbidity and Complex Polypharmacy to Improve the Process of Medication Reviews: Qualitative Interview and Focus Group Study
Background: Structured medication reviews (SMRs) are an essential component of medication optimization, especially for patients with multimorbidity and polypharmacy. However, the process remains challenging due to the complexities of patient data, time constraints, and the need for coordination among health care professionals (HCPs). This study explores HCPs’ perspectives on the integration of artificial intelligence (AI)–assisted tools to enhance the SMR process, with a focus on the potential b
Developing an AI-Assisted Tool That Identifies Patients With Multimorbidity and Complex Polypharmacy to Improve the Process of Medication Reviews: Qualitative Interview and Focus Group Study
-
Journal of Medical Internet Research
-
Intervention in Health Misinformation Using Large Language Models for Automated Detection, Thematic Analysis, and Inoculation: Case Study on COVID-19
Background: The rapid growth of social media as an information channel has enabled the swift spread of inaccurate or false health information, significantly impacting public health. This widespread dissemination of misinformation has caused confusion, eroded trust in health authorities, led to noncompliance with health guidelines, and encouraged risky health behaviors. Understanding the dynamics of misinformation on social media is essential for devising effective public health communication str
Intervention in Health Misinformation Using Large Language Models for Automated Detection, Thematic Analysis, and Inoculation: Case Study on COVID-19
-
STAT

-
STAT+: OpenAI invites you to upload medical records to ChatGPT
You’re reading the web edition of STAT’s Health Tech newsletter, our guide to how technology is transforming the life sciences. Sign up to get it delivered in your inbox every Tuesday and Thursday. Millions of people, including me and possibly you, are already asking ChatGPT questions about health. Still others are dumping otherwise inscrutable medical records downloaded from patient portals into the generative AI bot, hoping to glean new insights. Now, OpenAI will encourage this behavior
STAT+: OpenAI invites you to upload medical records to ChatGPT
You’re reading the web edition of STAT’s Health Tech newsletter, our guide to how technology is transforming the life sciences. Sign up to get it delivered in your inbox every Tuesday and Thursday.
Millions of people, including me and possibly you, are already asking ChatGPT questions about health. Still others are dumping otherwise inscrutable medical records downloaded from patient portals into the generative AI bot, hoping to glean new insights.
Now, OpenAI will encourage this behavior with a new health specific tab in the service that the company says has better security and privacy protections so users feel safe pouring sensitive medical data into the bot. In addition to uploading files, users can hook up data from products like Apple Health and Weight Watchers or obtain medical records from providers through b.well’s network. OpenAI promises that it won’t train its models on the data you put into ChatGPT Health. (Reminder: Data that you upload to a consumer service is not covered by HIPAA.)
Continue to STAT+ to read the full story…


© Adobe