Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
Internal Deployment Gaps in AI Regulation
arXiv:2601.08005v1 Announce Type: new Abstract: Frontier AI regulations primarily focus on systems deployed to external users, where deployment is more visible and subject to outside scrutiny. However, high-stakes applications can occur internally when companies deploy highly capable systems within their own organizations, such as for automating R\&D, accelerating critical business processes, and handling sensitive proprietary data. This paper examines how frontier AI regulations in the Uni
-
cs.AI, q-bio.NC updates on arXiv.org
-
Semantic Laundering in AI Agent Architectures: Why Tool Boundaries Do Not Confer Epistemic Warrant
arXiv:2601.08333v1 Announce Type: new Abstract: LLM-based agent architectures systematically conflate information transport mechanisms with epistemic justification mechanisms. We formalize this class of architectural failures as semantic laundering: a pattern where propositions with absent or weak warrant are accepted by the system as admissible by crossing architecturally trusted interfaces. We show that semantic laundering constitutes an architectural realization of the Gettier problem: propo
Semantic Laundering in AI Agent Architectures: Why Tool Boundaries Do Not Confer Epistemic Warrant
-
cs.AI, q-bio.NC updates on arXiv.org
-
WaterCopilot: An AI-Driven Virtual Assistant for Water Management
arXiv:2601.08559v1 Announce Type: new Abstract: Sustainable water resource management in transboundary river basins is challenged by fragmented data, limited real-time access, and the complexity of integrating diverse information sources. This paper presents WaterCopilot-an AI-driven virtual assistant developed through collaboration between the International Water Management Institute (IWMI) and Microsoft Research for the Limpopo River Basin (LRB) to bridge these gaps through a unified, interac
WaterCopilot: An AI-Driven Virtual Assistant for Water Management
-
cs.AI, q-bio.NC updates on arXiv.org
-
Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock
arXiv:2601.08673v1 Announce Type: new Abstract: Recent reports of large language models (LLMs) exhibiting behaviors such as deception, threats, or blackmail are often interpreted as evidence of alignment failure or emergent malign agency. We argue that this interpretation rests on a conceptual error. LLMs do not reason morally; they statistically internalize the record of human social interaction, including laws, contracts, negotiations, conflicts, and coercive arrangements. Behaviors commonly
Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock
-
cs.AI, q-bio.NC updates on arXiv.org
-
All Required, In Order: Phase-Level Evaluation for AI-Human Dialogue in Healthcare and Beyond
arXiv:2601.08690v1 Announce Type: new Abstract: Conversational AI is starting to support real clinical work, but most evaluation methods miss how compliance depends on the full course of a conversation. We introduce Obligatory-Information Phase Structured Compliance Evaluation (OIP-SCE), an evaluation method that checks whether every required clinical obligation is met, in the right order, with clear evidence for clinicians to review. This makes complex rules practical and auditable, helping cl
All Required, In Order: Phase-Level Evaluation for AI-Human Dialogue in Healthcare and Beyond
-
cs.AI, q-bio.NC updates on arXiv.org
-
Imaging-anchored Multiomics in Cardiovascular Disease: Integrating Cardiac Imaging, Bulk, Single-cell, and Spatial Transcriptomics
arXiv:2601.07871v1 Announce Type: cross Abstract: Cardiovascular disease arises from interactions between inherited risk, molecular programmes, and tissue-scale remodelling that are observed clinically through imaging. Health systems now routinely generate large volumes of cardiac MRI, CT and echocardiography together with bulk, single-cell and spatial transcriptomics, yet these data are still analysed in separate pipelines. This review examines joint representations that link cardiac imaging p
Imaging-anchored Multiomics in Cardiovascular Disease: Integrating Cardiac Imaging, Bulk, Single-cell, and Spatial Transcriptomics
-
cs.AI, q-bio.NC updates on arXiv.org
-
GI-Bench: A Panoramic Benchmark Revealing the Knowledge-Experience Dissociation of Multimodal Large Language Models in Gastrointestinal Endoscopy Against Clinical Standards
arXiv:2601.08183v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) show promise in gastroenterology, yet their performance against comprehensive clinical workflows and human benchmarks remains unverified. To systematically evaluate state-of-the-art MLLMs across a panoramic gastrointestinal endoscopy workflow and determine their clinical utility compared with human endoscopists. We constructed GI-Bench, a benchmark encompassing 20 fine-grained lesion categories. Twelve ML
GI-Bench: A Panoramic Benchmark Revealing the Knowledge-Experience Dissociation of Multimodal Large Language Models in Gastrointestinal Endoscopy Against Clinical Standards
-
cs.AI, q-bio.NC updates on arXiv.org
-
Moral Lenses, Political Coordinates: Towards Ideological Positioning of Morally Conditioned LLMs
arXiv:2601.08634v1 Announce Type: cross Abstract: While recent research has systematically documented political orientation in large language models (LLMs), existing evaluations rely primarily on direct probing or demographic persona engineering to surface ideological biases. In social psychology, however, political ideology is also understood as a downstream consequence of fundamental moral intuitions. In this work, we investigate the causal relationship between moral values and political posi
Moral Lenses, Political Coordinates: Towards Ideological Positioning of Morally Conditioned LLMs
-
cs.AI, q-bio.NC updates on arXiv.org
-
ISLA: A U-Net for MRI-based acute ischemic stroke lesion segmentation with deep supervision, attention, domain adaptation, and ensemble learning
arXiv:2601.08732v1 Announce Type: cross Abstract: Accurate delineation of acute ischemic stroke lesions in MRI is a key component of stroke diagnosis and management. In recent years, deep learning models have been successfully applied to the automatic segmentation of such lesions. While most proposed architectures are based on the U-Net framework, they primarily differ in their choice of loss functions and in the use of deep supervision, residual connections, and attention mechanisms. Moreover,
ISLA: A U-Net for MRI-based acute ischemic stroke lesion segmentation with deep supervision, attention, domain adaptation, and ensemble learning
-
cs.AI, q-bio.NC updates on arXiv.org
-
DeKeyNLU: Enhancing Natural Language to SQL Generation through Task Decomposition and Keyword Extraction
arXiv:2509.14507v2 Announce Type: replace Abstract: Natural Language to SQL (NL2SQL) provides a new model-centric paradigm that simplifies database access for non-technical users by converting natural language queries into SQL commands. Recent advancements, particularly those integrating Retrieval-Augmented Generation (RAG) and Chain-of-Thought (CoT) reasoning, have made significant strides in enhancing NL2SQL performance. However, challenges such as inaccurate task decomposition and keyword ex
DeKeyNLU: Enhancing Natural Language to SQL Generation through Task Decomposition and Keyword Extraction
-
cs.AI, q-bio.NC updates on arXiv.org
-
Generative Digital Twins: Vision-Language Simulation Models for Executable Industrial Systems
arXiv:2512.20387v4 Announce Type: replace Abstract: We propose a Vision-Language Simulation Model (VLSM) that unifies visual and textual understanding to synthesize executable FlexScript from layout sketches and natural-language prompts, enabling cross-modal reasoning for industrial simulation systems. To support this new paradigm, the study constructs the first large-scale dataset for generative digital twins, comprising over 120,000 prompt-sketch-code triplets that enable multimodal learning
Generative Digital Twins: Vision-Language Simulation Models for Executable Industrial Systems
-
cs.AI, q-bio.NC updates on arXiv.org
-
Explainable Molecular Property Prediction: Aligning Chemical Concepts with Predictions via Language Models
arXiv:2405.16041v4 Announce Type: replace-cross Abstract: Providing explainable molecular property predictions is critical for many scientific domains, such as drug discovery and material science. Though transformer-based language models have shown great potential in accurate molecular property prediction, they neither provide chemically meaningful explanations nor faithfully reveal the molecular structure-property relationships. In this work, we develop a framework for explainable molecular pr
Explainable Molecular Property Prediction: Aligning Chemical Concepts with Predictions via Language Models
-
cs.AI, q-bio.NC updates on arXiv.org
-
Hallucination, reliability, and the role of generative AI in science
arXiv:2504.08526v2 Announce Type: replace-cross Abstract: Generative AI increasingly supports scientific inference, from protein structure prediction to weather forecasting. Yet its distinctive failure mode, hallucination, raises epistemic alarm bells. I argue that this failure mode can be addressed by shifting from data-centric to phenomenon-centric assessment. Through case studies of AlphaFold and GenCast, I show how scientific workflows discipline generative models through theory-guided trai
Hallucination, reliability, and the role of generative AI in science
-
cs.AI, q-bio.NC updates on arXiv.org
-
Efficient and Reproducible Biomedical Question Answering using Retrieval Augmented Generation
arXiv:2505.07917v2 Announce Type: replace-cross Abstract: Biomedical question-answering (QA) systems require effective retrieval and generation components to ensure accuracy, efficiency, and scalability. This study systematically examines a Retrieval-Augmented Generation (RAG) system for biomedical QA, evaluating retrieval strategies and response time trade-offs. We first assess state-of-the-art retrieval methods, including BM25, BioBERT, MedCPT, and a hybrid approach, alongside common data sto
Efficient and Reproducible Biomedical Question Answering using Retrieval Augmented Generation
-
cs.AI, q-bio.NC updates on arXiv.org
-
Aligning Trustworthy AI with Democracy: A Dual Taxonomy of Opportunities and Risks
arXiv:2505.13565v2 Announce Type: replace-cross Abstract: Artificial Intelligence (AI) poses both significant risks and valuable opportunities for democratic governance. This paper introduces a dual taxonomy to evaluate AI's complex relationship with democracy: the AI Risks to Democracy (AIRD) taxonomy, which identifies how AI can undermine core democratic principles such as autonomy, fairness, and trust; and the AI's Positive Contributions to Democracy (AIPD) taxonomy, which highlights AI's po
Aligning Trustworthy AI with Democracy: A Dual Taxonomy of Opportunities and Risks
-
cs.AI, q-bio.NC updates on arXiv.org
-
Foundations of LLM Knowledge Materialization: Termination, Reproducibility, Robustness
arXiv:2510.06780v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) encode substantial factual knowledge, yet measuring and systematizing this knowledge remains challenging. Converting it into structured format, for example through recursive extraction approaches such as the GPTKB methodology (Hu et al., 2025b), is still underexplored. Key open questions include whether such extraction can terminate, whether its outputs are reproducible, and how robust they are to variations.
Foundations of LLM Knowledge Materialization: Termination, Reproducibility, Robustness
-
STAT

-
STAT+: On Day 2 of JPM, Gilead lays outs it next test, a VC looks to raise funds, and one firm has FDA whiplash
This is the online version of The Readout, STAT’s flagship biotech newsletter. Sign up to get it in your inbox. You’re back. We’re sort of back. It’s Day 2 of JPM and we’re definitely not exhausted or delirious yet.This is Elaine Chen, Adam Feuerstein, Matt Herper, and Allison DeAngelis again. We’ve got a lot more news today, so let’s get to it. The next test for Kite Pharma — and Gilead It’s anito-cel, the CAR-T therapy for multiple myeloma that Gilead is developing in partnership wi
STAT+: On Day 2 of JPM, Gilead lays outs it next test, a VC looks to raise funds, and one firm has FDA whiplash
This is the online version of The Readout, STAT’s flagship biotech newsletter. Sign up to get it in your inbox.
You’re back. We’re sort of back. It’s Day 2 of JPM and we’re definitely not exhausted or delirious yet.
This is Elaine Chen, Adam Feuerstein, Matt Herper, and Allison DeAngelis again. We’ve got a lot more news today, so let’s get to it.
The next test for Kite Pharma — and Gilead
It’s anito-cel, the CAR-T therapy for multiple myeloma that Gilead is developing in partnership with Arcellx. Gilead submitted the therapy to the FDA sometime before the end of December, Cindy Perettie, executive vice president of Kite Pharma, the cell therapy division of Gilead, told STAT at a Gilead media breakfast.
Continue to STAT+ to read the full story…


© Alex Hogan/STAT
-
TechCrunch
-
Microsoft announces glut of new data centers but says it won’t let your electricity bill go up
The tech giant has pledged to be a "good neighbor" as it continues to invest in AI infrastructure throughout the country.
Microsoft announces glut of new data centers but says it won’t let your electricity bill go up
-
Journal of Medical Internet Research
-
Assessment of the Diagnostic Performance and Clinical Impact of AI in Hepatic Steatosis: Systematic Review and Meta-Analysis
Background: The global rise of metabolic associated fatty liver disease (MAFLD) reflects the urgent need for accurate, non-invasive diagnostic approaches. The invasive nature of liver biopsy and the limited sensitivity of ultrasound (US) in detecting early steatosis highlight a critical diagnostic gap. Artificial intelligence (AI) has emerged as a transformative tool, enabling the automated detection and grading of hepatic steatosis (HS) from medical imaging data. Objective: This systematic revi
Assessment of the Diagnostic Performance and Clinical Impact of AI in Hepatic Steatosis: Systematic Review and Meta-Analysis
-
InfoQ

-
Google Introduces Conductor, a Context-Driven Development Extension for Gemini CLI
Google has released Conductor, a new preview extension for Gemini CLI that introduces a structured, context-driven approach to AI-assisted software development. The extension is designed to address a common limitation of chat-based coding tools: the loss of project context across sessions. By Robert Krzaczyński
Google Introduces Conductor, a Context-Driven Development Extension for Gemini CLI
Google has released Conductor, a new preview extension for Gemini CLI that introduces a structured, context-driven approach to AI-assisted software development. The extension is designed to address a common limitation of chat-based coding tools: the loss of project context across sessions.
By Robert Krzaczyński