Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
DarkPatterns-LLM: A Multi-Layer Benchmark for Detecting Manipulative and Harmful AI Behavior
arXiv:2512.22470v1 Announce Type: new Abstract: The proliferation of Large Language Models (LLMs) has intensified concerns about manipulative or deceptive behaviors that can undermine user autonomy, trust, and well-being. Existing safety benchmarks predominantly rely on coarse binary labels and fail to capture the nuanced psychological and social mechanisms constituting manipulation. We introduce \textbf{DarkPatterns-LLM}, a comprehensive benchmark dataset and diagnostic framework for fine-grai
-
cs.AI, q-bio.NC updates on arXiv.org
-
Why AI Safety Requires Uncertainty, Incomplete Preferences, and Non-Archimedean Utilities
arXiv:2512.23508v1 Announce Type: new Abstract: How can we ensure that AI systems are aligned with human values and remain safe? We can study this problem through the frameworks of the AI assistance and the AI shutdown games. The AI assistance problem concerns designing an AI agent that helps a human to maximise their utility function(s). However, only the human knows these function(s); the AI assistant must learn them. The shutdown problem instead concerns designing AI agents that: shut down w
Why AI Safety Requires Uncertainty, Incomplete Preferences, and Non-Archimedean Utilities
-
cs.AI, q-bio.NC updates on arXiv.org
-
Interpretable Link Prediction in AI-Driven Cancer Research: Uncovering Co-Authorship Patterns
arXiv:2512.22181v1 Announce Type: cross Abstract: Artificial intelligence (AI) is transforming cancer diagnosis and treatment. The intricate nature of this disease necessitates the collaboration of diverse stakeholders with varied expertise to ensure the effectiveness of cancer research. Despite its importance, forming effective interdisciplinary research teams remains challenging. Understanding and predicting collaboration patterns can help researchers, organizations, and policymakers optimize
Interpretable Link Prediction in AI-Driven Cancer Research: Uncovering Co-Authorship Patterns
-
cs.AI, q-bio.NC updates on arXiv.org
-
LLM-Guided Exemplar Selection for Few-Shot Wearable-Sensor Human Activity Recognition
arXiv:2512.22385v1 Announce Type: cross Abstract: In this paper, we propose an LLM-Guided Exemplar Selection framework to address a key limitation in state-of-the-art Human Activity Recognition (HAR) methods: their reliance on large labeled datasets and purely geometric exemplar selection, which often fail to distinguish similar weara-ble sensor activities such as walking, walking upstairs, and walking downstairs. Our method incorporates semantic reasoning via an LLM-generated knowledge prior t
LLM-Guided Exemplar Selection for Few-Shot Wearable-Sensor Human Activity Recognition
-
cs.AI, q-bio.NC updates on arXiv.org
-
MedGemma vs GPT-4: Open-Source and Proprietary Zero-shot Medical Disease Classification from Images
arXiv:2512.23304v1 Announce Type: cross Abstract: Multimodal Large Language Models (LLMs) introduce an emerging paradigm for medical imaging by interpreting scans through the lens of extensive clinical knowledge, offering a transformative approach to disease classification. This study presents a critical comparison between two fundamentally different AI architectures: the specialized open-source agent MedGemma and the proprietary large multimodal model GPT-4 for diagnosing six different disease
MedGemma vs GPT-4: Open-Source and Proprietary Zero-shot Medical Disease Classification from Images
-
cs.AI, q-bio.NC updates on arXiv.org
-
Generating Verifiable Chain of Thoughts from Exection-Traces
arXiv:2512.00127v2 Announce Type: replace-cross Abstract: Teaching language models to reason about code execution remains a fundamental challenge. While Chain-of-Thought (CoT) prompting has shown promise, current synthetic training data suffers from a critical weakness: the reasoning steps are often plausible-sounding explanations generated by teacher models, not verifiable accounts of what the code actually does. This creates a troubling failure mode where models learn to mimic superficially c
Generating Verifiable Chain of Thoughts from Exection-Traces
-
Cell Death Discovery nature.com science feeds
-
Modeling hepatocellular carcinoma and its microenvironment on a chip
Cell Death Discovery, Published online: 29 December 2025; doi:10.1038/s41420-025-02917-8Modeling hepatocellular carcinoma and its microenvironment on a chip
Modeling hepatocellular carcinoma and its microenvironment on a chip
Cell Death Discovery, Published online: 29 December 2025; doi:10.1038/s41420-025-02917-8
Modeling hepatocellular carcinoma and its microenvironment on a chip-
(Multiomics OR Omics) AND (Lung OR gastric OR Hepatocellular)
-
Metabolic signatures in gastroenteropancreatic neuroendocrine neoplasms: unraveling diagnostic and prognostic insights
Front Endocrinol (Lausanne). 2025 Dec 11;16:1676021. doi: 10.3389/fendo.2025.1676021. eCollection 2025.ABSTRACTGastroenteropancreatic neuroendocrine neoplasms (GEP-NENs) are a heterogeneous group of tumors characterized by diverse biological behaviors and variable clinical outcomes. Recent advances have highlighted the important role of metabolic reprogramming in tumorigenesis, progression, and therapeutic resistance in GEP-NENs. In this review, we synthesize the current evidence on metabolic bi
Metabolic signatures in gastroenteropancreatic neuroendocrine neoplasms: unraveling diagnostic and prognostic insights
Front Endocrinol (Lausanne). 2025 Dec 11;16:1676021. doi: 10.3389/fendo.2025.1676021. eCollection 2025.
ABSTRACT
Gastroenteropancreatic neuroendocrine neoplasms (GEP-NENs) are a heterogeneous group of tumors characterized by diverse biological behaviors and variable clinical outcomes. Recent advances have highlighted the important role of metabolic reprogramming in tumorigenesis, progression, and therapeutic resistance in GEP-NENs. In this review, we synthesize the current evidence on metabolic biomarkers and altered metabolic pathways-particularly those involving glucose, lipid, and amino acid metabolism. Key biomarkers such as GLUT-1, FASN, and enzymes involved in ferroptosis, cholesterol biosynthesis, and amino acid catabolism demonstrate strong associations with tumor aggressiveness, hypoxia, and mTOR signaling. Moreover, metabolomic profiling and functional studies suggest that metabolic markers may inform prognosis and predict response to targeted therapies such as Everolimus. Although promising, the clinical translation of these markers is still limited and requires further validation in large, subtype-specific cohorts. Our findings highlight the importance of integrating metabolic profiling into the diagnostic and therapeutic landscape of GEP-NENs. Future research should prioritize biomarker standardization, multi-omics integration, and the development of metabolism-based therapeutic strategies tailored to tumor subtype and differentiation grade.
PMID:41458541 | PMC:PMC12738315 | DOI:10.3389/fendo.2025.1676021
-
Journal of Medical Internet Research
-
Evaluating Peer Online Forums to Support Health: Ethical and Practical Challenges
Many people use peer online forums to seek support for health-related problems. More research is needed to understand the impacts of forum use, and how these are generated. However, there are significant ethical and practical challenges with the methods available to do the required research. We examine the key challenges associated with conducting each of the most commonly used online data collection methods: surveys, interviews, forum post analysis; and triangulation of these methods. Based on
Evaluating Peer Online Forums to Support Health: Ethical and Practical Challenges
-
npj Digital Medicine
-
Context matching is not reasoning when performing generalized clinical evaluation of generative language models
npj Digital Medicine, Published online: 27 December 2025; doi:10.1038/s41746-025-02253-2Context matching is not reasoning when performing generalized clinical evaluation of generative language models
Context matching is not reasoning when performing generalized clinical evaluation of generative language models
npj Digital Medicine, Published online: 27 December 2025; doi:10.1038/s41746-025-02253-2
Context matching is not reasoning when performing generalized clinical evaluation of generative language models-
MRD
-
Cancer in a drop: Liquid biopsy highlights from the World Conference on Lung Cancer (WCLC) 2025
J Liq Biopsy. 2025 Nov 29;10:100449. doi: 10.1016/j.jlb.2025.100449. eCollection 2025 Dec.ABSTRACTThe role of liquid biopsy in oncological care continues to expand, with multiple studies presented at the International Association for the Study of Lung Cancer (IASLC) 2025 World Conference on Lung Cancer (WCLC 2025). This review summarizes recent advances in liquid biopsy for thoracic oncology, encompassing both non-small cell lung cancer (NSCLC) and small cell lung cancer (SCLC). In early detecti
Cancer in a drop: Liquid biopsy highlights from the World Conference on Lung Cancer (WCLC) 2025
J Liq Biopsy. 2025 Nov 29;10:100449. doi: 10.1016/j.jlb.2025.100449. eCollection 2025 Dec.
ABSTRACT
The role of liquid biopsy in oncological care continues to expand, with multiple studies presented at the International Association for the Study of Lung Cancer (IASLC) 2025 World Conference on Lung Cancer (WCLC 2025). This review summarizes recent advances in liquid biopsy for thoracic oncology, encompassing both non-small cell lung cancer (NSCLC) and small cell lung cancer (SCLC). In early detection and screening, proteomic profiling has identified potential biomarkers predictive of future lung cancer risk. The integration of proteomics with clinical and imaging data can improve pulmonary nodule malignancy prediction. In resectable NSCLC, tumour-informed whole-genome sequencing (WGS) assay demonstrated high sensitivity for minimal residual disease (MRD) detection, with MRD clearance following neoadjuvant osimertinib or chemo-immunotherapy associated with favorable outcomes. In advanced NSCLC, longitudinal liquid biopsy analyses reveal dynamic subclonal evolution driving early treatment resistance. Circulating tumor DNA (ctDNA) clearance following targeted therapy in MET exon 14 skipping and BRAF-mutated tumors was associated with improved clinical outcomes. Emerging biomarkers such as ctDNA tumour fraction and circulating microRNA signatures are promising for radiotherapy stratification and prediction of immunotherapy-related toxicities. In SCLC, MRD monitoring enables earlier detection of disease progression and supports ctDNA-guided selection of patients for consolidation immunotherapy following chemotherapy. Overall, these advances demonstrate the expanding role of liquid biopsy in improving early detection, guiding treatment, and improving disease monitoring in lung cancer.
PMID:41438843 | PMC:PMC12720026 | DOI:10.1016/j.jlb.2025.100449
-
TechCrunch
-
The year data centers went from backend to center stage
Data centers are no longer the boring tech issue they once were.
The year data centers went from backend to center stage
-
cs.AI, q-bio.NC updates on arXiv.org
-
AI Needs Physics More Than Physics Needs AI
arXiv:2512.16344v1 Announce Type: new Abstract: Artificial intelligence (AI) is commonly depicted as transformative. Yet, after more than a decade of hype, its measurable impact remains modest outside a few high-profile scientific and commercial successes. The 2024 Nobel Prizes in Chemistry and Physics recognized AI's potential, but broader assessments indicate the impact to date is often more promotional than technical. We argue that while current AI may influence physics, physics has signific
AI Needs Physics More Than Physics Needs AI
-
cs.AI, q-bio.NC updates on arXiv.org
-
Distributional AGI Safety
arXiv:2512.16856v1 Announce Type: new Abstract: AI safety and alignment research has predominantly been focused on methods for safeguarding individual AI systems, resting on the assumption of an eventual emergence of a monolithic Artificial General Intelligence (AGI). The alternative AGI emergence hypothesis, where general capability levels are first manifested through coordination in groups of sub-AGI individual agents with complementary skills and affordances, has received far less attention.
Distributional AGI Safety
-
cs.AI, q-bio.NC updates on arXiv.org
-
AI4EOSC: a Federated Cloud Platform for Artificial Intelligence in Scientific Research
arXiv:2512.16455v1 Announce Type: cross Abstract: In this paper, we describe a federated compute platform dedicated to support Artificial Intelligence in scientific workloads. Putting the effort into reproducible deployments, it delivers consistent, transparent access to a federation of physically distributed e-Infrastructures. Through a comprehensive service catalogue, the platform is able to offer an integrated user experience covering the full Machine Learning lifecycle, including model deve
AI4EOSC: a Federated Cloud Platform for Artificial Intelligence in Scientific Research
-
cs.AI, q-bio.NC updates on arXiv.org
-
Plausibility as Failure: How LLMs and Humans Co-Construct Epistemic Error
arXiv:2512.16750v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as epistemic partners in everyday reasoning, yet their errors remain predominantly analyzed through predictive metrics rather than through their interpretive effects on human judgment. This study examines how different forms of epistemic failure emerge, are masked, and are tolerated in human AI interaction, where failure is understood as a relational breakdown shaped by model-generated plausibil
Plausibility as Failure: How LLMs and Humans Co-Construct Epistemic Error
-
cs.AI, q-bio.NC updates on arXiv.org
-
Constitutional Law and AI Governance: Constraints on Model Licensing and Research Classification
arXiv:2509.05361v2 Announce Type: replace-cross Abstract: Transformative AI systems may pose unprecedented catastrophic risks, but the U.S. Constitution places significant constraints on the government's ability to govern this technology. This paper examines how the First Amendment, administrative law, and the Fourteenth Amendment shape the legal vulnerability of two regulatory proposals: model licensing and AI research classification. While the First Amendment may provide some degree of protec
Constitutional Law and AI Governance: Constraints on Model Licensing and Research Classification
-
cs.AI, q-bio.NC updates on arXiv.org
-
First, do NOHARM: towards clinically safe large language models
arXiv:2512.01241v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are routinely used by physicians and patients for medical advice, yet their clinical safety profiles remain poorly characterized. We present NOHARM (Numerous Options Harm Assessment for Risk in Medicine), a benchmark using 100 real primary care-to-specialist consultation cases to measure frequency and severity of harm from LLM-generated medical recommendations. NOHARM covers 10 specialties, with 12,747 expert
First, do NOHARM: towards clinically safe large language models
-
cs.AI, q-bio.NC updates on arXiv.org
-
DrugRAG: Enhancing Pharmacy LLM Performance Through A Novel Retrieval-Augmented Generation Pipeline
arXiv:2512.14896v1 Announce Type: cross Abstract: Objectives: To evaluate large language model (LLM) performance on pharmacy licensure-style question-answering (QA) tasks and develop an external knowledge integration method to improve their accuracy. Methods: We benchmarked eleven existing LLMs with varying parameter sizes (8 billion to 70+ billion) using a 141-question pharmacy dataset. We measured baseline accuracy for each model without modification. We then developed a three-step retrieva
DrugRAG: Enhancing Pharmacy LLM Performance Through A Novel Retrieval-Augmented Generation Pipeline
-
cs.AI, q-bio.NC updates on arXiv.org
-
Leveraging LLMs for Structured Data Extraction from Unstructured Patient Records
arXiv:2512.13700v1 Announce Type: new Abstract: Manual chart review remains an extremely time-consuming and resource-intensive component of clinical research, requiring experts to extract often complex information from unstructured electronic health record (EHR) narratives. We present a secure, modular framework for automated structured feature extraction from clinical notes leveraging locally deployed large language models (LLMs) on institutionally approved, Health Insurance Portability and Ac