Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
SciHorizon-GENE: Benchmarking LLM for Life Sciences Inference from Gene Knowledge to Functional Understanding
arXiv:2601.12805v1 Announce Type: cross Abstract: Large language models (LLMs) have shown growing promise in biomedical research, particularly for knowledge-driven interpretation tasks. However, their ability to reliably reason from gene-level knowledge to functional understanding, However, their ability to reliably reason from gene-level knowledge to functional understanding, a core requirement for knowledge-enhanced cell atlas interpretation, remains largely underexplored. To address this gap
-
cs.AI, q-bio.NC updates on arXiv.org
-
Zero-shot adaptable task planning for autonomous construction robots: a comparative study of lightweight single and multi-AI agent systems
arXiv:2601.14091v1 Announce Type: cross Abstract: Robots are expected to play a major role in the future construction industry but face challenges due to high costs and difficulty adapting to dynamic tasks. This study explores the potential of foundation models to enhance the adaptability and generalizability of task planning in construction robots. Four models are proposed and implemented using lightweight, open-source large language models (LLMs) and vision language models (VLMs). These model
Zero-shot adaptable task planning for autonomous construction robots: a comparative study of lightweight single and multi-AI agent systems
-
cs.AI, q-bio.NC updates on arXiv.org
-
OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment
arXiv:2601.01576v2 Announce Type: replace-cross Abstract: Evaluating novelty is critical yet challenging in peer review, as reviewers must assess submissions against a vast, rapidly evolving literature. This report presents OpenNovelty, an LLM-powered agentic system for transparent, evidence-based novelty analysis. The system operates through four phases: (1) extracting the core task and contribution claims to generate retrieval queries; (2) retrieving relevant prior work based on extracted que
OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment
-
cs.AI, q-bio.NC updates on arXiv.org
-
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
arXiv:2505.17217v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) often exhibit gender bias, resulting in unequal treatment of male and female subjects across different contexts. To address this issue, we propose a novel data generation framework that fosters exploratory thinking in LLMs. Our approach prompts models to generate story pairs featuring male and female protagonists in structurally identical, morally ambiguous scenarios, then elicits and compares their moral jud
Mitigating Gender Bias via Fostering Exploratory Thinking in LLMs
-
(Multiomics OR Omics) AND (Pancreatic)
-
Complement-secreting CAFs are associated with better prognosis in pancreatic cancer: single-cell multiomics
Gut. 2026 Jan 13:gutjnl-2025-335683. doi: 10.1136/gutjnl-2025-335683. Online ahead of print.ABSTRACTBACKGROUND: Accumulating evidence has demonstrated that distinct tumour-promoting and tumour-restraining cancer-associated fibroblast (CAF) subtypes coexist in pancreatic ductal adenocarcinoma.OBJECTIVE: To develop targeted CAF therapeutic strategies by reprogramming tumour-promoting CAF subtypes.DESIGN: We leveraged multiomics technologies to systematically identify and characterise CAF subtypes
Complement-secreting CAFs are associated with better prognosis in pancreatic cancer: single-cell multiomics
Gut. 2026 Jan 13:gutjnl-2025-335683. doi: 10.1136/gutjnl-2025-335683. Online ahead of print.
ABSTRACT
BACKGROUND: Accumulating evidence has demonstrated that distinct tumour-promoting and tumour-restraining cancer-associated fibroblast (CAF) subtypes coexist in pancreatic ductal adenocarcinoma.
OBJECTIVE: To develop targeted CAF therapeutic strategies by reprogramming tumour-promoting CAF subtypes.
DESIGN: We leveraged multiomics technologies to systematically identify and characterise CAF subtypes transcriptionally, epigenetically and spatially and correlate them with clinicopathological features.
RESULTS: We found that complement-secreting CAFs (csCAFs), initially identified by our group and inflammatory CAFs (iCAFs) share significant overlap in their transcriptional profiles and chromatin accessibility. iCAFs specifically express transcription factors from the heme and oxidative homeostasis pathway and the activator protein 1 family, which are both involved in cellular response to oxidative stress. Notably, the composition of csCAFs among all CAFs declined during pancreatic carcinogenesis, while trajectory analysis showed that csCAFs could potentially differentiate into iCAFs. Spatially resolved analysis indicated that tumour regions with a higher csCAF composition were associated with lower levels of TGF-β ligands, fewer M2 tumour-associated macrophages and increased levels of lipid mediators. Additionally, we identified a spatially defined CXCL12-CXCR4 ligand-receptor interaction between csCAFs and T cells, but in distinct patterns between different metastatic organs. Patients with a higher composition of csCAFs have significantly longer overall survival and recurrence-free survival through multiplex immunohistochemistry and bulk RNA-seq deconvolution.
CONCLUSION: Our study demonstrates that csCAFs may represent an early-stage iCAF subtype and suggests a promising strategy for reprogramming iCAFs into csCAFs.
PMID:41534892 | DOI:10.1136/gutjnl-2025-335683
-
npj Digital Medicine
-
Geometric multi-instance learning for weakly supervised gastric cancer segmentation
npj Digital Medicine, Published online: 13 January 2026; doi:10.1038/s41746-025-02287-6Geometric multi-instance learning for weakly supervised gastric cancer segmentation
Geometric multi-instance learning for weakly supervised gastric cancer segmentation
npj Digital Medicine, Published online: 13 January 2026; doi:10.1038/s41746-025-02287-6
Geometric multi-instance learning for weakly supervised gastric cancer segmentation-
cs.AI, q-bio.NC updates on arXiv.org
-
Interoperability in AI Safety Governance: Ethics, Regulations, and Standards
arXiv:2601.06153v1 Announce Type: cross Abstract: This policy report draws on country studies from China, South Korea, Singapore, and the United Kingdom to identify effective tools and key barriers to interoperability in AI safety governance. It offers practical recommendations to support a globally informed yet locally grounded governance ecosystem. Interoperability is a central goal of AI governance, vital for reducing risks, fostering innovation, enhancing competitiveness, promoting standard
Interoperability in AI Safety Governance: Ethics, Regulations, and Standards
-
cs.AI, q-bio.NC updates on arXiv.org
-
FairMedQA: Benchmarking Bias in Large Language Models for Medical Question Answering
arXiv:2505.19562v2 Announce Type: replace Abstract: Large language models (LLMs) are approaching expert-level performance in medical question answering (QA), demonstrating strong potential to improve public healthcare. However, underlying biases related to sensitive attributes such as sex and race pose life-critical risks. The extent to which such sensitive attributes affect diagnosis remains an open question and requires comprehensive empirical investigation. Additionally, even the latest Coun
FairMedQA: Benchmarking Bias in Large Language Models for Medical Question Answering
-
cs.AI, q-bio.NC updates on arXiv.org
-
Streamlining evidence based clinical recommendations with large language models
arXiv:2505.10282v2 Announce Type: replace-cross Abstract: Clinical evidence underpins informed healthcare decisions, yet integrating it into real-time practice remains challenging due to intensive workloads, complex procedures, and time constraints. This study presents Quicker, an LLM-powered system that automates evidence synthesis and generates clinical recommendations following standard guideline development workflows. Quicker delivers an end-to-end pipeline from clinical questions to recomm
Streamlining evidence based clinical recommendations with large language models
-
Omics In Lung
-
The role of PCMT1 in prognosis tumor immune microenvironment and therapeutic responses across cancers
Discov Oncol. 2026 Jan 5. doi: 10.1007/s12672-025-04366-2. Online ahead of print.ABSTRACTBACKGROUND: Emerging evidence highlights the overexpression of Protein-L-isoaspartate (D-aspartate) O-methyltransferase (PCMT1) in multiple malignancies. However, its pan-cancer prognostic significance, tumor immune microenvironment (TIME) interactions, and therapeutic implications remain underexplored.METHODS: Multi-omics data were integrated from UCSC Xena, GTEx, UALCAN, and published cohorts. PCMT1 expres
The role of PCMT1 in prognosis tumor immune microenvironment and therapeutic responses across cancers
Discov Oncol. 2026 Jan 5. doi: 10.1007/s12672-025-04366-2. Online ahead of print.
ABSTRACT
BACKGROUND: Emerging evidence highlights the overexpression of Protein-L-isoaspartate (D-aspartate) O-methyltransferase (PCMT1) in multiple malignancies. However, its pan-cancer prognostic significance, tumor immune microenvironment (TIME) interactions, and therapeutic implications remain underexplored.
METHODS: Multi-omics data were integrated from UCSC Xena, GTEx, UALCAN, and published cohorts. PCMT1 expression patterns were systematically analyzed across 33 cancer types. Associations between PCMT1 and clinical outcomes, immune infiltration, immune checkpoint genes (ICGs), tumor mutation burden (TMB), microsatellite instability (MSI), and drug sensitivity were evaluated using bioinformatics pipelines.
RESULTS: Our pan-cancer analysis revealed differential expression patterns of PCMT1 across various malignancies, with significant upregulation in 20 cancer types and downregulation in 3 cancer types. Notably, PCMT1 overexpression was predominantly observed in epithelial-origin tumors, such as ACC (adrenocortical carcinoma), BRCA (breast invasive carcinoma), COAD (colon adenocarcinoma), and LUAD (lung adenocarcinoma). Survival analysis demonstrated that elevated PCMT1 expression was significantly correlated with unfavorable prognosis in multiple epithelial tumors, particularly in BRCA, esophageal carcinoma (ESCA), head and neck squamous cell carcinoma (HNSC), liver hepatocellular carcinoma (LIHC), and mesothelioma (MESO). Furthermore, comprehensive analysis identified significant associations between PCMT1 expression and various tumor microenvironment features, including immune scores, six distinct immune cell types, four immunosuppressive cell populations, cancer-associated fibroblasts (CAFs)-related markers, and immunosuppressive factors. PCMT1 expression also showed significant correlations with tumor mutation burden (TMB), microsatellite instability (MSI), DNA stemness score (DNAss), and RNA stemness score (RNAss). Particularly noteworthy was the strong positive correlation between PCMT1 expression and CAFs infiltration, along with their associated factors. These findings were further validated in independent immunotherapy cohorts, where PCMT1 consistently demonstrated immunosuppressive characteristics.
CONCLUSION: Multi-omics analysis suggests that PCMT1 may serve as a potential prognostic biomarker and a novel immunotherapy target for pan-cancer.
PMID:41491065 | DOI:10.1007/s12672-025-04366-2
-
cs.AI, q-bio.NC updates on arXiv.org
-
Digital Twin AI: Opportunities and Challenges from Large Language Models to World Models
arXiv:2601.01321v1 Announce Type: new Abstract: Digital twins, as precise digital representations of physical systems, have evolved from passive simulation tools into intelligent and autonomous entities through the integration of artificial intelligence technologies. This paper presents a unified four-stage framework that systematically characterizes AI integration across the digital twin lifecycle, spanning modeling, mirroring, intervention, and autonomous management. By synthesizing existing
Digital Twin AI: Opportunities and Challenges from Large Language Models to World Models
-
cs.AI, q-bio.NC updates on arXiv.org
-
OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment
arXiv:2601.01576v1 Announce Type: cross Abstract: Evaluating novelty is critical yet challenging in peer review, as reviewers must assess submissions against a vast, rapidly evolving literature. This report presents OpenNovelty, an LLM-powered agentic system for transparent, evidence-based novelty analysis. The system operates through four phases: (1) extracting the core task and contribution claims to generate retrieval queries; (2) retrieving relevant prior work based on extracted queries via
OpenNovelty: An LLM-powered Agentic System for Verifiable Scholarly Novelty Assessment
-
cs.AI, q-bio.NC updates on arXiv.org
-
PathFound: An Agentic Multimodal Model Activating Evidence-seeking Pathological Diagnosis
arXiv:2512.23545v1 Announce Type: cross Abstract: Recent pathological foundation models have substantially advanced visual representation learning and multimodal interaction. However, most models still rely on a static inference paradigm in which whole-slide images are processed once to produce predictions, without reassessment or targeted evidence acquisition under ambiguous diagnoses. This contrasts with clinical diagnostic workflows that refine hypotheses through repeated slide observations
PathFound: An Agentic Multimodal Model Activating Evidence-seeking Pathological Diagnosis
-
npj Digital Medicine
-
Context matching is not reasoning when performing generalized clinical evaluation of generative language models
npj Digital Medicine, Published online: 27 December 2025; doi:10.1038/s41746-025-02253-2Context matching is not reasoning when performing generalized clinical evaluation of generative language models
Context matching is not reasoning when performing generalized clinical evaluation of generative language models
npj Digital Medicine, Published online: 27 December 2025; doi:10.1038/s41746-025-02253-2
Context matching is not reasoning when performing generalized clinical evaluation of generative language models-
npj Digital Medicine
-
A novel evaluation benchmark for medical LLMs illuminating safety and effectiveness in clinical domains
npj Digital Medicine, Published online: 26 December 2025; doi:10.1038/s41746-025-02277-8A novel evaluation benchmark for medical LLMs illuminating safety and effectiveness in clinical domains
A novel evaluation benchmark for medical LLMs illuminating safety and effectiveness in clinical domains
npj Digital Medicine, Published online: 26 December 2025; doi:10.1038/s41746-025-02277-8
A novel evaluation benchmark for medical LLMs illuminating safety and effectiveness in clinical domains-
cs.AI, q-bio.NC updates on arXiv.org
-
aiXiv: A Next-Generation Open Access Ecosystem for Scientific Discovery Generated by AI Scientists
arXiv:2508.15126v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled AI agents to autonomously generate scientific proposals, conduct experiments, author papers, and perform peer reviews. Yet this flood of AI-generated research content collides with a fragmented and largely closed publication ecosystem. Traditional journals and conferences rely on human peer review, making them difficult to scale and often reluctant to accept AI-generated research con
aiXiv: A Next-Generation Open Access Ecosystem for Scientific Discovery Generated by AI Scientists
-
cs.AI, q-bio.NC updates on arXiv.org
-
Why Text Prevails: Vision May Undermine Multimodal Medical Decision Making
arXiv:2512.13747v1 Announce Type: cross Abstract: With the rapid progress of large language models (LLMs), advanced multimodal large language models (MLLMs) have demonstrated impressive zero-shot capabilities on vision-language tasks. In the biomedical domain, however, even state-of-the-art MLLMs struggle with basic Medical Decision Making (MDM) tasks. We investigate this limitation using two challenging datasets: (1) three-stage Alzheimer's disease (AD) classification (normal, mild cognitive i
Why Text Prevails: Vision May Undermine Multimodal Medical Decision Making
-
Journal of Medical Internet Research
-
Automated Multitier Tagging of Chinese Online Health Education Resources Using a Large Language Model: Development and Validation Study
Background: Precision health promotion, which aims to tailor health messages to individual needs, is hampered by the lack of structured metadata in vast digital health resource libraries. This bottleneck prevents scalable, personalized content delivery and exacerbates information overload for the public. Objective: This study aimed to develop, deploy, and validate an automated tagging system using a large language model (LLM) to create the foundational metadata infrastructure required for tailor
Automated Multitier Tagging of Chinese Online Health Education Resources Using a Large Language Model: Development and Validation Study
-
(Multiomics OR Omics) AND (Lung OR gastric OR Hepatocellular)
-
Exploring the role of lipid metabolism genes in gastric cancer prognosis and tumor immune microenvironment
J Int Med Res. 2025 Dec;53(12):3000605251403252. doi: 10.1177/03000605251403252. Epub 2025 Dec 11.ABSTRACTBackgroundGastric cancer remains a major global health challenge due to its high mortality rate and complex pathophysiological mechanisms. Emerging evidence highlights that dysregulated lipid metabolism contributes to gastric cancer progression and prognosis, but the associations between lipid metabolism-associated genes, gastric cancer patient survival, and tumor immune microenvironment rem
Exploring the role of lipid metabolism genes in gastric cancer prognosis and tumor immune microenvironment
J Int Med Res. 2025 Dec;53(12):3000605251403252. doi: 10.1177/03000605251403252. Epub 2025 Dec 11.
ABSTRACT
BackgroundGastric cancer remains a major global health challenge due to its high mortality rate and complex pathophysiological mechanisms. Emerging evidence highlights that dysregulated lipid metabolism contributes to gastric cancer progression and prognosis, but the associations between lipid metabolism-associated genes, gastric cancer patient survival, and tumor immune microenvironment remodeling are not fully elucidated.MethodsWe analyzed publicly available omics and clinical data, including RNA sequencing data from 371 gastric cancer samples in The Cancer Genome Atlas database and 433 gastric cancer samples in the Gene Expression Omnibus database. We first curated the top 100 lipid metabolism-associated genes based on relevance scores. Then, univariate Cox regression was used to identify genes significantly associated with overall survival. Consensus clustering was applied to these survival-related genes to define gastric cancer molecular subtypes. Copy number variation analysis was performed to assess genomic alterations of these genes in tumor samples. A prognostic risk model was constructed using least absolute shrinkage and selection operator regression and validated via multivariate Cox regression. Immune infiltration analysis using CIBERSORT and ESTIMATE algorithms was conducted to explore associations between lipid metabolism-associated genes and tumor immune microenvironment characteristics.ResultsA total of 3911 differentially expressed genes were identified between gastric cancer and adjacent normal tissues. Among the top 100 lipid metabolism-associated genes, 43 were significantly linked to patient survival, most of which were considered as poor prognostic factors. Copy number variation analysis revealed frequent copy number gains of these genes in tumor samples. Consensus clustering stratified patients into two molecular subtypes (LMAGcluster A and LMAGcluster B), with LMAGcluster A showing significantly worse survival outcomes (median survival: 2.6 years vs. 8.3 years in LMAGcluster B, p < 0.001). LMAGcluster A was also characterized by elevated infiltration of pro-tumor immune cells, such as regulatory T cells and follicular helper T cells. The prognostic model based on 14 key lipid metabolism-associated genes exhibited robust predictive performance, with area under the receiver operating characteristic curve values of 0.702-0.761 in The Cancer Genome Atlas cohort and 0.621-0.638 in the Gene Expression Omnibus cohort for 1-, 3-, and 5-year survival.ConclusionLipid metabolism-associated genes are closely associated with gastric cancer prognosis and tumor immune microenvironment remodeling. The identified gene-based molecular subtypes and prognostic model provide novel insights into gastric cancer progression, and the 14 key genes may serve as potential biomarkers and therapeutic targets.
PMID:41381057 | DOI:10.1177/03000605251403252
-
cs.AI, q-bio.NC updates on arXiv.org
-
The AI Productivity Index (APEX)
arXiv:2509.25721v4 Announce Type: replace-cross Abstract: We present an extended version of the AI Productivity Index (APEX-v1-extended), a benchmark for assessing whether frontier models are capable of performing economically valuable tasks in four jobs: investment banking associate, management consultant, big law associate, and primary care physician (MD). This technical report details the extensions to APEX-v1, including an increase in the held-out evaluation set from n = 50 to n = 100 cases