❌

Reading view

Multi-agent Self-triage System with Medical Flowcharts

arXiv:2511.12439v2 Announce Type: replace Abstract: Online health resources and large language models (LLMs) are increasingly used as a first point of contact for medical decision-making, yet their reliability in healthcare remains limited by low accuracy, lack of transparency, and susceptibility to unverified information. We introduce a proof-of-concept conversational self-triage system that guides LLMs with 100 clinically validated flowcharts from the American Medical Association, providing a structured and auditable framework for patient decision support. The system leverages a multi-agent framework consisting of a retrieval agent, a decision agent, and a chat agent to identify the most relevant flowchart, interpret patient responses, and deliver personalized, patient-friendly recommendations, respectively. Performance was evaluated at scale using synthetic datasets of simulated conversations. The system achieved 95.29% top-3 accuracy in flowchart retrieval (N=2,000) and 99.10% accuracy in flowchart navigation across varied conversational styles and conditions (N=37,200). By combining the flexibility of free-text interaction with the rigor of standardized clinical protocols, this approach demonstrates the feasibility of transparent, accurate, and generalizable AI-assisted self-triage, with potential to support informed patient decision-making while improving healthcare resource utilization.
  •  
  •  

AutoSurvey2: Empowering Researchers with Next Level Automated Literature Surveys

arXiv:2510.26012v3 Announce Type: replace Abstract: The rapid growth of research literature, particularly in large language models (LLMs), has made producing comprehensive and current survey papers increasingly difficult. This paper introduces autosurvey2, a multi-stage pipeline that automates survey generation through retrieval-augmented synthesis and structured evaluation. The system integrates parallel section generation, iterative refinement, and real-time retrieval of recent publications to ensure both topical completeness and factual accuracy. Quality is assessed using a multi-LLM evaluation framework that measures coverage, structure, and relevance in alignment with expert review standards. Experimental results demonstrate that autosurvey2 consistently outperforms existing retrieval-based and automated baselines, achieving higher scores in structural coherence and topical relevance while maintaining strong citation fidelity. By combining retrieval, reasoning, and automated evaluation into a unified framework, autosurvey2 provides a scalable and reproducible solution for generating long-form academic surveys and contributes a solid foundation for future research on automated scholarly writing. All code and resources are available at https://github.com/annihi1ation/auto_research.
  •  

Multi-agent Self-triage System with Medical Flowcharts

arXiv:2511.12439v1 Announce Type: new Abstract: Online health resources and large language models (LLMs) are increasingly used as a first point of contact for medical decision-making, yet their reliability in healthcare remains limited by low accuracy, lack of transparency, and susceptibility to unverified information. We introduce a proof-of-concept conversational self-triage system that guides LLMs with 100 clinically validated flowcharts from the American Medical Association, providing a structured and auditable framework for patient decision support. The system leverages a multi-agent framework consisting of a retrieval agent, a decision agent, and a chat agent to identify the most relevant flowchart, interpret patient responses, and deliver personalized, patient-friendly recommendations, respectively. Performance was evaluated at scale using synthetic datasets of simulated conversations. The system achieved 95.29% top-3 accuracy in flowchart retrieval (N=2,000) and 99.10% accuracy in flowchart navigation across varied conversational styles and conditions (N=37,200). By combining the flexibility of free-text interaction with the rigor of standardized clinical protocols, this approach demonstrates the feasibility of transparent, accurate, and generalizable AI-assisted self-triage, with potential to support informed patient decision-making while improving healthcare resource utilization.
  •  

Embracing the Future of Medical Education With Large Language Model–Based Virtual Patients: Scoping Review

Background: In recent years, large language models (LLMs) have experienced rapid development. LLM-based virtual patients have begun to gain attention, offering new opportunities for simulations in medical education. Objective: This study aims to systematically analyze the current applications, research trends, and challenges of LLM-based virtual patients in medical education and to explore potential future directions for development. Methods: This study adheres to the PRISMA-ScR (Preferred Reporting Items for Systematic Reviews and Meta-Analyses extension for Scoping Reviews) guidelines. Five databases (Web of Science Core Collection, PubMed, IEEE Xplore, Embase, and Scopus) were searched from January 1, 2018, to June 24, 2025, to identify studies related to the application of LLM-based virtual patients in medical education. A comprehensive analysis of LLM-based virtual patients from research design to application and evaluation was conducted. Results: A total of 28 studies were included in this scoping review. Analysis revealed that 92.9% (26/28) of the studies were published in the past 2 years, indicating that LLM-based virtual patient research is still in its early stages. The research primarily focuses on medical training and spans a wide range of medical disciplines. When using LLMs, advanced technologies such as social robots, virtual reality, and mixed reality are used to present LLM-based virtual patients. Combining these technologies with various supplementary tools enhances the realism of LLM-based virtual patients and improves user interaction. The evaluation of LLM-based virtual patients mainly emphasizes user experience. However, evaluation methods lack standardization, and only 13% (3/23) of studies used validated tools in assessing LLM-based virtual patients, while only 21.7% (5/23) of studies objectively measured learning outcomes facilitated by LLM-based virtual patients. All included studies expressed a positive attitude toward LLM-based virtual patients; however, they overlook privacy and security considerations in practical applications. Conclusions: LLM-based virtual patients hold significant innovation potential in medical education and are still in the early stages of development. They are primarily applied in medical training and show promise in communication skills training, although they cannot replace real-world interactions. Moreover, the heterogeneity of research designs, the absence of nonverbal cues in interactions, and concerns regarding privacy and security limit their broader implementation. Future research should focus on improving the reliability, realism, safety, and scientific efficacy of LLM-based virtual patients. Trial Registration: Open Science Framework Registries 10.17605/OSF.IO/DMC9Q; https://osf.io/DMC9Q/overview
  •  

MemeArena: Automating Context-Aware Unbiased Evaluation of Harmfulness Understanding for Multimodal Large Language Models

arXiv:2510.27196v1 Announce Type: cross Abstract: The proliferation of memes on social media necessitates the capabilities of multimodal Large Language Models (mLLMs) to effectively understand multimodal harmfulness. Existing evaluation approaches predominantly focus on mLLMs' detection accuracy for binary classification tasks, which often fail to reflect the in-depth interpretive nuance of harmfulness across diverse contexts. In this paper, we propose MemeArena, an agent-based arena-style evaluation framework that provides a context-aware and unbiased assessment for mLLMs' understanding of multimodal harmfulness. Specifically, MemeArena simulates diverse interpretive contexts to formulate evaluation tasks that elicit perspective-specific analyses from mLLMs. By integrating varied viewpoints and reaching consensus among evaluators, it enables fair and unbiased comparisons of mLLMs' abilities to interpret multimodal harmfulness. Extensive experiments demonstrate that our framework effectively reduces the evaluation biases of judge agents, with judgment results closely aligning with human preferences, offering valuable insights into reliable and comprehensive mLLM evaluations in multimodal harmfulness understanding. Our code and data are publicly available at https://github.com/Lbotirx/MemeArena.
  •  

Cell-free epigenomes enhanced fragmentomics-based model for early detection of lung cancer

Clin Transl Med. 2025 Feb;15(2):e70225. doi: 10.1002/ctm2.70225.

ABSTRACT

BACKGROUND: Lung cancer is a leading cause of cancer mortality, highlighting the need for innovative non-invasive early detection methods. Although cell-free DNA (cfDNA) analysis shows promise, its sensitivity in early-stage lung cancer patients remains a challenge. This study aimed to integrate insights from epigenetic modifications and fragmentomic features of cfDNA using machine learning to develop a more accurate lung cancer detection model.

METHODS: To address this issue, a multi-centre prospective cohort study was conducted, with participants harbouring suspicious malignant lung nodules and healthy volunteers recruited from two clinical centres. Plasma cfDNA was analysed for its epigenetic and fragmentomic profiles using chromatin immunoprecipitation sequencing, reduced representation bisulphite sequencing and low-pass whole-genome sequencing. Machine learning algorithms were then employed to integrate the multi-omics data, aiding in the development of a precise lung cancer detection model.

RESULTS: Cancer-related changes in cfDNA fragmentomics were significantly enriched in specific genes marked by cell-free epigenomes. A total of 609 genes were identified, and the corresponding cfDNA fragmentomic features were utilised to construct the ensemble model. This model achieved a sensitivity of 90.4% and a specificity of 83.1%, with an AUC of 0.94 in the independent validation set. Notably, the model demonstrated exceptional sensitivity for stage I lung cancer cases, achieving 95.1%. It also showed remarkable performance in detecting minimally invasive adenocarcinoma, with a sensitivity of 96.2%, highlighting its potential for early detection in clinical settings.

CONCLUSIONS: With feature selection guided by multiple epigenetic sequencing approaches, the cfDNA fragmentomics-based machine learning model demonstrated outstanding performance in the independent validation cohort. These findings highlight its potential as an effective non-invasive strategy for the early detection of lung cancer.

KEYPOINTS: Our study elucidated the regulatory relationships between epigenetic modifications and their effects on fragmentomic features. Identifying epigenetically regulated genes provided a critical foundation for developing the cfDNA fragmentomics-based machine learning model. The model demonstrated exceptional clinical performance, highlighting its substantial potential for translational application in clinical practice.

PMID:39909829 | PMC:PMC11798665 | DOI:10.1002/ctm2.70225

  •  

High-resolution spatially resolved proteomics of complex tissues based on microfluidics and transfer learning

PLATO, a high-resolution and high-throughput spatial mass spectrometry proteomics platform, identifies distinct tumor subtypes and key dysregulated proteins in human breast cancer.
  •  

Investigation of the Molecular Mechanism of Asthma in Meishan Pigs Using Multi-Omics Analysis

Animals (Basel). 2025 Jan 13;15(2):200. doi: 10.3390/ani15020200.

ABSTRACT

Asthma has been extensively studied in humans and animals, but the molecular mechanisms underlying asthma in Meishan pigs, a breed with distinct genetic and physiological characteristics, remain elusive. Understanding these mechanisms could provide insights into veterinary medicine and human asthma research. We investigated asthma pathogenesis in Meishan pigs through transcriptomic and metabolomic analyses of blood samples taken during autumn and winter. Asthma in Meishan pigs is related to inflammation, mitochondrial oxidative phosphorylation, and tricarboxylic acid (TCA) cycle disorders. Related genes include CXCL10, CCL8, CCL22, CCL21, OLR1, and ACKR1, while metabolites include succinic acid, riboflavin-5-phosphate, and fumaric acid. Transcriptomic sequencing was performed on panting and normal Meishan pigs, and differentially expressed genes underwent functional enrichment screening. Metabolomic analysis revealed differential metabolites and pathways between groups. Combined analyses indicated that lung inflammation is influenced by genetic, allergenic, and environmental factors disrupting oxidative phosphorylation in lung mitochondria, affecting the TCA cycle. Mitochondrial reactive oxygen species, glutathione S-transferases, arginase 1 and RORC in immune regulation, the Notch pathway, YPEL4 in cell proliferation, and MARCKS in airway mucus secretion play roles in asthma pathogenesis. This study highlights that many cytokines and signaling pathways contribute to asthma. Further studies are needed to elucidate their complex interactions.

PMID:39858200 | PMC:PMC11759154 | DOI:10.3390/ani15020200

  •  

Identifying specific functional roles for senescence across cell types

A dual recombinase-mediated genetic system for cell-type-specific lineage tracing, ablation, and gene manipulation of senescent cells reveals distinct roles of senescence across cell types.
  •  

Tumour vasculature at single-cell resolution

Nature, Published online: 10 July 2024; doi:10.1038/s41586-024-07698-1

An atlas of tumour vasculature shows that tumour angiogenesis is initiated from venous endothelial cells and extended towards arterial endothelial cells.
  •  

Hypoxia-induced epigenetic regulation of miR-485-3p promotes stemness and chemoresistance in pancreatic ductal adenocarcinoma via SLC7A11-mediated ferroptosis

Cell Death Discovery, Published online: 29 May 2024; doi:10.1038/s41420-024-02035-x

Hypoxia-induced epigenetic regulation of miR-485-3p promotes stemness and chemoresistance in pancreatic ductal adenocarcinoma via SLC7A11-mediated ferroptosis
  •  
  •  
❌