❌

Normal view

WisPaper: Your AI Scholar Search Engine

arXiv:2512.06879v1 Announce Type: cross Abstract: Researchers struggle to efficiently locate and manage relevant literature within the exponentially growing body of scientific publications. We present \textsc{WisPaper}, an intelligent academic retrieval and literature management platform that addresses this challenge through three integrated capabilities: (1) \textit{Scholar Search}, featuring both quick keyword-based and deep agentic search modes for efficient paper discovery; (2) \textit{Library}, a customizable knowledge base for systematic literature organization; and (3) \textit{AI Feeds}, an intelligent recommendation system that automatically delivers relevant new publications based on user interests. Unlike existing academic tools, \textsc{WisPaper} provides a closed-loop workflow that seamlessly connects literature discovery, management, and continuous tracking of research frontiers. Our multilingual and multidisciplinary system significantly reduces the time researchers from diverse backgrounds spend on paper screening and management, enabling them to focus on their core research activities. The platform is publicly accessible and serves researchers across academia and industry.

AutoSurvey2: Empowering Researchers with Next Level Automated Literature Surveys

arXiv:2510.26012v3 Announce Type: replace Abstract: The rapid growth of research literature, particularly in large language models (LLMs), has made producing comprehensive and current survey papers increasingly difficult. This paper introduces autosurvey2, a multi-stage pipeline that automates survey generation through retrieval-augmented synthesis and structured evaluation. The system integrates parallel section generation, iterative refinement, and real-time retrieval of recent publications to ensure both topical completeness and factual accuracy. Quality is assessed using a multi-LLM evaluation framework that measures coverage, structure, and relevance in alignment with expert review standards. Experimental results demonstrate that autosurvey2 consistently outperforms existing retrieval-based and automated baselines, achieving higher scores in structural coherence and topical relevance while maintaining strong citation fidelity. By combining retrieval, reasoning, and automated evaluation into a unified framework, autosurvey2 provides a scalable and reproducible solution for generating long-form academic surveys and contributes a solid foundation for future research on automated scholarly writing. All code and resources are available at https://github.com/annihi1ation/auto_research.

Large Language Model Benchmarks in Medical Tasks

arXiv:2410.21348v3 Announce Type: replace-cross Abstract: With the increasing application of large language models (LLMs) in the medical domain, evaluating these models' performance using benchmark datasets has become crucial. This paper presents a comprehensive survey of various benchmark datasets employed in medical LLM tasks. These datasets span multiple modalities including text, image, and multimodal benchmarks, focusing on different aspects of medical knowledge such as electronic health records (EHRs), doctor-patient dialogues, medical question-answering, and medical image captioning. The survey categorizes the datasets by modality, discussing their significance, data structure, and impact on the development of LLMs for clinical tasks such as diagnosis, report generation, and predictive decision support. Key benchmarks include MIMIC-III, MIMIC-IV, BioASQ, PubMedQA, and CheXpert, which have facilitated advancements in tasks like medical report generation, clinical summarization, and synthetic data generation. The paper summarizes the challenges and opportunities in leveraging these benchmarks for advancing multimodal medical intelligence, emphasizing the need for datasets with a greater degree of language diversity, structured omics data, and innovative approaches to synthesis. This work also provides a foundation for future research in the application of LLMs in medicine, contributing to the evolving field of medical artificial intelligence.

Stereo-seq V2: Spatial mapping of total RNA on FFPE sections with high resolution

Stereo-seq V2 facilitates single-cell-resolution spatial RNA mapping in FFPE samples through random primer capture, uncovering ncRNAs, host-pathogen transcriptome profiling, and spatial immune repertoires in situ.

Stereo-seq V2: Spatial mapping of total RNA on FFPE sections with high resolution

Cell. 2025 Aug 22:S0092-8674(25)00922-5. doi: 10.1016/j.cell.2025.08.008. Online ahead of print.

ABSTRACT

Performing total RNA profiling on formalin-fixed, paraffin-embedded (FFPE) samples, the predominant sample conservation method in clinical practice, remains challenging for current spatial transcriptomics techniques. Here, we introduce Stereo-seq V2, which employs random primers to capture and sequence RNAs in situ on FFPE sections and provides single-cell resolution. The random-priming-based strategy offers unbiased transcript capturing and uniform gene body coverage, which increase the sensitivity to marker genes, the efficiency of non-polyadenylation (poly(A)) RNA profiling, and immune repertoire coverage. We demonstrated the robust performance of Stereo-seq V2 on clinical FFPE samples using triple-negative breast cancer (TNBC) sections and identified tumor-specific alternative splicing events. In a Mycobacterium tuberculosis (Mtb)-infected mouse model, we monitored gene expression dynamics of host and pathogen transcriptomes simultaneously by utilizing Stereo-seq V2. We also assembled immune repertoires and identified Mtb-specific BCR clones, which could also be observed in human tuberculous lung samples. These results highlight Stereo-seq V2's potential in biomedical research and personalized medicine.

PMID:40882628 | DOI:10.1016/j.cell.2025.08.008

Molecular characterization of breast cancer and multiple primary malignancies: the latest application using unmarked quantitative proteomics

Int J Surg. 2025 Jul 22. doi: 10.1097/JS9.0000000000002999. Online ahead of print.

ABSTRACT

BACKGROUND: Breast cancer remains the most prevalent malignancy among women, and patients presenting with both breast and lung cancer pose significant challenges in clinical diagnosis and treatment. Currently, comprehensive multi-omics analyses for such multiple malignancies are lacking.

METHODS: An integrated multi-omics analysis was performed, incorporating quantitative proteomics and radiomics data from patients with single primary breast cancer as well as those with multiple primary tumors (breast and lung cancer).

RESULTS: Quantitative proteomics analysis revealed four distinct molecular signatures (Types I-IV). Patients with single breast cancer exhibited driving pathways primarily linked to cell proliferation (e.g., HER2), whereas those with multiple breast cancers showed enrichment in ER-related and proliferative pathways. In contrast, patients with multiple lung cancers displayed pathways associated with immune response and immune escape. Additionally, immune subtyping identified three distinct immune landscapes (Types I-III). Radiomic analysis demonstrated strong correlations between these molecular/immune subtypes and imaging findings. Patients with high imaging information scores exhibited pronounced tumor heterogeneity and reduced immune infiltration.

CONCLUSIONS: This study provides new insights into the molecular pathogenesis of multiple primary malignancies, particularly breast and lung cancer.

PMID:40694032 | DOI:10.1097/JS9.0000000000002999

A multiomics dataset of paired CT image and plasma cell-free DNA end motif for patients with pulmonary nodules

Sci Data. 2025 Apr 1;12(1):545. doi: 10.1038/s41597-025-04912-1.

ABSTRACT

Diagnosing lung cancer at a curable stage offers the opportunity for a favorable prognosis. The emerging epigenomics analysis on plasma cell-free DNA (cfDNA), including 5-methylcytosine (5mC) and 5-hydroxymethylcytosine (5hmC) modifications, has acted as a promising approach facilitating the identification of lung cancer. And, integrating 5mC biomarker with chest computed tomography (CT) image features could optimize the diagnosis of lung cancer, exceeding the performance of models built on single feature. However, the clinical applicability of integrated markers might be limited by the potential risk of overfitting due to small sample size. Hence, we prospectively collected peripheral blood sample and the paired chest CT images of 2032 patients with indeterminate pulmonary nodules across 5 centers, and constructed a large-scale, multi-institutional, multiomics database that encompass CT imaging data and plasma cfDNA fragmentomic in 5mC-, 5hmC-enriched regions. To our best knowledge, this dataset is the first radio-epigenomic dataset with the largest sample size, and provides multi-dimensional insights for early diagnosis of lung cancer, facilitating the individuated management for lung cancer.

PMID:40169596 | PMC:PMC11961589 | DOI:10.1038/s41597-025-04912-1

Investigation of the Molecular Mechanism of Asthma in Meishan Pigs Using Multi-Omics Analysis

Animals (Basel). 2025 Jan 13;15(2):200. doi: 10.3390/ani15020200.

ABSTRACT

Asthma has been extensively studied in humans and animals, but the molecular mechanisms underlying asthma in Meishan pigs, a breed with distinct genetic and physiological characteristics, remain elusive. Understanding these mechanisms could provide insights into veterinary medicine and human asthma research. We investigated asthma pathogenesis in Meishan pigs through transcriptomic and metabolomic analyses of blood samples taken during autumn and winter. Asthma in Meishan pigs is related to inflammation, mitochondrial oxidative phosphorylation, and tricarboxylic acid (TCA) cycle disorders. Related genes include CXCL10, CCL8, CCL22, CCL21, OLR1, and ACKR1, while metabolites include succinic acid, riboflavin-5-phosphate, and fumaric acid. Transcriptomic sequencing was performed on panting and normal Meishan pigs, and differentially expressed genes underwent functional enrichment screening. Metabolomic analysis revealed differential metabolites and pathways between groups. Combined analyses indicated that lung inflammation is influenced by genetic, allergenic, and environmental factors disrupting oxidative phosphorylation in lung mitochondria, affecting the TCA cycle. Mitochondrial reactive oxygen species, glutathione S-transferases, arginase 1 and RORC in immune regulation, the Notch pathway, YPEL4 in cell proliferation, and MARCKS in airway mucus secretion play roles in asthma pathogenesis. This study highlights that many cytokines and signaling pathways contribute to asthma. Further studies are needed to elucidate their complex interactions.

PMID:39858200 | PMC:PMC11759154 | DOI:10.3390/ani15020200

Deep whole-genome analysis of 494 hepatocellular carcinomas

Nature, Published online: 14 February 2024; doi:10.1038/s41586-024-07054-3

The Chinese Liver Cancer Atlas project depicts a panoramic genomic landscape of hepatocellular carcinoma, covering candidate coding and non-coding drivers, mutational signatures, extrachromosomal circular DNA, subclonal catastrophic events and detailed evolutionary history.

Novel DNA methylation biomarkers in stool and blood for early detection of colorectal cancer and precancerous lesions

Early detection and prevention of precancerous lesions can significantly reduce the morbidity and mortality of colorectal cancer (CRC). Here, we developed new candidate CpG site biomarkers for CRC and evaluate...
❌