Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
multiMentalRoBERTa: A Fine-tuned Multiclass Classifier for Mental Health Disorder
arXiv:2511.04698v1 Announce Type: cross Abstract: The early detection of mental health disorders from social media text is critical for enabling timely support, risk assessment, and referral to appropriate resources. This work introduces multiMentalRoBERTa, a fine-tuned RoBERTa model designed for multiclass classification of common mental health conditions, including stress, anxiety, depression, post-traumatic stress disorder (PTSD), suicidal ideation, and neutral discourse. Drawing on multiple
-
cs.AI, q-bio.NC updates on arXiv.org
-
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
arXiv:2510.22780v2 Announce Type: replace Abstract: AI agents are continually optimized for tasks related to human work, such as software engineering and professional writing, signaling a pressing trend with significant impacts on the human workforce. However, these agent developments have often not been grounded in a clear understanding of how humans execute work, to reveal what expertise agents possess and the roles they can play in diverse workflows. In this work, we study how agents do huma
How Do AI Agents Do Human Work? Comparing AI and Human Workflows Across Diverse Occupations
-
cs.AI, q-bio.NC updates on arXiv.org
-
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG
arXiv:2412.05447v3 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) is one of the leading and most widely used techniques for enhancing LLM retrieval capabilities, but it still faces significant limitations in commercial use cases. RAG primarily relies on the query-chunk text-to-text similarity in the embedding space for retrieval and can fail to capture deeper semantic relationships across chunks, is highly sensitive to chunking strategies, and is prone to hallucinat
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG
-
Nature Medicine
-
Author Correction: Global burden of chikungunya virus infections and the potential benefit of vaccination campaigns
Nature Medicine, Published online: 10 November 2025; doi:10.1038/s41591-025-04065-zAuthor Correction: Global burden of chikungunya virus infections and the potential benefit of vaccination campaigns
Author Correction: Global burden of chikungunya virus infections and the potential benefit of vaccination campaigns
Nature Medicine, Published online: 10 November 2025; doi:10.1038/s41591-025-04065-z
Author Correction: Global burden of chikungunya virus infections and the potential benefit of vaccination campaigns-
cs.AI, q-bio.NC updates on arXiv.org
-
A Mega-Study of Digital Twins Reveals Strengths, Weaknesses and Opportunities for Further Improvement
arXiv:2509.19088v3 Announce Type: replace-cross Abstract: Digital representations of individuals ("digital twins") promise to transform social science and decision-making. Yet it remains unclear whether such twins truly mirror the people they emulate. We conducted 19 preregistered studies with a representative U.S. panel and their digital twins, each constructed from rich individual-level data, enabling direct comparisons between human and twin behavior across a wide range of domains and stimul
A Mega-Study of Digital Twins Reveals Strengths, Weaknesses and Opportunities for Further Improvement
-
TechCrunch
-
VC Jennifer Neundorfer explains how founders can stand out in a crowded AI market
January Ventures co-founder Jennifer Neundorfer discussed this AI-driven funding market on the Equity podcast during TechCrunch Disrupt.
VC Jennifer Neundorfer explains how founders can stand out in a crowded AI market
-
cs.AI, q-bio.NC updates on arXiv.org
-
Fine-Tuning MedGemma for Clinical Captioning to Enhance Multimodal RAG over Malaysia CPGs
arXiv:2510.15418v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation systems are essential for providing fact-based guidance from Malaysian Clinical Practice Guidelines. However, their effectiveness with image-based queries is limited, as general Vision-Language Model captions often lack clinical specificity and factual grounding. This study proposes and validates a framework to specialize the MedGemma model for generating high-fidelity captions that serve as superior querie
Fine-Tuning MedGemma for Clinical Captioning to Enhance Multimodal RAG over Malaysia CPGs
-
Nature Biotechnology - Issue - nature.com science feeds
-
Publisher Correction: Deep-learning-based virtual screening of antibacterial compounds
Nature Biotechnology, Published online: 07 November 2025; doi:10.1038/s41587-025-02941-0Publisher Correction: Deep-learning-based virtual screening of antibacterial compounds
Publisher Correction: Deep-learning-based virtual screening of antibacterial compounds
Nature Biotechnology, Published online: 07 November 2025; doi:10.1038/s41587-025-02941-0
Publisher Correction: Deep-learning-based virtual screening of antibacterial compounds-
cs.AI, q-bio.NC updates on arXiv.org
-
No-Human in the Loop: Agentic Evaluation at Scale for Recommendation
arXiv:2511.03051v1 Announce Type: new Abstract: Evaluating large language models (LLMs) as judges is increasingly critical for building scalable and trustworthy evaluation pipelines. We present ScalingEval, a large-scale benchmarking study that systematically compares 36 LLMs, including GPT, Gemini, Claude, and Llama, across multiple product categories using a consensus-driven evaluation protocol. Our multi-agent framework aggregates pattern audits and issue codes into ground-truth labels via s
No-Human in the Loop: Agentic Evaluation at Scale for Recommendation
-
cs.AI, q-bio.NC updates on arXiv.org
-
Explaining Decisions in ML Models: a Parameterized Complexity Analysis (Part I)
arXiv:2511.03545v1 Announce Type: new Abstract: This paper presents a comprehensive theoretical investigation into the parameterized complexity of explanation problems in various machine learning (ML) models. Contrary to the prevalent black-box perception, our study focuses on models with transparent internal mechanisms. We address two principal types of explanation problems: abductive and contrastive, both in their local and global variants. Our analysis encompasses diverse ML models, includin
Explaining Decisions in ML Models: a Parameterized Complexity Analysis (Part I)
-
cs.AI, q-bio.NC updates on arXiv.org
-
Digital Transformation Chatbot (DTchatbot): Integrating Large Language Model-based Chatbot in Acquiring Digital Transformation Needs
arXiv:2511.02842v1 Announce Type: cross Abstract: Many organisations pursue digital transformation to enhance operational efficiency, reduce manual efforts, and optimise processes by automation and digital tools. To achieve this, a comprehensive understanding of their unique needs is required. However, traditional methods, such as expert interviews, while effective, face several challenges, including scheduling conflicts, resource constraints, inconsistency, etc. To tackle these issues, we inve
Digital Transformation Chatbot (DTchatbot): Integrating Large Language Model-based Chatbot in Acquiring Digital Transformation Needs
-
cs.AI, q-bio.NC updates on arXiv.org
-
Generative Artificial Intelligence in Bioinformatics: A Systematic Review of Models, Applications, and Methodological Advances
arXiv:2511.03354v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI) has become a transformative approach in bioinformatics that often enables advancements in genomics, proteomics, transcriptomics, structural biology, and drug discovery. To systematically identify and evaluate these growing developments, this review proposed six research questions (RQs), according to the preferred reporting items for systematic reviews and meta-analysis methods. The objective is to evalu
Generative Artificial Intelligence in Bioinformatics: A Systematic Review of Models, Applications, and Methodological Advances
-
cs.AI, q-bio.NC updates on arXiv.org
-
REFA: Reference Free Alignment for multi-preference optimization
arXiv:2412.16378v4 Announce Type: replace-cross Abstract: To mitigate reward hacking from response verbosity, modern preference optimization methods are increasingly adopting length normalization (e.g., SimPO, ORPO, LN-DPO). While effective against this bias, we demonstrate that length normalization itself introduces a failure mode: the URSLA shortcut. Here models learn to satisfy the alignment objective by prematurely truncating low-quality responses rather than learning from their semantic co
REFA: Reference Free Alignment for multi-preference optimization
-
Nature Biotechnology - Issue - nature.com science feeds
-
Drugmakers share data to feed voracious foundation models
Nature Biotechnology, Published online: 06 November 2025; doi:10.1038/s41587-025-02901-8Big pharma shares its machine learning models with biotechs, but awaits definitive data on success of artificial intelligence-generated drugs.
Drugmakers share data to feed voracious foundation models
Nature Biotechnology, Published online: 06 November 2025; doi:10.1038/s41587-025-02901-8
Big pharma shares its machine learning models with biotechs, but awaits definitive data on success of artificial intelligence-generated drugs.-
Nature Biotechnology - Issue - nature.com science feeds
-
Site-specific DNA insertion into the human genome with engineered recombinases
Nature Biotechnology, Published online: 06 November 2025; doi:10.1038/s41587-025-02895-3Engineered DNA recombinases efficiently and specifically insert genetic cargos without the use of landing pads.
Site-specific DNA insertion into the human genome with engineered recombinases
Nature Biotechnology, Published online: 06 November 2025; doi:10.1038/s41587-025-02895-3
Engineered DNA recombinases efficiently and specifically insert genetic cargos without the use of landing pads.-
STAT

-
STAT+: What’s FDA plotting for therapy chatbot regulation?
You’re reading the web edition of STAT’s Health Tech newsletter, our guide to how technology is transforming the life sciences. Sign up to get it delivered in your inbox every Tuesday and Thursday. What to know about the FDA’s therapy bots meeting The Food and Drug Administration is considering whether and how to regulate therapy chatbots that are based on large language models. Today, the agency’s Digital Health Advisory Committee is meeting to consider the topic. In a new story, I explai
STAT+: What’s FDA plotting for therapy chatbot regulation?
You’re reading the web edition of STAT’s Health Tech newsletter, our guide to how technology is transforming the life sciences. Sign up to get it delivered in your inbox every Tuesday and Thursday.
What to know about the FDA’s therapy bots meeting
The Food and Drug Administration is considering whether and how to regulate therapy chatbots that are based on large language models. Today, the agency’s Digital Health Advisory Committee is meeting to consider the topic. In a new story, I explain what’s going on, including some fresh insider intel.
The FDA wants to provide more clarity to developers of generative AI medical devices about what needs regulatory green light and how to get it. The agency is also also worried about LLM-based therapy bots that can provide unpredictable outputs. Regulators are aware about the growing concerns around general purpose bots like ChatGPT, which have been linked to delusions and allegedly to suicides.
Continue to STAT+ to read the full story…


© Sarah Silbiger/Getty Images
-
npj Digital Medicine
-
Improving dataset transparency in dermatologic Artificial Intelligence using a dataset nutrition label
npj Digital Medicine, Published online: 05 November 2025; doi:10.1038/s41746-025-02125-9Biased and poorly documented dermatology datasets pose risks to the development of safe and generalizable artificial intelligence (AI) tools. We created a Dataset Nutrition Label (DNL) for multiple dermatology datasets to support transparent and responsible data use. The DNL offers a structured, digestible summary of key attributes, including metadata, limitations, and risks, enabling data users to better ass
Improving dataset transparency in dermatologic Artificial Intelligence using a dataset nutrition label
npj Digital Medicine, Published online: 05 November 2025; doi:10.1038/s41746-025-02125-9
Biased and poorly documented dermatology datasets pose risks to the development of safe and generalizable artificial intelligence (AI) tools. We created a Dataset Nutrition Label (DNL) for multiple dermatology datasets to support transparent and responsible data use. The DNL offers a structured, digestible summary of key attributes, including metadata, limitations, and risks, enabling data users to better assess suitability and proactively address potential sources of bias in datasets.-
npj Digital Medicine
-
Evaluating clinical AI summaries with large language models as judges
npj Digital Medicine, Published online: 05 November 2025; doi:10.1038/s41746-025-02005-2Evaluating clinical AI summaries with large language models as judges
Evaluating clinical AI summaries with large language models as judges
npj Digital Medicine, Published online: 05 November 2025; doi:10.1038/s41746-025-02005-2
Evaluating clinical AI summaries with large language models as judges-
MRD
-
Liquid biopsy in gastrointestinal oncology: clinical applications and translational integration of ctDNA, CTCs, and sEVs
Oncol Rev. 2025 Oct 20;19:1702932. doi: 10.3389/or.2025.1702932. eCollection 2025.ABSTRACTBACKGROUND AND AIMS: Liquid biopsy offers a minimally invasive tool to detect actionable mutations, monitor minimal residual disease (MRD), and guide therapy in gastrointestinal (GI) cancers. We critically review the clinical utility of circulating tumor DNA (ctDNA), circulating tumor cells (CTCs), and small extracellular vesicles (sEVs) across GI malignancies and propose a framework for their integration i
Liquid biopsy in gastrointestinal oncology: clinical applications and translational integration of ctDNA, CTCs, and sEVs
Oncol Rev. 2025 Oct 20;19:1702932. doi: 10.3389/or.2025.1702932. eCollection 2025.
ABSTRACT
BACKGROUND AND AIMS: Liquid biopsy offers a minimally invasive tool to detect actionable mutations, monitor minimal residual disease (MRD), and guide therapy in gastrointestinal (GI) cancers. We critically review the clinical utility of circulating tumor DNA (ctDNA), circulating tumor cells (CTCs), and small extracellular vesicles (sEVs) across GI malignancies and propose a framework for their integration into clinical practice.
METHODS: We synthesized evidence from over 200 studies, including prospective trials and translational research, to assess diagnostic accuracy, prognostic value, and clinical actionability of each biomarker type in esophageal, gastric, colorectal, pancreatic, hepatocellular, and biliary cancers.
RESULTS: ctDNA has shown strong potential for MRD detection and treatment monitoring, particularly in colorectal and pancreatic cancer. CTCs offer insights into metastatic risk and therapeutic resistance, while sEVs provide molecular cargo relevant to immunomodulation and disease progression. Emerging microfluidics and AI-driven multi-omics approaches may overcome current limitations.
CONCLUSION: The integration of liquid biopsy technologies into GI oncology holds promise for early detection and precision therapy. We propose a five-phase clinical roadmap and outine the key research gaps that need to be addressed before widespread implementation in routine care.
PMID:41190015 | PMC:PMC12580207 | DOI:10.3389/or.2025.1702932
-
Journal of Medical Internet Research
-
Combining International Standards to Develop Clinical Decision Support for Parent Smoking Cessation in Pediatrics
Smoking has severe health consequences, and secondhand smoke (SHS) exposure among children increases the risk of sudden infant death syndrome, chronic respiratory diseases, such as asthma, and lung cancer in adulthood. For many parents, pediatricians are the primary source of interaction with the healthcare system. Nevertheless, in pediatric settings, appropriate tobacco treatments are rarely, if ever, provided to parents who smoke. To best address tobacco use among parents, it is ideal to devel