Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
Sci-Reasoning: A Dataset Decoding AI Innovation Patterns
arXiv:2601.04577v1 Announce Type: new Abstract: While AI innovation accelerates rapidly, the intellectual process behind breakthroughs -- how researchers identify gaps, synthesize prior work, and generate insights -- remains poorly understood. The lack of structured data on scientific reasoning hinders systematic analysis and development of AI research agents. We introduce Sci-Reasoning, the first dataset capturing the intellectual synthesis behind high-quality AI research. Using community-vali
-
cs.AI, q-bio.NC updates on arXiv.org
-
Beyond Monolithic Architectures: A Multi-Agent Search and Knowledge Optimization Framework for Agentic Search
arXiv:2601.04703v1 Announce Type: new Abstract: Agentic search has emerged as a promising paradigm for complex information seeking by enabling Large Language Models (LLMs) to interleave reasoning with tool use. However, prevailing systems rely on monolithic agents that suffer from structural bottlenecks, including unconstrained reasoning outputs that inflate trajectories, sparse outcome-level rewards that complicate credit assignment, and stochastic search noise that destabilizes learning. To a
Beyond Monolithic Architectures: A Multi-Agent Search and Knowledge Optimization Framework for Agentic Search
-
cs.AI, q-bio.NC updates on arXiv.org
-
Decision-Aware Trust Signal Alignment for SOC Alert Triage
arXiv:2601.04486v1 Announce Type: cross Abstract: Detection systems that utilize machine learning are progressively implemented at Security Operations Centers (SOCs) to help an analyst to filter through high volumes of security alerts. Practically, such systems tend to reveal probabilistic results or confidence scores which are ill-calibrated and hard to read when under pressure. Qualitative and survey based studies of SOC practice done before reveal that poor alert quality and alert overload g
Decision-Aware Trust Signal Alignment for SOC Alert Triage
-
cs.AI, q-bio.NC updates on arXiv.org
-
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
arXiv:2601.04758v1 Announce Type: cross Abstract: The Patent Trial and Appeal Board (PTAB) of the USPTO adjudicates thousands of ex parte appeals each year, requiring the integration of technical understanding and legal reasoning. While large language models (LLMs) are increasingly applied in patent and legal practice, their use has remained limited to lightweight tasks, with no established means of systematically evaluating their capacity for structured legal reasoning in the patent domain. In
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
-
cs.AI, q-bio.NC updates on arXiv.org
-
Smart IoT-Based Wearable Device for Detection and Monitoring of Common Cow Diseases Using a Novel Machine Learning Technique
arXiv:2601.04761v1 Announce Type: cross Abstract: Manual observation and monitoring of individual cows for disease detection present significant challenges in large-scale farming operations, as the process is labor-intensive, time-consuming, and prone to reduced accuracy. The reliance on human observation often leads to delays in identifying symptoms, as the sheer number of animals can hinder timely attention to each cow. Consequently, the accuracy and precision of disease detection are signifi
Smart IoT-Based Wearable Device for Detection and Monitoring of Common Cow Diseases Using a Novel Machine Learning Technique
-
cs.AI, q-bio.NC updates on arXiv.org
-
Belief in Authority: Impact of Authority in Multi-Agent Evaluation Framework
arXiv:2601.04790v1 Announce Type: cross Abstract: Multi-agent systems utilizing large language models often assign authoritative roles to improve performance, yet the impact of authority bias on agent interactions remains underexplored. We present the first systematic analysis of role-based authority bias in free-form multi-agent evaluation using ChatEval. Applying French and Raven's power-based theory, we classify authoritative roles into legitimate, referent, and expert types and analyze thei
Belief in Authority: Impact of Authority in Multi-Agent Evaluation Framework
-
cs.AI, q-bio.NC updates on arXiv.org
-
Atlas 2 -- Foundation models for clinical deployment
arXiv:2601.05148v1 Announce Type: cross Abstract: Pathology foundation models substantially advanced the possibilities in computational pathology -- yet tradeoffs in terms of performance, robustness, and computational requirements remained, which limited their clinical deployment. In this report, we present Atlas 2, Atlas 2-B, and Atlas 2-S, three pathology vision foundation models which bridge these shortcomings by showing state-of-the-art performance in prediction performance, robustness, and
Atlas 2 -- Foundation models for clinical deployment
-
cs.AI, q-bio.NC updates on arXiv.org
-
PsychEval: A Multi-Session and Multi-Therapy Benchmark for High-Realism AI Psychological Counselor
arXiv:2601.01802v3 Announce Type: replace Abstract: To develop a reliable AI for psychological assessment, we introduce \texttt{PsychEval}, a multi-session, multi-therapy, and highly realistic benchmark designed to address three key challenges: \textbf{1) Can we train a highly realistic AI counselor?} Realistic counseling is a longitudinal task requiring sustained memory and dynamic goal tracking. We propose a multi-session benchmark (spanning 6-10 sessions across three distinct stages) that de
PsychEval: A Multi-Session and Multi-Therapy Benchmark for High-Realism AI Psychological Counselor
-
Journal of Medical Internet Research
-
Developing an AI-Assisted Tool That Identifies Patients With Multimorbidity and Complex Polypharmacy to Improve the Process of Medication Reviews: Qualitative Interview and Focus Group Study
Background: Structured medication reviews (SMRs) are an essential component of medication optimization, especially for patients with multimorbidity and polypharmacy. However, the process remains challenging due to the complexities of patient data, time constraints, and the need for coordination among health care professionals (HCPs). This study explores HCPs’ perspectives on the integration of artificial intelligence (AI)–assisted tools to enhance the SMR process, with a focus on the potential b
Developing an AI-Assisted Tool That Identifies Patients With Multimorbidity and Complex Polypharmacy to Improve the Process of Medication Reviews: Qualitative Interview and Focus Group Study
-
Journal of Medical Internet Research
-
Intervention in Health Misinformation Using Large Language Models for Automated Detection, Thematic Analysis, and Inoculation: Case Study on COVID-19
Background: The rapid growth of social media as an information channel has enabled the swift spread of inaccurate or false health information, significantly impacting public health. This widespread dissemination of misinformation has caused confusion, eroded trust in health authorities, led to noncompliance with health guidelines, and encouraged risky health behaviors. Understanding the dynamics of misinformation on social media is essential for devising effective public health communication str
Intervention in Health Misinformation Using Large Language Models for Automated Detection, Thematic Analysis, and Inoculation: Case Study on COVID-19
-
STAT

-
Medicaid restrictions may lead to a million missed cancer screenings over two years: study
In less than a year, new Medicaid eligibility restrictions may lead millions of people to lose coverage and then miss potentially lifesaving cancer screenings like colonoscopies or mammograms. A new analysis estimates that Americans may miss more than a million cancer screenings for colorectal, breast, or lung cancer over the two years after the new policy takes effect. “I see patients every day that come to me with cancer and are asymptomatic, but their life gets turned upside down because t
Medicaid restrictions may lead to a million missed cancer screenings over two years: study
In less than a year, new Medicaid eligibility restrictions may lead millions of people to lose coverage and then miss potentially lifesaving cancer screenings like colonoscopies or mammograms. A new analysis estimates that Americans may miss more than a million cancer screenings for colorectal, breast, or lung cancer over the two years after the new policy takes effect.
“I see patients every day that come to me with cancer and are asymptomatic, but their life gets turned upside down because they are told they have cancer,” said Adrian Diaz, a surgical oncologist at the University of Chicago and one of the authors on the paper, published Thursday in JAMA Oncology. “In a positive way, we catch it early. It’s potentially treatable, curable. Seeing that number, over a million patients, who will not have that opportunity — I was taken aback.”


© ASHRAF SHAZLY/AFP via Getty Images
-
Journal of Medical Internet Research
-
A Web-Based Cancer Prevention Intervention for Rural Emerging Adults: Mixed Methods Development and Pilot-Testing Study
Background: The rapid growth of user-generated web-based health information increases the complexity of cancer information seeking. One promising strategy for promoting high-quality cancer information consumption is through targeted interventions that are intentionally designed to reach individuals in the web-based spaces they occupy. However, there is a paucity of evidence-based information on the best strategies for designing and implementing web-based health behavior change interventions to i
A Web-Based Cancer Prevention Intervention for Rural Emerging Adults: Mixed Methods Development and Pilot-Testing Study
-
(Multiomics OR Omics) AND (Lung OR gastric OR Hepatocellular)
-
Leveraging Genetic Instrumental Variables and Sequencing Analysis to Identify a Prognostic Signature Based on Epithelial Cell Markers in Lung Adenocarcinoma
Thorac Cancer. 2026 Jan;17(1):e70244. doi: 10.1111/1759-7714.70244.ABSTRACTMAIN PROBLEM: The treatment and prognosis of lung adenocarcinoma (LUAD) remain challenging. The study aimed to identify prognostic genes and construct a prognostic model for LUAD.METHODS: After identifying malignant alveolar type II (AT2) cells using InferCNV, we applied CytoTRACE, pseudo-time analysis, Mendelian randomization (MR), and univariate Cox regression analysis to identify prognostic genes. A prognostic model wa
Leveraging Genetic Instrumental Variables and Sequencing Analysis to Identify a Prognostic Signature Based on Epithelial Cell Markers in Lung Adenocarcinoma
Thorac Cancer. 2026 Jan;17(1):e70244. doi: 10.1111/1759-7714.70244.
ABSTRACT
MAIN PROBLEM: The treatment and prognosis of lung adenocarcinoma (LUAD) remain challenging. The study aimed to identify prognostic genes and construct a prognostic model for LUAD.
METHODS: After identifying malignant alveolar type II (AT2) cells using InferCNV, we applied CytoTRACE, pseudo-time analysis, Mendelian randomization (MR), and univariate Cox regression analysis to identify prognostic genes. A prognostic model was then developed using an optimized subset of these genes, selected through the least absolute shrinkage and selection operator (LASSO) algorithm. Further analyses included Gene Ontology enrichment analysis and the construction of a protein-protein interaction (PPI) network.
RESULTS: Pseudo-time analysis identified 3526 dynamically expressed genes during malignant AT2 cell dedifferentiation. Subsequent multi-omics integration refined the gene selection, yielding four prognostic genes for the final predictive model. The resulting model achieved area under the receiver operating characteristic (ROC) curve (AUC) values of 0.649, 0.675, and 0.654 for predicting 1, 2, and 3-year overall survival (OS) in the training set, respectively, and was successfully validated in two external cohorts at the corresponding time points. Moreover, survival analysis demonstrated that patients in the high-risk group had significantly poorer OS than those in the low-risk group, both in the training set and the validation sets (p < 0.01).
CONCLUSIONS: The study developed a novel signature based on genes dynamically expressed during malignant AT2 cell dedifferentiation, capable of predicting the prognosis of LUAD patients, and offered four accurate prognostic biomarkers (ADM, MARK4, PARVA, and RPS6KA1).
PMID:41500831 | DOI:10.1111/1759-7714.70244
-
npj Digital Medicine
-
An autonomous agentic workflow for clinical detection of cognitive concerns using large language models
npj Digital Medicine, Published online: 07 January 2026; doi:10.1038/s41746-025-02324-4An autonomous agentic workflow for clinical detection of cognitive concerns using large language models
An autonomous agentic workflow for clinical detection of cognitive concerns using large language models
npj Digital Medicine, Published online: 07 January 2026; doi:10.1038/s41746-025-02324-4
An autonomous agentic workflow for clinical detection of cognitive concerns using large language models