Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
GlobalDentBench: A Multinational Benchmark for Evaluating LLM Clinical Reasoning in Dentistry with Expert Calibration
arXiv:2605.24636v2 Announce Type: new Abstract: While large language models (LLMs) hold transformative potential for medicine, their reasoning robustness and safety in real-world clinical scenarios remain critically underexplored, particularly in dentistry. Here we introduce GlobalDentBench, the first multinational dental benchmark, featuring a taxonomy that encompasses 14 dental specialties across 88 countries and regions spanning six continents. The benchmark comprises 8,978 expert-validated
-
cs.AI, q-bio.NC updates on arXiv.org
-
DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs
arXiv:2605.25188v1 Announce Type: new Abstract: Multi-agent LLM systems improve reasoning by combining outputs from multiple agents, but interaction-heavy methods can introduce error propagation and high communication overhead. When agents exchange raw responses or reasoning traces, incorrect intermediate reasoning may be adopted and amplified, leading to confident but wrong consensus; multi-round communication also increases token consumption, latency, and inference cost. In this paper, we pro
DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs
-
cs.AI, q-bio.NC updates on arXiv.org
-
FrontierOR: Benchmarking LLMs' Capacity for Efficient Algorithm Design in Large-Scale Optimization
arXiv:2605.25246v2 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for optimization modeling and solver-code generation, yet practical operations research and optimization problems often require a harder capability: designing scalable algorithms that exploit problem structure and outperform direct formulation-and-solve baselines. Existing benchmarks are limited to small or simplified examples far below real-world scale and complexity. We introduce FrontierOR, amo
FrontierOR: Benchmarking LLMs' Capacity for Efficient Algorithm Design in Large-Scale Optimization
-
cs.AI, q-bio.NC updates on arXiv.org
-
Towards end-to-end LLM-based censoring-aware survival analysis
arXiv:2605.25399v1 Announce Type: new Abstract: Objective: Survival analysis is central to medical prediction, yet large language models (LLMs) are rarely used as end-to-end survival models because censoring prevents straightforward supervised fine-tuning. Here we present LLMSurvival, a framework that enables censoring-aware survival analysis with unmodified LLMs operating directly on tabular clinical data. Materials and Methods: LLMSurvival reformulates time-to-event prediction as pairwise r
Towards end-to-end LLM-based censoring-aware survival analysis
-
cs.AI, q-bio.NC updates on arXiv.org
-
AgentHijack: Benchmarking Computer Use Agent Robustness to Common Environment Corruptions
arXiv:2605.25707v1 Announce Type: new Abstract: Autonomous computer use agents that powered by multimodal large language models (MLLMs) are emerging as capable assistants for completing complex digital workflows. However, real-world execution environments are far from ideal: pop-ups, resolution changes, and competing applications frequently interfere with agent perception and control. We introduce AgentHijack, a benchmark designed to evaluate the robustness of computer-use agents under common c
AgentHijack: Benchmarking Computer Use Agent Robustness to Common Environment Corruptions
-
cs.AI, q-bio.NC updates on arXiv.org
-
Agent Learning via Early Experience
arXiv:2510.08558v3 Announce Type: replace Abstract: A long-term goal of language agents is to learn and improve through their own experience, ultimately outperforming humans in complex, real-world tasks. However, training agents from experience data with reinforcement learning remains difficult in many environments, which either lack verifiable rewards (e.g., websites) or require inefficient long-horizon rollouts (e.g., multi-turn tool use). As a result, most current agents rely on supervised f
Agent Learning via Early Experience
-
cs.AI, q-bio.NC updates on arXiv.org
-
Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
arXiv:2602.10090v3 Announce Type: replace Abstract: Recent advances in large language model (LLM) have empowered autonomous agents to perform multi-turn interactions with tools and environments. However, scaling such agent training is limited by the lack of diverse and reliable environments. In this paper, we propose Agent World Model (AWM), a fully synthetic environment generation pipeline. Using this pipeline, we scale to 1,000 environments covering everyday scenarios, in which agents can int
Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
-
cs.AI, q-bio.NC updates on arXiv.org
-
Reliable AI Needs to Externalize Implicit Knowledge: A Human-AI Collaboration Perspective
arXiv:2605.02010v2 Announce Type: replace Abstract: This position paper argues that reliable AI requires infrastructure for human validation of implicit knowledge. AI learns from both explicit knowledge (papers, documentation, structured databases) and implicit knowledge (reasoning patterns, debugging processes, intermediate steps). Implicit knowledge remains unexternalized because documentation cost exceeds perceived value -- yet AI learns from it indiscriminately, acquiring both beneficial pa
Reliable AI Needs to Externalize Implicit Knowledge: A Human-AI Collaboration Perspective
-
cs.AI, q-bio.NC updates on arXiv.org
-
Bridging Evolutionary Algorithms and Reinforcement Learning: A Comprehensive Survey on Hybrid Algorithms
arXiv:2401.11963v5 Announce Type: replace-cross Abstract: Evolutionary Reinforcement Learning (ERL), which integrates Evolutionary Algorithms (EAs) and Reinforcement Learning (RL) for optimization, has demonstrated remarkable performance advancements. By fusing both approaches, ERL has emerged as a promising research direction. This survey offers a comprehensive overview of the diverse research branches in ERL. Specifically, we systematically summarize recent advancements in related algorithms
Bridging Evolutionary Algorithms and Reinforcement Learning: A Comprehensive Survey on Hybrid Algorithms
-
cs.AI, q-bio.NC updates on arXiv.org
-
Reward-free Alignment for Conflicting Objectives
arXiv:2602.02495v3 Announce Type: replace-cross Abstract: Direct alignment methods are increasingly used to align large language models (LLMs) with human preferences. However, many real-world alignment problems involve multiple conflicting objectives, where naive aggregation of preferences can lead to unstable training and poor trade-offs. In particular, weighted loss methods may fail to identify update directions that simultaneously improve all objectives, and existing multi-objective approach
Reward-free Alignment for Conflicting Objectives
-
cs.AI, q-bio.NC updates on arXiv.org
-
SPA-Cache: Singular Proxies for Adaptive Caching in Diffusion Language Models
arXiv:2602.02544v2 Announce Type: replace-cross Abstract: While Diffusion Language Models (DLMs) offer a flexible, arbitrary-order alternative to the autoregressive paradigm, their non-causal nature precludes standard KV caching, forcing costly hidden state recomputation at every decoding step. Existing DLM caching approaches reduce this cost by selective hidden state updates; however, they are still limited by (i) costly token-wise update identification heuristics and (ii) rigid, uniform budge
SPA-Cache: Singular Proxies for Adaptive Caching in Diffusion Language Models
-
cs.AI, q-bio.NC updates on arXiv.org
-
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
arXiv:2605.02900v2 Announce Type: replace-cross Abstract: Embodied Artificial Intelligence (Embodied AI) integrates perception, cognition, planning, and interaction into agents that operate in open-world, safety-critical environments. As these systems gain autonomy and enter domains such as transportation, healthcare, and industrial or assistive robotics, ensuring their safety becomes both technically challenging and socially indispensable. Unlike digital AI systems, embodied agents must act un
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
-
Nature - Issue - nature.com science feeds
-
Bottom-Up Synthesis of Molecular Nanodiamond from Nanographene
Nature, Published online: 26 May 2026; doi:10.1038/s41586-026-10669-3Bottom-Up Synthesis of Molecular Nanodiamond from Nanographene
Bottom-Up Synthesis of Molecular Nanodiamond from Nanographene
Nature, Published online: 26 May 2026; doi:10.1038/s41586-026-10669-3
Bottom-Up Synthesis of Molecular Nanodiamond from Nanographene-
Omics in Gastric
-
Machine learning-based identification of key genes underlying sex differences in hepatocellular carcinoma and targeted drug screening
Biomed Rep. 2026 Apr 24;24(6):74. doi: 10.3892/br.2026.2147. eCollection 2026 Jun.ABSTRACTHepatocellular carcinoma (HCC) shows a marked predominance in men, yet the molecular basis for this sex disparity remains unclear. The present study leveraged multi-omics data and machine learning algorithms to identify key genes associated with sex-specific differences in HCC and to screen for putative candidate compounds, aiming to provide new insights for sex-specific therapy. The mRNA expression data of
Machine learning-based identification of key genes underlying sex differences in hepatocellular carcinoma and targeted drug screening
Biomed Rep. 2026 Apr 24;24(6):74. doi: 10.3892/br.2026.2147. eCollection 2026 Jun.
ABSTRACT
Hepatocellular carcinoma (HCC) shows a marked predominance in men, yet the molecular basis for this sex disparity remains unclear. The present study leveraged multi-omics data and machine learning algorithms to identify key genes associated with sex-specific differences in HCC and to screen for putative candidate compounds, aiming to provide new insights for sex-specific therapy. The mRNA expression data of male and female patients with HCC and paracancerous tissues were obtained from the GEO and TCGA databases. To mitigate overfitting, data were partitioned into independent training and testing sets. Candidate genes were screened by differential expression analysis and weighted gene co-expression network analysis. A total of four complementary algorithms, random forest, support vector machines, generalized linear models and extreme gradient boosting were used to identify key genes with high predictive capability. CYP17A1 and IRX3 were identified as the top differentially expressed core genes associated with HCC in men. Pan-cancer analysis showed that CYP17A1 was lowly expressed in the majority of tumors, but significantly highly expressed in HCC, rectal adenocarcinoma and gastric cancer (P<0.001). Functional cell-based assays showed that knockout of CYP17A1 inhibited the proliferation, migration and invasion ability of HCC cells (P<0.001). Immunohistochemistry showed that CYP17A1 protein expression was significantly increased in HCC tissues from male patients when compared with that in paracancerous tissues (P<0.001), whereas there was no significant difference in female patient tissues (P>0.05). Notably, while IRX3 was identified computationally, its functional role remains to be experimentally validated. Molecular docking predicted a potential interaction between the natural compound Saikosaponin A and the CYP17A1 protein, and cellular assays revealed that it dose-dependently inhibits HCC cell malignant phenotypes. The present study suggests that CYP17A1 is associated with sex differences in HCC, potentially via the androgen signaling axis. Furthermore, IRX3 emerges as a novel hypothesis-generating candidate gene. Finally, the findings of the present study highlight Saikosaponin A as a putative therapeutic candidate for male patients with HCC, warranting further target-dependency investigations.
PMID:42125766 | PMC:PMC13158723 | DOI:10.3892/br.2026.2147
-
Cell
-
Phages communicate across species to shape microbial ecosystems
Gallego-del-Sol et al. show that arbitrium-coding phages can sense non-cognate peptide signals from other phages to regulate lysis-lysogeny decisions. This crosstalk affects lysis-lysogeny outcomes of phage infections, mixed lysogenic communities, and polylysogens. Our results demonstrate that crosstalk is an important mechanism that drives phage interactions in microbial communities.
Phages communicate across species to shape microbial ecosystems
-
(Multiomics OR Omics) AND (Pancreatic)
-
FCGR2B (+) Macrophages as a Critical Node Linking Ferroptosis and Immunosuppression: A Multiomics Framework for Prognosis and Therapy in High-Grade Serous Ovarian Cancer
Hum Mutat. 2026 Apr 6;2026:8027584. doi: 10.1155/humu/8027584. eCollection 2026.ABSTRACTBACKGROUND: High-grade serous ovarian cancer (HGSOC) is characterized by a complex tumor microenvironment and poor prognosis, yet the roles of specific tumor-associated macrophages (TAMs) subpopulations in driving disease progression remain elusive.METHODS: This study evaluated the prognostic relevance of FCGR2B in HGSOC. Single-cell RNA sequencing identified FCGR2B + TAMs as a distinct macrophage subpopulati
FCGR2B (+) Macrophages as a Critical Node Linking Ferroptosis and Immunosuppression: A Multiomics Framework for Prognosis and Therapy in High-Grade Serous Ovarian Cancer
Hum Mutat. 2026 Apr 6;2026:8027584. doi: 10.1155/humu/8027584. eCollection 2026.
ABSTRACT
BACKGROUND: High-grade serous ovarian cancer (HGSOC) is characterized by a complex tumor microenvironment and poor prognosis, yet the roles of specific tumor-associated macrophages (TAMs) subpopulations in driving disease progression remain elusive.
METHODS: This study evaluated the prognostic relevance of FCGR2B in HGSOC. Single-cell RNA sequencing identified FCGR2B + TAMs as a distinct macrophage subpopulation with unique transcriptional features. Integrative analyses combining single-cell and bulk differentially expressed genes, macrophage-associated modules, and ferroptosis-related gene sets identified 26 candidate prognostic genes, from which a four-gene signature (CRYAB, PLAUR, EREG, and C5AR1) was derived to construct the prognostic risk model. The model was validated in an independent cohort. Immune infiltration, single-cell trajectory, copy number variation, and drug-gene associations were analyzed to explore the molecular and therapeutic implications of risk stratification.
RESULTS: HGSOC patients classified as high risk exhibited poorer survival outcomes, increased infiltration of M2-like macrophages, elevated expression of immune checkpoints, and enrichment of immune- and ferroptosis-related pathways. Trajectory and copy number variation analyses revealed stage-specific gene expression patterns and amplification-associated regulation. Drug-gene association analyses further suggested that high-risk patients may be more responsive to targeted therapies and proteasome inhibitors, whereas low-risk patients may benefit from conventional chemotherapy.
CONCLUSION: FCGR2B + TAMs are closely linked to HGSOC progression, and the proposed prognostic model based on FCGR2B + TAMs provides predictive value and potential therapeutic insights for patient stratification.
PMID:41953398 | PMC:PMC13054137 | DOI:10.1155/humu/8027584
-
cs.AI, q-bio.NC updates on arXiv.org
-
Enhancing behavioral nudges with large language model-based iterative personalization: A field experiment on electricity and hot-water conservation
arXiv:2604.03881v1 Announce Type: cross Abstract: Nudging is widely used to promote behavioral change, but its effectiveness is often limited when recipients must repeatedly translate feedback into workable next steps under changing circumstances. Large language models (LLMs) may help reduce part of this cognitive work by generating personalized guidance and updating it iteratively across intervention rounds. We developed an LLM agent for iterative personalization and tested it in a three-arm r
Enhancing behavioral nudges with large language model-based iterative personalization: A field experiment on electricity and hot-water conservation
-
cs.AI, q-bio.NC updates on arXiv.org
-
Uncertainty as a Planning Signal: Multi-Turn Decision Making for Goal-Oriented Conversation
arXiv:2604.03924v1 Announce Type: cross Abstract: Goal-oriented conversational systems require making sequential decisions under uncertainty about the user's intent, where the algorithm must balance information acquisition and target commitment over multiple turns. Existing approaches address this challenge from different perspectives: structured methods enable multi-step planning but rely on predefined schemas, while LLM-based approaches support flexible interactions but lack long-horizon deci
Uncertainty as a Planning Signal: Multi-Turn Decision Making for Goal-Oriented Conversation
-
cs.AI, q-bio.NC updates on arXiv.org
-
How AI Aggregation Affects Knowledge
arXiv:2604.04906v1 Announce Type: cross Abstract: Artificial intelligence (AI) changes social learning when aggregated outputs become training data for future predictions. To study this, we extend the DeGroot model by introducing an AI aggregator that trains on population beliefs and feeds synthesized signals back to agents. We define the learning gap as the deviation of long-run beliefs from the efficient benchmark, allowing us to capture how AI aggregation affects learning. Our main result id
How AI Aggregation Affects Knowledge
-
cs.AI, q-bio.NC updates on arXiv.org
-
Autonomous Agents for Scientific Discovery: Orchestrating Scientists, Language, Code, and Physics
arXiv:2510.09901v2 Announce Type: replace Abstract: Computing has long served as a cornerstone of scientific discovery. Recently, a paradigm shift has emerged with the rise of large language models (LLMs), introducing autonomous systems, referred to as agents, that accelerate discovery across varying levels of autonomy. These language agents provide a flexible and versatile framework that orchestrates interactions with human scientists, natural language, computer language and code, and physics.