Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
A Sober Look at Agentic Misalignment in Automated Workflows
arXiv:2605.24197v1 Announce Type: new Abstract: We study a class of emergent misalignment in multi-agent systems (MAS), with a focus on automated workflows, which we refer to agentic misalignment. Although these systems can solve complex tasks, they often fail because agents act according to implicit proxy utilities that do not align with the intended human goals. We formally define these behaviors and analyze them within a Bayesian framework, showing that generic utilities naturally lead to po
-
cs.AI, q-bio.NC updates on arXiv.org
-
A Large-Scale Dataset and Benchmark: Do Protein-Ligand Models Learn Binding Sites or Just Binding Likelihood?
arXiv:2605.24045v1 Announce Type: cross Abstract: Protein-ligand modeling underpins computational drug discovery and molecular design. Existing protein-ligand benchmarks typically evaluate whether a protein and ligand interact and how strongly they bind, through tasks such as binary binding prediction and affinity regression. However, these evaluations provide limited evidence of whether models can localize binding sites or identify the non-covalent interactions underlying molecular recognition
A Large-Scale Dataset and Benchmark: Do Protein-Ligand Models Learn Binding Sites or Just Binding Likelihood?
-
cs.AI, q-bio.NC updates on arXiv.org
-
FoodMonitor: Benchmarking MLLMs for Explainable Compliance Analysis
arXiv:2605.24503v1 Announce Type: cross Abstract: As AI-powered compliance monitoring becomes increasingly important in public governance and industrial safety, the ability to provide verifiable evidence and traceable accountability signals is essential. However, existing video anomaly detection datasets focus on event-level binary classification, lacking the rule-driven, explainable analysis required for real-world compliance scenarios. We introduce FoodMonitor, a benchmark for explainable com
FoodMonitor: Benchmarking MLLMs for Explainable Compliance Analysis
-
cs.AI, q-bio.NC updates on arXiv.org
-
DisDop: Distillation with Domain Priors for Open-Vocabulary Aerial Object Detection
arXiv:2605.24639v1 Announce Type: cross Abstract: With the widespread application of drones in recent years, object detection of aerial images has attracted increasing attention, especially open-vocabulary aerial detection which is not restricted to predefined categories. Due to the scarcity of drone's viewpoint images and their significant differences from natural images, it is difficult to achieve satisfying results by directly applying vanilla open-vocabulary detection methods designed for n
DisDop: Distillation with Domain Priors for Open-Vocabulary Aerial Object Detection
-
cs.AI, q-bio.NC updates on arXiv.org
-
AME-TS: Anchored Mixture-of-Experts for Time Series Forecasting
arXiv:2605.25166v1 Announce Type: cross Abstract: Time series forecasting models are increasingly scaled through large Transformer backbones, yet most existing approaches process all series through a shared dense computation path despite substantial heterogeneity in temporal structure. Mixture-of-Experts (MoE) offers a natural alternative by enabling conditional computation, but standard MoE routing leaves expert specialization weakly identified and often unstable during downstream adaptation.
AME-TS: Anchored Mixture-of-Experts for Time Series Forecasting
-
cs.AI, q-bio.NC updates on arXiv.org
-
BC Protocol: Structured Dual-Expert Dialogue for Eliciting High-Quality Chain-of-Thought Post-Training Data
arXiv:2605.25549v1 Announce Type: cross Abstract: High-quality expert chain-of-thought (CoT) data is one of the core bottlenecks in large language model (LLM) post-training. Existing data production methods each have structural limitations: crowdsourced annotation lacks deep reasoning paths; expert solo writing is constrained by the "expert blind spot" -- experts structurally skip reasoning steps they consider obvious; RLHF only produces preference signals rather than reasoning chains. This p
BC Protocol: Structured Dual-Expert Dialogue for Eliciting High-Quality Chain-of-Thought Post-Training Data
-
cs.AI, q-bio.NC updates on arXiv.org
-
When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills
arXiv:2605.25832v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as proposal generators for evolutionary robot design, yet most loops remain memoryless: simulator results shape the next population but are not preserved as reusable design knowledge. We present Auto-Robotist, a self-evolving LLM agent that distills morphology-search traces into an explicit natural-language skill library. Each skill stores a structural archetype, evidence-grounded positive and n
When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills
-
cs.AI, q-bio.NC updates on arXiv.org
-
QUIET: A Multi-Blank Cascaded Story Cloze Benchmark for LLM Creative Generation Capability
arXiv:2605.25955v1 Announce Type: cross Abstract: Large language models (LLMs) face a dual challenge in creative capability evaluation: existing benchmarks (e.g., Story Cloze Test, HellaSwag) measure models' discriminative ability over narrative continuation using multiple-choice recognition paradigms, rather than directly measuring creative generation capability; rubric-based scoring and LLM-as-Judge methods rely on subjective dimension assessment or natural language model outputs, and cannot
QUIET: A Multi-Blank Cascaded Story Cloze Benchmark for LLM Creative Generation Capability
-
cs.AI, q-bio.NC updates on arXiv.org
-
Creative Quality Alignment: Expert Tacit Knowledge Transfer via Chain-of-Thought Fine-Tuning
arXiv:2605.25977v1 Announce Type: cross Abstract: This paper provides an empirical implementation of the creative quality metric proposed in Calibrated Surprise (Zou & Xu, 2026a). The question this paper addresses is: does this mathematical claim hold at the engineering level? To make the answer as general as possible, we deliberately choose the strictest engineering conditions: low data cost and a small base model. Training data comes from approximately 100 expert chain-of-thought (CoT)
Creative Quality Alignment: Expert Tacit Knowledge Transfer via Chain-of-Thought Fine-Tuning
-
cs.AI, q-bio.NC updates on arXiv.org
-
AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild
arXiv:2605.22715v2 Announce Type: replace-cross Abstract: As wearable and mobile devices become increasingly embedded in daily life, they offer a practical way to continuously sense human motion in the wild. But inertial signals are highly dependent on the sensing setup, including body location, mounting position, sensor orientation, device hardware, and sampling protocol. This setup dependence makes it difficult to learn motion representations that transfer across devices and datasets, and lim
AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild
-
Nature - Issue - nature.com science feeds
-
Clinical application of base editing for treating β-thalassaemia
Nature, Published online: 08 April 2026; doi:10.1038/s41586-026-10342-9A clinical phase 1 trial of a single infusion of CS-101, CD34+ cells modified using a transformer base editor to reactivate fetal haemoglobin production, led to early and enduring transfusion independence in patients with β-thalassaemia.
Clinical application of base editing for treating β-thalassaemia
Nature, Published online: 08 April 2026; doi:10.1038/s41586-026-10342-9
A clinical phase 1 trial of a single infusion of CS-101, CD34+ cells modified using a transformer base editor to reactivate fetal haemoglobin production, led to early and enduring transfusion independence in patients with β-thalassaemia.-
cs.AI, q-bio.NC updates on arXiv.org
-
SelfGrader: Stable Jailbreak Detection for Large Language Models using Token-Level Logits
arXiv:2604.01473v1 Announce Type: cross Abstract: Large Language Models (LLMs) are powerful tools for answering user queries, yet they remain highly vulnerable to jailbreak attacks. Existing guardrail methods typically rely on internal features or textual responses to detect malicious queries, which either introduce substantial latency or suffer from the randomness in text generation. To overcome these limitations, we propose SelfGrader, a lightweight guardrail method that formulates jailbreak
SelfGrader: Stable Jailbreak Detection for Large Language Models using Token-Level Logits
-
MRD
-
Surgery-centered integrated strategies for personalized hepatocellular carcinoma care
Cancer Biol Med. 2026 Mar 30:j.issn.2095-3941.2026.0045. doi: 10.20892/j.issn.2095-3941.2026.0045. Online ahead of print.ABSTRACTHepatocellular carcinoma (HCC) remains a major global health burden characterized by late-stage diagnosis and high postoperative recurrence rates. This review presents a surgery-centered precision management framework integrating 3 synergistic components: early detection, precision surgery, and recurrence prevention. Early detection strategies incorporate multiparamete
Surgery-centered integrated strategies for personalized hepatocellular carcinoma care
Cancer Biol Med. 2026 Mar 30:j.issn.2095-3941.2026.0045. doi: 10.20892/j.issn.2095-3941.2026.0045. Online ahead of print.
ABSTRACT
Hepatocellular carcinoma (HCC) remains a major global health burden characterized by late-stage diagnosis and high postoperative recurrence rates. This review presents a surgery-centered precision management framework integrating 3 synergistic components: early detection, precision surgery, and recurrence prevention. Early detection strategies incorporate multiparameter risk models including the gender, age, AFP-L3, AFP, and DCP (GALAD) as well as age, sex, AFP, and PIVKA-II (ASAP) scores, alongside circulating tumor DNA methylation-based liquid biopsy, thus enabling tumor identification at stages amenable to curative resection. Precision surgery optimizes patient selection through refined staging systems including the Chinese liver cancer staging (CNLC), and functional assessments including the albumin-bilirubin (ALBI) grade, whereas conversion therapy and minimally invasive approaches extend surgical eligibility to selected patients with intermediate-stage disease. To mitigate the risk of postoperative recurrence, distinguishing between early and late recurrence patterns and monitoring minimal residual disease are critical strategies. Perioperative systemic therapies, particularly immune checkpoint inhibitor-based combinations, show promise for eradicating micrometastatic disease. This integrated framework provides a cohesive, evidence-based approach to personalized HCC management aimed at maximizing curative potential and long-term survival.
PMID:41913379 | DOI:10.20892/j.issn.2095-3941.2026.0045
-
Nature - Issue - nature.com science feeds
-
Parasites trigger epithelial cell crosstalk to drive gut–brain signalling
Nature, Published online: 25 March 2026; doi:10.1038/s41586-026-10281-5Paracrine signalling between tuft cells and enterochromaffin cells is a key mode of immune–sensory and gut–brain communication, and accounts for the pattern of gastrointestinal symptoms that occurs during parasite infections.
Parasites trigger epithelial cell crosstalk to drive gut–brain signalling
Nature, Published online: 25 March 2026; doi:10.1038/s41586-026-10281-5
Paracrine signalling between tuft cells and enterochromaffin cells is a key mode of immune–sensory and gut–brain communication, and accounts for the pattern of gastrointestinal symptoms that occurs during parasite infections.-
Omics In Lung
-
Single-cell multiomics uncovers an endothelial mechanosensitive PIEZO1-IL-33 axis driving pulmonary fibrosis
Nat Commun. 2026 Mar 20;17(1):2655. doi: 10.1038/s41467-026-70193-w.ABSTRACTPulmonary fibrosis represents a progressive interstitial lung disease marked by excessive extracellular matrix deposition and architectural distortion. Vascular endothelial cells critically contribute to fibrogenesis through paracrine secretion of pro-fibrotic mediators, yet their mechanobiological regulation remains elusive. Using integrated single-cell multi-omics profiling of human pulmonary fibrosis specimens and exp
Single-cell multiomics uncovers an endothelial mechanosensitive PIEZO1-IL-33 axis driving pulmonary fibrosis
Nat Commun. 2026 Mar 20;17(1):2655. doi: 10.1038/s41467-026-70193-w.
ABSTRACT
Pulmonary fibrosis represents a progressive interstitial lung disease marked by excessive extracellular matrix deposition and architectural distortion. Vascular endothelial cells critically contribute to fibrogenesis through paracrine secretion of pro-fibrotic mediators, yet their mechanobiological regulation remains elusive. Using integrated single-cell multi-omics profiling of human pulmonary fibrosis specimens and experimental fibrosis models induced by bleomycin or silica, we identify mechanosensitive Piezo1 upregulation in Endothelial cells as a hallmark of fibrotic progression. Endothelial-specific Piezo1 knockout significantly attenuates Bleomycin-induced fibrotic remodeling in male mice, establishing its pathogenic necessity. Mechanistically, PIEZO1 activation promotes pulmonary fibrosis development via CAPN2-mediated STAT3 phosphorylation, which may regulate the secretion of the pro-fibrotic molecule interleukin-33. These findings suggest that the endothelial PIEZO1-CAPN2-STAT3-IL33 axis is a potential therapeutic target for PF intervention.
PMID:41862476 | PMC:PMC13004862 | DOI:10.1038/s41467-026-70193-w
-
cs.AI, q-bio.NC updates on arXiv.org
-
Towards unified brain-to-text decoding across speech production and perception
arXiv:2603.12628v1 Announce Type: new Abstract: Speech production and perception are the main ways humans communicate daily. Prior brain-to-text decoding studies have largely focused on a single modality and alphabetic languages. Here, we present a unified brain-to-sentence decoding framework for both speech production and perception in Mandarin Chinese. The framework exhibits strong generalization ability, enabling sentence-level decoding when trained only on single-character data and supporti
Towards unified brain-to-text decoding across speech production and perception
-
cs.AI, q-bio.NC updates on arXiv.org
-
CRAFT-GUI: Curriculum-Reinforced Agent For GUI Tasks
arXiv:2508.11360v2 Announce Type: replace Abstract: As autonomous agents become adept at understanding and interacting with graphical user interface (GUI) environments, a new era of automated task execution is emerging. Recent studies have demonstrated that Reinforcement Learning (RL) can effectively enhance agents' performance in dynamic interactive GUI environments. However, these methods face two key limitations: (1) they overlook the significant variation in difficulty across different GUI
CRAFT-GUI: Curriculum-Reinforced Agent For GUI Tasks
-
cs.AI, q-bio.NC updates on arXiv.org
-
Fibration Policy Optimization
arXiv:2603.08239v1 Announce Type: cross Abstract: Large language models are increasingly trained as heterogeneous systems spanning multiple domains, expert partitions, and agentic pipelines, yet prevalent proximal objectives operate at a single scale and lack a principled mechanism for coupling token-level, trajectory-level, and higher-level hierarchical stability control. To bridge this gap, we derive the Aggregational Policy Censoring Objective (APC-Obj), the first exact unconstrained reformu
Fibration Policy Optimization
-
Cell Death Discovery nature.com science feeds
-
Oscillatory shear stress-driven endothelial-to-mesenchymal transition: a critical mechanical signal transduction mechanism in atherosclerosis progression
Cell Death Discovery, Published online: 10 March 2026; doi:10.1038/s41420-026-03000-6Oscillatory shear stress-driven endothelial-to-mesenchymal transition: a critical mechanical signal transduction mechanism in atherosclerosis progression
Oscillatory shear stress-driven endothelial-to-mesenchymal transition: a critical mechanical signal transduction mechanism in atherosclerosis progression
Cell Death Discovery, Published online: 10 March 2026; doi:10.1038/s41420-026-03000-6
Oscillatory shear stress-driven endothelial-to-mesenchymal transition: a critical mechanical signal transduction mechanism in atherosclerosis progression-
cs.AI, q-bio.NC updates on arXiv.org
-
MIND: Unified Inquiry and Diagnosis RL with Criteria Grounded Clinical Supports for Psychiatric Consultation
arXiv:2603.03677v1 Announce Type: cross Abstract: Large language models (LLMs) have advanced medical dialogue systems, yet psychiatric consultation poses substantially higher demands due to subjective ambiguity and comorbidity complexity: an agent must continuously extract psychopathological cues from incomplete and inconsistent patient reports in multi-turn interactions and perform rigorous differential diagnostic reasoning. However, existing methods face two fundamental challenges. First, wit