Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
arXiv:2605.24834v1 Announce Type: cross Abstract: Large language model (LLM) safety classifiers such as Llama Guard are effective at detecting overtly harmful prompts but remain vulnerable to adversarial jailbreak attacks that disguise malicious intent through role-play scenarios, fictional framing, and indirect requests. We present Reflect-Guard, a method that augments LLM-based safety classifiers with chain-of-thought self-reflection capabilities through parameter-efficient fine-tuning. Our a
-
cs.AI, q-bio.NC updates on arXiv.org
-
PathMem: Toward Cognition-Aligned Memory Transformation for Pathology MLLMs
arXiv:2603.09943v2 Announce Type: replace Abstract: Computational pathology demands both visual pattern recognition and dynamic integration of structured domain knowledge, including taxonomy, grading criteria, and clinical evidence. In practice, diagnostic reasoning requires linking morphological evidence with formal diagnostic and grading criteria. Although multimodal large language models (MLLMs) demonstrate strong vision language reasoning capabilities, they lack explicit mechanisms for stru
PathMem: Toward Cognition-Aligned Memory Transformation for Pathology MLLMs
-
Nature - Issue - nature.com science feeds
-
High-fidelity identification of guest species in porous materials
Nature, Published online: 20 May 2026; doi:10.1038/s41586-026-10527-2A reconstruction method based on Gaussian-apodized single-sideband electron ptychography removes artefacts to enable the high-fidelity identification of guest species in porous materials.
High-fidelity identification of guest species in porous materials
Nature, Published online: 20 May 2026; doi:10.1038/s41586-026-10527-2
A reconstruction method based on Gaussian-apodized single-sideband electron ptychography removes artefacts to enable the high-fidelity identification of guest species in porous materials.-
Omics In Lung
-
Pulmonary-Intestinal Axis: Shared Genetic Basis and Mediating Factors Identified Through Multi-Omics Analysis
Int J Chron Obstruct Pulmon Dis. 2026 Apr 7;21:561645. doi: 10.2147/COPD.S561645. eCollection 2026.ABSTRACTBACKGROUND: Chronic obstructive pulmonary disease (COPD) is a systemic condition with comorbidities beyond the lung (eg, cardiovascular and metabolic disorders), and gastrointestinal (GI) disorders are also common. The shared genetic basis of COPD-GI comorbidity and its mediating factors remain unclear. We hypothesized that COPD and GI diseases share pleiotropic genetic architecture implica
Pulmonary-Intestinal Axis: Shared Genetic Basis and Mediating Factors Identified Through Multi-Omics Analysis
Int J Chron Obstruct Pulmon Dis. 2026 Apr 7;21:561645. doi: 10.2147/COPD.S561645. eCollection 2026.
ABSTRACT
BACKGROUND: Chronic obstructive pulmonary disease (COPD) is a systemic condition with comorbidities beyond the lung (eg, cardiovascular and metabolic disorders), and gastrointestinal (GI) disorders are also common. The shared genetic basis of COPD-GI comorbidity and its mediating factors remain unclear. We hypothesized that COPD and GI diseases share pleiotropic genetic architecture implicating lipid-metabolic pathways, with smoking mediating part of the association.
METHODS: We analyzed publicly available European-ancestry GWAS summary statistics for COPD (Global Biobank Meta-analysis Initiative), 15 GI diseases (FinnGen), and smoking phenotypes (UK Biobank). Genetic correlation was estimated using linkage disequilibrium score regression (LDSC) and high-definition likelihood (HDL). Multi-trait analysis of GWAS (MTAG) boosted COPD discovery by leveraging genetically correlated GI traits. We integrated locus-to-gene mapping with multi-tissue expression quantitative trait loci (eQTL) and plasma protein quantitative trait loci (pQTL) evidence to prioritize shared loci, genes, and proteins. Bidirectional two-sample Mendelian randomization (MR) tested causal directions, and two-step mediation MR evaluated smoking.
RESULTS: COPD showed significant genetic correlation with nine GI diseases. We identified six comorbidity-associated loci (three with CADD > 12.37) and 13 unique candidate pleiotropic genes; APOE was supported by proteomic evidence. Enrichment analyses highlighted lipid-metabolism pathways. MR suggested COPD increases risk of gastroesophageal reflux disease (GERD), irritable bowel syndrome (IBS), acute appendicitis, and gastric ulcer, while diverticular disease showed reverse causality toward COPD. Smoking partially mediated the COPD effect on GERD, acute appendicitis, and gastric ulcer.
CONCLUSION: COPD and multiple GI disorders share a distributed pleiotropic genetic basis within the broader systemic comorbidity spectrum of COPD. Multi-omics evidence supports a genomic pulmonary-intestinal axis in which lipid metabolism and smoking-related mechanisms contribute to COPD and GI comorbidity, providing targets for risk stratification and potential intervention.
PMID:41978582 | PMC:PMC13070119 | DOI:10.2147/COPD.S561645
-
cs.AI, q-bio.NC updates on arXiv.org
-
KLong: Training LLM Agent for Extremely Long-horizon Tasks
arXiv:2602.17547v2 Announce Type: replace Abstract: This paper introduces KLong, an open-source LLM agent trained to solve extremely long-horizon tasks. The principle is to first cold-start the model via trajectory-splitting SFT, then scale it via progressive RL training. Specifically, we first activate basic agentic abilities of a base model with a comprehensive SFT recipe. Then, we introduce Research-Factory, an automated pipeline that generates high-quality training data by collecting resear
KLong: Training LLM Agent for Extremely Long-horizon Tasks
-
cs.AI, q-bio.NC updates on arXiv.org
-
LLMs Judge Themselves: A Game-Theoretic Framework for Human-Aligned Evaluation
arXiv:2510.15746v2 Announce Type: replace-cross Abstract: Ideal or real - that is the question.In this work, we explore whether principles from game theory can be effectively applied to the evaluation of large language models (LLMs). This inquiry is motivated by the growing inadequacy of conventional evaluation practices, which often rely on fixed-format tasks with reference answers and struggle to capture the nuanced, subjective, and open-ended nature of modern LLM behavior. To address these c
LLMs Judge Themselves: A Game-Theoretic Framework for Human-Aligned Evaluation
-
Oncogene - Issue - nature.com science feeds
-
S100A6 promotes liver metastasis by activating FGFR3 signaling in <i>BAP1</i>-deficient uveal melanoma
Oncogene, Published online: 04 April 2026; doi:10.1038/s41388-026-03766-0S100A6 promotes liver metastasis by activating FGFR3 signaling in BAP1-deficient uveal melanoma
S100A6 promotes liver metastasis by activating FGFR3 signaling in <i>BAP1</i>-deficient uveal melanoma
Oncogene, Published online: 04 April 2026; doi:10.1038/s41388-026-03766-0
S100A6 promotes liver metastasis by activating FGFR3 signaling in BAP1-deficient uveal melanoma-
cs.AI, q-bio.NC updates on arXiv.org
-
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook
arXiv:2604.02029v1 Announce Type: new Abstract: Latent space is rapidly emerging as a native substrate for language-based models. While modern systems are still commonly understood through explicit token-level generation, an increasing body of work shows that many critical internal processes are more naturally carried out in continuous latent space than in human-readable verbal traces. This shift is driven by the structural limitations of explicit-space computation, including linguistic redunda
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook
-
cs.AI, q-bio.NC updates on arXiv.org
-
Do Emotions in Prompts Matter? Effects of Emotional Framing on Large Language Models
arXiv:2604.02236v1 Announce Type: new Abstract: Emotional tone is pervasive in human communication, yet its influence on large language model (LLM) behaviour remains unclear. Here, we examine how first-person emotional framing in user-side queries affect LLM performance across six benchmark domains, including mathematical reasoning, medical question answering, reading comprehension, commonsense reasoning and social inference. Across models and tasks, static emotional prefixes usually produce on
Do Emotions in Prompts Matter? Effects of Emotional Framing on Large Language Models
-
cs.AI, q-bio.NC updates on arXiv.org
-
Reproducible, Explainable, and Effective Evaluations of Agentic AI for Software Engineering
arXiv:2604.01437v1 Announce Type: cross Abstract: With the advancement of Agentic AI, researchers are increasingly leveraging autonomous agents to address challenges in software engineering (SE). However, the large language models (LLMs) that underpin these agents often function as black boxes, making it difficult to justify the superiority of Agentic AI approaches over baselines. Furthermore, missing information in the evaluation design description frequently renders the reproduction of result
Reproducible, Explainable, and Effective Evaluations of Agentic AI for Software Engineering
-
cs.AI, q-bio.NC updates on arXiv.org
-
Bias Is a Subspace, Not a Coordinate: A Geometric Rethinking of Post-hoc Debiasing in Vision-Language Models
arXiv:2511.18123v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have become indispensable for multimodal reasoning, yet their representations often encode and amplify demographic biases, resulting in biased associations and misaligned predictions in downstream tasks. Such behavior undermines fairness and distorts the intended alignment between vision and language. Recent post-hoc approaches attempt to mitigate bias by replacing the most attribute-correlated embedding coo
Bias Is a Subspace, Not a Coordinate: A Geometric Rethinking of Post-hoc Debiasing in Vision-Language Models
-
(Multiomics OR Omics) AND (Lung OR gastric OR Hepatocellular)
-
Two-step clinical care pathway to predict MASLD-related advanced fibrosis and long-term outcomes in type 2 diabetes
Gut. 2026 Feb 9;75(3):576-587. doi: 10.1136/gutjnl-2025-337506.ABSTRACTBACKGROUND: Current guidelines recommend a two-step approach for risk stratification of metabolic dysfunction-associated steatotic liver disease (MASLD), starting with Fibrosis-4 index (FIB-4) followed by liver stiffness measurement (LSM) using vibration-controlled transient elastography (VCTE).OBJECTIVE: To evaluate this approach for predicting advanced fibrosis and liver-related events (LREs) in patients with type 2 diabete
Two-step clinical care pathway to predict MASLD-related advanced fibrosis and long-term outcomes in type 2 diabetes
Gut. 2026 Feb 9;75(3):576-587. doi: 10.1136/gutjnl-2025-337506.
ABSTRACT
BACKGROUND: Current guidelines recommend a two-step approach for risk stratification of metabolic dysfunction-associated steatotic liver disease (MASLD), starting with Fibrosis-4 index (FIB-4) followed by liver stiffness measurement (LSM) using vibration-controlled transient elastography (VCTE).
OBJECTIVE: To evaluate this approach for predicting advanced fibrosis and liver-related events (LREs) in patients with type 2 diabetes (T2D).
DESIGN: A prospective liver biopsy cohort of T2D patients with histologically confirmed MASLD from seven centres in China was used to assess diagnostic performance for advanced fibrosis. The international VCTE-Prognosis cohort, including T2D patients with MASLD who underwent VCTE at 16 centres in the USA, Europe and Asia, with longitudinal follow-up, was used to assess LREs, defined as hepatic decompensation or hepatocellular carcinoma.
RESULTS: 4781 participants were included. In the liver biopsy cohort (n=352; 22.2% with advanced fibrosis), applying LSM thresholds of <8 kPa and >12 kPa after FIB-4 classified patients into 63.4% low-risk, 9.4% intermediate-risk and 27.3% high-risk, with a correct classification rate of 71%. In the VCTE-Prognosis cohort (n=4429; median follow-up 51.3 (IQR 27.4-70.7) months), 140 (3.2%) patients developed LREs (110 (2.5%) with hepatic decompensation and 59 (1.3%) with hepatocellular carcinoma). The two-step approach classified 72.6%, 6.8% and 20.6% of patients into low-risk, intermediate-risk and high-risk groups, with corresponding 5-year cumulative LRE incidences of 0.7%, 0.9% and 11.8%. Refining classification of intermediate FIB-4 patients using LSM <10 kPa (low-risk) and >15 kPa (high-risk) reduced the intermediate-risk group to 5.6% while preserving predictive accuracy.
CONCLUSION: The non-invasive two-step approach of FIB-4 followed by LSM effectively stratifies MASLD-related advanced fibrosis and LREs risk in T2D. Applying LSM cut-offs of 10 and 15 kPa further optimises risk stratification for future LREs.
PMID:41911049 | DOI:10.1136/gutjnl-2025-337506
-
Journal of Medical Internet Research
-
Effect of a Digital-Driven Physician-Pharmacist Collaborative Model for Diabetes in Primary Health Care: Cluster Randomized Trial
Background: Evidence-based physician-pharmacist collaborative clinics have demonstrated significant short-term benefits for patients with type 2 diabetes (T2D), but their long-term effectiveness remains unclear, especially in primary health care settings. Objective: This study aimed to explore the long-term effectiveness and cost-effectiveness of a novel, digital-driven, multifaceted physician-pharmacist collaborative model for managing patients with T2D in underresourced settings. Methods: We c
Effect of a Digital-Driven Physician-Pharmacist Collaborative Model for Diabetes in Primary Health Care: Cluster Randomized Trial
-
cs.AI, q-bio.NC updates on arXiv.org
-
Towards Realistic Personalization: Evaluating Long-Horizon Preference Following in Personalized User-LLM Interactions
arXiv:2603.04191v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly serving as personal assistants, where users share complex and diverse preferences over extended interactions. However, assessing how well LLMs can follow these preferences in realistic, long-term situations remains underexplored. This work proposes RealPref, a benchmark for evaluating realistic preference-following in personalized user-LLM interactions. RealPref features 100 user profiles, 1300 persona
Towards Realistic Personalization: Evaluating Long-Horizon Preference Following in Personalized User-LLM Interactions
-
cs.AI, q-bio.NC updates on arXiv.org
-
Boosting Meta-Learning for Few-Shot Text Classification via Label-guided Distance Scaling
arXiv:2603.02267v1 Announce Type: cross Abstract: Few-shot text classification aims to recognize unseen classes with limited labeled text samples. Existing approaches focus on boosting meta-learners by developing complex algorithms in the training stage. However, the labeled samples are randomly selected during the testing stage, so they may not provide effective supervision signals, leading to misclassification. To address this issue, we propose a \textbf{L}abel-guided \textbf{D}istance \textb
Boosting Meta-Learning for Few-Shot Text Classification via Label-guided Distance Scaling
-
cs.AI, q-bio.NC updates on arXiv.org
-
xLLM Technical Report
arXiv:2510.14686v2 Announce Type: replace-cross Abstract: We introduce xLLM, an intelligent and efficient Large Language Model (LLM) inference framework designed for high-performance, large-scale enterprise-grade serving, with deep optimizations for diverse AI accelerators. To address these challenges, xLLM builds a novel decoupled service-engine architecture. At the service layer, xLLM-Service features an intelligent scheduling module that efficiently processes multimodal requests and co-locat
xLLM Technical Report
-
cs.AI, q-bio.NC updates on arXiv.org
-
Behavior Learning (BL): Learning Hierarchical Optimization Structures from Data
arXiv:2602.20152v1 Announce Type: cross Abstract: Inspired by behavioral science, we propose Behavior Learning (BL), a novel general-purpose machine learning framework that learns interpretable and identifiable optimization structures from data, ranging from single optimization problems to hierarchical compositions. It unifies predictive performance, intrinsic interpretability, and identifiability, with broad applicability to scientific domains involving optimization. BL parameterizes a composi