Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
Xuanwu: Evolving General Multimodal Models into an Industrial-Grade Foundation for Content Ecosystems
arXiv:2603.29211v1 Announce Type: new Abstract: In recent years, multimodal large models have continued to improve on general benchmarks. However, in real-world content moderation and adversarial settings, mainstream models still suffer from degraded generalization and catastrophic forgetting because of limited fine-grained visual perception and insufficient modeling of long-tail noise. In this paper, we present Xuanwu VL-2B as a case study of how general multimodal models can be developed into
-
cs.AI, q-bio.NC updates on arXiv.org
-
The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning
arXiv:2603.29025v1 Announce Type: cross Abstract: Large language models systematically fail when a salient surface cue conflicts with an unstated feasibility constraint. We study this through a diagnose-measure-bridge-treat framework. Causal-behavioral analysis of the ``car wash problem'' across six models reveals approximately context-independent sigmoid heuristics: the distance cue exerts 8.7 to 38 times more influence than the goal, and token-level attribution shows patterns more consistent
The Model Says Walk: How Surface Heuristics Override Implicit Constraints in LLM Reasoning
-
Nature - Issue - nature.com science feeds
-
Electric dipole moment drives the dynamics of the TNFR1 complex I signalosome
Nature, Published online: 01 April 2026; doi:10.1038/s41586-026-10304-1Long-range interactions mediated by protein electric dipole moments have a role in driving the assembly and disassembly of super-signalling complex I for promoting NF-κB signalling.
Electric dipole moment drives the dynamics of the TNFR1 complex I signalosome
Nature, Published online: 01 April 2026; doi:10.1038/s41586-026-10304-1
Long-range interactions mediated by protein electric dipole moments have a role in driving the assembly and disassembly of super-signalling complex I for promoting NF-κB signalling.-
Omics In Lung
-
Integrative Multi-omics Analysis of Buti Huatan Tang in Chronic Obstructive Pulmonary Disease
J Vis Exp. 2026 Mar 13;(229). doi: 10.3791/70383.ABSTRACTThis study utilized a multi-omics and computational biology framework to investigate the therapeutic potential of the Traditional Chinese Medicine (TCM) formula Buti Huatan Tang (BTHTT) against chronic obstructive pulmonary disease (COPD). Significant physiological improvements were observed in a rat model following BTHTT intervention. Histological analysis showed a reversal of lung pathological damage, while biochemical assays, and transc
Integrative Multi-omics Analysis of Buti Huatan Tang in Chronic Obstructive Pulmonary Disease
J Vis Exp. 2026 Mar 13;(229). doi: 10.3791/70383.
ABSTRACT
This study utilized a multi-omics and computational biology framework to investigate the therapeutic potential of the Traditional Chinese Medicine (TCM) formula Buti Huatan Tang (BTHTT) against chronic obstructive pulmonary disease (COPD). Significant physiological improvements were observed in a rat model following BTHTT intervention. Histological analysis showed a reversal of lung pathological damage, while biochemical assays, and transcriptomics confirmed the normalization of IL-1β and IL-1R2 levels. Additionally, metabolic profiling revealed that BTHTT corrected disruptions in T3 and T4 thyroid hormone levels. A negative correlation was observed between the IL-1β/IL-1R2 axis and these thyroid hormones, indicating that their regulation is associated with the formula's therapeutic effect. Beyond direct measurements, machine learning algorithms identified ten COPD signature genes from clinical databases. Pathway enrichment analysis suggests that BTHTT may act through cytokine-cytokine-receptor interactions and thyroid hormone synthesis pathways. Furthermore, while 283 components were identified in vivo, compounds such as tanshinone IIA and cryptotanshinone are currently considered candidate active substances. Their role as primary drivers is supported by a model in which they stably bind to IL-1R2; this inference is based on molecular docking and molecular dynamics (MD) simulations rather than direct experimental isolation. Overall, the data support a model in which BTHTT exerts a multi-target effect on COPD by modulating inflammation and metabolic homeostasis. This integrated approach provides a refined scientific basis for the clinical application of BTHTT and highlights specific pathways for future experimental validation.
PMID:41911070 | DOI:10.3791/70383
-
(Multiomics OR Omics) AND (Lung OR gastric OR Hepatocellular)
-
Integrative Multi-omics Analysis of Buti Huatan Tang in Chronic Obstructive Pulmonary Disease
J Vis Exp. 2026 Mar 13;(229). doi: 10.3791/70383.ABSTRACTThis study utilized a multi-omics and computational biology framework to investigate the therapeutic potential of the Traditional Chinese Medicine (TCM) formula Buti Huatan Tang (BTHTT) against chronic obstructive pulmonary disease (COPD). Significant physiological improvements were observed in a rat model following BTHTT intervention. Histological analysis showed a reversal of lung pathological damage, while biochemical assays, and transc
Integrative Multi-omics Analysis of Buti Huatan Tang in Chronic Obstructive Pulmonary Disease
J Vis Exp. 2026 Mar 13;(229). doi: 10.3791/70383.
ABSTRACT
This study utilized a multi-omics and computational biology framework to investigate the therapeutic potential of the Traditional Chinese Medicine (TCM) formula Buti Huatan Tang (BTHTT) against chronic obstructive pulmonary disease (COPD). Significant physiological improvements were observed in a rat model following BTHTT intervention. Histological analysis showed a reversal of lung pathological damage, while biochemical assays, and transcriptomics confirmed the normalization of IL-1β and IL-1R2 levels. Additionally, metabolic profiling revealed that BTHTT corrected disruptions in T3 and T4 thyroid hormone levels. A negative correlation was observed between the IL-1β/IL-1R2 axis and these thyroid hormones, indicating that their regulation is associated with the formula's therapeutic effect. Beyond direct measurements, machine learning algorithms identified ten COPD signature genes from clinical databases. Pathway enrichment analysis suggests that BTHTT may act through cytokine-cytokine-receptor interactions and thyroid hormone synthesis pathways. Furthermore, while 283 components were identified in vivo, compounds such as tanshinone IIA and cryptotanshinone are currently considered candidate active substances. Their role as primary drivers is supported by a model in which they stably bind to IL-1R2; this inference is based on molecular docking and molecular dynamics (MD) simulations rather than direct experimental isolation. Overall, the data support a model in which BTHTT exerts a multi-target effect on COPD by modulating inflammation and metabolic homeostasis. This integrated approach provides a refined scientific basis for the clinical application of BTHTT and highlights specific pathways for future experimental validation.
PMID:41911070 | DOI:10.3791/70383
-
npj Digital Medicine
-
A multicenter randomized clinical trial of portable transcranial alternating current stimulation for major depressive disorder
npj Digital Medicine, Published online: 28 March 2026; doi:10.1038/s41746-026-02575-9A multicenter randomized clinical trial of portable transcranial alternating current stimulation for major depressive disorder
A multicenter randomized clinical trial of portable transcranial alternating current stimulation for major depressive disorder
npj Digital Medicine, Published online: 28 March 2026; doi:10.1038/s41746-026-02575-9
A multicenter randomized clinical trial of portable transcranial alternating current stimulation for major depressive disorder-
cs.AI, q-bio.NC updates on arXiv.org
-
DreamAudio: Customized Text-to-Audio Generation with Diffusion Models
arXiv:2509.06027v2 Announce Type: replace-cross Abstract: With the development of large-scale diffusion-based and language-modeling-based generative models, impressive progress has been achieved in text-to-audio generation. Despite producing high-quality outputs, existing text-to-audio models mainly aim to generate semantically aligned sound and fall short of controlling fine-grained acoustic characteristics of specific sounds. As a result, users who need specific sound content may find it diff
DreamAudio: Customized Text-to-Audio Generation with Diffusion Models
-
cs.AI, q-bio.NC updates on arXiv.org
-
SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
arXiv:2602.12670v3 Announce Type: replace Abstract: Agent Skills are structured packages of procedural knowledge that augment LLM agents at inference time. Despite rapid adoption, there is no standard way to measure whether they actually help. We present SkillsBench, a benchmark of 86 tasks across 11 domains paired with curated Skills and deterministic verifiers. Each task is evaluated under three conditions: no Skills, curated Skills, and self-generated Skills. We test 7 agent-model configurat
SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
-
cs.AI, q-bio.NC updates on arXiv.org
-
Deep Expert Injection for Anchoring Retinal VLMs with Domain-Specific Knowledge
arXiv:2603.07131v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) show immense potential for automated ophthalmic diagnosis. However, their clinical deployment is severely hindered by lacking domain-specific knowledge. In this work, we identify two structural deficiencies hindering reliable medical reasoning: 1) the Perception Gap, where general-purpose visual encoders fail to resolve fine-grained pathological cues (e.g., microaneurysms); and 2) the Reasoning Gap, where spa
Deep Expert Injection for Anchoring Retinal VLMs with Domain-Specific Knowledge
-
cs.AI, q-bio.NC updates on arXiv.org
-
Think, Speak, Decide: Language-Augmented Multi-Agent Reinforcement Learning for Economic Decision-Making
arXiv:2511.12876v3 Announce Type: replace Abstract: Economic decision-making depends not only on structured signals such as prices and taxes, but also on unstructured language, including peer dialogue and media narratives. While multi-agent reinforcement learning (MARL) has shown promise in optimizing economic decisions, it struggles with the semantic ambiguity and contextual richness of language. We propose LAMP (Language-Augmented Multi-Agent Policy), a framework that integrates language into
Think, Speak, Decide: Language-Augmented Multi-Agent Reinforcement Learning for Economic Decision-Making
-
cs.AI, q-bio.NC updates on arXiv.org
-
SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
arXiv:2602.12670v2 Announce Type: replace Abstract: Agent Skills are structured packages of procedural knowledge that augment LLM agents at inference time. Despite rapid adoption, there is no standard way to measure whether they actually help. We present SkillsBench, a benchmark of 86 tasks across 11 domains paired with curated Skills and deterministic verifiers. Each task is evaluated under three conditions: no Skills, curated Skills, and self-generated Skills. We test 7 agent-model configurat
SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
-
Oncogene - Issue - nature.com science feeds
-
Arginine methylation-dependent stabilization of SUV39H1 promotes breast cancer growth
Oncogene, Published online: 07 March 2026; doi:10.1038/s41388-026-03712-0Arginine methylation-dependent stabilization of SUV39H1 promotes breast cancer growth
Arginine methylation-dependent stabilization of SUV39H1 promotes breast cancer growth
Oncogene, Published online: 07 March 2026; doi:10.1038/s41388-026-03712-0
Arginine methylation-dependent stabilization of SUV39H1 promotes breast cancer growth-
cs.AI, q-bio.NC updates on arXiv.org
-
Towards Realistic Personalization: Evaluating Long-Horizon Preference Following in Personalized User-LLM Interactions
arXiv:2603.04191v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly serving as personal assistants, where users share complex and diverse preferences over extended interactions. However, assessing how well LLMs can follow these preferences in realistic, long-term situations remains underexplored. This work proposes RealPref, a benchmark for evaluating realistic preference-following in personalized user-LLM interactions. RealPref features 100 user profiles, 1300 persona
Towards Realistic Personalization: Evaluating Long-Horizon Preference Following in Personalized User-LLM Interactions
-
cs.AI, q-bio.NC updates on arXiv.org
-
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
arXiv:2603.03818v1 Announce Type: cross Abstract: Continual learning is a long-standing challenge in robot policy learning, where a policy must acquire new skills over time without catastrophically forgetting previously learned ones. While prior work has extensively studied continual learning in relatively small behavior cloning (BC) policy models trained from scratch, its behavior in modern large-scale pretrained Vision-Language-Action (VLA) models remains underexplored. In this work, we found
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
-
cs.AI, q-bio.NC updates on arXiv.org
-
ToolRLA: Multiplicative Reward Decomposition for Tool-Integrated Agents
arXiv:2603.01620v2 Announce Type: replace Abstract: Tool-integrated agents that interleave reasoning with API calls are promising for complex tasks, yet aligning them for high-stakes, domain-specific deployment remains challenging: existing reinforcement learning approaches rely on coarse binary rewards that cannot distinguish tool selection errors from malformed parameters. We present ToolRLA, a three-stage post-training pipeline (SFT $\rightarrow$ GRPO $\rightarrow$ DPO) for domain-specific t
ToolRLA: Multiplicative Reward Decomposition for Tool-Integrated Agents
-
cs.AI, q-bio.NC updates on arXiv.org
-
Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play
arXiv:2509.25541v2 Announce Type: replace-cross Abstract: Although reinforcement learning (RL) has emerged as a promising approach for improving vision-language models (VLMs) and multimodal large language models (MLLMs), current methods rely heavily on manually curated datasets and costly human verification, which limits scalable self-improvement in multimodal systems. To address this challenge, we propose Vision-Zero, a label-free, domain-agnostic multi-agent self-play framework for self-evolv
Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play
-
cs.AI, q-bio.NC updates on arXiv.org
-
UniG2U-Bench: Do Unified Models Advance Multimodal Understanding?
arXiv:2603.03241v1 Announce Type: cross Abstract: Unified multimodal models have recently demonstrated strong generative capabilities, yet whether and when generation improves understanding remains unclear. Existing benchmarks lack a systematic exploration of the specific tasks where generation facilitates understanding. To this end, we introduce UniG2U-Bench, a comprehensive benchmark categorizing generation-to-understanding (G2U) evaluation into 7 regimes and 30 subtasks, requiring varying de
UniG2U-Bench: Do Unified Models Advance Multimodal Understanding?
-
cs.AI, q-bio.NC updates on arXiv.org
-
xLLM Technical Report
arXiv:2510.14686v2 Announce Type: replace-cross Abstract: We introduce xLLM, an intelligent and efficient Large Language Model (LLM) inference framework designed for high-performance, large-scale enterprise-grade serving, with deep optimizations for diverse AI accelerators. To address these challenges, xLLM builds a novel decoupled service-engine architecture. At the service layer, xLLM-Service features an intelligent scheduling module that efficiently processes multimodal requests and co-locat
xLLM Technical Report
-
cs.AI, q-bio.NC updates on arXiv.org
-
Synthesizing Multimodal Geometry Datasets from Scratch and Enabling Visual Alignment via Plotting Code
arXiv:2602.18745v1 Announce Type: cross Abstract: Multimodal geometry reasoning requires models to jointly understand visual diagrams and perform structured symbolic inference, yet current vision--language models struggle with complex geometric constructions due to limited training data and weak visual--symbolic alignment. We propose a pipeline for synthesizing complex multimodal geometry problems from scratch and construct a dataset named \textbf{GeoCode}, which decouples problem generation in
Synthesizing Multimodal Geometry Datasets from Scratch and Enabling Visual Alignment via Plotting Code
-
cs.AI, q-bio.NC updates on arXiv.org
-
Towards Reliable Negative Sampling for Recommendation with Implicit Feedback via In-Community Popularity
arXiv:2602.18759v1 Announce Type: cross Abstract: Learning from implicit feedback is a fundamental problem in modern recommender systems, where only positive interactions are observed and explicit negative signals are unavailable. In such settings, negative sampling plays a critical role in model training by constructing negative items that enable effective preference learning and ranking optimization. However, designing reliable negative sampling strategies remains challenging, as they must si