Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
Safety-Aligned 3D Object Detection: Single-Vehicle, Cooperative, and End-to-End Perspectives
arXiv:2604.03325v1 Announce Type: cross Abstract: Perception plays a central role in connected and autonomous vehicles (CAVs), underpinning not only conventional modular driving stacks, but also cooperative perception systems and recent end-to-end driving models. While deep learning has greatly improved perception performance, its statistical nature makes perfect predictions difficult to attain. Meanwhile, standard training objectives and evaluation benchmarks treat all perception errors equall
-
cs.AI, q-bio.NC updates on arXiv.org
-
RDEx-CMOP: Feasibility-Aware Indicator-Guided Differential Evolution for Fixed-Budget Constrained Multiobjective Optimization
arXiv:2604.03708v1 Announce Type: cross Abstract: Constrained multiobjective optimisation requires fast feasibility attainment together with stable convergence and diversity preservation under strict evaluation budgets. This report documents RDEx-CMOP, the differential evolution variant used in the IEEE CEC 2025 numerical optimisation competition (C06 special session) constrained multiobjective track. RDEx-CMOP integrates an {\epsilon}-level feasibility schedule, a SPEA2-style indicator-driven
RDEx-CMOP: Feasibility-Aware Indicator-Guided Differential Evolution for Fixed-Budget Constrained Multiobjective Optimization
-
cs.AI, q-bio.NC updates on arXiv.org
-
XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language Models
arXiv:2510.15148v2 Announce Type: replace-cross Abstract: Omni-modal large language models (OLLMs) aim to unify audio, vision, and text understanding within a single framework. While existing benchmarks primarily evaluate general cross-modal question-answering ability, it remains unclear whether OLLMs achieve modality-invariant reasoning or exhibit modality-specific biases. We introduce XModBench, a large-scale tri-modal benchmark explicitly designed to measure cross-modal consistency. XModBenc
XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language Models
-
cs.AI, q-bio.NC updates on arXiv.org
-
RASA: Routing-Aware Safety Alignment for Mixture-of-Experts Models
arXiv:2602.04448v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) language models introduce unique challenges for safety alignment due to their sparse routing mechanisms, which can enable degenerate optimization behaviors under standard full-parameter fine-tuning. In our preliminary experiments, we observe that naively applying full-parameter safety fine-tuning to MoE models can reduce attack success rates through routing or expert dominance effects, rather than by directly rep
RASA: Routing-Aware Safety Alignment for Mixture-of-Experts Models
-
Nature Cancer
-
Leptomeningeal metastatic cancer cells induce a permissive choroid plexus vasculature through extracellular-vesicle-derived 5-HIAA signaling
Nature Cancer, Published online: 03 April 2026; doi:10.1038/s43018-026-01145-yHuang, Hou, Yang et al. demonstrate that leptomeningeal metastatic cells favor the formation of a premetastatic niche by remodeling the choroid plexus vasculature through the serotonin metabolite 5-hydroxyindoleacetic acid, which signals into endothelial cells through the aryl hydrocarbon receptor.
Leptomeningeal metastatic cancer cells induce a permissive choroid plexus vasculature through extracellular-vesicle-derived 5-HIAA signaling
Nature Cancer, Published online: 03 April 2026; doi:10.1038/s43018-026-01145-y
Huang, Hou, Yang et al. demonstrate that leptomeningeal metastatic cells favor the formation of a premetastatic niche by remodeling the choroid plexus vasculature through the serotonin metabolite 5-hydroxyindoleacetic acid, which signals into endothelial cells through the aryl hydrocarbon receptor.-
cs.AI, q-bio.NC updates on arXiv.org
-
InfiniteVL: Synergizing Linear and Sparse Attention for Highly-Efficient, Unlimited-Input Vision-Language Models
arXiv:2512.08829v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly tasked with ultra-long multimodal understanding. While linear architectures offer constant computation and memory footprints, they often struggle with high-frequency visual perception compared to standard Transformers. To bridge this gap, we introduce \textbf{InfiniteVL}. We first develop a hybrid base model called \textbf{InfiniteVL-Base} that interleaves a small fraction of Full Attention
InfiniteVL: Synergizing Linear and Sparse Attention for Highly-Efficient, Unlimited-Input Vision-Language Models
-
Omics in Hepatocellular
-
Proposed Role of Circadian Clock Genes in Pathogenesis of HCC: Molecular Subtyping and Characterization
Biomedicines. 2026 Mar 12;14(3):645. doi: 10.3390/biomedicines14030645.ABSTRACTBackground: Hepatocellular carcinoma (HCC) stands as a prevalent global health issue with increasing incidence and mortality rates. Hepatocellular carcinoma (HCC) exhibits profound molecular and clinical heterogeneity, which limits the effectiveness of current therapeutic strategies. Circadian rhythm disruption has been implicated in metabolic reprogramming, proliferation, and immune modulation in cancer, but its role
Proposed Role of Circadian Clock Genes in Pathogenesis of HCC: Molecular Subtyping and Characterization
Biomedicines. 2026 Mar 12;14(3):645. doi: 10.3390/biomedicines14030645.
ABSTRACT
Background: Hepatocellular carcinoma (HCC) stands as a prevalent global health issue with increasing incidence and mortality rates. Hepatocellular carcinoma (HCC) exhibits profound molecular and clinical heterogeneity, which limits the effectiveness of current therapeutic strategies. Circadian rhythm disruption has been implicated in metabolic reprogramming, proliferation, and immune modulation in cancer, but its role in shaping HCC heterogeneity remains poorly defined. Methods: Four public HCC transcriptomic cohorts (TCGA-LIHC, CHCC, LIRI, LICA) were integrated using RMA normalization and ComBat for batch correction. Consensus clustering based on 31 core circadian clock genes (CCGs) identified robust molecular subtypes. Multi-omics characterization-including genomic alterations, pathway activity (GSEA/GSVA), immune microenvironment profiling (CIBERSORT, EPIC, MCP-counter, xCell), and drug-sensitivity prediction (pRRophetic/oncoPredict)-was performed to delineate subtype-specific biological properties. A nine-gene CCG-based RiskScore model was constructed using LASSO Cox regression to internally validate subtype robustness and intra-subtype risk stratification. Results: Using consensus clustering of 31 core CCGs in TCGA-LIHC and three independent validation cohorts (CHCC, LIRI, LICA), we identified three reproducible subtypes-Cluster-1 (metabolic-quiescent), Cluster-2 (transition-intermediate), and Cluster-3 (proliferation-inflammatory)-which were recapitulated across cohorts and showed distinct overall survival (Cluster-3 worst; log-rank p values significant across datasets). Multi-omic characterization revealed that Cluster-3 exhibits the highest tumor mutational burden and CNV burden with enrichment of TP53/AXIN1/TERT alterations, strong activation of cell-cycle, E2F, and G2M programs, and an immune-hot yet immunosuppressed microenvironment enriched for TAMs, Tregs and MDSCs. By contrast, Cluster-1 shows relative genomic stability, dominant hepatic metabolic signatures (fatty-acid oxidation, bile-acid and xenobiotic metabolism) and an immune-cold phenotype. Single-cell mapping linked ALAS1 expression to malignant hepatocytes predominating in Cluster-1, whereas NONO and CSNK1D localized to stromal (CAFs/TECs) and both malignant/immune compartments respectively in Cluster-3, providing a cellular mechanism for subtype-specific metabolism, angiogenesis and immune modulation. Finally, a nine-gene CCG-based RiskScore validated prognostic stratification and drug-sensitivity predictions indicated subtype-specific therapeutic vulnerabilities (notably increased predicted TKI sensitivity in Cluster-3). Conclusion: In conclusion, this study proposes a robust circadian rhythm-based molecular classification of hepatocellular carcinoma, revealing three biologically and clinically distinct subtypes characterized by divergent genomic alterations, metabolic programs, immune microenvironment states, and prognostic patterns. By integrating bulk and single-cell transcriptomic data, we identify subtype-specific roles of key circadian regulators-including ALAS1, NONO, and CSNK1D-in shaping tumor metabolism, proliferation, stromal remodeling, and immune suppression. These findings highlight circadian dysregulation as a potential upstream factor associated with HCC heterogeneity and provide a conceptual framework for developing subtype-tailored mechanistic studies and circadian-informed therapeutic strategies.
PMID:41898292 | PMC:PMC13024568 | DOI:10.3390/biomedicines14030645
-
cs.AI, q-bio.NC updates on arXiv.org
-
Think Before You Drive: World Model-Inspired Multimodal Grounding for Autonomous Vehicles
arXiv:2512.03454v3 Announce Type: replace-cross Abstract: Interpreting natural-language commands to localize target objects is critical for autonomous driving (AD). Existing visual grounding (VG) methods for autonomous vehicles (AVs) typically struggle with ambiguous, context-dependent instructions, as they lack reasoning over 3D spatial relations and anticipated scene evolution. Grounded in the principles of world models, we propose ThinkDeeper, a framework that reasons about future spatial st
Think Before You Drive: World Model-Inspired Multimodal Grounding for Autonomous Vehicles
-
npj Digital Medicine
-
Performance of DeepSeek in the generation of in-training examination questions in radiology resident education
npj Digital Medicine, Published online: 24 March 2026; doi:10.1038/s41746-026-02568-8Performance of DeepSeek in the generation of in-training examination questions in radiology resident education
Performance of DeepSeek in the generation of in-training examination questions in radiology resident education
npj Digital Medicine, Published online: 24 March 2026; doi:10.1038/s41746-026-02568-8
Performance of DeepSeek in the generation of in-training examination questions in radiology resident education-
cs.AI, q-bio.NC updates on arXiv.org
-
SynPlanResearch-R1: Encouraging Tool Exploration for Deep Research with Synthetic Plans
arXiv:2603.07853v1 Announce Type: new Abstract: Research Agents enable models to gather information from the web using tools to answer user queries, requiring them to dynamically interleave internal reasoning with tool use. While such capabilities can in principle be learned via reinforcement learning with verifiable rewards (RLVR), we observe that agents often exhibit poor exploration behaviors, including premature termination and biased tool usage. As a result, RLVR alone yields limited impro
SynPlanResearch-R1: Encouraging Tool Exploration for Deep Research with Synthetic Plans
-
cs.AI, q-bio.NC updates on arXiv.org
-
CCR-Bench: A Comprehensive Benchmark for Evaluating LLMs on Complex Constraints, Control Flows, and Real-World Cases
arXiv:2603.07886v1 Announce Type: cross Abstract: Enhancing the ability of large language models (LLMs) to follow complex instructions is critical for their deployment in real-world applications. However, existing evaluation methods often oversimplify instruction complexity as a mere additive combination of atomic constraints, failing to adequately capture the high-dimensional complexity arising from the intricate interplay of content and format, logical workflow control, and real-world applica
CCR-Bench: A Comprehensive Benchmark for Evaluating LLMs on Complex Constraints, Control Flows, and Real-World Cases
-
cs.AI, q-bio.NC updates on arXiv.org
-
Adaptation of Agentic AI: A Survey of Post-Training, Memory, and Skills
arXiv:2512.16301v3 Announce Type: replace Abstract: Large language model (LLM) agents are moving beyond prompting alone. ChatGPT marked the rise of general-purpose LLM assistants, DeepSeek showed that on-policy reinforcement learning with verifiable rewards can improve reasoning and tool use, and OpenClaw highlights a newer direction in which agents accumulate persistent memory and reusable skills. Yet the research landscape remains fragmented across post-training, retrieval, memory, and skill
Adaptation of Agentic AI: A Survey of Post-Training, Memory, and Skills
-
cs.AI, q-bio.NC updates on arXiv.org
-
Role-Aware Conditional Inference for Spatiotemporal Ecosystem Carbon Flux Prediction
arXiv:2603.03531v1 Announce Type: cross Abstract: Accurate prediction of terrestrial ecosystem carbon fluxes (e.g., CO$_2$, GPP, and CH$_4$) is essential for understanding the global carbon cycle and managing its impacts. However, prediction remains challenging due to strong spatiotemporal heterogeneity: ecosystem flux responses are constrained by slowly varying regime conditions, while short-term fluctuations are driven by high-frequency dynamic forcings. Most existing learning-based approache
Role-Aware Conditional Inference for Spatiotemporal Ecosystem Carbon Flux Prediction
-
cs.AI, q-bio.NC updates on arXiv.org
-
Dual-Modality Multi-Stage Adversarial Safety Training: Robustifying Multimodal Web Agents Against Cross-Modal Attacks
arXiv:2603.04364v1 Announce Type: cross Abstract: Multimodal web agents that process both screenshots and accessibility trees are increasingly deployed to interact with web interfaces, yet their dual-stream architecture opens an underexplored attack surface: an adversary who injects content into the webpage DOM simultaneously corrupts both observation channels with a consistent deceptive narrative. Our vulnerability analysis on MiniWob++ reveals that attacks including a visual component far out
Dual-Modality Multi-Stage Adversarial Safety Training: Robustifying Multimodal Web Agents Against Cross-Modal Attacks
-
cs.AI, q-bio.NC updates on arXiv.org
-
DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inference
arXiv:2602.18846v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved remarkable multimodal understanding and reasoning capabilities, yet remain computationally expensive due to dense visual tokenization. Existing efficiency approaches either merge redundant visual tokens or drop them progressively in language backbone, often trading accuracy for speed. In this work, we propose DUET-VLM, a versatile plug-and-play dual compression framework that consists of (a) vision-onl
DUET-VLM: Dual stage Unified Efficient Token reduction for VLM Training and Inference
-
cs.AI, q-bio.NC updates on arXiv.org
-
ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training
arXiv:2602.12691v2 Announce Type: replace-cross Abstract: We study how to improve large foundation vision-language-action (VLA) systems through online reinforcement learning (RL) in real-world settings. Central to this process is the value function, which provides learning signals to guide VLA learning from experience. In practice, the value function is estimated from trajectory fragments collected from different data sources, including historical policies and intermittent human interventions.