Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
Dynamic Dual-Granularity Skill Bank for Agentic RL
arXiv:2603.28716v2 Announce Type: replace Abstract: Agentic RL can benefit substantially from reusable experience, yet existing skill-based methods mainly extract trajectory-level guidance and often lack principled mechanisms for maintaining an evolving skill memory. We propose D2Skill, a dynamic dual-granularity skill bank for agentic RL that organizes reusable experience into task skills for high-level guidance and step skills for fine-grained decision support and error correction. D2Skill jo
-
cs.AI, q-bio.NC updates on arXiv.org
-
SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment
arXiv:2604.08988v3 Announce Type: replace Abstract: Current LLM-based agents demonstrate strong performance in episodic task execution but remain constrained by static toolsets and episodic amnesia, failing to accumulate experience across task boundaries. This paper formalizes the Self-Evolving Agent (SEA) from the perspective of digital embodiment and continuous cross-task evolution, introduces the Evolutionary Flywheel as its minimal sufficient architecture, and presents SEA-Eval -- the first
SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment
-
cs.AI, q-bio.NC updates on arXiv.org
-
JAEGER: Joint 3D Audio-Visual Grounding and Reasoning in Simulated Physical Environments
arXiv:2602.18527v2 Announce Type: replace-cross Abstract: Current audio-visual large language models (AV-LLMs) are predominantly restricted to 2D perception, relying on RGB video and monaural audio. This design choice introduces a fundamental dimensionality mismatch that precludes reliable source localization and spatial reasoning in complex 3D environments. We address this limitation by presenting JAEGER, a framework that extends AV-LLMs to 3D space, to enable joint spatial grounding and reaso
JAEGER: Joint 3D Audio-Visual Grounding and Reasoning in Simulated Physical Environments
-
cs.AI, q-bio.NC updates on arXiv.org
-
Data Difficulty and the Generalization--Extrapolation Tradeoff in LLM Fine-Tuning
arXiv:2605.12906v2 Announce Type: replace-cross Abstract: Data selection during supervised fine-tuning (SFT) can critically change the behavior of large language models (LLMs). Although existing work has studied the effect of selecting data based on heuristics such as perplexity, difficulty, or length, the reported findings are often inconsistent or context-dependent. In this work, we systematically study the role of data difficulty in fine-tuning from both empirical and theoretical perspective
Data Difficulty and the Generalization--Extrapolation Tradeoff in LLM Fine-Tuning
-
Omics in Gastric
-
FDX1 as a predictive biomarker and therapeutic target for lymph node metastasis in gastric cancer
Clin Exp Med. 2026 May 10. doi: 10.1007/s10238-026-02160-0. Online ahead of print.ABSTRACTThe prognostic values of cuproptosis-related genes (CRGs) in gastric cancer with lymph node metastasis (GCLM), especially in the tumor immune microenvironment (TIME), remain unclear. We analyzed the expression, mutation, immunity, drug sensitivity, and prognostic value of CRGs in GCLM using TCGA and GEO cohorts. Consensus clustering was performed to identify CRG subtypes, with differences characterized by m
FDX1 as a predictive biomarker and therapeutic target for lymph node metastasis in gastric cancer
Clin Exp Med. 2026 May 10. doi: 10.1007/s10238-026-02160-0. Online ahead of print.
ABSTRACT
The prognostic values of cuproptosis-related genes (CRGs) in gastric cancer with lymph node metastasis (GCLM), especially in the tumor immune microenvironment (TIME), remain unclear. We analyzed the expression, mutation, immunity, drug sensitivity, and prognostic value of CRGs in GCLM using TCGA and GEO cohorts. Consensus clustering was performed to identify CRG subtypes, with differences characterized by multi-omics analysis. A CRG-based prognostic risk score and immune score were constructed for individualized assessment, and the role of CRGs was validated through in vitro and in vivo experiments. Consensus clustering revealed that CRGs were significantly enriched in biological processes related to mitosis and energy metabolism, as well as in immune-related and cancer-associated pathways. Four distinct CRG subtypes were identified, showing marked differences in expression profiles, prognosis, genetic alterations, TIME, and chemotherapeutic drug sensitivity. We developed an exploratory CRG-based prognostic risk score for preliminary individualized assessment, and the functional relevance of CRGs in GCLM was further validated through in vitro experiments. Among these, FDX1, LIAS, DLAT, MTF1, and GLS were identified as key determinants of overall survival in patients with GCLM, with FDX1 emerging as a potential independent prognostic factor. Notably, upregulation of FDX1 significantly suppressed lymph node metastasis of gastric cancer cells in a mouse popliteal lymph node metastasis model. Our data uncovers FDX1 might be a potential favorable prognostic factors in GCLM patients. These findings may improve our understanding of CRGs in GCLM and provide new in-sights for assessing prognosis and developing more effective treatment strategies.
PMID:42107026 | DOI:10.1007/s10238-026-02160-0
-
cs.AI, q-bio.NC updates on arXiv.org
-
ActionNex: A Virtual Outage Manager for Cloud
arXiv:2604.03512v1 Announce Type: new Abstract: Outage management in large-scale cloud operations remains heavily manual, requiring rapid triage, cross-team coordination, and experience-driven decisions under partial observability. We present \textbf{ActionNex}, a production-grade agentic system that supports end-to-end outage assistance, including real-time updates, knowledge distillation, and role- and stage-conditioned next-best action recommendations. ActionNex ingests multimodal operationa
ActionNex: A Virtual Outage Manager for Cloud
-
cs.AI, q-bio.NC updates on arXiv.org
-
Schema-Aware Planning and Hybrid Knowledge Toolset for Reliable Knowledge Graph Triple Verification
arXiv:2604.04190v1 Announce Type: new Abstract: Knowledge Graphs (KGs) serve as a critical foundation for AI systems, yet their automated construction inevitably introduces noise, compromising data trustworthiness. Existing triple verification methods, based on graph embeddings or language models, often suffer from single-source bias by relying on either internal structural constraints or external semantic evidence, and usually follow a static inference paradigm. As a result, they struggle with
Schema-Aware Planning and Hybrid Knowledge Toolset for Reliable Knowledge Graph Triple Verification
-
cs.AI, q-bio.NC updates on arXiv.org
-
Diagonal-Tiled Mixed-Precision Attention for Efficient Low-Bit MXFP Inference
arXiv:2604.03950v1 Announce Type: cross Abstract: Transformer-based large language models (LLMs) have demonstrated remarkable performance across a wide range of real-world tasks, but their inference cost remains prohibitively high due to the quadratic complexity of attention and the memory bandwidth limitations of high-precision operations. In this work, we present a low-bit mixed-precision attention kernel using the microscaling floating-point (MXFP) data format, utilizing the computing capabi
Diagonal-Tiled Mixed-Precision Attention for Efficient Low-Bit MXFP Inference
-
cs.AI, q-bio.NC updates on arXiv.org
-
ROSClaw: A Hierarchical Semantic-Physical Framework for Heterogeneous Multi-Agent Collaboration
arXiv:2604.04664v1 Announce Type: cross Abstract: The integration of large language models (LLMs) with embodied agents has improved high-level reasoning capabilities; however, a critical gap remains between semantic understanding and physical execution. While vision-language-action (VLA) and vision-language-navigation (VLN) systems enable robots to perform manipulation and navigation tasks from natural language instructions, they still struggle with long-horizon sequential and temporally struct
ROSClaw: A Hierarchical Semantic-Physical Framework for Heterogeneous Multi-Agent Collaboration
-
cs.AI, q-bio.NC updates on arXiv.org
-
Mind Your HEARTBEAT! Claw Background Execution Inherently Enables Silent Memory Pollution
arXiv:2603.23064v3 Announce Type: replace-cross Abstract: We identify a critical security vulnerability in mainstream Claw personal AI agents: untrusted content encountered during heartbeat-driven background execution can silently pollute agent memory and subsequently influence user-facing behavior without the user's awareness. This vulnerability arises from an architectural design shared across the Claw ecosystem: heartbeat background execution runs in the same session as user-facing conversat
Mind Your HEARTBEAT! Claw Background Execution Inherently Enables Silent Memory Pollution
-
cs.AI, q-bio.NC updates on arXiv.org
-
Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation
arXiv:2604.02324v1 Announce Type: cross Abstract: Language models (LMs) are increasingly extended with new learnable vocabulary tokens for domain-specific tasks, such as Semantic-ID tokens in generative recommendation. The standard practice initializes these new tokens as the mean of existing vocabulary embeddings, then relies on supervised fine-tuning to learn their representations. We present a systematic analysis of this strategy: through spectral and geometric diagnostics, we show that mean
Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation
-
npj Digital Medicine
-
Multidimensional evaluation of large language models in radiology report readability
npj Digital Medicine, Published online: 01 April 2026; doi:10.1038/s41746-026-02589-3Multidimensional evaluation of large language models in radiology report readability
Multidimensional evaluation of large language models in radiology report readability
npj Digital Medicine, Published online: 01 April 2026; doi:10.1038/s41746-026-02589-3
Multidimensional evaluation of large language models in radiology report readability-
cs.AI, q-bio.NC updates on arXiv.org
-
ATP-Bench: Towards Agentic Tool Planning for MLLM Interleaved Generation
arXiv:2603.29902v1 Announce Type: new Abstract: Interleaved text-and-image generation represents a significant frontier for Multimodal Large Language Models (MLLMs), offering a more intuitive way to convey complex information. Current paradigms rely on either image generation or retrieval augmentation, yet they typically treat the two as mutually exclusive paths, failing to unify factuality with creativity. We argue that the next milestone in this field is Agentic Tool Planning, where the model
ATP-Bench: Towards Agentic Tool Planning for MLLM Interleaved Generation
-
cs.AI, q-bio.NC updates on arXiv.org
-
MultiGen: Level-Design for Editable Multiplayer Worlds in Diffusion Game Engines
arXiv:2603.06679v2 Announce Type: replace Abstract: Video world models have shown immense promise for interactive simulation and entertainment, but current systems still struggle with two important aspects of interactivity: user control over the environment for reproducible, editable experiences, and shared inference where players hold influence over a common world. To address these limitations, we introduce an explicit external memory into the system, a persistent state operating independent o
MultiGen: Level-Design for Editable Multiplayer Worlds in Diffusion Game Engines
-
cs.AI, q-bio.NC updates on arXiv.org
-
QuestA: Expanding Reasoning Capacity in LLMs via Question Augmentation
arXiv:2507.13266v4 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has emerged as a central paradigm for training large language models (LLMs) in reasoning tasks. Yet recent studies question RL's ability to incentivize reasoning capacity beyond the base model. This raises a key challenge: how can RL be adapted to solve harder reasoning problems more effectively? To address this challenge, we propose a simple yet effective strategy via Question Augmentation: introduce partial
QuestA: Expanding Reasoning Capacity in LLMs via Question Augmentation
-
Omics in Gastric
-
Targeting sialic acid metabolism: a therapeutic strategy against gastric cancer driven by WZ35
Cell Oncol (Dordr). 2026 Mar 23;49(2):60. doi: 10.1007/s13402-026-01194-6.ABSTRACTGlycolytic reprogramming is closely associated with the occurrence and progression of gastric cancer. Specifically, the energy derived from glucose metabolism and the cellular proteins by its intermediate products influence gastric cancer development. However, as an important branch of glucose metabolism, sialic acid metabolism and its mediated sialylation modifications remain insufficiently studied in gastric canc
Targeting sialic acid metabolism: a therapeutic strategy against gastric cancer driven by WZ35
Cell Oncol (Dordr). 2026 Mar 23;49(2):60. doi: 10.1007/s13402-026-01194-6.
ABSTRACT
Glycolytic reprogramming is closely associated with the occurrence and progression of gastric cancer. Specifically, the energy derived from glucose metabolism and the cellular proteins by its intermediate products influence gastric cancer development. However, as an important branch of glucose metabolism, sialic acid metabolism and its mediated sialylation modifications remain insufficiently studied in gastric cancer, and their specific relationship with malignant tumor progression requires further exploration. This study employed a multi‑omics approach, integrating metabolomics, single‑cell RNA sequencing, and bulk RNA sequencing analyses, to investigate the metabolic landscape of gastric cancer and its associated alterations. The results indicated that sialic acid is a characteristic metabolite in malignant gastric cancer tissues. It modulates biological functions such as immune response, proliferative activity, and metabolic remodeling within gastric cancer tissues by influencing sialylation modifications. Furthermore, we identified the drug WZ35, which can inhibit the malignant proliferation of gastric cancer by targeting both sialic acid metabolism and sialylated protein modifications. We put forward a conjecture that the metabolism and modification of sialic acid promote the malignant development of gastric cancer, and we discovered that the drug WZ35 has an inhibitory effect on the sialic acid metabolism of gastric cancer.
GRAPHICAL ABSTRACT:
PMID:41870836 | PMC:PMC13009457 | DOI:10.1007/s13402-026-01194-6
-
cs.AI, q-bio.NC updates on arXiv.org
-
PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments
arXiv:2603.23231v1 Announce Type: new Abstract: Empowering large language models with long-term memory is crucial for building agents that adapt to users' evolving needs. However, prior evaluations typically interleave preference-related dialogues with irrelevant conversations, reducing the task to needle-in-a-haystack retrieval while ignoring relationships between events that drive the evolution of user preferences. Such settings overlook a fundamental characteristic of real-world personalizat
PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments
-
cs.AI, q-bio.NC updates on arXiv.org
-
A Synchronous EEG-fNIRS BCI: A Proof-of-Concept for Multimodal Avalanche Analysis of Motor Cognition in Older Adults
arXiv:2603.23358v1 Announce Type: new Abstract: This proof-of-concept study introduces a novel multimodal framework combining synchronized EEG-fNIRS modalities with neuronal avalanche analysis to identify early network dysfunction in Alzheimer's disease. The approach leverages complementary neural signals to examine motor network dynamics during execution and imagery tasks within an interactive task environment. Preliminary analysis of a small pilot cohort (N=4 subjects, including one with Mild
A Synchronous EEG-fNIRS BCI: A Proof-of-Concept for Multimodal Avalanche Analysis of Motor Cognition in Older Adults
-
cs.AI, q-bio.NC updates on arXiv.org
-
Mind Your HEARTBEAT! Claw Background Execution Inherently Enables Silent Memory Pollution
arXiv:2603.23064v2 Announce Type: cross Abstract: We identify a critical security vulnerability in mainstream Claw personal AI agents: untrusted content encountered during heartbeat-driven background execution can silently pollute agent memory and subsequently influence user-facing behavior without the user's awareness. This vulnerability arises from an architectural design shared across the Claw ecosystem: heartbeat background execution runs in the same session as user-facing conversation, so
Mind Your HEARTBEAT! Claw Background Execution Inherently Enables Silent Memory Pollution
-
(Multiomics OR Omics) AND (Lung OR gastric OR Hepatocellular)
-
Targeting sialic acid metabolism: a therapeutic strategy against gastric cancer driven by WZ35
Cell Oncol (Dordr). 2026 Mar 23;49(2):60. doi: 10.1007/s13402-026-01194-6.NO ABSTRACTPMID:41870836 | DOI:10.1007/s13402-026-01194-6
Targeting sialic acid metabolism: a therapeutic strategy against gastric cancer driven by WZ35
Cell Oncol (Dordr). 2026 Mar 23;49(2):60. doi: 10.1007/s13402-026-01194-6.
NO ABSTRACT
PMID:41870836 | DOI:10.1007/s13402-026-01194-6