Normal view
-
cs.AI, q-bio.NC updates on arXiv.org
-
LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition
arXiv:2605.24005v1 Announce Type: new Abstract: The evolution of Large Language Model (LLM) reasoning is bottlenecked by the scarcity of high-quality process data. While self-alignment via endogenous rewards offers a solution, mining valid supervision faces three challenges: (1) Label Noise via Mimetic Bias, where rewards prioritize statistical likelihood over logical truth, creating a "correctness illusion" that masks compounding errors; (2) Coarse-Grained Supervision, where sparse global outc
-
cs.AI, q-bio.NC updates on arXiv.org
-
STREAM: A Data-Centric Framework for Mining High-Value Task-Oriented Dialogues from Streaming Media
arXiv:2605.25162v1 Announce Type: cross Abstract: Large language models for vertical domains are bottlenecked by the scarcity of complex, domain-specific task-oriented dialogues. Existing data acquisition pipelines face a persistent trilemma: expert annotation is expensive, real-world service conversations are constrained by privacy and commercial restrictions, and static corpora quickly become temporally stale. We propose Stream, a data-centric framework that leverages publicly available strea
STREAM: A Data-Centric Framework for Mining High-Value Task-Oriented Dialogues from Streaming Media
-
cs.AI, q-bio.NC updates on arXiv.org
-
Extreme Region Policy Distillation
arXiv:2605.25582v1 Announce Type: cross Abstract: Reinforcement learning for large language models faces a fundamental trade-off between sample efficiency and asymptotic performance: strictly on-policy methods discard trajectories after a single update, while off-policy reuse introduces distribution mismatch that existing trust-region techniques mitigate primarily by enforcing conservative optimization, often leaving rich training signals underutilized. To investigate this, we perform extensive
Extreme Region Policy Distillation
-
cs.AI, q-bio.NC updates on arXiv.org
-
OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation
arXiv:2605.25829v1 Announce Type: cross Abstract: Recent vision-language-action (VLA) models and world action models (WAMs) advance robotic manipulation by enriching intermediate representations with auxiliary spatial features or future visual-state prediction. However, these representations largely remain within the observation space and do not share the rigid-body geometry of the action space, forcing the action decoder to implicitly recover this geometry. We propose OASIS, a visuomotor polic
OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation
-
cs.AI, q-bio.NC updates on arXiv.org
-
PowerFlow: Unlocking the Dual Nature of LLMs via Principled Distribution Matching
arXiv:2603.18363v2 Announce Type: replace-cross Abstract: Unsupervised Reinforcement Learning from Internal Feedback (RLIF) has emerged as a promising paradigm for eliciting the latent capabilities of Large Language Models (LLMs) without external supervision. However, current methods rely on heuristic intrinsic rewards, which often lack a well-defined theoretical optimization target and are prone to degenerative biases. In this work, we introduce PowerFlow, a principled framework that reformula
PowerFlow: Unlocking the Dual Nature of LLMs via Principled Distribution Matching
-
cs.AI, q-bio.NC updates on arXiv.org
-
SURGE: Surrogate Gradient Adaptation in Binary Neural Networks
arXiv:2605.10989v3 Announce Type: replace-cross Abstract: The training of Binary Neural Networks (BNNs) is fundamentally based on gradient approximation for non-differentiable binarization operations (e.g., sign function). However, prevailing methods including the Straight-Through Estimator (STE) and its improved variants, rely on hand-crafted designs that suffer from gradient mismatch problem and information loss induced by fixed-range gradient clipping. To address this, we propose SURrogate G
SURGE: Surrogate Gradient Adaptation in Binary Neural Networks
-
cs.AI, q-bio.NC updates on arXiv.org
-
ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison
arXiv:2605.20278v2 Announce Type: replace-cross Abstract: Long-form image captioning exposes a reward granularity problem in RL: captions are judged as whole sequences, while the important errors occur at the level of individual visual claims. A good dense caption should be both faithful and informative, avoiding hallucination without omitting salient details. Yet pairwise preferences, reference-based metrics, and holistic scalar rewards compress these local errors into a single sequence-level
ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison
-
cs.AI, q-bio.NC updates on arXiv.org
-
AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild
arXiv:2605.22715v2 Announce Type: replace-cross Abstract: As wearable and mobile devices become increasingly embedded in daily life, they offer a practical way to continuously sense human motion in the wild. But inertial signals are highly dependent on the sensing setup, including body location, mounting position, sensor orientation, device hardware, and sampling protocol. This setup dependence makes it difficult to learn motion representations that transfer across devices and datasets, and lim
AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild
-
cs.AI, q-bio.NC updates on arXiv.org
-
SSDAU: Structured Semantic Data Augmentation for Joint Entity and Relation Extraction
arXiv:2605.23440v2 Announce Type: replace-cross Abstract: Joint Entity and Relation Extraction (JERE) is highly susceptible to weak generalization due to low-quality training data. Data augmentation is a common strategy to enhance model generalization across different domains. However, existing data augmentation methods often overlook text relevance and may disrupt semantic structures and dependencies, making it difficult to generate effective augmented data for improving model generalization.
SSDAU: Structured Semantic Data Augmentation for Joint Entity and Relation Extraction
-
(Multiomics OR Omics) AND (Pancreatic)
-
Multi-omics Analysis Reveals the Protection of a Quadruple Probiotic Mixture in Experimental Autoimmune Hepatitis
Probiotics Antimicrob Proteins. 2026 May 23. doi: 10.1007/s12602-026-11062-2. Online ahead of print.ABSTRACTAutoimmune hepatitis (AIH) is a chronic progressive inflammatory liver disease with a rising global incidence. The treatment of AIH remains challenging because first-line drugs show limited efficacy and systemic side effects. Gut microbiota plays a crucial role in the pathogenesis of AIH, leading to growing interest in developing probiotic-based therapies. In this study, we used multi-omic
Multi-omics Analysis Reveals the Protection of a Quadruple Probiotic Mixture in Experimental Autoimmune Hepatitis
Probiotics Antimicrob Proteins. 2026 May 23. doi: 10.1007/s12602-026-11062-2. Online ahead of print.
ABSTRACT
Autoimmune hepatitis (AIH) is a chronic progressive inflammatory liver disease with a rising global incidence. The treatment of AIH remains challenging because first-line drugs show limited efficacy and systemic side effects. Gut microbiota plays a crucial role in the pathogenesis of AIH, leading to growing interest in developing probiotic-based therapies. In this study, we used multi-omics analysis to investigate the therapeutic effects of a quadruple probiotic mixture (Probiotic-quad) consisting of Bifidobacterium infantis, Lactobacillus acidophilus, Enterococcus faecalis, and Bacillus cereus in a well-established chronic AIH murine model. Our results showed that Probiotic-quad treatment significantly alleviated AIH progression, as evidenced by lower serum liver enzyme levels, ameliorated hepatic inflammatory infiltration and histopathological damage. Metagenomic sequencing results showed that gut dysbiosis in AIH mice was partially reversed after Probiotic-quad administration. Additionally, the integrity of the intestinal epithelial barrier was restored, accompanied by a reduction in serum lipopolysaccharide levels. Untargeted metabolomic and transcriptomic analysis revealed that Probiotic-quad treatment was linked to alterations in hepatic metabolism, including the citrate cycle and tryptophan metabolism, and was associated with reduced activation of the NF-κB and NOD-like receptor signaling pathways. These findings suggest that Probiotic-quad treatment ameliorates AIH severity and is potentially associated with changes in hepatic immune responses, metabolism, gut microbiota, and intestinal barrier function, highlighting its potential as an adjuvant therapy for AIH.
PMID:42176246 | DOI:10.1007/s12602-026-11062-2
-
AAAS: Table of Contents
-
Nodeless superconducting gap and electron-boson coupling in (La,Pr,Sm)3Ni2O7 films
Science, Ahead of Print.
Nodeless superconducting gap and electron-boson coupling in (La,Pr,Sm)3Ni2O7 films
-
Journal of Medical Internet Research
-
Child Vaccination Status and Behavioral and Social Drivers of Vaccination Among Their Caregivers in the Philippines: Cross-Sectional Survey Study Comparison of Household, Mobile, and Online Modes
Background: The World Health Organization recommends that countries routinely collect data on the behavioral and social drivers (BeSD) of vaccination to inform public health interventions that increase vaccine uptake. There is a need to identify data collection methods that can rapidly and inexpensively collect representative data, particularly in low- and middle-income countries. Objective: This study aimed to understand BeSD drivers of vaccination in the Philippines and assess the trade-offs b
Child Vaccination Status and Behavioral and Social Drivers of Vaccination Among Their Caregivers in the Philippines: Cross-Sectional Survey Study Comparison of Household, Mobile, and Online Modes
-
Nature - Issue - nature.com science feeds
-
Superconductivity and electronic structures of nickelate thin film superstructures
Nature, Published online: 08 April 2026; doi:10.1038/s41586-026-10352-7Engineered Ruddlesden–Popper nickelate superstructures show that specific Fermi surface features enable ambient-pressure superconductivity, linking structural configuration, electronic structure and superconducting behaviour. .
Superconductivity and electronic structures of nickelate thin film superstructures
Nature, Published online: 08 April 2026; doi:10.1038/s41586-026-10352-7
Engineered Ruddlesden–Popper nickelate superstructures show that specific Fermi surface features enable ambient-pressure superconductivity, linking structural configuration, electronic structure and superconducting behaviour. .-
cs.AI, q-bio.NC updates on arXiv.org
-
TableVision: A Large-Scale Benchmark for Spatially Grounded Reasoning over Complex Hierarchical Tables
arXiv:2604.03660v1 Announce Type: new Abstract: Structured tables are essential for conveying high-density information in professional domains such as finance, healthcare, and scientific research. Despite the progress in Multimodal Large Language Models (MLLMs), reasoning performance remains limited for complex tables with hierarchical layouts. In this paper, we identify a critical Perception Bottleneck through quantitative analysis. We find that as task complexity scales, the number of involve
TableVision: A Large-Scale Benchmark for Spatially Grounded Reasoning over Complex Hierarchical Tables
-
cs.AI, q-bio.NC updates on arXiv.org
-
Memory Intelligence Agent
arXiv:2604.04503v2 Announce Type: new Abstract: Deep research agents (DRAs) integrate LLM reasoning with external tools. Memory systems enable DRAs to leverage historical experiences, which are essential for efficient reasoning and autonomous evolution. Existing methods rely on retrieving similar trajectories from memory to aid reasoning, while suffering from key limitations of ineffective memory evolution and increasing storage and retrieval costs. To address these problems, we propose a novel
Memory Intelligence Agent
-
cs.AI, q-bio.NC updates on arXiv.org
-
Learning Additively Compositional Latent Actions for Embodied AI
arXiv:2604.03340v1 Announce Type: cross Abstract: Latent action learning infers pseudo-action labels from visual transitions, providing an approach to leverage internet-scale video for embodied AI. However, most methods learn latent actions without structural priors that encode the additive, compositional structure of physical motion. As a result, latents often entangle irrelevant scene details or information about future observations with true state changes and miscalibrate motion magnitude. W
Learning Additively Compositional Latent Actions for Embodied AI
-
cs.AI, q-bio.NC updates on arXiv.org
-
DIRECT: Video Mashup Creation via Hierarchical Multi-Agent Planning and Intent-Guided Editing
arXiv:2604.04875v1 Announce Type: cross Abstract: Video mashup creation represents a complex video editing paradigm that recomposes existing footage to craft engaging audio-visual experiences, demanding intricate orchestration across semantic, visual, and auditory dimensions and multiple levels. However, existing automated editing frameworks often overlook the cross-level multimodal orchestration to achieve professional-grade fluidity, resulting in disjointed sequences with abrupt visual transi
DIRECT: Video Mashup Creation via Hierarchical Multi-Agent Planning and Intent-Guided Editing
-
cs.AI, q-bio.NC updates on arXiv.org
-
TSPO: Breaking the Double Homogenization Dilemma in Multi-turn Search Policy Optimization
arXiv:2601.22776v2 Announce Type: replace Abstract: Multi-turn tool-integrated reasoning enables Large Language Models (LLMs) to solve complex tasks through iterative information retrieval. However, current reinforcement learning (RL) frameworks for search-augmented reasoning predominantly rely on sparse outcome-level rewards, leading to a "Double Homogenization Dilemma." This manifests as (1) Process homogenization, where the thinking, reasoning, and tooling involved in generation are ignored.
TSPO: Breaking the Double Homogenization Dilemma in Multi-turn Search Policy Optimization
-
cs.AI, q-bio.NC updates on arXiv.org
-
Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation
arXiv:2604.02368v3 Announce Type: replace Abstract: As Large Language Models (LLMs) exhibit plateauing performance on conventional benchmarks, a pivotal challenge persists: evaluating their proficiency in complex, open-ended tasks characterizing genuine expert-level cognition. Existing frameworks suffer from narrow domain coverage, reliance on generalist tasks, or self-evaluation biases. To bridge this gap, we present XpertBench, a high-fidelity benchmark engineered to assess LLMs across authen
Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation
-
cs.AI, q-bio.NC updates on arXiv.org
-
Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving
arXiv:2603.13842v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving is typically built upon imitation learning (IL), yet its performance is constrained by the quality of human demonstrations. To overcome this limitation, recent methods incorporate reinforcement learning (RL) through sequential fine-tuning. However, such a paradigm remains suboptimal: sequential RL fine-tuning can introduce policy drift and often leads to a performance ceiling due to its dependence on the pre