❌

Reading view

Owl-AuraID 1.0: An Intelligent System for Autonomous Scientific Instrumentation and Scientific Data Analysis

arXiv:2603.29828v1 Announce Type: new Abstract: Scientific discovery increasingly depends on high-throughput characterization, yet automation is hindered by proprietary GUIs and the limited generalizability of existing API-based systems. We present Owl-AuraID, a software-hardware collaborative embodied agent system that adopts a GUI-native paradigm to operate instruments through the same interfaces as human experts. Its skill-centric framework integrates Type-1 (GUI operation) and Type-2 (data analysis) skills into end-to-end workflows, connecting physical sample handling with scientific interpretation. Owl-AuraID demonstrates broad coverage across ten categories of precision instruments and diverse workflows, including multimodal spectral analysis, microscopic imaging, and crystallographic analysis, supporting modalities such as FTIR, NMR, AFM, and TGA. Overall, Owl-AuraID provides a practical, extensible foundation for autonomous laboratories and illustrates a path toward evolving laboratory intelligence through reusable operational and analytical skills. The code are available at https://github.com/OpenOwlab/AuraID.
  •  

Trem1 regulates neutrophil metabolism and recruitment in lung ischemia-reperfusion injury

Redox Biol. 2026 Jan 14;92:104026. doi: 10.1016/j.redox.2026.104026. Online ahead of print.

ABSTRACT

Primary graft dysfunction (PGD) caused by ischemia-reperfusion injury (IRI) is a major complication after lung transplantation, yet its underlying mechanisms remain unclear. Triggering receptor expressed on myeloid cells 1 (Trem1) is an important mediator of inflammation, but its role in neutrophil function and metabolic reprogramming during lung IRI is not well understood. In this study, we used a murine orthotopic lung transplantation model with cold ischemia and reperfusion, and Trem1 knockout (Trem1-/-) and myeloid-specific Trem1 conditional knockout mice (LysmCreTrem1fl) to explore the role of Trem1 in neutrophil recruitment, neutrophil extracellular trap (NET) formation, and metabolism. Our results show that Trem1 expression increases in both mouse and human lungs after reperfusion and correlates with neutrophil infiltration and lung injury. Trem1 deficiency significantly reduced neutrophil and macrophage recruitment, NET formation, and tissue damage. Multi-omics analysis revealed that Trem1 deletion suppressed oxidative phosphorylation (OXPHOS) and induced a metabolic shift in neutrophils toward glycolysis. In clinical samples, the abundance of TREM1+ neutrophils was correlated with PGD severity and OXPHOS activity. These findings identify Trem1 as a key regulator of neutrophil metabolism and recruitment in lung IRI, and suggest that targeting Trem1 may provide a novel therapeutic strategy to mitigate PGD and improve lung transplant outcomes.

PMID:41861599 | DOI:10.1016/j.redox.2026.104026

  •  

Cheers: Decoupling Patch Details from Semantic Representations Enables Unified Multimodal Comprehension and Generation

arXiv:2603.12793v1 Announce Type: cross Abstract: A recent cutting-edge topic in multimodal modeling is to unify visual comprehension and generation within a single model. However, the two tasks demand mismatched decoding regimes and visual representations, making it non-trivial to jointly optimize within a shared feature space. In this work, we present Cheers, a unified multimodal model that decouples patch-level details from semantic representations, thereby stabilizing semantics for multimodal understanding and improving fidelity for image generation via gated detail residuals. Cheers includes three key components: (i) a unified vision tokenizer that encodes and compresses image latent states into semantic tokens for efficient LLM conditioning, (ii) an LLM-based Transformer that unifies autoregressive decoding for text generation and diffusion decoding for image generation, and (iii) a cascaded flow matching head that decodes visual semantics first and then injects semantically gated detail residuals from the vision tokenizer to refine high-frequency content. Experiments on popular benchmarks demonstrate that Cheers matches or surpasses advanced UMMs in both visual understanding and generation. Cheers also achieves 4x token compression, enabling more efficient high-resolution image encoding and generation. Notably, Cheers outperforms the Tar-1.5B on the popular benchmarks GenEval and MMBench, while requiring only 20% of the training cost, indicating effective and efficient (i.e., 4x token compression) unified multimodal modeling. We will release all code and data for future research.
  •  

SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition

arXiv:2511.21471v2 Announce Type: replace Abstract: Spatial cognition is fundamental to real-world multimodal intelligence, allowing models to effectively interact with the physical environment. While multimodal large language models (MLLMs) have made significant strides, existing benchmarks often oversimplify spatial cognition, reducing it to a single-dimensional metric, which fails to capture the hierarchical structure and interdependence of spatial abilities. To address this gap, we propose a hierarchical spatial cognition framework that decomposes spatial intelligence into five progressively complex levels from basic observation to high-level planning. Building upon this taxonomy, we construct SpatialBench, a large-scale, fine-grained benchmark covering 15 tasks aligned with these cognitive levels. To provide a unified evaluation across heterogeneous tasks, we further introduce a high-level capability-oriented metric that reliably assesses a model's overall spatial reasoning ability. Extensive experiments over massive MLLMs reveal distinct performance stratification across cognitive levels: models exhibit strong perceptual grounding yet remain limited in symbolic reasoning, causal inference, and planning. Additional human tests demonstrate that humans perform selective, goal-directed abstraction, while MLLMs tend to over-attend to surface details without coherent spatial intent. Our work establishes the first systematic framework for measuring hierarchical spatial cognition in MLLMs, laying the foundation for future spatially intelligent systems.
  •  
❌