❌

Normal view

Viral gene replication enhances AAV vector quality and reduces manufacturing costs

Liu and colleagues developed a robust in cellulo plasmid DNA replication system in human cells for replicating plasmid-borne adeno-associated virus (AAV) Rep/Cap genes during recombinant AAV (rAAV) production. This new approach not only enables a 10- to 20-fold plasmid reduction to significantly lower manufacturing costs but also substantially enhances rAAV potency, titer, and purity.

The DreAM-plus integrative RNA switch enhances transient AAV expression and reduces side effects of gene editing

This study developed a multi-layer inducible RNA switch that achieves transient expression of gene-delivery vectors in hepatic and non-hepatic tissues. As an exemplary application, this RNA switch triggers pulsive expression of gene editors that reduces the off-target effects and immunotoxicity of gene editing.

Lineage-specific pulmonary transcriptome landscape of coronavirus infection unveils universal immunotherapy for viral pneumonia

In the infection courses of different SARS-CoV-2 variants, disease outcomes and signatures were delineated by physiological changes, viral load, pathology, and pulmonary transcriptome analysis. This multi-dimensional landscape of disease outcomes and underlying mechanisms might provide important clues for immunotherapy of SARS-CoV-2 infection and pneumonia caused by other respiratory viruses.

Time-Frequency Geometric Cross-Attention for Chunked Vision-Language-Action Models

arXiv:2609.09925v1 Announce Type: new Abstract: Modern vision-language-action (VLA) policies predict a whole chunk of actions: one to two seconds of coordinated motion emitted in a single forward pass. Yet an action chunk is essentially a short multivariate trajectory, but inside these models it is a sequence of generic per-timestep hidden tokens decoded by a linear head. This under-serves two motion structures. First, frequency: a chunk superimposes a smooth global trend and fine corrective motion across time scales, and a single token entangles them. Second, cross-phase geometry: motions of different phases (reach, contact, grasp adjustment, settling) unfold along very different, near-orthogonal directions in representation space, yet are tightly related for the task and arise across the time axis. Dot-product attention scores alignment by an inner product, so it favors aligned tokens and is least sensitive near orthogonality, leaving such relationships for the network to recover through a detour. We introduce Time-Frequency Geometric Cross-Attention (TFGCA), a drop-in module repairing both blind spots. TFGCA uses a per-dimension learnable stationary wavelet transform to decompose the action chunk into time-frequency tokens, and each time token retrieves information from them via a cross-attention that fuses the dot product (similarity) with the wedge-product magnitude (sensitive to near-orthogonality) through a learnable weight. A zero-initialized residual reproduces the base behavior at initialization, so it can be dropped onto a pretrained VLA and fine-tuned jointly. Relative to the same-source base, TFGCA improves in-distribution LIBERO by +1.5 on average, the OOD LIBERO-Plus by +6.3, the randomized average under RoboTwin domain randomization by +28.5, and the overall success rate on three real-robot AgiBot A2 tasks by +11.67 points, with larger gains out of distribution.

AgentHijack: Visual Patch Attacks on Multimodal Computer-Use Agents

arXiv:2609.09212v1 Announce Type: cross Abstract: This paper presents an end-to-end evaluation framework for image-triggered command injection against computer-use agents (CUAs). The goal is to test whether a local visual patch can induce verifiable environmental consequences along the full chain of screenshot input, VLM generation, action parsing, and environment execution. We train and deploy patches on author-controlled GitHub Pages pages and a locally deployed CSDN clone, and evaluate them in real environments across five open-source or publicly available GUI-agent or vision-language-model (VLM) backends. Our experiment aggregates 600 instance-level online cases, with T-ASR, TAPR, and E2E-ASR reaching 84.5%, 47.0%, and 20.3%, respectively. Trajectory analysis further shows that in some successful cases the agent first executes a malicious terminal command and then continues the original benign task. These results indicate that optimized local visual signals can affect not only VLM outputs but also propagate through the execution pipeline of open CUAs and create real environmental risk.

Geometry Conditioning in an Embodied SLM: Training Controls and Robustness Diagnostics in a 0.8B Hybrid Model

10 September 2026 at 12:00
arXiv:2609.09213v1 Announce Type: cross Abstract: We study how physical-state inputs affect a 0.8B hybrid language model adapted for manipulation with 6.2M trainable parameters. Six conditions are trained on three LIBERO-Spatial tasks and evaluated over three seeds and 540 held-out rollouts. Conditioning recurrent decay gates on geometric increments yields 28.9% success, compared with 36.7% when those increments are shuffled during training and 24.4% without explicit object/goal geometry. Both geometry policies receive correct inputs at evaluation. A token adapter using the same increments scores 27.8%; differences vary across seeds and remain inconclusive. Token-clock conditioning scores 11.1%, including one seed that fails to converge. In separate robustness tests, a state-only relative-coordinate policy retains 7/10 success under frame relabeling, whereas all four tested visual policies fall to at most 3/20 after a 5 cm object displacement. These results show no reliable advantage from training-time geometric alignment under this recipe and illustrate the gap between coordinate invariance and physical-layout generalization. Episode records, seed-level analyses, and figure-generation code accompany the paper.

Harbor Adapters and Harbor-Index: Infrastructure and a Curated Meta-Dataset for Large-Scale Agentic Evaluation

arXiv:2609.04298v2 Announce Type: replace Abstract: Evaluating agents on the growing number of agentic benchmarks is challenging because they often require complex environments and agent integrations. We introduce Harbor Adapters, a unified evaluation infrastructure for agentic benchmarks. Our work makes three contributions. First, we develop benchmark adapters that port more than 80 benchmarks to evaluate arbitrary agents, and validate them through rigorous code review and parity experiments. Second, we conduct a large-scale evaluation of 8 models spanning capability tiers across 54 benchmarks; every model is run with Terminus-2 and with one of 3 native harnesses. This enables a broader analysis of agent capabilities and failure modes than was previously possible. Third, we introduce Harbor-Index, a curated set of 82 difficult, diverse, and high-quality tasks spanning 29 benchmarks, refined from the adapted suite through difficulty filtering, AI and human audit, and an audit-and-fix loop. Harbor-Index preserves the challenge and breadth of large-scale agentic evaluations while being affordable to run; no evaluated model-harness configuration exceeds 30% pass rate, and the strongest (GPT-5.5 with Codex) reaches 28.0%. We release the adapters, evaluation results, in-depth analysis, and Harbor-Index as open-source artifacts to support more reliable and comprehensive evaluation of language-model agents.

FiberTune: Preserving Action-Fiber Visual Residuals in Vision-Language-Action Fine-Tuning

arXiv:2606.08653v2 Announce Type: replace-cross Abstract: Action-supervised fine-tuning of vision-language-action (VLA) policies fits demonstrations effectively but constrains only the directions that change predicted actions, leaving visual structure consistent across action-equivalent states free to collapse. We formalize this as residual visual collapse along local action fibers and propose FiberTune, a training-time objective that preserves teacher-structured visual residuals without adding inference-time overhead. FiberTune uses an online action probe to estimate action-predictive feature directions, filters them from intermediate visual-token representations, and aligns the resulting probe-filtered residuals to a frozen visual teacher while regularizing their effective rank. Under identical training conditions, FiberTune improves over task-loss-only fine-tuning in every one of six controlled simulation settings spanning two benchmarks and two architectures (pi_0.5 and OpenVLA-OFT), as well as on physical SO-101 pick-place; representative gains include +10.7 percentage points SR(5) on long-horizon CALVIN ABC-to-D and physical SO-101 task success rising from 72.7% to 78.1%. Residual diagnostics show that these gains coincide with increased probe-filtered residual teacher alignment and effective rank, consistent with the action-fiber motivation.

LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation

arXiv:2608.30935v2 Announce Type: replace-cross Abstract: Embodied navigation requires agents to translate heterogeneous goals and visual observations into actions across tasks, environments, and robot embodiments. Modern vision-language models (VLMs) already encode spatial priors for visual grounding, spatial reasoning, and pointing, but these capabilities are rarely elicited directly for robot control. Existing navigation systems instead rely on task- or embodiment-specific components, fragmenting perception, reasoning, and action while offering limited generalization. Here we present LightNav-0, a compact generalist embodied navigation model that elicits the spatial intelligence of a pretrained VLM and aligns it with navigation, without task-specific prediction heads. LightNav-0 represents diverse navigation tasks through a unified token interface: dual-channel pointing expresses task-, scene-, and embodiment-agnostic spatial intent, while a residual vector-quantized action tokenizer maps this intent to precise, embodiment-specific trajectories. Together with temporally aware visual history compression, ER mid-training, supervised fine-tuning, and reinforcement learning, this formulation supports instruction following, open-vocabulary object navigation, and visual tracking within a single model. The navigation training corpus spans 2K+ scenes and 4K+ hours of embodied navigation data. LightNav-ER, the embodied-reasoning checkpoint used to initialize LightNav-0, attains the highest complete-set average across 8 embodied-reasoning benchmarks, while LightNav-0 achieves state-of-the-art monocular success rates across all 10 public navigation simulation settings. Real-world evaluations further demonstrate zero-shot generalization across robot embodiments, diverse scenes, and static and dynamic targets. These results establish compact VLMs as a unified and transferable backbone for generalist embodied navigation.

The redox architecture of gestational diabetes mellitus: from cellular stress engine to epigenetic and mitochondrial rewiring

Free Radic Biol Med. 2026 Sep 9;256:441-460. doi: 10.1016/j.freeradbiomed.2026.09.006. Online ahead of print.

ABSTRACT

Gestational diabetes mellitus (GDM) is a common pregnancy complication with a rising global prevalence, posing serious short-term and long-term health threats to both mothers and offspring. This review repositions GDM as a systemic disorder in which oxidative stress acts as a proposed mechanistic hub, linking upstream risk factors to downstream pathophysiology. We first examine how "upstream" factors-including genetic susceptibility, pre-conception status, and environmental exposures-converge to promote a state of pathological redox imbalance. We then examine key mechanistic pathways through which oxidative stress is thought to contribute to systemic insulin resistance and pancreatic β-cell failure, highlighting novel pathways involving intercellular communication via tunneling nanotubes and exosomes. Furthermore, we explore the downstream cascade, where oxidative stress may program maternal accelerated biological aging and multi-organ offspring disease trajectories through nuclear epigenetic programming and mitochondrial dysfunction programming, leaving what has been termed a persistent "metabolic memory". Consequently, this review evaluates emerging strategies that target oxidative stress for early prediction and precision intervention. Early prediction models based on direct redox biomarkers and multi-omics signatures hold potential to shift diagnosis from late-gestation oral glucose tolerance test (OGTT) to first-trimester risk stratification. Current supporting evidence draws from human epidemiological associations, ex vivo placental analyses, and experimental models. However, direct causal and interventional validation in pregnant women remains limited. Integrating targeted redox risk stratification and precision interventions into a life-course clinical framework may help interrupt the intergenerational transmission of metabolic disease initiated by GDM.

PMID:42716407 | DOI:10.1016/j.freeradbiomed.2026.09.006

Denisovans from southwestern China and their subsistence strategies

Nature, Published online: 09 September 2026; doi:10.1038/s41586-026-10997-4

Evidence from Bianfu Cave shows specialized hunting, expedient stone-tool production and extensive bone use of Denisovans, providing new insights into their ecology, behaviour and cultural legacy in eastern Asia.

Perioperative Modulation of the Gut-Liver Axis in Liver Surgery: Clinical Evidence and Future Directions

J Vis Exp. 2026 Sep 1;(235). doi: 10.3791/73747.

ABSTRACT

Liver resection and liver transplantation remain cornerstone treatments for many hepatobiliary diseases, yet postoperative infection, impaired liver regeneration, and post-hepatectomy liver failure (PHLF) remain serious complications. Perioperative stressors can disrupt the gut-liver axis by altering the intestinal microbiota, epithelial barrier integrity, microbial metabolites, bile acid signaling, and host immunity. This review examines how these alterations relate to clinical outcomes and evaluates evidence for microbiota-targeted interventions, including probiotics, synbiotics, nutritional optimization, antibiotic stewardship, bile acid modulation, and emerging multiomics strategies. We distinguish liver resection from living-donor and deceased-donor liver transplantation because the patient populations, graft or remnant anatomy, ischemia-reperfusion exposures, immune status, and outcome definitions differ. Clinical evidence most consistently supports selected pro-/synbiotic strategies for reducing postoperative infection in higher-risk settings, whereas microbiome-based prediction of PHLF, fecal microbiota transplantation (FMT), bile acid-directed therapy, and precision multiomics-guided pathways remain investigational. Future work should use transparent literature identification, standardized perioperative protocols, risk-defined populations, external validation, and prospective multicenter trials. A better understanding of gut-liver interactions may help preserve beneficial host-microbial signals while limiting translocation and inflammation during recovery.

PMID:42683887 | DOI:10.3791/73747

❌