❌

Normal view

Nonsense-mediated mRNA decay inhibition reshapes the cancer immunopeptidome

Immunity. 2026 Apr 8:S1074-7613(26)00075-0. doi: 10.1016/j.immuni.2026.02.005. Online ahead of print.

ABSTRACT

DNA mutations are a well-characterized source of neoepitopes in immunotherapy. Here, we examined the contribution of dysregulated RNA processing to neoantigen production. Leveraging multi-omics and checkpoint inhibitor (CPI) response data from >1,000 patients, we identified reduced activity of the nonsense-mediated mRNA decay (NMD) pathway kinase SMG1 as a predictor of improved CPI response. NMD inhibition through SMG1 targeting stabilized transcripts containing premature termination codons, most of which were of non-mutational origin. This reshaped the major histocompatibility complex class I (MHC class I)-bound immunopeptidome and increased neoantigen abundance to levels comparable to high mutation burden tumors. Functionally, NMD inhibition drove antigen-dependent T cell-mediated tumor cell killing in vitro, promoted activation of tissue-resident T cells in patient-derived models ex vivo, and improved CPI efficacy in vivo. Our findings establish NMD inhibition as a strategy to harness a previously inaccessible source of canonical and non-canonical neoantigens, with the potential to increase tumor immunogenicity across cancers.

PMID:41956098 | DOI:10.1016/j.immuni.2026.02.005

Individual and Combined Effects of English as a Second Language and Typos on LLM Performance

arXiv:2604.04723v1 Announce Type: cross Abstract: Large language models (LLMs) are used globally, and because much of their training data is in English, they typically perform best on English inputs. As a result, many non-native English speakers interact with them in English as a second language (ESL), and these inputs often contain typographical errors. Prior work has largely studied the effects of ESL variation and typographical errors separately, even though they often co-occur in real-world use. In this study, we use the Trans-EnV framework to transform standard English inputs into eight ESL variants and apply MulTypo to inject typos at three levels: low, moderate, and severe. We find that combining ESL variation and typos generally leads to larger performance drops than either factor alone, though the combined effect is not simply additive. This pattern is clearest on closed-ended tasks, where performance degradation can be characterized more consistently across ESL variants and typo levels, while results on open-ended tasks are more mixed. Overall, these findings suggest that evaluations on clean standard English may overestimate real-world model performance, and that evaluating ESL variation and typographical errors in isolation does not fully capture model behavior in realistic settings.

Genomic history of early dogs in Europe

Nature, Published online: 25 March 2026; doi:10.1038/s41586-026-10112-7

Genome-wide analysis shows European dogs existed by 14,200 years ago, were already genetically distinct, received less Neolithic Southwest Asian admixture than humans did and contributed substantially to later European dogs.

A Review of the Role of Zeqi Decoction in the Treatment of Non-Small Cell Lung Cancer

18 March 2026 at 18:00

J Multidiscip Healthc. 2026 Mar 11;19:584071. doi: 10.2147/JMDH.S584071. eCollection 2026.

ABSTRACT

Non-small cell lung cancer (NSCLC) is one of the malignant tumors with the highest incidence and mortality rates. Zeqi Decoction has the functions of "promoting diuresis and reducing swelling, resolving phlegm and dispersing nodules", embodying the unique approach of traditional Chinese medicine in treating lung cancer by "strengthening the body's resistance and eliminating pathogenic factors". Modern research shows that Zeqi Decoction exerts anti-NSCLC effects through multiple pathways and targets. In terms of the material basis of its efficacy, its active ingredients (such as diterpene esters and flavonoids contained in Zeqi) have the ability to directly inhibit the proliferation, invasion and migration of tumor cells and induce apoptosis. In terms of the mechanism of action, basic experiments have revealed that Zeqi Decoction can down-regulate the S100A9/STAT3 signaling pathway, inhibit the immunosuppressive activity of myelium-derived suppressor cells (MDSCs), reshape the tumor microenvironment, thereby enhancing the cytotoxic function of CD8⁺T cells, and can also regulate the EGFR/PI3K/Akt pathway to affect PD-L1 expression. Intervene in tumor immune escape; In terms of clinical transformation, the combination of Zexi Decoction with chemotherapy and targeted therapy can improve patients' symptoms such as cough and pleural effusion, prolong progression-free survival, and alleviate the toxic and side effects of Western medical treatment. In addition, Zexi Decoction also shows potential value in reversing drug resistance such as gemcitabine. At present, there are still problems such as the lack of standardized protocols and unclear molecular mechanisms in the research. In the future, it is necessary to combine new technologies such as network pharmacology and multi-omics analysis to deepen the research on the pharmacological material basis, dose-effect relationship and evidence-based medicine of Zeqi Decoction, so as to promote the clinical application and transformation of the combination of traditional Chinese and Western medicine in the treatment of NSCLC.

PMID:41847115 | PMC:PMC12991379 | DOI:10.2147/JMDH.S584071

R1-Code-Interpreter: LLMs Reason with Code via Supervised and Multi-stage Reinforcement Learning

arXiv:2505.21668v3 Announce Type: replace Abstract: Practical guidance on training Large Language Models (LLMs) to leverage Code Interpreter across diverse tasks remains lacking. We present R1-Code-Interpreter, an extension of a text-only LLM trained via multi-turn supervised fine-tuning (SFT) and reinforcement learning (RL) to autonomously generate multiple code queries during step-by-step reasoning. Unlike prior RL + tool-use efforts focused on narrow domains such as math or retrieval, we curate 144 diverse reasoning and planning tasks and show that training a general-purpose Code Interpreter across them presents significant challenges due to task heterogeneity and scarcity of effective samples. To address this, we introduce a multi-stage curriculum learning approach that partitions training samples by measured improvement potential. The RL training prioritizes samples with higher potential and gradually shifts to lower-potential ones, increasing the average RL gains from merely +3.4% to +9.3% across Qwen-2.5 models (3/7/14B). Our final model, R1-CI-14B, improves average accuracy on the 37 test tasks from 44.1% to 72.4%, outperforming text-only GPT-4o (58.6%) and GPT-4o with Code Interpreter (70.9%). Notably, R1-CI-14B also exhibits emergent self-checking behavior through code generation. Datasets, Codes, and Models are available at https://github.com/yongchao98/R1-Code-Interpreter and https://huggingface.co/yongchao98.

Closing the Gap Between Text and Speech Understanding in LLMs

arXiv:2510.13632v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) can be adapted to extend their text capabilities to speech inputs. However, these speech-adapted LLMs consistently underperform their text-based counterparts--and even cascaded pipelines--on language understanding tasks. We term this shortfall the text-speech understanding gap: the performance drop observed when a speech-adapted LLM processes spoken inputs relative to when the original text-based LLM processes the equivalent text. Recent approaches to narrowing this gap either rely on large-scale speech synthesis of text corpora, which is costly and heavily dependent on synthetic data, or on large-scale proprietary speech datasets, which are not reproducible. As a result, there remains a need for more data-efficient alternatives for closing the text-speech understanding gap. In this work, we analyze the gap as driven by two factors: (i) forgetting of text capabilities during adaptation, and (ii) cross-modal misalignment between speech and text. Based on this analysis, we introduce SALAD--Sample-efficient Alignment with Learning through Active selection and cross-modal Distillation--which combines cross-modal distillation with targeted synthetic data to improve alignment while mitigating forgetting. Applied to 3B and 7B LLMs, SALAD achieves competitive performance with a strong open-weight model across broad-domain benchmarks in knowledge, language understanding, and reasoning, while training on over an order of magnitude less speech data from public corpora.

EBPO: Empirical Bayes Shrinkage for Stabilizing Group-Relative Policy Optimization

arXiv:2602.05165v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has proven effective for enhancing the reasoning capabilities of Large Language Models (LLMs). However, dominant approaches like Group Relative Policy Optimization (GRPO) face critical stability challenges: they suffer from high estimator variance under computational constraints (small group sizes) and vanishing gradient signals in saturated failure regimes where all responses yield identical zero rewards. To address this, we propose Empirical Bayes Policy Optimization (EBPO), a novel framework that regularizes local group-based baselines by borrowing strength from the policy's accumulated global statistics. Instead of estimating baselines in isolation, EBPO employs a shrinkage estimator that dynamically balances local group statistics with a global prior updated via Welford's online algorithm. Theoretically, we demonstrate that EBPO guarantees strictly lower Mean Squared Error (MSE), bounded entropy decay, and non-vanishing penalty signals in failure scenarios compared to GRPO. Empirically, EBPO consistently outperforms GRPO and other established baselines across diverse benchmarks, including AIME and OlympiadBench. Notably, EBPO exhibits superior training stability, achieving high-performance gains even with small group sizes, and benefits significantly from difficulty-stratified curriculum learning.

What Do We Mean by 'Pilot Study': Early Findings from a Meta-Review of Pilot Study Reporting at CHI

arXiv:2602.13488v1 Announce Type: cross Abstract: Pilot studies (PS) are ubiquitous in HCI research. CHI papers routinely reference 'pilot studies', 'pilot tests', or 'preliminary studies' to justify design decisions, verify procedures, or motivate methodological choices. Yet despite their frequency, the role of pilot studies in HCI remains conceptually vague and empirically underexamined. Unlike fields such as medicine, nursing, and education, where pilot and feasibility studies have well-established definitions, guidelines, reporting standards and even a dedicated research journal, the CHI community lacks a shared understanding of what constitutes a pilot study, why they are conducted, and how they should be reported. Many papers reference pilots 'in passing', without details about design, outcomes, or how the pilot informed the main study. This variability suggests a methodological blind spot in our community.
❌