❌

Reading view

Finch: Benchmarking Finance & Accounting across Spreadsheet-Centric Enterprise Workflows

arXiv:2512.13168v4 Announce Type: replace Abstract: We introduce FinWorkBench (a.k.a. Finch), a benchmark for evaluating agents on real-world, enterprise-grade finance and accounting workflows that interleave data entry, structuring, formatting, web search, cross-file retrieval, calculation, modeling, validation, translation, visualization, and reporting. Finch is built from authentic enterprise workspaces from Enron (15,000 files and 500,000 emails) and other financial institutions spanning 2000 to 2025, preserving the in-the-wild messiness of multimodal artifacts such as tables and charts across diverse domains including budgeting, trading, and asset management. We propose a workflow construction process that combines LLM-assisted mining of workflows from authentic enterprise environments with expert annotation. Specifically, we use LLM-assisted, expert-verified derivation of workflows from real-world email threads and spreadsheet version histories, followed by meticulous workflow annotation requiring more than 700 hours of expert effort. This process yields 172 composite workflows with 384 tasks, involving 1,710 spreadsheets with 27 million cells, along with PDFs and other artifacts, capturing the intrinsically messy, long-horizon, knowledge-intensive, and collaborative nature of enterprise work. We conduct both human and automated evaluations of frontier AI systems, including GPT 5.1, Claude Sonnet/Opus 4.5, Gemini 3 Pro, Grok 4, and Qwen 3 Max. GPT 5.1 Pro spends an average of 16.8 minutes per workflow yet passes only 38.4% of workflows. Comprehensive case studies further highlight the challenges that real-world enterprise workflows pose for AI agents.
  •  

Pathogenesis and immune regulation of rheumatoid arthritis-associated interstitial lung disease: from basic research to clinical implications

Front Immunol. 2026 Mar 13;17:1770348. doi: 10.3389/fimmu.2026.1770348. eCollection 2026.

ABSTRACT

Interstitial lung disease (ILD) is one of the most common extra-articular manifestations of rheumatoid arthritis (RA). Some patients with RA-ILD may develop progressive pulmonary fibrosis, leading to severe impairment of lung function and respiratory failure, which impacts quality of life and can even be life-threatening. This review identified genetic susceptibility, environmental factors, and immune dysregulation as key contributors to the etiology and pathogenesis of RA-ILD. We highlight that autoantibodies, adaptive immune abnormalities, and tertiary lymphoid organ formation significantly drive pulmonary inflammation and fibrosis, while pro-inflammatory cytokines and epithelial-mesenchymal transition (EMT) further contribute to lung tissue injury. Current treatment options, including glucocorticoids, immunosuppressants, and antifibrotic agents such as nintedanib and pirfenidone, are often limited by substantial side effects. Additionally, emerging therapies like JAK inhibitors, CAR-T cells, and the upcoming phosphodiesterase-4B inhibitor, nerandomilast, show promise, but no curative treatment exists to date. Future research could focus on multi-omics technologies and conducting multicenter clinical trials to establish therapeutic targets and advance precision medicine for RA-ILD.

PMID:41909710 | PMC:PMC13021622 | DOI:10.3389/fimmu.2026.1770348

  •  

AttestLLM: Efficient Attestation Framework for Billion-scale On-device LLMs

arXiv:2509.06326v2 Announce Type: replace-cross Abstract: As on-device LLMs(e.g., Apple on-device Intelligence) are widely adopted to reduce network dependency, improve privacy, and enhance responsiveness, verifying the legitimacy of models running on local devices becomes critical. Existing attestation techniques are not suitable for billion-parameter Large Language Models (LLMs), struggling to remain both time- and memory-efficient while addressing emerging threats in the LLM era. In this paper, we present AttestLLM, the first-of-its-kind attestation framework to protect the hardware-level intellectual property (IP) of device vendors by ensuring that only authorized LLMs can execute on target platforms. AttestLLM leverages an algorithm/software/hardware co-design approach to embed robust watermarking signatures onto the activation distributions of LLM building blocks. It also optimizes the attestation protocol within the Trusted Execution Environment (TEE), providing efficient verification without compromising inference throughput. Extensive proof-of-concept evaluations on LLMs from Llama, Qwen, and Phi families for on-device use cases demonstrate AttestLLM's attestation reliability, fidelity, and efficiency. Furthermore, AttestLLM enforces model legitimacy and exhibits resilience against model replacement and forgery attacks.
  •  
❌