arXiv Machine Learning

SwiftRepertoire: Few-Shot Immune-Signature Synthesis via Dynamic Kernel Codes

arXiv:2602. 01051v5 Announce Type: replace Abstract: Repertoire-level analysis of T cell receptors offers a biologically grounded signal for disease detection and immune monitoring, yet practical deployment is impeded by label sparsity, cohort heterogeneity, and the computational burden of adapting large encoders to new tasks.

arXiv AI
Jul 23

SubQuad: Near-Quadratic-Free Structure Inference with Distribution-Balanced Objectives in Adaptive Receptor framework

arXiv:2602. 17330v5 Announce Type: replace-cross Abstract: Comparative analysis of adaptive immune repertoires at population scale is hampered by two practical bottlenecks: the near-quadratic cost of pairwise affinity evaluations and dataset imbalances that obscure clinically important minority clonotypes.

By Rong Fu, Zijian Zhang, Kun Liu, Jiekai Wu, Xianda Li, Simon Fong
arXiv Machine Learning
4d ago

Estimating the Causal Effects of T Cell Receptors

The paper introduces a method for estimating the causal effects of T cell receptor (TCR) sequences on patient outcomes using observational TCR sequencing and clinical data. It corrects for unobserved confounders by leveraging the pre-selection TCR repertoire generated through V(D)J recombination as a natural experiment, and employs permutation‑invariant neural networks to scale to millions of sequences. The approach is validated on semisynthetic data and applied to COVID‑19 severity, identifying TCRs that are observed in patients, bind SARS‑CoV‑2 antigens in vitro, and positively influence clinical outcomes.

By Eli N. Weinstein, Elizabeth B. Wood, David M. Blei
arXiv AI
4d ago

Explainability from Training with Applications to TCR-Epitope Prediction

The paper introduces Explainability from Training (EFT), a model‑agnostic method that tracks how deep learning models learn and organize evidence during training. EFT is applied to four leading T cell receptor‑epitope prediction models, revealing distinct learning trajectories for CNNs and transformers, conflicts between TCR alpha and beta chain evidence, and differences in feature preferences when using real versus predicted structural data. The authors also present a new benchmark, TCR‑XAI2, comprising 388 experimentally resolved TCR‑epitope structures and several predicted models to evaluate these insights.

By Jiarui Li, Zixiang Yin, Samuel Landry, Zhengming Ding, Ramgopal Mettu
arXiv Machine Learning
Jun 9

Integrating gene regulatory priors into Transformer attention with scTransformer for interpretable scRNA-seq analysis

arXiv:2606. 09558v1 Announce Type: cross Abstract: Motivation: Transformer-based models are increasingly applied to large-scale single-cell transcriptomics, showing strong performance through self-supervised learning on millions of cells.

By Mikele Milia, Louis Fabrice Tshimanga, Henning Mueller, Manfredo Atzori, Barbara Di Camillo
arXiv AI
Jun 30

Data-Efficient Multimodal Alignment for Histopathology-based Molecular Prediction

arXiv:2606. 29949v1 Announce Type: cross Abstract: H&E-stained whole-slide images offer cohort-scale availability and rich spatial context but lack molecular specificity, whereas bulk RNA-seq provides transcriptome-wide resolution at high cost with limited archival availability.

By Dominik Winter, Dominik Vonficht, Lo\"ic Le Bescond, Christian Gebbe, Marco Rosati, Richard J. Chen, Markus Schick, Ross Stewart, Nicolas Brieu