arXiv:2506. 01883v3 Announce Type: replace-cross Abstract: Training deep learning models on single-cell datasets with hundreds of millions of cells requires loading data from disk, as these datasets exceed available memory.
By Davide D'Ascenzo, Sebastiano Cultrera di Montesano
arXiv:2609.39215v1 Announce Type: cross
Abstract: Time series anomaly detection (TSAD) is increasingly deployed in streaming settings, where data arrive sequentially and may exhibit non-stationarity....
By Magali Parrino, Antoine Ajenjo, Emmanuel Remy, Pierre Stephan, Pierre Senellart, Paul Boniol
arXiv:2608.30923v1 Announce Type: cross
Abstract: Stream learning is commonly evaluated through predictive performance and adaptation to concept drift. However, sustained operation of a stream learne...
By Sebastian Buschj\"ager, Nuwan Gunasekara, Heitor Murilo Gomes
arXiv:2505. 06835v5 Announce Type: replace Abstract: Sliced optimal transport (SOT), or sliced Wasserstein (SW) distance, is widely recognized for its statistical and computational scalability.
By Khai Nguyen
arXiv:2607. 28880v1 Announce Type: cross Abstract: Modern multimedia machine learning workloads increasingly store large-scale datasets in cloud object storage services such as AWS S3.
By Debopam Sanyal, Hongjie Chen, Alexey Tumanov, Joshua Kimball
arXiv:2608. 00720v1 Announce Type: cross Abstract: Mapping neural networks to FPGAs enables low-latency, energy-efficient inference, particularly for lookup table (LUT)-based models that eliminate multipliers and map directly to reconfigurable fabric.
By Oliver Cassidy, Marta Andronic, George A. Constantinides
arXiv:2502. 06434v2 Announce Type: replace-cross Abstract: Dataset pruning (DP) and dataset distillation (DD) fundamentally differ in their outputs: DP selects original image subsets, while DD generates synthetic images.
By Lingao Xiao, Songhua Liu, Yang He, Xinchao Wang
StreamTTT is a streaming vision-language model that balances real-time perception with long-term memory by writing long-range history into fast weights outside the attention context, while keeping a short sliding key-value cache for recent evidence. The model is trained on both offline long-video QA and a new real-time QA corpus, and it outperforms SimpleStream-4B on OVO-Bench by 1.4 points in real-time perception and 3.7 points in backward tracing. StreamTTT-4B also competes with the larger SimpleStream-8B on the StreamingBench Real-Time Visual Understanding subset.
By Joya Chen, Zeyun Zhong, Mike Zheng Shou
Maia 200 is a software‑defined dataflow system that delivers high performance AI acceleration, achieving 10,145 Tflop/s in FP4 and 5,072 Tflop/s in FP8 within a 750 W TDP and 7 TB/s HBM bandwidth. It exemplifies a new class of Software Defined Locally Accessed Dataflow Architectures (SDLA), which program dataflow engines to orchestrate specialized memories and data‑movement engines, shifting focus from thread‑centric to data‑movement‑centric design. The system offers significant cost and energy savings while supporting massive parallelism for AI inference workloads, positioning it as a compelling solution for next‑generation high‑performance computing.
By Sherry Xu, Marco Heddes, Jackson Peng, Tom Savell, Monica Tang, Prashant Ranjan, Jesse Benson, Ofer Dekel, Saurabh Dighe, Anupama Kurpad, Artour Levin, Matthew Mattina, George Petre, Cheng Tang, Yuan Yu, Li Zhang, Torsten Hoefler
arXiv:2601. 07048v5 Announce Type: replace-cross Abstract: Approximate nearest neighbor search (ANNS) is a core problem in machine learning and information retrieval applications.
By Hunter McCoy, Zikun Wang, Prashant Pandey
arXiv:2609.07956v1 Announce Type: new
Abstract: Tabular Foundation Models (TFMs) have recently demonstrated strong predictive performance through in-context learning, but their deployment in high-thr...
By Vitor Crista, Afonso Louren\c{c}o, Diogo Martinho, Goreti Marreiros