arXiv Machine Learning

From system models to class models: An in-context learning paradigm

arXiv:2308. 13380v3 Announce Type: replace-cross Abstract: Is it possible to understand the intricacies of a dynamical system not solely from its input/output pattern, but also by observing the behavior of other systems within the same class?

arXiv AI
Jul 28

Extracting Algorithms in Pre-trained LLMs: A Case on Hidden Markov Models

arXiv:2607. 22646v1 Announce Type: new Abstract: Large language models (LLMs) display a striking ability to predict next observations from Hidden Markov Models (HMMs) via in-context learning (ICL), but the algorithm underlying this capability remains undetermined: prior work has proposed several candidates without consensus, and none has been grounded in the model's internal activations.

By Yijia Dai, Zhaolin Gao, Yahya Sattar, Jennifer J. Sun, Sarah Dean
Hugging Face Trending Papers
Jul 13

Invariant Learning Dynamics of Transformers in Inductive Reasoning Tasks

We present a theoretical framework to explain the emergence of inductive reasoning abilities in Transformer language models. While previous works on Transformer learning dynamics have so far been mostly tied to specific tasks, we study a generalized class of inductive tasks that unifies several synthetic tasks known in the literature, including in-context n-grams and multi-hop reasoning.

arXiv Machine Learning
Aug 27

SAMpLE: A SystemC-AMS Machine LEarning-based Framework for Virtual Prototyping

SAMpLE is an open‑source SystemC‑AMS framework that treats machine learning models as first‑class Timed Dataflow components via a standardized plug‑and‑play interface. It offers a native C++ backend for online training of lightweight models and an offline backend that runs externally developed models without re‑implementation. By using ONNX as a model exchange format, SAMpLE enables the integration and evaluation of diverse ML solutions within a single, reproducible simulation workflow.

By Andrei Mihai Albu, Sara Vinco
arXiv AI
Sep 10

Memory in Deep Time-Series Models

arXiv:2609.06006v1 Announce Type: cross Abstract: Deep learning for time series has progressed through successive architectural paradigms, from recurrent networks and transformers to structured state...

By Minh Hoang Nguyen, Huu Hiep Nguyen, Manh Nguyen, Van Dai Do, Dung Nguyen, Hung Le