arXiv Machine Learning

Resolution limits for process comparison from event data

The paper examines how event logs used in process mining can fail to reveal concurrent versus sequential activities, using a hospital example where blood tests and imaging may occur simultaneously or in alternating order. It demonstrates that standard stochastic language approaches only expose the assumptions of their discovery algorithms, often misrepresenting concurrency. The authors argue that the key to distinguishing concurrent behavior lies in the choice of recorded data—such as precise start and end times or object‑centric ordering—rather than simply increasing sample size.

Hugging Face Trending Papers
Sep 17

Resolution limits for process comparison from event data

The paper examines how event data from hospital processes can obscure whether activities occur concurrently or sequentially. It shows that standard event‑log approaches, based on stochastic language, often fail to distinguish concurrency because any log can be explained by a model with no concurrent events. The authors argue that the key to resolving this ambiguity lies in the choice of what is recorded—such as precise start and end times or object‑centric ordering—rather than simply collecting more data.

arXiv Machine Learning
Sep 22

Concurrency-Aware Process Model Forecasting with Causal Nets

The paper introduces a new approach to process model forecasting that uses causal nets instead of traditional directly-follows graphs, enabling explicit representation of concurrency. It forecasts time series of relation and binding counts, reconstructs future process models with AND/XOR semantics, and evaluates them using a protocol that handles partial traces for conformance checking. Experiments on four event logs show that the forecasted models achieve conformance close to re‑mined models and outperform static discovery baselines, though filtering infrequent bindings improves metrics at the cost of losing concurrent behavior.

By Yongbo Yu, Jari Peeperkorn, Johannes De Smedt, Jochen De Weerdt
arXiv AI
Sep 2

Automated Event Log Generation from Unstructured Text Using Finetuned LLMs

The paper introduces a scalable framework that uses finetuned large language models (LLMs) to translate unstructured textual resources into structured event logs for process mining. By creating a new text-to-log dataset and finetuning LLMs on it, the authors demonstrate that the resulting models produce high‑fidelity event logs, outperforming few‑shot or zero‑shot prompting methods. This approach enables previously unused organizational data, such as incident tickets and manuals, to be incorporated into process mining workflows.

By Maximilian Seeth, Gabriel Marques Tavares, Daniel Schuster
arXiv AI
Aug 28

Learning to Predict, Discover, and Reason in High-Dimensional Event Sequences

The paper proposes a new framework for automated fault diagnostics in modern vehicles by treating diagnostic trouble codes (DTCs) as a high‑dimensional language. It introduces Transformer‑based models for predictive maintenance, scalable causal discovery methods, and a multi‑agent system that automatically generates Boolean error‑pattern rules. The approach aims to replace costly manual grouping of DTCs with scalable, data‑driven techniques.

By Hugo Math
arXiv AI
Jun 24

A global log for medical AI

arXiv:2510. 04033v2 Announce Type: replace Abstract: Modern computer systems rely on syslog, a universal protocol that records critical events across heterogeneous infrastructure.

By Ayush Noori, Aaron E. Boussina, Hai Ho Bich, James Anibal, Julia Maslinski, Manuel Burger, Martin Faltys, Adam Rodman, Alan Karthikesalingam, Alessandro Blasimme, Annelia Itwaru, Ben Kaplan, Bilal A. Mateen, Christopher A. Longhurst, Daniel Yang, Dave deBronkart, Effy Vayena, Fedor Sergeev, Gauden Galea, Ha Thi Hai Duong, Harold F. Wolf III, Jacob Waxman, Joerg C. Schefold, Joshua C. Mandel, Juliana Rotich, Kenneth D. Mandl, Lily Poursoltan, Maryam Mustafa, Melissa Miles, Nigam H. Shah, Noa Dagan, Pavan Bodanki, Peter Lee, Philipp Koralus, Prathamesh Parchure, Prem Timsina, Ran D. Balicer, Robert Korom, Scott Mahoney, Seth Hain, Tien Yin Wong, Trevor Mundel, Vivek Natarajan, Ankit Sakhuja, Benjamin Glicksberg, C. Louise Thwaites, Gunnar R\"atsch, Karandeep Singh, David A. Clifton, Isaac S. Kohane, Marinka Zitnik