Resolution limits for process comparison from event data
cs.DB, cs.LG, stat.ML
Submitted: 2026-09-17
Updated: 2026-09-17
Comments: 35 pages, 4 figures; 13-page supplementary material as an ancillary file
Code: https://github.com/AntonyRLee/proc-posets
License: http://creativecommons.org/licenses/by/4.0/
The gist: One hospital runs bloods and imaging at the same time.
Terminology
Abstract
One hospital runs bloods and imaging at the same time. Another runs them one after the other, in either order, equally often. Knowing which actually happened, and how it is recorded in data, is critical for all operational managers. In process mining, the standard approach is to construct an event log, and attempt to discover concurrent and sequential processes in a data-driven way. We show this standard approach, built on the stochastic language of an event log, reports only the assumptions of its discovery algorithm, because every such log is explained equally well by a model with no concurrency at all. Further, before any data is acquired, we characterise when data can and cannot distinguish concurrent behaviour. Where it cannot, the distinction is recoverable from evidence the stochastic language discards, such as the times at which activities start and end, or object-centric records that fix an order within an execution. The remedy is therefore a choice of what is recorded, rather than a larger sample. This impacts decision making, as planning resource for truly concurrent services is very different from sequential services.
Sources
- OCEL (Object-Centric Event Log) 2.0 Specification
- Advancements and Challenges in Object-Centric Process Mining: A Systematic Literature Review
- Process Comparison Using Object-Centric Process Cubes
Related papers
- Vibe Coding on Trial: Operating Characteristics of Unanimous LLM Juries
- Human-Level Text-to-SQL via Reinforcement Learning on Verified Data, Without Pipeline Engineering
- Bridging Business Intent and Data: A Benchmark for Automatic Relational Data Product Generation
- DexterSQL: Deep Schema Exploration and Rule-based Correction for Text-to-SQL Generation
- MaDI-Bench: An End-to-End Data Integration Benchmark
- Eigenius: A Typed Knowledge-Graph DBMS with Epistemic Stratification and Institution-Mediated Reasoning