A Study of Quality and Accuracy Trade-offs in Process Mining

Zan Huang,Akhil Kumar

doi:10.1287/ijoc.1100.0444

Abstract

In recent years, many algorithms have been proposed to extract process models from process execution logs. The process models describe the ordering relationships between tasks in a process in terms of standard constructs like sequence, parallel, choice, and loop. Most algorithms assume that each trace in a log represents a correct execution sequence based on a model. In practice, logs are often noisy, and algorithms designed for correct logs are not able to handle noisy logs. In this paper we share our key insights from a study of noise in process logs both real and synthetic. We found that all process logs can be explained by a block-structured model with two special self-loop and optional structures, making it trivial to build a fully accurate process model for any given log, even one with inaccurate data or noise present in it. However, such a model suffers from low quality. By controlling the use of self-loop and optional structures of tasks and blocks of tasks, we can balance the quality and accuracy trade-off to derive high-quality process models that explain a given percentage of traces in the log. Finally, new quality metrics and a novel quality-based algorithm for model extraction from noisy logs are described. The results of the experiments with the algorithm on real and synthetic data are reported and analyzed at length.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A Study of Quality and Accuracy Trade-offs in Process Mining

Abstract

Talk to us

Similar Papers

More From: INFORMS Journal on Computing

Lead the way for us

Journal: INFORMS Journal on Computing	Publication Date: May 1, 2012
Citations: 27

Similar Papers

Process mining on noisy logs — Can log sanitization help to improve performance?
Hsin-Jung Cheng ... Akhil Kumar
Decision Support Systems | VOL. 79
Hsin-Jung Cheng, et. al.Hsin-Jung Cheng ... Akhil Kumar
21 Aug 2015
Decision Support Systems | VOL. 79

A principled approach to mining from noisy logs using Heuristics Miner
Philip Weber ... Peter Tino
-
Philip Weber, et. al.Philip Weber ... Peter Tino
01 Apr 2013
01 Apr 2013

A novel approach to process mining: Intentional process models discovery
Ghazaleh Khodabandelou ... Charlotte Hug
-
Ghazaleh Khodabandelou, et. al.Ghazaleh Khodabandelou ... Charlotte Hug
01 May 2014
01 May 2014

An Automatic Business Process Modeling Method Based on Markov Transition Matrix in BPM
Li Yan ... Feng Yu-Qiang
-
Li Yan, et. al.Li Yan ... Feng Yu-Qiang
01 Jan 2006
01 Jan 2006

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A Study of Quality and Accuracy Trade-offs in Process Mining

Abstract

Talk to us

Similar Papers

More From: INFORMS Journal on Computing