Link Prediction for Event Logs in the Process Industry

  • 2025-08-12 17:22:29
  • Anastasia Zhukova, Thomas Walton, Christian E. Matt, Bela Gipp
  • 0

Abstract

Knowledge management (KM) is vital in the process industry for optimizingoperations, ensuring safety, and enabling continuous improvement througheffective use of operational data and past insights. A key challenge in thisdomain is the fragmented nature of event logs in shift books, where relatedrecords, e.g., entries documenting issues related to equipment or processes andthe corresponding solutions, may remain disconnected. This fragmentationhinders the recommendation of previous solutions to the users. To address thisproblem, we investigate record linking (RL) as link prediction, commonlystudied in graph-based machine learning, by framing it as a cross-documentcoreference resolution (CDCR) task enhanced with natural language inference(NLI) and semantic text similarity (STS) by shifting it into the causalinference (CI). We adapt CDCR, traditionally applied in the news domain, intoan RL model to operate at the passage level, similar to NLI and STS, whileaccommodating the process industry's specific text formats, which containunstructured text and structured record attributes. Our RL model outperformedthe best versions of NLI- and STS-driven baselines by 28% (11.43 points) and27% (11.21 points), respectively. Our work demonstrates how domain adaptationof the state-of-the-art CDCR models, enhanced with reasoning capabilities, canbe effectively tailored to the process industry, improving data quality andconnectivity in shift logs.

 

Quick Read (beta)

loading the full paper ...