Combining explainable artificial intelligence and information visualization holds great potential for users to understand and reason about complex multidimensional sequential data. This work proposes a semi-supervised two-step approach for extracting long- and short-term patterns in low-dimensional representations of sequential data. First, unsupervised sequence clustering is used to identify long-term patterns. Second, these long-term patterns serve as supervisory information for training a self-attention-based sequence classification model. The resulting feature embedding is used to identify short-term patterns. The approach is validated on a self-generated dataset consisting of heart-shaped paths with different sampling rates, rotations, scales, and translations. The results demonstrate the approach’s effectiveness for clustering semantically similar paths and/or path sequences. This detection of both global long-term patterns and local short-term patterns facilitates the understanding and reasoning about complex multidimensional sequential data.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Explainable Long and Short-Term Pattern Detection in Projected Sequential Data

  • Matthias Bittner,
  • Andreas Hinterreiter,
  • Klaus Eckelt,
  • Marc Streit

摘要

Combining explainable artificial intelligence and information visualization holds great potential for users to understand and reason about complex multidimensional sequential data. This work proposes a semi-supervised two-step approach for extracting long- and short-term patterns in low-dimensional representations of sequential data. First, unsupervised sequence clustering is used to identify long-term patterns. Second, these long-term patterns serve as supervisory information for training a self-attention-based sequence classification model. The resulting feature embedding is used to identify short-term patterns. The approach is validated on a self-generated dataset consisting of heart-shaped paths with different sampling rates, rotations, scales, and translations. The results demonstrate the approach’s effectiveness for clustering semantically similar paths and/or path sequences. This detection of both global long-term patterns and local short-term patterns facilitates the understanding and reasoning about complex multidimensional sequential data.