Abstract <p>This paper presents a machine learning (ML)-based heuristic for finding the optimum sub-system size for the CUDA implementation of the parallel partition algorithm. Computational experiments for different system of linear algebraic equation (SLAE) sizes are conducted, and the optimum subsystem size for each of them is found empirically. To estimate a model for the subsystem size, we perform the k-nearest neighbors (kNN) classification method. Statistical analysis of the results is done. By comparing the predicted values with the actual data, the algorithm is deemed to be acceptably good. Next, the heuristic is expanded to work for the recursive parallel partition algorithm as well. An algorithm for determining the optimum subsystem size for each recursive step is formulated. A kNN model for predicting the optimum number of recursive steps for a particular SLAE size is built.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

ML-Based Optimum Subsystem Size for the GPU Implementation of the Tridiagonal Partition Method

  • M. Veneva

摘要

Abstract

This paper presents a machine learning (ML)-based heuristic for finding the optimum sub-system size for the CUDA implementation of the parallel partition algorithm. Computational experiments for different system of linear algebraic equation (SLAE) sizes are conducted, and the optimum subsystem size for each of them is found empirically. To estimate a model for the subsystem size, we perform the k-nearest neighbors (kNN) classification method. Statistical analysis of the results is done. By comparing the predicted values with the actual data, the algorithm is deemed to be acceptably good. Next, the heuristic is expanded to work for the recursive parallel partition algorithm as well. An algorithm for determining the optimum subsystem size for each recursive step is formulated. A kNN model for predicting the optimum number of recursive steps for a particular SLAE size is built.