StatApriori: an efficient algorithm for searching statistically significant association rules

作者:Wilhelmiina Hämäläinen

摘要

Searching statistically significant association rules is an important but neglected problem. Traditional association rules do not capture the idea of statistical dependence and the resulting rules can be spurious, while the most significant rules may be missing. This leads to erroneous models and predictions which often become expensive. The problem is computationally very difficult, because the significance is not a monotonic property. However, in this paper, we prove several other properties, which can be used for pruning the search space. The properties are implemented in the StatApriori algorithm, which searches statistically significant, non-redundant association rules. Empirical experiments have shown that StatApriori is very efficient, but in the same time it finds good quality rules.

论文关键词:Association rule, Statistical significance, Dependence, StatApriori, Search algorithm

论文评审过程:

论文官网地址:https://doi.org/10.1007/s10115-009-0229-8