Experimental study for the comparison of classifier combination methods

作者:

Highlights:

摘要

In this paper, we compare the performances of classifier combination methods (bagging, modified random subspace method, classifier selection, parametric fusion) to logistic regression in consideration of various characteristics of input data. Four factors used to simulate the logistic model are: (a) combination function among input variables, (b) correlation between input variables, (c) variance of observation, and (d) training data set size. In view of typically unknown combination function among input variables, we use a Taguchi design to improve the practicality of our study results by letting it as an uncontrollable factor. Our experimental study results indicate the following: when training set size is large, performances of logistic regression and bagging are not significantly different. However, when training set size is small, the performance of logistic regression is worse than bagging. When training data set size is small and correlation is strong, both modified random subspace method and bagging perform better than the other three methods. When correlation is weak and variance is small, both parametric fusion and classifier selection algorithm appear to be the worst at our disappointment.

论文关键词:Bagging,Random subspace method,Classifier selection,Parametric fusion

论文评审过程:Received 7 June 2005, Revised 8 May 2006, Accepted 29 June 2006, Available online 6 September 2006.

论文官网地址:https://doi.org/10.1016/j.patcog.2006.06.027