论文信息 - Solving Multiclass Learning Problems via Error-Correcting Output Codes

Solving Multiclass Learning Problems via Error-Correcting Output Codes

Multiclass learning problems involve finding a definition for an unknown function f(x) whose range is a discrete set containing k > 2 values (i.e., k "classes"). The definition is acquired by studying collections of training examples of the form (xi, f(xi)). Existing approaches to multiclass learning problems include direct application of multiclass algorithms such as the decision-tree algorithms C4.5 and CART, application of binary concept learning algorithms to learn individual binary functions for each of the k classes, and application of binary concept learning algorithms with distributed output representations. This paper compares these three approaches to a new technique in which error-correcting codes are employed as a distributed output representation. We show that these output representations improve the generalization performance of both C4.5 and backpropagation on a wide range of multiclass learning tasks. We also demonstrate that this approach is robust with respect to changes in the size of the training sample, the assignment of distributed representations to particular classes, and the application of overfitting avoidance techniques such as decision-tree pruning. Finally, we show that--like the other methods--the error-correcting code technique can provide reliable class probability estimates. Taken together, these results demonstrate that error-correcting output codes provide a general-purpose method for improving the performance of inductive learning programs on multiclass problems.

Thomas G. Dietterich | Ghulum Bakiri | Ghulum Bakiri

[1] F ROSENBLATT,et al. The perceptron: a probabilistic model for information storage and organization in the brain. , 1958, Psychological review.

[2] Dwijendra K. Ray-Chaudhuri,et al. Binary mixture flow with free energy lattice Boltzmann methods , 2022, arXiv.org.

[3] W. W. Peterson,et al. Error-Correcting Codes. , 1962 .

[4] J. W. Machanik,et al. FUNCTION MODELING EXPERIMENTS. , 1963 .

[5] G. W. Snedecor. Statistical Methods , 1964 .

[6] Statistical methods , 1980 .

[7] Leslie G. Valiant,et al. A theory of the learnable , 1984, STOC '84.

[8] Geoffrey E. Hinton,et al. Learning internal representations by error propagation , 1986 .

[9] English Text,et al. Parallel Networks that Learn to Pronounce , 1987 .

[10] Terrence J. Sejnowski,et al. Parallel Networks that Learn to Pronounce English Text , 1987, Complex Syst..

[11] Waibel. A novel objective function for improved phoneme recognition using time delay neural networks , 1989 .