Towards Gaussian Bayesian Network Fusion

Data sets are growing in complexity thanks to the increasing facilities we have nowadays to both generate and store data. This poses many challenges to machine learning that are leading to the proposal of new methods and paradigms, in order to be able to deal with what is nowadays referred to as Big Data. In this paper we propose a method for the aggregation of different Bayesian network structures that have been learned from separate data sets, as a first step towards mining data sets that need to be partitioned in an horizontal way, i.e. with respect to the instances, in order to be processed. Considerations that should be taken into account when dealing with this situation are discussed. Scalable learning of Bayesian networks is slowly emerging, and our method constitutes one of the first insights into Gaussian Bayesian network aggregation from different sources. Tested on synthetic data it obtains good results that surpass those from individual learning. Future research will be focused on expanding the method and testing more diverse data sets.

[1]  Serafín Moral,et al.  Qualitative combination of Bayesian networks , 2003, Int. J. Intell. Syst..

[2]  Bruce Abramson,et al.  The Topological Fusion of Bayes Nets , 1992, UAI.

[3]  Tsai-Hung Fan,et al.  Regression analysis for massive datasets , 2007, Data Knowl. Eng..

[4]  Kobra Etminani,et al.  DemocraticOP: A Democratic way of aggregating Bayesian network parameters , 2013, Int. J. Approx. Reason..

[5]  Concha Bielza,et al.  Learning an L1-Regularized Gaussian Bayesian Network in the Equivalence Class Space , 2010, IEEE Transactions on Systems, Man, and Cybernetics, Part B (Cybernetics).

[6]  Matthew Richardson,et al.  Learning with Knowledge from Multiple Experts , 2003, ICML.

[7]  Tom Burr,et al.  Causation, Prediction, and Search , 2003, Technometrics.

[8]  Pedrito Maynard-Reid,et al.  Aggregating Learned Probabilistic Beliefs , 2001, UAI.

[9]  José M. Peña,et al.  On Local Optima in Learning Bayesian Networks , 2003, UAI.

[10]  Constantin F. Aliferis,et al.  The max-min hill-climbing Bayesian network structure learning algorithm , 2006, Machine Learning.

[11]  David Heckerman,et al.  Learning Gaussian Networks , 1994, UAI.

[12]  Marco Scutari,et al.  Learning Bayesian Networks with the bnlearn R Package , 2009, 0908.3817.

[13]  Concha Bielza,et al.  Bayesian network modeling of the consensus between experts: An application to neuron classification , 2014 .

[14]  G. Schwarz Estimating the Dimension of a Model , 1978 .

[15]  Judea Pearl,et al.  Probabilistic reasoning in intelligent systems - networks of plausible inference , 1991, Morgan Kaufmann series in representation and reasoning.