(Semi-)Predictive Discretization During Model Selection
Author(s)
Steck, Harald; Jaakkola, Tommi S.
DownloadAIM-2003-002.ps (4.100Mb)
Additional downloads
Metadata
Show full item recordAbstract
In this paper, we present an approach to discretizing multivariate continuous data while learning the structure of a graphical model. We derive the joint scoring function from the principle of predictive accuracy, which inherently ensures the optimal trade-off between goodness of fit and model complexity (including the number of discretization levels). Using the so-called finest grid implied by the data, our scoring function depends only on the number of data points in the various discretization levels. Not only can it be computed efficiently, but it is also independent of the metric used in the continuous space. Our experiments with gene expression data show that discretization plays a crucial role regarding the resulting network structure.
Date issued
2003-02-25Other identifiers
AIM-2003-002
Series/Report no.
AIM-2003-002
Keywords
AI, Discretization, Graphical models