Who Learns Better Bayesian Network Structures: Constraint-Based, Score-based or Hybrid Algorithms?

Scutari, Marco; Graafland, Catharina Elisabeth; Gutiérrez, José Manuel

Statistics > Methodology

arXiv:1805.11908v2 (stat)

[Submitted on 30 May 2018 (v1), revised 1 Aug 2018 (this version, v2), latest version 31 Jul 2019 (v3)]

Title:Who Learns Better Bayesian Network Structures: Constraint-Based, Score-based or Hybrid Algorithms?

Authors:Marco Scutari, Catharina Elisabeth Graafland, José Manuel Gutiérrez

View PDF

Abstract:The literature groups algorithms to learn the structure of Bayesian networks from data in three separate classes: constraint-based algorithms, which use conditional independence tests to learn the dependence structure of the data; score-based algorithms, which use goodness-of-fit scores as objective functions to maximise; and hybrid algorithms that combine both approaches. Famously, Cowell (2001) showed that algorithms in the first two classes learn the same structures when the topological ordering of the network is known and we use entropy to assess conditional independence and goodness of fit.
In this paper we address the complementary question: how do these classes of algorithms perform outside of the assumptions above? We approach this question by recognising that structure learning is defined by the combination of a statistical criterion and an algorithm that determines how the criterion is applied to the data. Removing the confounding effect of different choices for the statistical criterion, we find using both simulated and real-world data that constraint-based algorithms do not appear to be more efficient or more sensitive to errors than score-based algorithms; and that hybrid algorithms are not faster or more accurate than constraint-based algorithms. This suggests that commonly held beliefs on structure learning in the literature are strongly influenced by the choice of particular statistical criteria rather than just properties of the algorithms themselves.

Comments:	12 pages, 5 figures
Subjects:	Methodology (stat.ME); Machine Learning (stat.ML)
Cite as:	arXiv:1805.11908 [stat.ME]
	(or arXiv:1805.11908v2 [stat.ME] for this version)
	https://doi.org/10.48550/arXiv.1805.11908
Journal reference:	Journal of Machine Learning Research (72, Proceedings Track, PGM 2018), 416-427

Submission history

From: Marco Scutari [view email]
[v1] Wed, 30 May 2018 11:42:44 UTC (1,907 KB)
[v2] Wed, 1 Aug 2018 10:00:45 UTC (1,963 KB)
[v3] Wed, 31 Jul 2019 14:36:11 UTC (4,965 KB)

Statistics > Methodology

Title:Who Learns Better Bayesian Network Structures: Constraint-Based, Score-based or Hybrid Algorithms?

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Methodology

Title:Who Learns Better Bayesian Network Structures: Constraint-Based, Score-based or Hybrid Algorithms?

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators