Skip to main navigation Skip to search Skip to main content

Discussion of "Bayesian Cluster Analysis: Point Estimation and Credible Balls" by Sara Wade and Zoubin Ghahramani

Publication: Scientific journalJournal articlepeer-review

Abstract

Clustering is widely studied in statistics and machine learning, with applications in a variety of fields. As opposed to popular algorithms such as agglomerative hierarchical clustering or k-means which return a single clustering solution, Bayesian nonparametric models provide a posterior over the entire space of partitions, allowing one to assess statistical properties, such as uncertainty on the number of clusters. However, an important problem is how to summarize the posterior; the huge dimension of partition space and difficulties in visualizing it add to this problem. In a Bayesian analysis, the posterior of a real-valued parameter of interest is often summarized by reporting a point estimate such as the posterior mean along with 95% credible intervals to characterize uncertainty. In this paper, we extend these ideas to develop appropriate point estimates and credible sets to summarize the posterior of the clustering structure based on decision and information theoretic techniques.
Original languageEnglish
Pages (from-to)601 - 603
JournalBayesian Analysis
Volume13
Issue number2
DOIs
Publication statusPublished - 2018

Austrian Classification of Fields of Science and Technology (ÖFOS)

  • 102022 Software development
  • 101029 Mathematical statistics
  • 101018 Statistics

Cite this