Dendrogram of mixing measures: Hierarchical clustering and model selection for finite mixture models
arxiv(2024)
摘要
We present a new way to summarize and select mixture models via the
hierarchical clustering tree (dendrogram) constructed from an overfitted latent
mixing measure. Our proposed method bridges agglomerative hierarchical
clustering and mixture modeling. The dendrogram's construction is derived from
the theory of convergence of the mixing measures, and as a result, we can both
consistently select the true number of mixing components and obtain the
pointwise optimal convergence rate for parameter estimation from the tree, even
when the model parameters are only weakly identifiable. In theory, it
explicates the choice of the optimal number of clusters in hierarchical
clustering. In practice, the dendrogram reveals more information on the
hierarchy of subpopulations compared to traditional ways of summarizing mixture
models. Several simulation studies are carried out to support our theory. We
also illustrate the methodology with an application to single-cell RNA sequence
analysis.
更多查看译文
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
数据免责声明
页面数据均来自互联网公开来源、合作出版商和通过AI技术自动分析结果,我们不对页面数据的有效性、准确性、正确性、可靠性、完整性和及时性做出任何承诺和保证。若有疑问,可以通过电子邮件方式联系我们:report@aminer.cn