☆ 4.7 Article

Outlier-Robust Subsampling Techniques for Persistent Homology

JOURNAL OF MACHINE LEARNING RESEARCH (2023)

期刊

JOURNAL OF MACHINE LEARNING RESEARCH

卷 24, 期 -, 页码 -

出版社

MICROTOME PUBL

关键词

landmarks; persistent homology; subsampling; outliers; noise

类别

Automation & Control Systems Computer Science, Artificial Intelligence

向作者/读者索取更多资源

Protocol

社区支持

Reagent

社区支持

智能总结 New
摘要

This article proposes a novel approach to select landmarks specifically for persistent homology (PH) that preserves coarse topological information of the original dataset. The method is tested on artificial datasets with different levels of noise and outperforms standard methods and a subsampling technique based on an outlier-robust version of the k-means algorithm in terms of robustness to outliers under low sampling densities.

In recent years, persistent homology (PH) has been successfully applied to real-world data in many different settings. Despite significant computational advances, PH algorithms do not yet scale to large datasets preventing interesting applications. One approach to address computational issues posed by PH is to select a set of landmarks by subsampling from the data. Currently, these landmark points are chosen either at random or using the maxmin algorithm. Neither is ideal as random selection tends to favour dense areas of the data while the maxmin algorithm is very sensitive to noise. Here, we propose a novel approach to select landmarks specifically for PH that preserves coarse topological information of the original dataset. Our method is motivated by the Mayer-Vietoris sequence and requires only local PH calculations thus enabling efficient computation. We test our landmarks on artificial data sets which contain different levels of noise and compare them to standard landmark selection techniques. We demonstrate that our landmark selection outperforms standard methods as well as a subsampling technique based on an outlier-robust version of the k-means algorithm for low sampling densities in noisy data with respect to robustness to outliers.

Outlier-Robust Subsampling Techniques for Persistent Homology

期刊

JOURNAL OF MACHINE LEARNING RESEARCH

出版社

MICROTOME PUBL

关键词

类别

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

Outlier-Robust Subsampling Techniques for Persistent Homology

期刊

JOURNAL OF MACHINE LEARNING RESEARCH

出版社

MICROTOME PUBL

关键词

类别

向作者/读者索取更多资源

Protocol

Reagent

作者

我是这篇论文的作者

评论

主要评分

次要评分

新颖性

重要性

科学严谨性

评价这篇论文

推荐

导出引文

分享论文