Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

Persistence Fisher Kernel: A Riemannian Manifold Kernel for Persistence Diagrams

About

Algebraic topology methods have recently played an important role for statistical analysis with complicated geometric structured data such as shapes, linked twist maps, and material data. Among them, \textit{persistent homology} is a well-known tool to extract robust topological features, and outputs as \textit{persistence diagrams} (PDs). However, PDs are point multi-sets which can not be used in machine learning algorithms for vector data. To deal with it, an emerged approach is to use kernel methods, and an appropriate geometry for PDs is an important factor to measure the similarity of PDs. A popular geometry for PDs is the \textit{Wasserstein metric}. However, Wasserstein distance is not \textit{negative definite}. Thus, it is limited to build positive definite kernels upon the Wasserstein distance \textit{without approximation}. In this work, we rely upon the alternative \textit{Fisher information geometry} to propose a positive definite kernel for PDs \textit{without approximation}, namely the Persistence Fisher (PF) kernel. Then, we analyze eigensystem of the integral operator induced by the proposed kernel for kernel machines. Based on that, we derive generalization error bounds via covering numbers and Rademacher averages for kernel machines with the PF kernel. Additionally, we show some nice properties such as stability and infinite divisibility for the proposed kernel. Furthermore, we also propose a linear time complexity over the number of points in PDs for an approximation of our proposed kernel with a bounded error. Throughout experiments with many different tasks on various benchmark datasets, we illustrate that the PF kernel compares favorably with other baseline kernels for PDs.

Tam Le, Makoto Yamada• 2018

Related benchmarks

TaskDatasetResultRank
Graph ClassificationMUTAG
Accuracy85.6
697
Graph ClassificationNCI1
Accuracy81.7
460
Graph ClassificationIMDB-B
Accuracy71.2
322
Graph ClassificationNCI109
Accuracy78.5
223
Graph ClassificationDD
Accuracy79.4
175
Graph ClassificationPTC
Accuracy62.4
167
Graph ClassificationIMDB MULTI
Accuracy48.6
109
Graph ClassificationPROTEIN
Accuracy75.2
48
Graph ClassificationREDDIT-5K
Accuracy56.2
26
Showing 9 of 9 rows

Other info

Follow for update