跳到正文
arXiv:cs.LG· Pramita Bagchi, Edward Bae, Atish Mitra, Alexander D. Silberman, \v{Z}iga Virk, Sushovan Majhi·· 7 小时前AI 评分27

两组持久图差异在哪?固定预算下的校准局部推断

Where Do Two Populations of Persistence Diagrams Differ? Calibrated Local Inference at a Fixed Budget

AI 导读

研究针对持久图群体的两样本检验无法定位差异区域的问题,提出在固定图数量下对局部均值对比做同时推断,用加性 landmark 响应估计、高斯乘子 bootstrap 校准同时置信区间,并允许组间协方差不等。

正文

View PDF HTML (experimental)

Abstract:Many two-sample tests for populations of persistence diagrams assess global differences without identifying the regions of the birth-death plane that contribute to them. We study simultaneous inference for local mean contrasts when the number of available diagrams is fixed. They are differences in expected weighted feature mass within $\ell_\infty$ neighborhoods at several centers and radii. We estimate these contrasts using additive landmark responses. A Gaussian multiplier bootstrap calibrates simultaneous confidence intervals while allowing unequal group covariances. The neighborhoods whose intervals exclude zero form a map with approximate family-wise error control, and selecting a subset of original intervals for display preserves their joint coverage guarantee. On the simultaneous coverage event, every reported neighborhood lies within twice its radius of the support of the mean-measure difference. A geometric result gives sufficient radius conditions for a displaced feature to produce a nonzero contrast. A comparison of sufficient detection thresholds quantifies the tradeoff between reducing the number of tested coordinates and reserving observations for an independent pilot. In simulations with 40 to 120 diagrams per class, the bands achieved 94%-98% simultaneous coverage under both the strict null and equal means with unequal covariances. In the latter setting, a permutation maximum and the pooled-t implementation of the two-stage persistence-image test of Moon and Lazar rejected in up to 32% and 26% of runs, respectively. In the fixed-budget simulations, spending a third of the observations on a pilot to choose landmarks or radii located changes less often than a prespecified grid at a single radius. On the MUTAG benchmark, the localized region concentrates on rings of fused-ring systems, an exploratory reading.
Comments: 45 pages, 9 figures. Appendices with proofs and additional experiments. Under review
Subjects: Methodology (stat.ME); Machine Learning (cs.LG); Statistics Theory (math.ST)
MSC classes: 62G10, 62G15, 55N31
Cite as: arXiv:2610.08292 [stat.ME]
  (or arXiv:2610.08292v1 [stat.ME] for this version)
  https://doi.org/10.48550/arXiv.2610.08292

arXiv-issued DOI via DataCite (pending registration)

Submission history

From: Sushovan Majhi [view email]
[v1] Tue, 6 Oct 2026 13:00:00 UTC (200 KB)

来源:arXiv:cs.LG · arxiv.org