跳到正文
arXiv:cs.LG· Robert R Nerem, Pranav Singh, Cheyenne Ward, Yusu Wang·· 4 小时前AI 评分40

NN-linkage:与 Lance-Williams 递推算法对齐的神经凝聚层次聚类树构建

Algorithmically Aligned Neural Agglomerative Tree Construction

AI 导读

研究者提出 NN-linkage,一种与 Lance-Williams(LW)递推算法对齐的神经网络模型,可学习任务特定、局部依赖的合并规则,同时保留经典 linkage 算法的递归结构与高效推理。

正文

View PDF HTML (experimental)

Abstract:Linkage algorithms for hierarchical clustering (HC) are a powerful and efficient framework for constructing clustering trees, yet it is often unclear which merge rule best suits a given dataset or task. In contrast, neural approaches can learn from data, but often fail to retain the efficiency and size generalization of classical algorithms. We introduce NN-linkage, a neural network (NN) model that can learn task-specific and locally dependent merge rules while retaining the recursive structure and efficient inference of classical linkage algorithms. In particular, our model is algorithmically aligned with the Lance-Williams (LW) recurrence, a parameterized framework for defining a broad, continuous family of linkage rules for agglomerative HC. Classical methods such as single linkage (SL), complete linkage (CL), and average linkage arise as discrete choices within this broader family. We show that NN-linkage is a universal approximator for continuous linkage functions, including LW recurrences, and, when paired with a transformer encoding, can also approximate globally dependent rules such as robust single-linkage. We further show that NN-linkage can exactly implement any symmetric constant-coefficient LW recurrence across all input sizes. On the empirical front, we evaluate NN-linkage in real-world applications, clock-tree routing and phylogenetic reconstruction, using both synthetic and real datasets, demonstrating its effectiveness over both classical algorithms and other neural approaches. By learning merge rules directly from target trees, NN-linkage extends efficient HC to scientific and engineering objectives not adequately captured by existing hand-designed linkage rules.
Subjects: Machine Learning (cs.LG)
Cite as: arXiv:2610.07271 [cs.LG]
  (or arXiv:2610.07271v1 [cs.LG] for this version)
  https://doi.org/10.48550/arXiv.2610.07271

arXiv-issued DOI via DataCite (pending registration)

Submission history

From: Robert Nerem [view email]
[v1] Mon, 5 Oct 2026 19:11:48 UTC (168 KB)

来源:arXiv:cs.LG · arxiv.org