arXiv:cs.LG· Qiye Lu, Jiang Ji, Liang Zhang·· 3 小时前
排序先验对齐:外部先验何时对信用风险建模有用?
Ranking Prior Alignment for Credit Risk Modeling: When Do External Priors Matter?
AI 导读
研究者提出 Ranking Prior Alignment,一种模型无关框架,通过温度缩放的 KL 散度损失将外部排序先验蒸馏进任意评分模型,统一了 MIL 注意力与 XGBoost 自定义目标两种实现。
正文
Abstract:Cold-start credit scoring -- deploying models with scarce labeled data, weak features, or minimal capacity -- is a recurring problem in financial machine learning. When a new lending product launches, labeled default data is scarce, feature pipelines are immature, and models must be deployed with minimal capacity to avoid overfitting. Standard defenses operate on the same limited data; what is needed is a source of external regularization grounded in domain knowledge.
We propose Ranking Prior Alignment, a model-agnostic framework that distills external ranking priors (from domain experts, teacher models, or LLMs) into any scoring model via a temperature-scaled KL divergence loss. The framework unifies neural (MIL attention) and tree-based (XGBoost custom objective) architectures through a single formulation: L = L_task + gamma(t) * KL(P_agent || P_model), where gamma(t) follows an exponential decay schedule. The method requires no external model at inference, and its tree-based instantiation tolerates annotation noise up to eta = 0.5.
On an industrial dataset of over 1.5M merchants, MIL alignment achieves 7/7 positive evaluation cells at 3K bags (1 ID + 3 OOT + 3 degradation metrics; peak Delta AUC = +0.020 on OOT-1), and XGBoost ablation achieves 9/9 positive metrics at 300 bags. Cross-dataset validation on public Amex shows 5/5 positive folds (avg Delta AUC = +0.041). Four model families (MIL, XGBoost, LightGBM, Logistic Regression) and four teacher architectures show that the framework is both model-agnostic and prior-source-independent.
We further observe that alignment gains exhibit an inverse-scaling pattern: benefits grow as data abundance N, model capacity C, and feature quality Q decrease, helping practitioners decide when to invest in prior annotation.
| Comments: | 8 pages, 5 figures, 14 tables |
| Subjects: | Machine Learning (cs.LG) |
| Cite as: | arXiv:2610.11146 [cs.LG] |
| (or arXiv:2610.11146v1 [cs.LG] for this version) | |
| https://doi.org/10.48550/arXiv.2610.11146 arXiv-issued DOI via DataCite (pending registration) |
Submission history
From: Qiye Lu [view email]
[v1]
Thu, 8 Oct 2026 03:11:02 UTC (374 KB)
来源:arXiv:cs.LG · arxiv.org