arXiv:cs.LG· Sreejeet Maity, Aritra Mitra·· 4 小时前AI 评分36
从不可靠轨迹中学习:抗 adversarial 的联邦 Q-Learning
Learning from Unreliable Trajectories: Adversarially-Robust Federated Q-Learning
AI 导读
研究提出 Robust Async-Fed-Q,一种基于 epoch 的联邦强化学习算法,通过方差缩减的 Bellman 最优算子估计与服务器端鲁棒聚合,在部分智能体发送任意损坏信息时仍保留协作的样本效率增益。
正文
Abstract:We study federated reinforcement learning in which multiple agents interact with a common Markov decision process and communicate through a central server to collaboratively learn the optimal state-action value function. Our goal is to understand whether the sample-efficiency benefits of collaboration can be retained when a fraction of the agents behave adversarially and transmit arbitrarily corrupted information. To address this problem, we introduce Robust Async-Fed-Q, an epoch-based federated learning algorithm that combines variance-reduced estimation of the Bellman optimality operator at the agents with robust aggregation at the server. We establish high-probability finite-time guarantees showing that the proposed method preserves the statistical gains of collaboration among the honest agents while tolerating adversarial corruption. In particular, the effect of the adversarial agents decreases as the amount of data collected by each honest agent grows and eventually vanishes in the infinite-sample limit. We complement these guarantees with information-theoretic lower bounds that characterize the unavoidable statistical cost of adversarial corruption, leading to the first nearly matching upper and lower bounds for adversarially robust federated reinforcement learning. We further extend our framework to accommodate single-trajectory Markovian sampling and heterogeneous partial coverage, where different agents may explore different regions of the state-action space and learning relies on their collective coverage. Finally, our epoch-based design substantially improves the best known communication complexity for federated Q-learning under asynchronous sampling.
| Subjects: | Machine Learning (cs.LG); Systems and Control (eess.SY) |
| Cite as: | arXiv:2610.06918 [cs.LG] |
| (or arXiv:2610.06918v1 [cs.LG] for this version) | |
| https://doi.org/10.48550/arXiv.2610.06918 arXiv-issued DOI via DataCite (pending registration) |
Submission history
From: Sreejeet Maity [view email]
[v1]
Fri, 2 Oct 2026 20:54:53 UTC (5,242 KB)
来源:arXiv:cs.LG · arxiv.org