强化学习
钢筋
计算机科学
动量(技术分析)
心理学
人工智能
社会心理学
业务
财务
作者
Sheng Yue,Xingyuan Hua,Yongheng Deng,Li-Li Chen,Ju Ren,Yaoxue Zhang
出处
期刊:
日期:2024-12-18
卷期号:33 (2): 865-880
被引量:1
标识
DOI:10.1109/tnet.2024.3510352
摘要
Federated Reinforcement Learning (FRL) is an attractive edge learning paradigm for decision-making applications, which has garnered significant interest recently. However, owing to the inherent spatio-temporal non-stationarity of local state-action distributions, current FRL approaches typically suffer from high interaction and communication costs. In this paper, we introduce a new FRL method, which incorporates momentum, importance sampling, and server-side adjustments, capable of controlling the gradient shifts induced by the non-stationary data. We prove that by proper selection of momentum parameters and interaction frequency, it can achieve $\tilde {\mathcal {O}}(H N^{-1}\epsilon ^{-3/2})$ and $\tilde {\mathcal {O}}(\epsilon ^{-1})$ interaction and communication complexities (N represents the agent number), where the interaction complexity achieves linear speedup with the number of agents, and the communication complexity aligns with the best achievable among existing first-order FL algorithms. Further, we leverage attention-based contextual representation extraction to enable the learning policy to adapt to heterogeneous tasks and environments. Extensive experiments demonstrate that our proposed method significantly outperforms existing baselines on a range of complex, high-dimensional single-task and multi-task benchmarks.
科研通智能强力驱动
Strongly Powered by AbleSci AI