Distributed Dynamic No-Regret Learning in Two-Network Zero-Sum Games

成果类型:
Article
署名作者:
Liao, Lan; Yuan, Deming; Ho, Daniel W. C.; Zheng, Wei Xing; Zhang, Baoyong; Yu, Zhan
署名单位:
Nanjing University of Science & Technology; City University of Hong Kong; Western Sydney University; Hong Kong Baptist University
刊物名称:
IEEE TRANSACTIONS ON AUTOMATIC CONTROL
ISSN/ISSBN:
0018-9286
DOI:
10.1109/TAC.2025.3586980
发表日期:
2026
关键词:
摘要:
This article considers the time-varying zero-sum game between two multiagent networks with different topologies. The two networks are modeled as adversarial players, where agents within each network communicate with their local neighbors while simultaneously gathering information from the opposing network through a bipartite interaction framework. The payoffs of the agents can be quantified by time-varying cost functions, and the objective is to design an online distributed algorithm to optimize the payoff for each network in the game. We use dynamic Nash equilibrium regret and duality gap as performance metrics and propose a projection-free algorithm called Online Distributed Frank-Wolfe in two-network (ODFW-TN). We assess the convergence of Algorithm ODFW-TN through regret analysis, and establish sublinear bounds for dynamic Nash equilibrium regret and duality gap of Algorithm ODFW-TN with respect to T, where T is the total number of games. Moreover, we validate the effectiveness of Algorithm ODFW-TN via a numerical experiment of a time-varying bilinear matrix game.