Convergence Rates for Localized Actor-Critic in Networked Markov Potential Games
Creators
Abstract
We introduce a class of networked Markov potential games where agents are associated with nodes in a network. Each agent has its own local potential function, and the reward of each agent depends only on the states and actions of agents within a κ-hop neighborhood. In this context, we propose a localized actor-critic algorithm. The algorithm is scalable since each agent uses only local information and does not need access to the global state. Further, the algorithm overcomes the curse of dimensionality through the use of function approximation. Our main results provide finite-sample guarantees up to a localization error and a function approximation error. Specifically, we achieve an O̅(ϵ⁻⁴) sample complexity measured by the averaged Nash regret. This is the first finite-sample bound for multi-agent competitive games that does not depend on the number of agents.
Additional Information
Attribution 4.0 International (CC BY 4.0)Attached Files
Submitted - 2303.04865.pdf
Files
2303.04865.pdf
Additional details
Identifiers
- Eprint ID
- 120098
- Resolver ID
- CaltechAUTHORS:20230316-204018535
Related works
- Describes
- http://arxiv.org/abs/2303.04865 (URL)
Dates
- Created
-
2023-03-16Created from EPrint's datestamp field
- Updated
-
2023-03-16Created from EPrint's last_modified field