Policy Gradient Methods Find the Nash Equilibrium in N-player General-sum Linear-quadratic Games. (arXiv:2107.13090v2 [math.OC] UPDATED) | allainews.com

Aug. 16, 2022, 1:12 a.m. | Ben Hambly, Renyuan Xu, Huining Yang

stat.ML updates on arXiv.org arxiv.org

We consider a general-sum N-player linear-quadratic game with stochastic
dynamics over a finite horizon and prove the global convergence of the natural
policy gradient method to the Nash equilibrium. In order to prove the
convergence of the method, we require a certain amount of noise in the system.
We give a condition, essentially a lower bound on the covariance of the noise
in terms of the model parameters, in order to guarantee convergence. We
illustrate our results with numerical experiments …

arxiv equilibrium games general gradient linear math nash equilibrium policy

More from arxiv.org / stat.ML updates on arXiv.org

Calabi-Yau Four/Five/Six-folds as $\mathbb{P}^n_\textbf{w}$ Hypersurfaces: Machine Learning, Approximation, and Generation 4 hours ago | arxiv.org

abstract approximation arxiv five +17

Bayesian Quantile Regression with Subset Selection: A Posterior Summarization Perspective 4 hours ago | arxiv.org

abstract arxiv bayesian distribution +16

The Projected Covariance Measure for assumption-lean variable significance testing 4 hours ago | arxiv.org

abstract arxiv covariance lean +14

A Heteroskedasticity-Robust Overidentifying Restriction Test with High-Dimensional Covariates 4 hours ago | arxiv.org

abstract arxiv econ.em errors +11

Adjoint Sensitivity Analysis on Multi-Scale Bioprocess Stochastic Reaction Network 4 hours ago | arxiv.org

abstract analysis arxiv challenges +15

Neural Networks Optimized by Genetic Algorithms in Cosmology 4 hours ago | arxiv.org

abstract algorithms applications artificial +14

Seeded graph matching for the correlated Gaussian Wigner model via the projected power method 1 day, 4 hours ago | arxiv.org

abstract agreement arxiv edge +10

Convergence and Complexity Guarantee for Inexact First-order Riemannian Optimization Algorithms 1 day, 4 hours ago | arxiv.org

abstract algorithms analyze arxiv +11

Mixture of partially linear experts 1 day, 4 hours ago | arxiv.org

abstract arxiv benefits computational +9

Lead Developer (AI)

@ Cere Network | San Francisco, US

View on ai-jobs.net

Research Engineer

@ Allora Labs | Remote

View on ai-jobs.net

Ecosystem Manager

@ Allora Labs | Remote

View on ai-jobs.net

Founding AI Engineer, Agents

@ Occam AI | New York

View on ai-jobs.net

AI Engineer Intern, Agents

@ Occam AI | US

View on ai-jobs.net

AI Research Scientist

@ Vara | Berlin, Germany and Remote

View on ai-jobs.net