Exponentially fast convergence to (strict) equilibrium via hedging

Cohen, Johanne; Héliou, Amélie; Mertikopoulos, Panayotis

Computer Science > Computer Science and Game Theory

arXiv:1607.08863 (cs)

[Submitted on 29 Jul 2016]

Title:Exponentially fast convergence to (strict) equilibrium via hedging

Authors:Johanne Cohen, Amélie Héliou, Panayotis Mertikopoulos

View PDF

Abstract:Motivated by applications to data networks where fast convergence is essential, we analyze the problem of learning in generic N-person games that admit a Nash equilibrium in pure strategies. Specifically, we consider a scenario where players interact repeatedly and try to learn from past experience by small adjustments based on local - and possibly imperfect - payoff information. For concreteness, we focus on the so-called "hedge" variant of the exponential weights algorithm where players select an action with probability proportional to the exponential of the action's cumulative payoff over time. When players have perfect information on their mixed payoffs, the algorithm converges locally to a strict equilibrium and the rate of convergence is exponentially fast - of the order of $\mathcal{O}(\exp(-a\sum_{j=1}^{t}\gamma_{j}))$ where $a>0$ is a constant and $\gamma_{j}$ is the algorithm's step-size. In the presence of uncertainty, convergence requires a more conservative step-size policy, but with high probability, the algorithm remains locally convergent and achieves an exponential convergence rate.

Comments:	14 pages
Subjects:	Computer Science and Game Theory (cs.GT); Machine Learning (cs.LG); Optimization and Control (math.OC)
Cite as:	arXiv:1607.08863 [cs.GT]
	(or arXiv:1607.08863v1 [cs.GT] for this version)
	https://doi.org/10.48550/arXiv.1607.08863

Submission history

From: Panayotis Mertikopoulos [view email]
[v1] Fri, 29 Jul 2016 16:16:49 UTC (209 KB)

Full-text links:

Access Paper:

view license

Current browse context:

math.OC

< prev | next >

new | recent | 2016-07

Change to browse by:

cs
cs.GT
cs.LG
math

References & Citations

DBLP - CS Bibliography

listing | bibtex

Johanne Cohen
Amélie Héliou
Panayotis Mertikopoulos

export BibTeX citation

Computer Science > Computer Science and Game Theory

Title:Exponentially fast convergence to (strict) equilibrium via hedging

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Science and Game Theory

Title:Exponentially fast convergence to (strict) equilibrium via hedging

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators