Learning With Episodic Hypothesis Testing In General Games: A Framework For Equilibrium Selection
2025 Β· Ruifan Yang, Manxi Wu
Abstract
We introduce a new hypothesis testing-based learning dynamics in which players update their strategies by combining hypothesis testing with utility-driven exploration. In this dynamics, each player forms beliefs about opponents' strategies and episodically tests these beliefs using empirical observations. Beliefs are resampled either when the hypothesis test is rejected or through exploration, where the probability of exploration decreases with the player's (transformed) utility. In general finite normal-form games, we show that the learning process converges to a set of approximate Nash equilibria and, more importantly, to a refinement that selects equilibria maximizing the minimum (transformed) utility across all players. Our result establishes convergence to equilibrium in general finite games and reveals a novel mechanism for equilibrium selection induced by the structure of the learning dynamics.
Authors
(none)
Tags
Stats
Related papers
- Episodic Logit-q Dynamics For Efficient Learning In Stochastic Teams (2022)0.00
- Finite-horizon Approximations And Episodic Equilibrium For Stochastic Games (2023)0.00
- Efficient Episodic Learning Of Nonstationary And Unknown Zero-sum Games Using Expert Game Ensembles (2021)3.58
- Bayesian Learning In Episodic Zero-sum Games (2026)0.00
- Higher-order Uncoupled Learning Dynamics And Nash Equilibrium (2025)0.00
- Decentralized Optimal Equilibrium Learning In Stochastic Games Via Single-bit Feedback (2026)0.00
- Exploration-exploitation In Multi-agent Competition: Convergence With Bounded Rationality (2021)0.00
- Learning In Multi-memory Games Triggers Complex Dynamics Diverging From Nash Equilibrium (2023)0.00