Dynamic social learning under graph constraints

We introduce a model of graph-constrained dynamic choice with reinforcement modeled by positively <inline-formula xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"><tex-math notation="LaTeX">$\alpha$</tex-math></inline-formula> -homogeneous rewards. We show that its empirical process, which can be written as a stochastic approximation recursion with Markov noise, has the same probability law as a certain vertex reinforced random walk. We use this equivalence to show that for <inline-formula xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"><tex-math notation="LaTeX">$\alpha &gt; 0$</tex-math></inline-formula> , the asymptotic outcome concentrates around the optimum in a certain limiting sense when “annealed” by letting <inline-formula xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink"><tex-math notation="LaTeX">$\alpha \uparrow \infty$</tex-math></inline-formula> slowly.

Paper

Similar papers

© 2026 NYSGPT2525 LLC