Statistical inference of the value function for reinforcement learning in infinite‐horizon settings. (22nd December 2021)