A new paper from MATS argues that market dynamics are mathematically equivalent to Markov Decision Processes. The author frames price recursion as a rational theory of reward, drawing parallels between Bayesian inference and machine learning backpropagation. This theoretical bridge offers a new lens for researchers studying reward hacking and alignment.