This little clip provides a visualization for the recursive formula of the value function v_pi(s) in Reinforcement Learning. It illustrates a simple way to read a backup diagram in order to obtain the corresponding formula.
This was done using manim, the free python library use by Grant Sanderson for his series 3blue1brown; do check them out!
In questa pagina del sito puoi guardare il video online RL: Value Function Formula Visualization della durata di ore minuti seconda in buona qualità , che l'utente ha caricato Naoshikuu 08 ottobre 2020, condividi il link con amici e conoscenti, su youtube questo video è già stato visto 2,522 volte e gli è piaciuto 87 spettatori. Buona visione!