Live data from Hacker News

The Little Book of Reinforcement Learning

github.com

1–10 of 29 posts

Re: The Little Book of Reinforcement Learning

#7
post #2

Is this riffing on Strunk and Whites: The Elements of Style? Often referred to as "The Little Book".

I'm assuming it's more in line with The Little Schemer series of books (https://felleisen.org/matthias/BTLS-index.html) or maybe the little book of deep learning (https://fleuret.org/francois/lbdl.html)?

Re: The Little Book of Reinforcement Learning

#10
Real biological operant behavior isn't exactly trial and error learning.

Many factors shape and guide initial responses.

What I've noticed in some descriptions of models is the use of optimization for reinforcement to shape responses. In real organisms behavior may be controlled by short or long term outcomes, and may oscillate between this "optimization" based on schedules. This produces variability in the trials which can adjust behavior. Are we seeing these reinforcement models do this?

Post reply on HN