Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
cbames89
on Dec 27, 2020
|
parent
|
context
|
favorite
| on:
Introduction to Reinforcement Learning (2015)
Technical point: Value functions that are a constant multiples of each other result in the same behavior.
bitL
on Dec 28, 2020
[–]
Making a constant multiplication mistake somewhere in the code doesn't imply the new value function would be a constant multiply of the optimal one.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: