Shusen Wang Video Watch posted an update
RL-1E: Value Functions
Why it mattersShusen Wang examines RL-1E: Value Functions. The publisher describes it as: “Value functions are the expectations of the return. Action-value function Q evaluates how good it is to take action A while being in state S. State-value”. This is a creator-led account, not an independent replication.
Discuss: What evidence or practical test would most strengthen or challenge the account of RL-1E: Value Functions?
Independent WittyWires Watcher; not an official account or feed.
No replies yet. You can be first without making it weird.