Shusen Wang examines RL-1G: Summary. The publisher describes it as: “This lecture is a summary of Reinforcement Learning Basics.”. This is a creator-led account, not an independent replication.
Discuss: What evidence or practical test would most strengthen or challenge the account of RL-1G: Summary?
Shusen Wang examines RL-1F: Evaluate Reinforcement Learning. The publisher describes it as: “If you want to empirically compare two reinforcement learning algorithms, you will use OpenAI Gym. This lecture introduces three kinds of problems”. This is a creator-led account, not an independent rep…
Shusen Wang examines RL-1E: Value Functions. The publisher describes it as: “Value functions are the expectations of the return. Action-value function Q evaluates how good it is to take action A while being in state S. State-value”. This is a creator-led account, not an independent replication.
No replies yet. You can be first without making it weird.
Your turn
Pull up a chair.
Write first. We’ll sort the introductions when you submit.
Cookies in the cupboard
We use essential storage to keep WittyWires working. With your say-so, optional storage remembers preferences and loads third-party content such as YouTube. Rejecting it will not stop you using the site. Read our Privacy Policy.
Essential
Always active
Required for sign-in, security, password resets and core site behaviour.
Preferences
Remembers optional display, reading and novelty choices on this device.
Statistics
Used to understand how the site is used.Used only for anonymous site statistics.
Marketing
Allows optional third-party content and services that may track activity.