/notes/

Recent Notes

  • Generalization Factors

    Sep 03, 2026

    • dl
  • Grokking

    Sep 03, 2026

    • dl
  • Knowledge Distillation

    Sep 03, 2026

    • dl

See 1929 more →

Home

❯

RL

❯

Reinforcement Learning

Reinforcement Learning

Aug 16, 20261 min read

  • rl

Notes from:

  • Understanding Deep Learning

Basics

  • RL Basics

  • Markov Process

  • Policy

  • Value Function

  • Bellman Equations

  • Tabular Reinforcement Learning

    • Dynamic Programming RL
    • Monte Carlo Methods
    • On-Policy vs. Off-Policy
    • Temporal Difference Methods
      • SARSA
      • Q-Learning
  • Fitted Q-Learning

    • Deep Q-Networks
    • Double Q-Learning
  • Policy Gradient Methods

    • REINFORCE
  • Actor-Critic Methods

  • Offline Reinforcement Learning


Graph View

Backlinks

  • Simulated Annealing
  • Deep Learning
  • RL Basics

Created with Quartz v5.0.0 © 2026

  • Main site