Skip to content

Deep Q-networks

Neural networks as value approximators, and the machinery required to make that stable.

  • From linear features to a neural QQ
  • Experience replay
  • Target networks
  • Double DQN and duelling architectures
  • What actually makes DQN unstable

Everything above this chapter in the sidebar.