Papers
Topics
Authors
Recent
Search
2000 character limit reached

On the Performance of Temporal Difference Learning With Neural Networks

Published 8 Dec 2023 in cs.LG | (2312.05397v1)

Abstract: Neural Temporal Difference (TD) Learning is an approximate temporal difference method for policy evaluation that uses a neural network for function approximation. Analysis of Neural TD Learning has proven to be challenging. In this paper we provide a convergence analysis of Neural TD Learning with a projection onto B(θ0,ω)B(\theta_0, \omega), a ball of fixed radius ω\omega around the initial point θ0\theta_0. We show an approximation bound of O(ϵ)+O~(1/m)O(\epsilon) + \tilde{O} (1/\sqrt{m}) where ϵ\epsilon is the approximation quality of the best neural network in B(θ0,ω)B(\theta_0, \omega) and mm is the width of all hidden layers in the network.

Citations (5)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.