Article citationsMore>>

Sutton, R. S., McAllester, D. A., Singh, S. P., & Mansour, Y. (2000). Policy Gradient Methods for Reinforcement Learning with Function Approximation. In M. I. Jordan, Y. Lecun, & S. A. Solla (Eds.), Advances in Neural Information Processing Systems (pp. 1057-1063). MIT Press.

has been cited by the following article:

SCIRP Newsletter
Copyright © 2006-2026 Scientific Research Publishing Inc. All Rights Reserved.
Top