1. Q-learning;Watkins;Mach. Learn.,1992
2. Hosu, I.A., and Rebedea, T. (2016). Playing atari games with deep reinforcement learning and human checkpoint replay. arXiv.
3. Mnih, V., Kavukcuoglu, K., Silver, D., Graves, A., Antonoglou, I., Wierstra, D., and Riedmiller, M. (2013). Playing atari with deep reinforcement learning. arXiv.
4. Lillicrap, T.P., Hunt, J.J., Pritzel, A., Heess, N., Erez, T., Tassa, Y., Silver, D., and Wierstra, D. (2015). Continuous control with deep reinforcement learning. arXiv.
5. Gu, S., Lillicrap, T., Sutskever, I., and Levine, S. (2016, January 20–22). Continuous deep q-learning with model-based acceleration. Proceedings of the International Conference on Machine Learning, PMLR, New York, NY, USA.