1. Littman, M.L.: Reinforcement learning improves behaviour from evaluative feedback. Nature 521 (7553), 445–451 (2015)
2. Kiumarsi, B., Vamvoudakis, K.G., Modares, H., Lewis, F.L.: Optimal and autonomous control using reinforcement learning: A survey. IEEE Trans. Neural Netw. Learn. Syst., 1–21 (2017)
3. Racanière, S., Weber, T., Reichert, D.P., Buesing, L., Guez, A., Rezende, D., Jimenez, A., Badia, P., Vinyals, O., Heess, N., Li, Y., Pascanu, R., Battaglia, P., Hassabis, D., Silver, D., Wierstra, D.: Imagination-augmented agents for deep reinforcement learning. In: Advances in Neural Information Processing Systems (NIPS), pp. 5694–5705 (2017)
4. Bellemare, M.G., Ostrovski, G., Guez, A., Thomas, P.S., Munos, R.: Increasing the action gap: New operators for reinforcement learning. In: Proceedings of Workshops at the AAAI Conference on Artificial Intelligence (AAAI), pp 1476–1483 (2016)
5. Lample, G., Chaplot, D.S.: Playing fps games with deep reinforcement learning. In: Proceedings of Workshops at the AAAI Conference on Artificial Intelligence (AAAI), pp 2140–2146 (2017)