TreeQN and ATreeC: differentiable tree planning for deep reinforcement learning

Combining deep model-free reinforcement learning with on-line planning is a promising approach to building on the successes of deep RL. On-line planning with look-ahead trees has proven successful in environments where transition models are known a priori. However, in complex environments where tran...

সম্পূর্ণ বিবরণ

গ্রন্থ-পঞ্জীর বিবরন
প্রধান লেখক: Farquhar, G, Rocktaeschel, T, Igl, M, Whiteson, S
বিন্যাস: Conference item
প্রকাশিত: International Conference on Learning Representations 2018