Gal Dalal, Assaf Hallak, Steven Dalton, Iuri Frosio, Shie Mannor, Gal Chechik. Improve Agents without Retraining: Parallel Tree Search with Off-Policy Correction. In Marc'Aurelio Ranzato, Alina Beygelzimer, Yann N. Dauphin, Percy Liang, Jennifer Wortman Vaughan, editors, Advances in Neural Information Processing Systems 34: Annual Conference on Neural Information Processing Systems 2021, NeurIPS 2021, December 6-14, 2021, virtual. pages 5518-5530, 2021. [doi]
@inproceedings{DalalHDFMC21, title = {Improve Agents without Retraining: Parallel Tree Search with Off-Policy Correction}, author = {Gal Dalal and Assaf Hallak and Steven Dalton and Iuri Frosio and Shie Mannor and Gal Chechik}, year = {2021}, url = {https://proceedings.neurips.cc/paper/2021/hash/2bd235c31c97855b7ef2dc8b414779af-Abstract.html}, researchr = {https://researchr.org/publication/DalalHDFMC21}, cites = {0}, citedby = {0}, pages = {5518-5530}, booktitle = {Advances in Neural Information Processing Systems 34: Annual Conference on Neural Information Processing Systems 2021, NeurIPS 2021, December 6-14, 2021, virtual}, editor = {Marc'Aurelio Ranzato and Alina Beygelzimer and Yann N. Dauphin and Percy Liang and Jennifer Wortman Vaughan}, }