One-shot high-fidelity imitation: Training large-scale deep nets with rlJan 1, 2018·T Le Paine,SG Colmenarejo,Z Wang,S Reed,Y Aytar,T Pfaff,MW Hoffman,G Barth-Maron,S Cabi,D Budden,N De Freitas· 0 min read PDF Cite VideoTypeJournal articlePublicationarXiv 2018Last updated on Jan 1, 2018 ← Grandmaster level in StarCraft II using multi-agent reinforcement learning Jan 1, 2019Playing hard exploration games by watching youtube Jan 1, 2018 →