{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/zero-shot-task-generalization-with-multi-task","title":"Zero-Shot Task Generalization with Multi-Task Deep Reinforcement Learning","arxiv_id":"1706.05064","date":"2017-06-15","proceeding":"ICML 2017 8","authors":["Junhyuk Oh","Satinder Singh","Honglak Lee","Pushmeet Kohli"],"abstract":"As a step towards developing zero-shot task generalization capabilities in\nreinforcement learning (RL), we introduce a new RL problem where the agent\nshould learn to execute sequences of instructions after learning useful skills\nthat solve subtasks. In this problem, we consider two types of generalizations:\nto previously unseen instructions and to longer sequences of instructions. For\ngeneralization over unseen instructions, we propose a new objective which\nencourages learning correspondences between similar subtasks by making\nanalogies. For generalization over sequential instructions, we present a\nhierarchical architecture where a meta controller learns to use the acquired\nskills for executing the instructions. To deal with delayed reward, we propose\na new neural architecture in the meta controller that learns when to update the\nsubtask, which makes learning more efficient. Experimental results on a\nstochastic 3D domain show that the proposed ideas are crucial for\ngeneralization to longer instructions as well as unseen instructions.","url_abs":"http://arxiv.org/abs/1706.05064v2","url_pdf":"http://arxiv.org/pdf/1706.05064v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"zero-shot-task-generalization-with-multi-task","repo_url":"https://github.com/seriousssam/zero-shot-task-generalization-implementation","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"pytorch","reach":{"status":"unanswered"}}],"tasks":[{"task_slug":"deep-reinforcement-learning","task_name":"Deep Reinforcement Learning"},{"task_slug":"reinforcement-learning","task_name":"Reinforcement Learning"},{"task_slug":"reinforcement-learning-1","task_name":"Reinforcement Learning (RL)"},{"task_slug":"reinforcement-learning-2","task_name":"reinforcement-learning"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=1706.05064","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}