{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/smooth-imitation-learning-for-online-sequence","title":"Smooth Imitation Learning for Online Sequence Prediction","arxiv_id":"1606.00968","date":"2016-06-03","proceeding":null,"authors":["Hoang M. Le","Andrew Kang","Yisong Yue","Peter Carr"],"abstract":"We study the problem of smooth imitation learning for online sequence\nprediction, where the goal is to train a policy that can smoothly imitate\ndemonstrated behavior in a dynamic and continuous environment in response to\nonline, sequential context input. Since the mapping from context to behavior is\noften complex, we take a learning reduction approach to reduce smooth imitation\nlearning to a regression problem using complex function classes that are\nregularized to ensure smoothness. We present a learning meta-algorithm that\nachieves fast and stable convergence to a good policy. Our approach enjoys\nseveral attractive properties, including being fully deterministic, employing\nan adaptive learning rate that can provably yield larger policy improvements\ncompared to previous approaches, and the ability to ensure stable convergence.\nOur empirical results demonstrate significant performance gains over previous\napproaches.","url_abs":"http://arxiv.org/abs/1606.00968v1","url_pdf":"http://arxiv.org/pdf/1606.00968v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"smooth-imitation-learning-for-online-sequence","repo_url":"https://github.com/hoangminhle/SIMILE","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"none","reach":{"status":"ok"}},{"paper_slug":"smooth-imitation-learning-for-online-sequence","repo_url":"https://github.com/lucianacendon/simile","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"none","reach":{"status":"ok","spdx":"MIT"}}],"tasks":[{"task_slug":"imitation-learning","task_name":"Imitation Learning"},{"task_slug":"prediction","task_name":"Prediction"},{"task_slug":"regression-1","task_name":"regression"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"syntology_url":"https://syntology.ai/paper/1606.00968","atlas_url":"https://app.syntology.ai/?focus=1606.00968","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"1606.00968"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-25T09:33:49+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/lucianacendon/simile","reach":{"status":"ok","spdx":"MIT"}},{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/hoangminhle/SIMILE","reach":{"status":"ok"}}],"summary":{"unverified":3},"by_repo_kind":{"listed":{"samples":3,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"8099bd273824efbe","entry":"denormalize","repo":"lucianacendon/simile","repo_kind":"listed","path":"Lib/utils.py","file_url":"https://github.com/lucianacendon/simile/blob/HEAD/Lib/utils.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"8099bd273824efbe"}},{"code_sha256_prefix":"1d45dfa597d895d1","entry":"normalize_test_data","repo":"lucianacendon/simile","repo_kind":"listed","path":"Lib/utils.py","file_url":"https://github.com/lucianacendon/simile/blob/HEAD/Lib/utils.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"1d45dfa597d895d1"}},{"code_sha256_prefix":"d06ac2cda5bfc7bc","entry":"normalize_train_data","repo":"lucianacendon/simile","repo_kind":"listed","path":"Lib/utils.py","file_url":"https://github.com/lucianacendon/simile/blob/HEAD/Lib/utils.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"d06ac2cda5bfc7bc"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}