{"about":{"non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","site":"https://codewithpapers.app","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page","syntology":{"site":"https://syntology.ai","developers":"https://syntology.ai/developers","mcp":{"server":"https://syntology.ai/mcp","transport":"streamable-http","server_card":"https://syntology.ai/.well-known/mcp/server-card.json","auth":{"type":"trial token, no account","trial_token":"https://syntology.ai/api/oauth/trial/token","method":"POST","docs":"https://syntology.ai/developers"}},"have":"https://syntology.ai/api/graph/have?x=<method, arXiv id or title> (free, answers coverage only)","paper_base":"https://syntology.ai/paper/","atlas_base":"https://app.syntology.ai/?focus="},"machine_readable":[{"url":"https://codewithpapers.app/llms.txt","what":"the machine catalog: every machine-readable file, counted"},{"url":"https://codewithpapers.app/index/manifest.json","what":"paper-to-code index by arXiv id, with Syntology's counts"},{"url":"https://codewithpapers.app/search/manifest.json","what":"site search index (titles, authors) and its files"},{"url":"https://codewithpapers.app/download","what":"bulk files: Syntology's layer, described there"},{"url":"https://codewithpapers.app/build_manifest.json","what":"the build record: inputs, counts, exclusions, probes"}]},"url":"/task/imitation-learning/papers/14","list_of":"/task/imitation-learning","task":"Imitation Learning","archive":{"snapshot":"2025-07-28"},"key_notes":{"n_ran_checked":"legacy name, kept unchanged so existing readers do not break: it counts the samples that ran with no instrument failure (honoured, violated, and ran with no contract checked); it does not mean a contract was checked, and the pages print it as 'K with no instrument failure', not 'K checked'","n_constructed":"a sub-count of the samples that ran, never subtracted from them and never a failure: an executed sample whose run returned an instance of its own class (fixture_out_type equals the entry name): the run built an object and did not compute a result (Syntology's RAN record, counts.constructed)"},"syntology_read_at":"2026-09-28T10:30:06+00:00","order":"archive","order_definition":"repositories listed in the archive (most first), then date (newest first), then slug","page":14,"pages_in_order":22,"rows_per_page":100,"rows":[1301,1400],"of":2122,"counts":{"archive_papers_tagged":2122,"with_a_code_link":691,"where_syntology_ran_a_sample":234,"not_listed_spam_title":0,"listed":2122,"listed_where_code_ran":234,"where_syntology_ran_a_sample_split":{"with_a_run_with_no_instrument_failure":191,"every_run_a_failure_of_syntologys_instrument":43,"listed_with_a_run_with_no_instrument_failure":191,"listed_every_run_a_failure_of_syntologys_instrument":43,"filter":{"states":["a run with no instrument failure","any run, instrument failures included"],"default":"a run with no instrument failure","note":"on the 'only where code ran' pages the default hides, in the browser, the rows where every run was a failure of Syntology's instrument; the second state shows them again. Rows are hidden, never re-ordered; these twins list every row"}},"definition":"distinct papers the archive tags; 'where Syntology ran a sample' counts papers with at least one harvested sample that ran, which is not a correctness claim"},"first_page":"/task/imitation-learning","prev":"/task/imitation-learning/papers/13","next":"/task/imitation-learning/papers/15","papers":[{"url":null,"slug":"imitating-task-and-motion-planning-with","title":"Imitating Task and Motion Planning with Visuomotor Transformers","date":"2023-05-25","arxiv_id":"2305.16309","repositories_listed":0,"syntology":null},{"url":null,"slug":"deep-reinforcement-learning-based-multi-1","title":"Deep Reinforcement Learning-based Multi-objective Path Planning on the Off-road Terrain Environment for Ground Vehicles","date":"2023-05-23","arxiv_id":"2305.13783","repositories_listed":0,"syntology":null},{"url":null,"slug":"end-to-end-stable-imitation-learning-via","title":"End-to-End Stable Imitation Learning via Autonomous Neural Dynamic Policies","date":"2023-05-22","arxiv_id":"2305.12886","repositories_listed":0,"syntology":null},{"url":null,"slug":"on-the-correspondence-between","title":"On the Correspondence between Compositionality and Imitation in Emergent Neural Communication","date":"2023-05-22","arxiv_id":"2305.12941","repositories_listed":0,"syntology":null},{"url":null,"slug":"replicating-complex-dialogue-policy-of-humans","title":"Replicating Complex Dialogue Policy of Humans via Offline Imitation Learning with Supervised Regularization","date":"2023-05-06","arxiv_id":"2305.03987","repositories_listed":0,"syntology":null},{"url":null,"slug":"an-imitation-learning-based-algorithm","title":"An Imitation Learning Based Algorithm Enabling Priori Knowledge Transfer in Modern Electricity Markets for Bayesian Nash Equilibrium Estimation","date":"2023-05-04","arxiv_id":"2305.06924","repositories_listed":0,"syntology":null},{"url":null,"slug":"scanpath-prediction-in-panoramic-videos-via","title":"Scanpath Prediction in Panoramic Videos via Expected Code Length Minimization","date":"2023-05-04","arxiv_id":"2305.02536","repositories_listed":0,"syntology":null},{"url":null,"slug":"calm-conditional-adversarial-latent-models","title":"CALM: Conditional Adversarial Latent Models for Directable Virtual Characters","date":"2023-05-02","arxiv_id":"2305.02195","repositories_listed":0,"syntology":null},{"url":null,"slug":"get-back-here-robust-imitation-by-return-to","title":"Get Back Here: Robust Imitation by Return-to-Distribution Planning","date":"2023-05-02","arxiv_id":"2305.01400","repositories_listed":0,"syntology":null},{"url":null,"slug":"learning-environment-for-the-air-domain-lead","title":"Learning Environment for the Air Domain (LEAD)","date":"2023-04-27","arxiv_id":"2304.14423","repositories_listed":0,"syntology":null},{"url":null,"slug":"programmatically-grounded-compositionally","title":"Programmatically Grounded, Compositionally Generalizable Robotic Manipulation","date":"2023-04-26","arxiv_id":"2304.13826","repositories_listed":0,"syntology":null},{"url":null,"slug":"causal-semantic-communication-for-digital","title":"Causal Semantic Communication for Digital Twins: A Generalizable Imitation Learning Approach","date":"2023-04-25","arxiv_id":"2304.12502","repositories_listed":0,"syntology":null},{"url":"/paper/learning-fine-grained-bimanual-manipulation","slug":"learning-fine-grained-bimanual-manipulation","title":"Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware","date":"2023-04-23","arxiv_id":"2304.13705","repositories_listed":0,"syntology":null},{"url":null,"slug":"behavior-retrieval-few-shot-imitation","title":"Behavior Retrieval: Few-Shot Imitation Learning by Querying Unlabeled Datasets","date":"2023-04-18","arxiv_id":"2304.08742","repositories_listed":0,"syntology":null},{"url":null,"slug":"affordances-from-human-videos-as-a-versatile","title":"Affordances from Human Videos as a Versatile Representation for Robotics","date":"2023-04-17","arxiv_id":"2304.08488","repositories_listed":0,"syntology":null},{"url":null,"slug":"mddl-a-framework-for-reinforcement-learning","title":"MDDL: A Framework for Reinforcement Learning-based Position Allocation in Multi-Channel Feed","date":"2023-04-17","arxiv_id":"2304.09087","repositories_listed":0,"syntology":null},{"url":null,"slug":"reward-free-policy-imitation-learning-for","title":"Reward-free Policy Imitation Learning for Conversational Search","date":"2023-04-17","arxiv_id":"2304.07988","repositories_listed":0,"syntology":null},{"url":null,"slug":"a-review-on-longitudinal-car-following-model","title":"Car-Following Models: A Multidisciplinary Review","date":"2023-04-14","arxiv_id":"2304.07143","repositories_listed":0,"syntology":null},{"url":null,"slug":"synthetically-generating-human-like-data-for","title":"Synthetically Generating Human-like Data for Sequential Decision Making Tasks via Reward-Shaped Imitation Learning","date":"2023-04-14","arxiv_id":"2304.07280","repositories_listed":0,"syntology":null},{"url":null,"slug":"for-pre-trained-vision-models-in-motor","title":"For Pre-Trained Vision Models in Motor Control, Not All Policy Learning Methods are Created Equal","date":"2023-04-10","arxiv_id":"2304.04591","repositories_listed":0,"syntology":null},{"url":null,"slug":"crisp-curriculum-inducing-primitive-informed","title":"CRISP: Curriculum Inducing Primitive Informed Subgoal Prediction for Hierarchical Reinforcement Learning","date":"2023-04-07","arxiv_id":"2304.03535","repositories_listed":0,"syntology":null},{"url":null,"slug":"end-to-end-manipulator-calligraphy-planning","title":"End-to-end Manipulator Calligraphy Planning via Variational Imitation Learning","date":"2023-04-06","arxiv_id":"2304.02801","repositories_listed":0,"syntology":null},{"url":null,"slug":"entl-embodied-navigation-trajectory-learner","title":"ENTL: Embodied Navigation Trajectory Learner","date":"2023-04-05","arxiv_id":"2304.02639","repositories_listed":0,"syntology":null},{"url":null,"slug":"quantum-imitation-learning","title":"Quantum Imitation Learning","date":"2023-04-04","arxiv_id":"2304.02480","repositories_listed":0,"syntology":null},{"url":null,"slug":"imitation-learning-from-nonlinear-mpc-via-the","title":"Imitation Learning from Nonlinear MPC via the Exact Q-Loss and its Gauss-Newton Approximation","date":"2023-04-03","arxiv_id":"2304.01782","repositories_listed":0,"syntology":null},{"url":null,"slug":"specification-guided-data-aggregation-for","title":"Specification-Guided Data Aggregation for Semantically Aware Imitation Learning","date":"2023-03-29","arxiv_id":"2303.17010","repositories_listed":0,"syntology":null},{"url":null,"slug":"efficient-deep-learning-of-robust-adaptive","title":"Efficient Deep Learning of Robust, Adaptive Policies using Tube MPC-Guided Data Augmentation","date":"2023-03-28","arxiv_id":"2303.15688","repositories_listed":0,"syntology":null},{"url":null,"slug":"embedding-contextual-information-through","title":"Embedding Contextual Information through Reward Shaping in Multi-Agent Learning: A Case Study from Google Football","date":"2023-03-25","arxiv_id":"2303.15471","repositories_listed":0,"syntology":null},{"url":null,"slug":"exploring-the-use-of-deep-learning-in-task","title":"Exploring the use of deep learning in task-flexible ILC","date":"2023-03-25","arxiv_id":"2303.14402","repositories_listed":0,"syntology":null},{"url":null,"slug":"interpretable-motion-planner-for-urban","title":"Interpretable Motion Planner for Urban Driving via Hierarchical Imitation Learning","date":"2023-03-24","arxiv_id":"2303.13986","repositories_listed":0,"syntology":null},{"url":null,"slug":"disturbance-injection-under-partial","title":"Disturbance Injection under Partial Automation: Robust Imitation Learning for Long-horizon Tasks","date":"2023-03-22","arxiv_id":"2303.12375","repositories_listed":0,"syntology":null},{"url":null,"slug":"penalty-based-imitation-learning-with-cross","title":"Penalty-Based Imitation Learning With Cross Semantics Generation Sensor Fusion for Autonomous Driving","date":"2023-03-21","arxiv_id":"2303.11888","repositories_listed":0,"syntology":null},{"url":null,"slug":"bridging-imitation-and-online-reinforcement","title":"Bridging Imitation and Online Reinforcement Learning: An Optimistic Tale","date":"2023-03-20","arxiv_id":"2303.11369","repositories_listed":0,"syntology":null},{"url":null,"slug":"implicit-and-explicit-commonsense-for-multi","title":"Implicit and Explicit Commonsense for Multi-sentence Video Captioning","date":"2023-03-14","arxiv_id":"2303.07545","repositories_listed":0,"syntology":null},{"url":null,"slug":"learning-to-transfer-in-hand-manipulations","title":"Learning to Transfer In-Hand Manipulations Using a Greedy Shape Curriculum","date":"2023-03-14","arxiv_id":"2303.12726","repositories_listed":0,"syntology":null},{"url":null,"slug":"sample-efficient-adversarial-imitation-1","title":"Sample-efficient Adversarial Imitation Learning","date":"2023-03-14","arxiv_id":"2303.07846","repositories_listed":0,"syntology":null},{"url":null,"slug":"conbat-control-barrier-transformer-for-safe","title":"ConBaT: Control Barrier Transformer for Safe Policy Learning","date":"2023-03-07","arxiv_id":"2303.04212","repositories_listed":0,"syntology":null},{"url":null,"slug":"decoupling-skill-learning-from-robotic","title":"Decoupling Skill Learning from Robotic Control for Generalizable Object Manipulation","date":"2023-03-07","arxiv_id":"2303.04016","repositories_listed":0,"syntology":null},{"url":null,"slug":"offline-imitation-learning-with-suboptimal","title":"Offline Imitation Learning with Suboptimal Demonstrations via Relaxed Distribution Matching","date":"2023-03-05","arxiv_id":"2303.02569","repositories_listed":0,"syntology":null},{"url":null,"slug":"interactive-text-generation","title":"Interactive Text Generation","date":"2023-03-02","arxiv_id":"2303.00908","repositories_listed":0,"syntology":null},{"url":null,"slug":"automated-task-time-interventions-to-improve","title":"Automated Task-Time Interventions to Improve Teamwork using Imitation Learning","date":"2023-03-01","arxiv_id":"2303.00413","repositories_listed":0,"syntology":null},{"url":null,"slug":"diffusion-model-augmented-behavioral-cloning","title":"Diffusion Model-Augmented Behavioral Cloning","date":"2023-02-26","arxiv_id":"2302.13335","repositories_listed":0,"syntology":null},{"url":"/paper/k-shap-policy-clustering-algorithm-for","slug":"k-shap-policy-clustering-algorithm-for","title":"K-SHAP: Policy Clustering Algorithm for Anonymous Multi-Agent State-Action Pairs","date":"2023-02-23","arxiv_id":"2302.11996","repositories_listed":0,"syntology":{"n":4,"n_ran":3,"n_constructed":0,"n_ran_checked":0,"n_instrument":3,"n_unverified":1,"n_honours":0,"n_violates":0,"n_no_contract":0,"n_pointer_only":4,"phrase":"3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified","sample_list":"/paper/k-shap-policy-clustering-algorithm-for#ran","syntology_url":"https://syntology.ai/paper/2302.11996","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2302.11996"}},"official":null}},{"url":null,"slug":"asking-for-help-failure-prediction-in","title":"Asking for Help: Failure Prediction in Behavioral Cloning through Value Approximation","date":"2023-02-08","arxiv_id":"2302.04334","repositories_listed":0,"syntology":null},{"url":null,"slug":"scaling-self-supervised-end-to-end-driving","title":"Scaling Vision-based End-to-End Driving with Multi-View Attention Learning","date":"2023-02-07","arxiv_id":"2302.03198","repositories_listed":0,"syntology":null},{"url":null,"slug":"a-strong-baseline-for-batch-imitation","title":"A Strong Baseline for Batch Imitation Learning","date":"2023-02-06","arxiv_id":"2302.02788","repositories_listed":0,"syntology":null},{"url":null,"slug":"ditto-offline-imitation-learning-with-world","title":"DITTO: Offline Imitation Learning with World Models","date":"2023-02-06","arxiv_id":"2302.03086","repositories_listed":0,"syntology":null},{"url":null,"slug":"aligning-robot-and-human-representations","title":"Aligning Robot and Human Representations","date":"2023-02-03","arxiv_id":"2302.01928","repositories_listed":0,"syntology":null},{"url":null,"slug":"synthesizing-physical-character-scene","title":"Synthesizing Physical Character-Scene Interactions","date":"2023-02-02","arxiv_id":"2302.00883","repositories_listed":0,"syntology":null},{"url":null,"slug":"behaviour-discriminator-a-simple-data","title":"Improving Behavioural Cloning with Positive Unlabeled Learning","date":"2023-01-27","arxiv_id":"2301.11734","repositories_listed":0,"syntology":null},{"url":null,"slug":"language-guided-task-adaptation-for-imitation","title":"Language-guided Task Adaptation for Imitation Learning","date":"2023-01-24","arxiv_id":"2301.09770","repositories_listed":0,"syntology":null},{"url":null,"slug":"smart-self-supervised-multi-task-pretraining","title":"SMART: Self-supervised Multi-task pretrAining with contRol Transformers","date":"2023-01-24","arxiv_id":"2301.09816","repositories_listed":0,"syntology":null},{"url":null,"slug":"graph-neural-networks-for-decentralized-multi-2","title":"Graph Neural Networks for Decentralized Multi-Agent Perimeter Defense","date":"2023-01-23","arxiv_id":"2301.09689","repositories_listed":0,"syntology":null},{"url":null,"slug":"deep-reinforcement-learning-for-power-trading","title":"Domain-adapted Learning and Imitation: DRL for Power Arbitrage","date":"2023-01-19","arxiv_id":"2301.08360","repositories_listed":0,"syntology":null},{"url":null,"slug":"direct-learning-from-sparse-and-shifting","title":"DIRECT: Learning from Sparse and Shifting Rewards using Discriminative Reward Co-Training","date":"2023-01-18","arxiv_id":"2301.07421","repositories_listed":0,"syntology":null},{"url":null,"slug":"nerf-in-the-palm-of-your-hand-corrective","title":"NeRF in the Palm of Your Hand: Corrective Augmentation for Robotics via Novel-View Synthesis","date":"2023-01-18","arxiv_id":"2301.08556","repositories_listed":0,"syntology":null},{"url":null,"slug":"adaptive-neural-networks-using-residual","title":"Adaptive Neural Networks Using Residual Fitting","date":"2023-01-13","arxiv_id":"2301.05744","repositories_listed":0,"syntology":null},{"url":null,"slug":"explaining-imitation-learning-through-frames","title":"Explaining Imitation Learning through Frames","date":"2023-01-03","arxiv_id":"2301.01088","repositories_listed":0,"syntology":null},{"url":null,"slug":"genetic-imitation-learning-by-reward","title":"Genetic Imitation Learning by Reward Extrapolation","date":"2023-01-03","arxiv_id":"2301.07182","repositories_listed":0,"syntology":null},{"url":null,"slug":"imitation-learning-as-state-matching-via","title":"Imitation Learning As State Matching via Differentiable Physics","date":"2023-01-01","arxiv_id":null,"repositories_listed":0,"syntology":null},{"url":null,"slug":"refteacher-a-strong-baseline-for-semi","title":"RefTeacher: A Strong Baseline for Semi-Supervised Referring Expression Comprehension","date":"2023-01-01","arxiv_id":null,"repositories_listed":0,"syntology":null},{"url":null,"slug":"bayesian-learning-for-dynamic-inference","title":"Bayesian Learning for Dynamic Inference","date":"2022-12-30","arxiv_id":"2301.00032","repositories_listed":0,"syntology":null},{"url":null,"slug":"behavioral-cloning-via-search-in-video","title":"Behavioral Cloning via Search in Video PreTraining Latent Space","date":"2022-12-27","arxiv_id":"2212.13326","repositories_listed":0,"syntology":null},{"url":null,"slug":"imitation-is-not-enough-robustifying","title":"Imitation Is Not Enough: Robustifying Imitation with Reinforcement Learning for Challenging Driving Scenarios","date":"2022-12-21","arxiv_id":"2212.11419","repositories_listed":0,"syntology":null},{"url":null,"slug":"i2d2-inductive-knowledge-distillation-with","title":"I2D2: Inductive Knowledge Distillation with NeuroLogic and Self-Imitation","date":"2022-12-19","arxiv_id":"2212.09246","repositories_listed":0,"syntology":null},{"url":null,"slug":"model-based-trajectory-stitching-for-improved-1","title":"Model-based trajectory stitching for improved behavioural cloning and its applications","date":"2022-12-08","arxiv_id":"2212.04280","repositories_listed":0,"syntology":null},{"url":null,"slug":"accelerating-self-imitation-learning-from","title":"Accelerating Self-Imitation Learning from Demonstrations via Policy Constraints and Q-Ensemble","date":"2022-12-07","arxiv_id":"2212.03562","repositories_listed":0,"syntology":null},{"url":null,"slug":"learning-graph-search-heuristics","title":"Learning Graph Search Heuristics","date":"2022-12-07","arxiv_id":"2212.03978","repositories_listed":0,"syntology":null},{"url":null,"slug":"efficient-learning-of-voltage-control","title":"Efficient Learning of Voltage Control Strategies via Model-based Deep Reinforcement Learning","date":"2022-12-06","arxiv_id":"2212.02715","repositories_listed":0,"syntology":null},{"url":null,"slug":"accelerating-interactive-human-like","title":"Accelerating Interactive Human-like Manipulation Learning with GPU-based Simulation and High-quality Demonstrations","date":"2022-12-05","arxiv_id":"2212.02126","repositories_listed":0,"syntology":null},{"url":null,"slug":"learning-to-optimize-in-model-predictive","title":"Learning to Optimize in Model Predictive Control","date":"2022-12-05","arxiv_id":"2212.02603","repositories_listed":0,"syntology":null},{"url":null,"slug":"learning-and-blending-robot-hugging-behaviors","title":"Learning and Blending Robot Hugging Behaviors in Time and Space","date":"2022-12-03","arxiv_id":"2212.01507","repositories_listed":0,"syntology":null},{"url":null,"slug":"embedding-synthetic-off-policy-experience-for","title":"Embedding Synthetic Off-Policy Experience for Autonomous Driving via Zero-Shot Curricula","date":"2022-12-02","arxiv_id":"2212.01375","repositories_listed":0,"syntology":null},{"url":null,"slug":"generalizable-human-robot-collaborative","title":"Generalizable Human-Robot Collaborative Assembly Using Imitation Learning and Force Control","date":"2022-12-02","arxiv_id":"2212.01434","repositories_listed":0,"syntology":null},{"url":null,"slug":"multi-task-imitation-learning-for-linear","title":"Multi-Task Imitation Learning for Linear Dynamical Systems","date":"2022-12-01","arxiv_id":"2212.00186","repositories_listed":0,"syntology":null},{"url":null,"slug":"safe-reinforcement-learning-with","title":"Safe Reinforcement Learning with Probabilistic Control Barrier Functions for Ramp Merging","date":"2022-12-01","arxiv_id":"2212.00618","repositories_listed":0,"syntology":null},{"url":null,"slug":"transfer-rl-via-the-undo-maps-formalism","title":"Transfer RL via the Undo Maps Formalism","date":"2022-11-26","arxiv_id":"2211.14469","repositories_listed":0,"syntology":null},{"url":null,"slug":"discovering-generalizable-spatial-goal","title":"Discovering Generalizable Spatial Goal Representations via Graph-based Active Reward Learning","date":"2022-11-24","arxiv_id":"2211.15339","repositories_listed":0,"syntology":null},{"url":null,"slug":"improving-multimodal-interactive-agents-with","title":"Improving Multimodal Interactive Agents with Reinforcement Learning from Human Feedback","date":"2022-11-21","arxiv_id":"2211.11602","repositories_listed":0,"syntology":null},{"url":null,"slug":"robotic-skill-acquisition-via-instruction","title":"Robotic Skill Acquisition via Instruction Augmentation with Vision-Language Models","date":"2022-11-21","arxiv_id":"2211.11736","repositories_listed":0,"syntology":null},{"url":null,"slug":"abc-adversarial-behavioral-cloning-for","title":"ABC: Adversarial Behavioral Cloning for Offline Mode-Seeking Imitation Learning","date":"2022-11-08","arxiv_id":"2211.04005","repositories_listed":0,"syntology":null},{"url":null,"slug":"learning-modular-robot-locomotion-from","title":"Learning Modular Robot Locomotion from Demonstrations","date":"2022-10-31","arxiv_id":"2210.17491","repositories_listed":0,"syntology":null},{"url":null,"slug":"imitating-opponent-to-win-adversarial-policy","title":"Imitating Opponent to Win: Adversarial Policy Imitation Learning in Two-player Competitive Games","date":"2022-10-30","arxiv_id":"2210.16915","repositories_listed":0,"syntology":null},{"url":null,"slug":"d-shape-demonstration-shaped-reinforcement","title":"D-Shape: Demonstration-Shaped Reinforcement Learning via Goal Conditioning","date":"2022-10-26","arxiv_id":"2210.14428","repositories_listed":0,"syntology":null},{"url":null,"slug":"cut-and-approximate-3d-shape-reconstruction","title":"Cut-and-Approximate: 3D Shape Reconstruction from Planar Cross-sections with Deep Reinforcement Learning","date":"2022-10-22","arxiv_id":"2210.12509","repositories_listed":0,"syntology":null},{"url":null,"slug":"differentiable-constrained-imitation-learning","title":"Differentiable Constrained Imitation Learning for Robot Motion Planning and Control","date":"2022-10-21","arxiv_id":"2210.11796","repositories_listed":0,"syntology":null},{"url":null,"slug":"learning-and-retrieval-from-prior-data-for","title":"Learning and Retrieval from Prior Data for Skill-based Imitation Learning","date":"2022-10-20","arxiv_id":"2210.11435","repositories_listed":0,"syntology":null},{"url":null,"slug":"nift-neural-interaction-field-and-template","title":"NIFT: Neural Interaction Field and Template for Object Manipulation","date":"2022-10-20","arxiv_id":"2210.10992","repositories_listed":0,"syntology":null},{"url":null,"slug":"robust-imitation-via-mirror-descent-inverse-1","title":"Robust Imitation via Mirror Descent Inverse Reinforcement Learning","date":"2022-10-20","arxiv_id":"2210.11201","repositories_listed":0,"syntology":null},{"url":null,"slug":"cnt-conditioning-on-noisy-targets-a-new","title":"CNT (Conditioning on Noisy Targets): A new Algorithm for Leveraging Top-Down Feedback","date":"2022-10-18","arxiv_id":"2210.09505","repositories_listed":0,"syntology":null},{"url":null,"slug":"hierarchical-model-based-imitation-learning","title":"Hierarchical Model-Based Imitation Learning for Planning in Autonomous Driving","date":"2022-10-18","arxiv_id":"2210.09539","repositories_listed":0,"syntology":null},{"url":null,"slug":"output-feedback-tube-mpc-guided-data","title":"Output Feedback Tube MPC-Guided Data Augmentation for Robust, Efficient Sensorimotor Policy Learning","date":"2022-10-18","arxiv_id":"2210.10127","repositories_listed":0,"syntology":null},{"url":null,"slug":"model-predictive-control-via-on-policy","title":"Model Predictive Control via On-Policy Imitation Learning","date":"2022-10-17","arxiv_id":"2210.09206","repositories_listed":0,"syntology":null},{"url":"/paper/robust-imitation-of-a-few-demonstrations-with","slug":"robust-imitation-of-a-few-demonstrations-with","title":"Robust Imitation of a Few Demonstrations with a Backwards Model","date":"2022-10-17","arxiv_id":"2210.09337","repositories_listed":0,"syntology":{"n":5,"n_ran":4,"n_constructed":0,"n_ran_checked":4,"n_instrument":0,"n_unverified":1,"n_honours":0,"n_violates":0,"n_no_contract":4,"n_pointer_only":0,"phrase":"4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified","sample_list":"/paper/robust-imitation-of-a-few-demonstrations-with#ran","syntology_url":"https://syntology.ai/paper/2210.09337","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2210.09337"}},"official":null}},{"url":null,"slug":"learning-based-motion-planning-in-dynamic","title":"Learning-based Motion Planning in Dynamic Environments Using GNNs and Temporal Encoding","date":"2022-10-16","arxiv_id":"2210.08408","repositories_listed":0,"syntology":null},{"url":null,"slug":"eliciting-compatible-demonstrations-for-multi","title":"Eliciting Compatible Demonstrations for Multi-Human Imitation Learning","date":"2022-10-14","arxiv_id":"2210.08073","repositories_listed":0,"syntology":null},{"url":null,"slug":"real-world-offline-reinforcement-learning","title":"Real World Offline Reinforcement Learning with Realistic Data Source","date":"2022-10-12","arxiv_id":"2210.06479","repositories_listed":0,"syntology":null},{"url":null,"slug":"graph-neural-network-policies-and-imitation","title":"Graph Neural Network Policies and Imitation Learning for Multi-Domain Task-Oriented Dialogues","date":"2022-10-11","arxiv_id":"2210.05252","repositories_listed":0,"syntology":null},{"url":"/paper/a-new-path-scaling-vision-and-language","slug":"a-new-path-scaling-vision-and-language","title":"A New Path: Scaling Vision-and-Language Navigation with Synthetic Instructions and Imitation Learning","date":"2022-10-06","arxiv_id":"2210.03112","repositories_listed":0,"syntology":null},{"url":null,"slug":"extraneousness-aware-imitation-learning-1","title":"Extraneousness-Aware Imitation Learning","date":"2022-10-04","arxiv_id":"2210.01379","repositories_listed":0,"syntology":null}],"record_sha256":"8003422096423c7783da9a5d7977db340a5929b69552b262766a9aa0abc902d9","record_changed_at":"2026-09-28","record_changed_at_basis":"first_hashed"}