{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/the-multi-agent-reinforcement-learning-in","title":"The Multi-Agent Reinforcement Learning in MalmÖ (MARLÖ) Competition","arxiv_id":"1901.08129","date":"2019-01-23","proceeding":null,"authors":["Diego Perez-Liebana","Katja Hofmann","Sharada Prasanna Mohanty","Noburu Kuno","Andre Kramer","Sam Devlin","Raluca D. Gaina","Daniel Ionita"],"abstract":"Learning in multi-agent scenarios is a fruitful research direction, but\ncurrent approaches still show scalability problems in multiple games with\ngeneral reward settings and different opponent types. The Multi-Agent\nReinforcement Learning in Malm\\\"O (MARL\\\"O) competition is a new challenge that\nproposes research in this domain using multiple 3D games. The goal of this\ncontest is to foster research in general agents that can learn across different\ngames and opponent types, proposing a challenge as a milestone in the direction\nof Artificial General Intelligence.","url_abs":"http://arxiv.org/abs/1901.08129v1","url_pdf":"http://arxiv.org/pdf/1901.08129v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"the-multi-agent-reinforcement-learning-in","repo_url":"https://github.com/crowdAI/marLo","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"none","reach":{"status":"ok","spdx":"MIT"}},{"paper_slug":"the-multi-agent-reinforcement-learning-in","repo_url":"https://github.com/maximecohen2/conception_solution_appli_ia_i4_tp_research","is_official":0,"mentioned_in_paper":0,"mentioned_in_github":1,"framework":"none","reach":{"status":"ok"}}],"tasks":[{"task_slug":"multi-agent-reinforcement-learning","task_name":"Multi-agent Reinforcement Learning"},{"task_slug":"reinforcement-learning","task_name":"Reinforcement Learning"},{"task_slug":"reinforcement-learning-1","task_name":"Reinforcement Learning (RL)"},{"task_slug":"reinforcement-learning-2","task_name":"reinforcement-learning"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=1901.08129","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"1901.08129"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/maximecohen2/conception_solution_appli_ia_i4_tp_research","reach":{"status":"ok"}},{"provenance":"external:paperswithcode_snapshot_2025-07-28","url":"https://github.com/crowdAI/marLo","reach":{"status":"ok","spdx":"MIT"}}],"summary":{"unverified":4},"by_repo_kind":{"listed":{"samples":4,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":0,"samples":[{"code_sha256_prefix":"9888ad094311172f","entry":"eval_performance","repo":"crowdAI/marLo","repo_kind":"listed","path":"marlo/experiments/evaluator.py","file_url":"https://github.com/crowdAI/marLo/blob/HEAD/marlo/experiments/evaluator.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"9888ad094311172f"}},{"code_sha256_prefix":"0cdcdcf9fe7fe29d","entry":"launch_minecraft_in_background","repo":"crowdAI/marLo","repo_kind":"listed","path":"marlo/launch_minecraft_in_background.py","file_url":"https://github.com/crowdAI/marLo/blob/HEAD/marlo/launch_minecraft_in_background.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"0cdcdcf9fe7fe29d"}},{"code_sha256_prefix":"b0a9e1e280bb73f0","entry":"run_evaluation_episodes","repo":"crowdAI/marLo","repo_kind":"listed","path":"marlo/experiments/evaluator.py","file_url":"https://github.com/crowdAI/marLo/blob/HEAD/marlo/experiments/evaluator.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"b0a9e1e280bb73f0"}},{"code_sha256_prefix":"d5bf0d80fd025e0a","entry":"threaded","repo":"crowdAI/marLo","repo_kind":"listed","path":"marlo/utils.py","file_url":"https://github.com/crowdAI/marLo/blob/HEAD/marlo/utils.py","link_basis":"harvester_set","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"MIT","inline_ok":true,"mcp_get_code":{"code_sha256":"d5bf0d80fd025e0a"}}]},"arxiv_metadata":null,"syntology_extracted_results":null}