{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/vision-position-multi-modal-beam-prediction","title":"Vision-Position Multi-Modal Beam Prediction Using Real Millimeter Wave Datasets","arxiv_id":"2111.07574","date":"2021-11-15","proceeding":null,"authors":["Gouranga Charan","Tawfik Osman","Andrew Hredzak","Ngwe Thawdar","Ahmed Alkhateeb"],"abstract":"Enabling highly-mobile millimeter wave (mmWave) and terahertz (THz) wireless communication applications requires overcoming the critical challenges associated with the large antenna arrays deployed at these systems. In particular, adjusting the narrow beams of these antenna arrays typically incurs high beam training overhead that scales with the number of antennas. To address these challenges, this paper proposes a multi-modal machine learning based approach that leverages positional and visual (camera) data collected from the wireless communication environment for fast beam prediction. The developed framework has been tested on a real-world vehicular dataset comprising practical GPS, camera, and mmWave beam training data. The results show the proposed approach achieves more than $\\approx$ 75\\% top-1 beam prediction accuracy and close to 100\\% top-3 beam prediction accuracy in realistic communication scenarios.","url_abs":"https://arxiv.org/abs/2111.07574v1","url_pdf":"https://arxiv.org/pdf/2111.07574v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"vision-position-multi-modal-beam-prediction","repo_url":"https://github.com/gourangc/Vision-Position-Beam-Prediction","is_official":1,"mentioned_in_paper":0,"mentioned_in_github":0,"framework":"pytorch","reach":null}],"tasks":[{"task_slug":"beam-prediction","task_name":"Beam Prediction"},{"task_slug":null,"task_name":"Position"},{"task_slug":"prediction","task_name":"Prediction"}],"methods":[{"method_slug":"gps","method_name":"GPS"}],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}