{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/models-of-visually-grounded-speech-signal-pay","title":"Models of Visually Grounded Speech Signal Pay Attention To Nouns: a Bilingual Experiment on English and Japanese","arxiv_id":"1902.03052","date":"2019-02-08","proceeding":null,"authors":["William N. Havard","Jean-Pierre Chevrot","Laurent Besacier"],"abstract":"We investigate the behaviour of attention in neural models of visually\ngrounded speech trained on two languages: English and Japanese. Experimental\nresults show that attention focuses on nouns and this behaviour holds true for\ntwo very typologically different languages. We also draw parallels between\nartificial neural attention and human attention and show that neural attention\nfocuses on word endings as it has been theorised for human attention. Finally,\nwe investigate how two visually grounded monolingual models can be used to\nperform cross-lingual speech-to-speech retrieval. For both languages, the\nenriched bilingual (speech-image) corpora with part-of-speech tags and forced\nalignments are distributed to the community for reproducible research.","url_abs":"http://arxiv.org/abs/1902.03052v1","url_pdf":"http://arxiv.org/pdf/1902.03052v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"models-of-visually-grounded-speech-signal-pay","repo_url":"https://github.com/William-N-Havard/VGS-dataset-metadata","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"none","reach":{"status":"unanswered"}}],"tasks":[{"task_slug":"retrieval","task_name":"Retrieval"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"syntology_url":null,"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}