{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/im2text-describing-images-using-1-million","title":"Im2Text: Describing Images Using 1 Million Captioned Photographs","arxiv_id":null,"date":"2011-12-01","proceeding":"NeurIPS 2011 12","authors":["Vicente Ordonez","Girish Kulkarni","Tamara L. Berg"],"abstract":"We develop and demonstrate automatic image description methods using a large captioned photo collection.  One contribution is our technique for the automatic collection of this new dataset -- performing a huge number of Flickr queries and then filtering the noisy results down to 1 million images with associated visually relevant captions.  Such a collection allows us to approach the extremely challenging problem of description generation using relatively simple non-parametric methods and produces surprisingly effective results. We also develop methods incorporating many state of the art, but fairly noisy, estimates of image content to produce even more pleasing results. Finally we introduce a new objective performance measure for image captioning.","url_abs":"http://papers.nips.cc/paper/4470-im2text-describing-images-using-1-million-captioned-photographs","url_pdf":"http://papers.nips.cc/paper/4470-im2text-describing-images-using-1-million-captioned-photographs.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[{"task_slug":"image-captioning","task_name":"Image Captioning"},{"task_slug":null,"task_name":"Image Description"}],"methods":[],"datasets_introduced":[{"slug":"sbu-captions-dataset","name":"SBU Captions Dataset","full_name":""}],"methods_introduced":[],"results":[],"syntology":{"syntology_url":null,"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}