{"url":"/dataset/crello","name":"Crello","full_name":"Crello dataset","description_markdown":"Crello dataset consists of design templates obtained from online design service, crello.com. The dataset contains designs for various display formats, such as social media posts, banner ads, blog headers, or printed posters, all in a vector format. In dataset construction, design templates and associated resources (e.g., linked images) from crello.com were first downloaded. After the initial data acquisition, the data structure was inspected and identified useful vector graphic information in each template. Next, mal-formed templates or those having more\r\nthan 50 elements were eliminated, resulting in 23,182 templates. The data was paritioned to 18,768 / 2,315 / 2,278 examples for train, validation, and test splits.\r\n\r\nSource: [https://arxiv.org/pdf/2108.01249v1.pdf](https://arxiv.org/pdf/2108.01249v1.pdf)\r\n\r\nImage source: [https://arxiv.org/pdf/2108.01249v1.pdf](https://arxiv.org/pdf/2108.01249v1.pdf)","description_withheld":null,"homepage":"https://github.com/CyberAgentAILab/canvas-vae/blob/main/docs/crello-dataset.md","introduced_date":"2021-08-28","introduced_date_note":null,"introduced_by":{"paper":"/paper/canvasvae-learning-to-generate-vector-graphic","title":"CanvasVAE: Learning to Generate Vector Graphic Documents","first_author":"Kota Yamaguchi","url":null},"license":{"name":"CDLA-Permissive-2.0","url":"https://cdla.dev/permissive-2-0/"},"modalities":[{"name":"Images","url":"/datasets/modality/images"}],"tasks":[],"languages":[],"variants":["Crello"],"data_loaders":[],"num_papers_in_archive":11,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}