{"url":"/dataset/entityseg","name":"EntitySeg","full_name":null,"description_markdown":"The **EntitySeg** dataset contains 33,227 images with high-quality mask annotations. Compared with existing dataets, there are three distinct properties in EntitySeg. First, 71.25% and 86.23% of the images are of high resolution with at least 2000px×2000px and 1000px×1000px which is more consistent with current digital imaging trends. Second, the dataset is open-world and is not limited to predefined classes. Third, the mask annotation along the boundaries are more accurate than existing datasets.\r\n\r\nSource: [Fine-Grained Entity Segmentation](https://arxiv.org/pdf/2211.05776v1.pdf)\r\n\r\nImage Source: [https://arxiv.org/pdf/2211.05776v1.pdf](https://arxiv.org/pdf/2211.05776v1.pdf)","description_withheld":null,"homepage":"http://luqi.info/entityv2.github.io/","introduced_date":"2022-11-10","introduced_date_note":null,"introduced_by":{"paper":"/paper/fine-grained-entity-segmentation","title":"High-Quality Entity Segmentation","first_author":"Lu Qi","url":null},"license":{"name":"Creative Commons Attribution-NonCommercial 4.0 International License","url":"https://github.com/dvlab-research/Entity/blob/main/LICENSE"},"modalities":[{"name":"Images","url":"/datasets/modality/images"}],"tasks":[{"name":"Semantic Segmentation","url":"/task/semantic-segmentation","datasets_with_task":"/datasets/task/semantic-segmentation"},{"name":"Image Segmentation","url":"/task/image-segmentation","datasets_with_task":"/datasets/task/image-segmentation"}],"languages":[],"variants":["EntitySeg"],"data_loaders":[],"num_papers_in_archive":9,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}