Home › Datasets

Datasets

archive 2025-07-28

12,172 datasets listed, ordered by the archive's paper count. Page 69 of 254: 48 shown of 12,172.

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter a dataset can carry several tags, so counts overlap

Modality 39

Task 500 shown of 3,717, by dataset count

Question Answering 413Semantic Segmentation 347Object Detection 332Image Classification 283Classification 222Text Classification 167Language Modelling 166Text Generation 162Visual Question Answering (VQA) 144Named Entity Recognition (NER) 130Pose Estimation 124Anomaly Detection 119Action Recognition 115Instance Segmentation 111Sentiment Analysis 104Text Summarization 98Domain Adaptation 96Speech Recognition 96Reading Comprehension 95Information Retrieval 932D Object Detection 92Image Retrieval 87Machine Translation 82Relation Extraction 822D Semantic Segmentation 81Natural Language Inference 81Image Captioning 79Image Generation 79Depth Estimation 77Node Classification 75Natural Language Understanding 72Autonomous Driving 70Code Generation 70Object Tracking 693D Object Detection 67Face Recognition 67Link Prediction 67Person Re-Identification 65Data Augmentation 63Multi-Task Learning 593D Reconstruction 58Common Sense Reasoning 56Recommendation Systems 56Video Understanding 56Optical Character Recognition (OCR) 55Graph Classification 54Abstractive Text Summarization 52Emotion Recognition 52Medical Image Segmentation 52Word Embeddings 523D Human Pose Estimation 50Hate Speech Detection 47Self-Supervised Learning 47Knowledge Graphs 46Novel View Synthesis 46Time Series Forecasting 46Coreference Resolution 45Few-Shot Learning 45Object Recognition 45Segmentation 45Entity Linking 44Image Clustering 44Misinformation 44Semantic Parsing 44Visual Reasoning 44Audio Classification 43Image Segmentation 43Image Super-Resolution 43Machine Reading Comprehension 43Multi-Object Tracking 43Scene Understanding 43Video Question Answering 43Zero-Shot Learning 43Temporal Action Localization 423D Semantic Segmentation 41Fine-Grained Image Classification 41Decision Making 40Face Detection 39Retrieval 38Video Captioning 38Panoptic Segmentation 37Text Retrieval 373D Pose Estimation 36Continual Learning 35Optical Flow Estimation 35Stance Detection 35Trajectory Prediction 35Video Retrieval 35Action Detection 34Automatic Speech Recognition (ASR) 34Dialogue Generation 34Unsupervised Domain Adaptation 34Action Classification 33Metric Learning 33Monocular Depth Estimation 33Music Generation 33Image-to-Image Translation 32Visual Question Answering 32Domain Generalization 31regression 31Activity Recognition 30Fake News Detection 30Multi-Label Classification 30Sign Language Recognition 30Skeleton Based Action Recognition 30Autonomous Vehicles 29Facial Expression Recognition (FER) 29Scene Text Recognition 29Visual Object Tracking 29Document Summarization 28Object Counting 28Visual Place Recognition 28Visual Tracking 28task 28Emotion Classification 27NER 27Neural Architecture Search 27Slot Filling 27Unsupervised Anomaly Detection 27Visual Localization 27DeepFake Detection 26Drug Discovery 26Mathematical Reasoning 26Music Information Retrieval 26Open-Domain Question Answering 26Part-Of-Speech Tagging 26Visual Odometry 262D Human Pose Estimation 25Intent Detection 25Question Generation 25Text-to-Image Generation 25Time Series Analysis 25Video Prediction 25Cross-Modal Retrieval 24Data-to-Text Generation 24Face Verification 24Few-Shot Image Classification 24Out-of-Distribution Detection 24Speech Enhancement 24Video Generation 246D Pose Estimation 23Crowd Counting 23Denoising 23Fairness 23Hand Pose Estimation 23Human-Object Interaction Detection 23Low-Light Image Enhancement 23Math Word Problem Solving 23Relation Classification 23Scene Classification 23Automatic Speech Recognition 22Cell Segmentation 22Handwriting Recognition 22Image Denoising 22Multimodal Deep Learning 22Robotic Grasping 22Speech Emotion Recognition 22Speech Synthesis 22Super-Resolution 22Task-Oriented Dialogue Systems 22Video Classification 22Graph Regression 21Imitation Learning 21Medical Diagnosis 21Molecular Property Prediction 21Simultaneous Localization and Mapping 21Text-To-SQL 213D Hand Pose Estimation 20Age Estimation 20Aspect-Based Sentiment Analysis (ABSA) 20Change Detection 20Language Identification 20Named Entity Recognition 20Reinforcement Learning (RL) 20Sound Event Detection 20Stereo Matching 20Style Transfer 20Text Simplification 20Traffic Prediction 20Animal Pose Estimation 19Graph Clustering 19Medical Image Classification 19Motion Synthesis 19Robot Navigation 19Semantic Textual Similarity 19Time Series Classification 19Video Object Segmentation 19Visual Navigation 193D Instance Segmentation 18Action Segmentation 18Document Classification 18Gaze Estimation 18Image Enhancement 18Image Restoration 18Logical Reasoning 18Multiple Object Tracking 18Multivariate Time Series Forecasting 18Object Localization 18Paraphrase Identification 18Quantization 18Text-To-Speech Synthesis 18Time Series 18Transductive Zero-Shot Classification 18Transfer Learning 18Zero-Shot Video Question Answer 183D Object Tracking 17Action Recognition In Videos 17Beat Tracking 17Binarization 17Deblurring 17Face Alignment 17Face Anti-Spoofing 17Fact Verification 17Image Dehazing 17Image Inpainting 17Lane Detection 17Lesion Segmentation 17Meta-Learning 17Motor Imagery Decoding (left-hand vs right-hand) 17Paraphrase Generation 17Pedestrian Detection 17RGB Salient Object Detection 17SMAC+ 17Salient Object Detection 17Scene Text Detection 17Semantic Similarity 17Sign Language Translation 17Topic Models 17Translation 17Video Anomaly Detection 17Active Learning 16Binary Classification 16Contrastive Learning 16Event Extraction 16Facial Landmark Detection 16Grammatical Error Correction 16Hand Gesture Recognition 16Image Manipulation Detection 16Image Quality Assessment 16Joint Entity and Relation Extraction 16Knowledge Graph Completion 16Learning with noisy labels 16Long-tail Learning 16Prompt Engineering 16Scene Recognition 16Self-Driving Cars 16Speech Separation 16Trajectory Forecasting 16Video Summarization 16Video Super-Resolution 16Weather Forecasting 16Word Sense Disambiguation 16Zero-shot Text Search 163D Action Recognition 153D Classification 153D Human Reconstruction 15Audio Source Separation 15Computed Tomography (CT) 15Cross-Lingual Transfer 15Dependency Parsing 15Gesture Recognition 15Hierarchical Multi-label Classification 15Human Activity Recognition 15Knowledge Base Question Answering 15License Plate Recognition 15Multi-Document Summarization 15Node Classification on Non-Homophilic (Heterophilic) Graphs 15Physical Simulations 15Saliency Detection 15Sentiment Classification 15Single-View 3D Reconstruction 15Token Classification 15Tumor Segmentation 15Video Frame Interpolation 15Video Inpainting 15Within-Session ERP 152D Pose Estimation 14Boundary Detection 14Code Search 14Emotion Recognition in Conversation 14Event Detection 14Handwritten Text Recognition 14Human Detection 14Indoor Localization 14Instruction Following 14Intent Classification 14Multi-class Classification 14Music Transcription 14Nested Named Entity Recognition 14Network Intrusion Detection 14Node Clustering 14Open Information Extraction 14Pose Tracking 14Sarcasm Detection 14Semantic Role Labeling 14Semi-Supervised Image Classification 14Sentence Classification 14Spoken Language Understanding 14Text-to-Video Generation 14Weakly Supervised Object Detection 14motion prediction 14Adversarial Robustness 13Clustering Algorithms Evaluation 13Community Detection 13Density Estimation 13Dialogue State Tracking 13Disentanglement 13Document Layout Analysis 13Downbeat Tracking 13Entity Disambiguation 13Hyperspectral Image Classification 13Image Registration 13Keypoint Detection 13Multi-Label Text Classification 13Multiple Instance Learning 13Open-Domain Dialog 13Semi-Supervised Semantic Segmentation 13Semi-Supervised Video Object Segmentation 13Small Object Detection 13UIE 13Video Object Detection 13Video Object Tracking 13Vision and Language Navigation 13Activity Detection 12COVID-19 Diagnosis 12Code Completion 12Conversational Response Selection 12Entity Resolution 12Entity Typing 12Fact Checking 12Fine-Grained Image Recognition 12Fraud Detection 12Graph Embedding 12Image Compression 12Image Reconstruction 12Imputation 12Key Information Extraction 12License Plate Detection 12Motion Forecasting 12Multiple-choice 12News Classification 12Object Detection In Aerial Images 12Point Cloud Registration 12Real-Time Semantic Segmentation 12Referring Expression Comprehension 12Relational Reasoning 12Robot Manipulation 12Speaker Verification 12Surface Normals Estimation 12Time Series Prediction 12Video Quality Assessment 12Video Segmentation 123D Face Reconstruction 11Action Anticipation 11Adversarial Attack 11Audio Generation 11Bias Detection 11Depth Completion 11Dimensionality Reduction 11Event-based vision 11Generalizable Person Re-identification 11Graph Matching 11Head Pose Estimation 11Material Recognition 11Mathematical Question Answering 11Medical Visual Question Answering 11Multimodal Emotion Recognition 11Multimodal Reasoning 11Opinion Mining 11Outlier Detection 11Referring Expression Segmentation 11Scene Graph Generation 11Sentence Embeddings 11Sequential Recommendation 11Speaker Diarization 11Stochastic Optimization 11Synthetic Data Generation 11Unsupervised Object Segmentation 11Unsupervised Person Re-Identification 11Virtual Try-on 11Visual Grounding 11Zero-Shot Composed Image Retrieval (ZS-CIR) 113D Depth Estimation 10Acoustic Scene Classification 10Answer Selection 10Audio Tagging 10Automatic Post-Editing 10Benchmarking 10Binary text classification 10Chatbot 10Citation Recommendation 10Code Translation 10Column Type Annotation 10Continuous Control 10Conversational Question Answering 10Conversational Response Generation 10Discourse Parsing 10Face Swapping 10Facial Attribute Classification 10Federated Learning 10Font Recognition 10Generalized Zero-Shot Learning 10Graph Learning 10Handwriting generation 10Human Part Segmentation 10LIDAR Semantic Segmentation 10Large Language Model 10Motion Estimation 10Multi-Label Image Classification 10Multiview Detection 10Music Classification 10Object Detection In Indoor Scenes 10Online Multi-Object Tracking 10Passage Retrieval 10Person Identification 10Remote Sensing Image Classification 10Robust classification 10Scene Segmentation 10Semantic SLAM 10Table Detection 10Table annotation 10Text-to-Code Generation 10Time Series Anomaly Detection 10Time Series Regression 10Topic Classification 10Unsupervised Object Detection 10Video Grounding 10Video Recognition 10Within-Session Motor Imagery (left hand vs. right hand) 103D Medical Imaging Segmentation 93D Shape Reconstruction 96D Pose Estimation using RGBD 9Abusive Language 9Action Quality Assessment 9Anomaly Classification 9AutoML 9Blind Super-Resolution 9Camera Localization 9Change Point Detection 9Chart Question Answering 9Code Repair 9Color Image Denoising 9Conditional Image Generation 9Cross-Lingual NER 9Defect Detection 9Dense Video Captioning 9Dialogue Understanding 9Edge Detection 9Explanation Generation 9Few-Shot Audio Classification 9Fine-Grained Visual Recognition 9Food Recognition 9Gait Recognition 9Homography Estimation 9Interactive Segmentation 9Intrusion Detection 9KG-to-Text Generation 9Learning-To-Rank 9Medical Report Generation 9Multi-agent Reinforcement Learning 9Multiple Choice Question Answering (MCQA) 9Music Source Separation 9Out of Distribution (OOD) Detection 9Pedestrian Attribute Recognition 9Person Search 9Program Repair 9Real-Time Object Detection 9Reinforcement Learning 9Representation Learning 9Robust Object Detection 9Saliency Prediction 9Scene Generation 9Shadow Removal 9Small Data Image Classification 9Sound Event Localization and Detection 9Surface Reconstruction 9Trajectory Planning 9Unsupervised Semantic Segmentation 9Vehicle Re-Identification 9Video Description 9Video Reconstruction 9Vision-Language Navigation 9Visual Dialog 9Within-Session Motor Imagery (right hand vs. feet) 9object-detection 93D Object Recognition 83D Object Reconstruction 8AI Agent 8Ad-hoc video search 8Anomaly Detection In Surveillance Videos 8Arithmetic Reasoning 8Automated Theorem Proving 8Brain Tumor Segmentation 8Breast Cancer Detection 8Causal Inference 8Chinese Reading Comprehension 8Click-Through Rate Prediction 8Code Classification 8Code Summarization 8Colorectal Polyps Characterization 8

Language 367

English 3,998Chinese 460German 199French 190Spanish 156Russian 143Japanese 111Arabic 109Italian 98Portuguese 93Hindi 81Vietnamese 71Korean 68Bengali 66Turkish 63Persian 60Dutch 51Tamil 49Indonesian 44Polish 44Czech 41Danish 35Finnish 35Telugu 35Romanian 34Urdu 34Thai 33Marathi 31Hungarian 29Multilingual 29Swedish 29Greek 26Gujarati 26Hebrew 25Mandarin Chinese 25Estonian 24Malayalam 23Ukrainian 23Bulgarian 22Basque 21Kannada 19Punjabi 19Catalan 18Croatian 18Slovak 18Swahili 18Lithuanian 17Latvian 16Norwegian 16Serbian 16Slovenian 16Kazakh 15Amharic 14Iranian Persian 14Albanian 12Kurdish 12Assamese 11Burmese 11Sinhala 11Tagalog 11Yoruba 11Armenian 10Azerbaijani 10Filipino 10Irish 10Macedonian 10Welsh 10American Sign Language 9Georgian 9Maltese 9Mongolian 9Sanskrit 9Breton 8Galician 8Hausa 8Igbo 8Odia 8Oriya (macrolanguage) 8Esperanto 7Nepali (individual language) 7Nepali (macrolanguage) 7Oromo 7Somali 7Uzbek 7Afrikaans 6Bambara 6Belarusian 6Central Khmer 6Guarani 6Icelandic 6Javanese 6Malagasy 6Nigerian Pidgin 6Serbo-Croatian 6Standard Arabic 6Sundanese 6Tibetan 6Western Panjabi 6Wolof 6Bosnian 5Central Kurdish 5Fon 5Ganda 5Haitian 5Latin 5Malay (individual language) 5Norwegian Nynorsk 5Quechua 5Scottish Gaelic 5Sindhi 5Tigrinya 5Aymara 4Bangala 4Bavarian 4Chechen 4Dhivehi 4Egyptian Arabic 4Ewe 4Kabyle 4Lingala 4Norwegian Bokmål 4Tatar 4Tetum 4Tswana 4Twi 4Upper Sorbian 4Xhosa 4Aragonese 3Bashkir 3Bishnupriya 3Cebuano 3Central Pashto 3Chuvash 3Erzya 3Faroese 3Fulah 3Goan Konkani 3Iloko 3Interlingue 3Kinyarwanda 3Kirghiz 3Lao 3Luo (Kenya and Tanzania) 3Maithili 3Modern Greek 3Nyanja 3Occitan (post 1500) 3Romansh 3Rundi 3Russia Buriat 3Sardinian 3South Azerbaijani 3Southern Pashto 3Swiss German 3Turkmen 3Uighur 3Yiddish 3Zulu 3Ancient Greek 2Argentine Sign Language 2Asturian 2Avaric 2Bangladeshi Sign Language 2Bhojpuri 2Central Bikol 2Cherokee 2Church Slavic 2Cornish 2Corsican 2Dimli (individual language) 2Eastern Mari 2German Sign Language 2Gothic 2Gulf Arabic 2Ido 2Inuktitut 2Jamaican Creole English 2Jejueo 2Kalaallisut 2Kalmyk 2Karachay-Balkar 2Komi 2Komi-Permyak 2Lezghian 2Limburgan 2Livvi 2Lojban 2Lombard 2Low German 2Lower Sorbian 2Luxembourgish 2Malay (macrolanguage) 2Manipuri 2Manx 2Maori 2Mazanderani 2Minangkabau 2Mingrelian 2Mirandese 2Moksha 2Mossi 2Naxi 2Neapolitan 2Newari 2Northern Frisian 2Northern Kurdish 2Northern Luri 2Northern Sami 2Old Spanish 2Ossetian 2Pampanga 2Piemontese 2Pushto 2Shona 2Sichuan Yi 2Sicilian 2Swati 2Swiss-German Sign Language 2Tai 2Tajik 2Tsonga 2Turkish Sign Language 2Tuvinian 2Udmurt 2Venda 2Venetian 2Volapük 2Walloon 2Waray (Philippines) 2Western Frisian 2Western Mari 2Wu Chinese 2Yakut 2Yue Chinese 2Abkhazian 1Achinese 1Adyghe 1Afar 1Akan 1Akkadian 1Akuntsu 1Ambonese Malay 1Ancient Hebrew 1Andaman Creole Hindi 1Apurinã 1Arpitan 1Assyrian Neo-Aramaic 1Banjar 1Bemba (Zambia) 1Bislama 1Bodo (India) 1Buginese 1Chamorro 1Chavacano 1Cheyenne 1Choctaw 1Chukot 1Congo Swahili 1Coptic 1Cree 1Creek 1Crimean Tatar 1Cusco Quechua 1Dogri (macrolanguage) 1Dzongkha 1Extremaduran 1Fiji Hindi 1Fijian 1French Sign Language 1Friulian 1Gagauz 1Gan Chinese 1Geez 1Gilaki 1Greek Sign Language 1Hakha Chin 1Hakka Chinese 1Halh Mongolian 1Hawaiian 1Herero 1Hiri Motu 1Interlingua (International Auxiliary Language Association) 1Inupiaq 1Kabardian 1Kanuri 1Kara-Kalpak 1Karelian 1Kashmiri 1Kashubian 1Khunsari 1Kikuyu 1Komi-Zyrian 1Kongo 1Krio 1Kuanyama 1Kupang Malay 1Kölsch 1Ladino 1Lak 1Latgalian 1Ligurian 1Literary Chinese 1Lozi 1Lunda 1Luo (Cameroon) 1Lushai 1Makasar 1Malayic Dayak 1Marshallese 1Mbyá Guaraní 1Min Dong Chinese 1Modern Greek (1453-) 1Moroccan Arabic 1Mundurukú 1Narom 1Nauru 1Navajo 1Nayini 1Ndonga 1Northern Pashto 1Novial 1Official Aramaic (700-300 BCE) 1Old English (ca. 450-1100) 1Old French 1Old Russian 1Old Turkish 1Pali 1Pangasinan 1Papiamento 1Pedi 1Pennsylvania German 1Pfaelzisch 1Picard 1Pitcairn-Norfolk 1Pontic 1Rajasthani 1Rusyn 1Samoan 1Sango 1Saterfriesisch 1Scots 1Silesian 1Skolt Sami 1Soi 1South Levantine Arabic 1Southern Sotho 1Sranan Tongo 1Swahili (macrolanguage) 1Swedish Sign Language 1Tahitian 1Tok Pisin 1Tonga (Tonga Islands) 1Tonga (Zambia) 1Tosk Albanian 1Tulu 1Tumbuka 1Tunisian Arabic 1Tupinambá 1Uab Meto 1Veps 1Vlaams 1Vlax Romani 1Votic 1Warlpiri 1Zaza 1Zeeuws 1Zhuang 1

All datasets 3265–3312 of 12,172

This is a multiscale dynamic human mobility flow dataset across the United States, with data starting from January 1st, 2019.
9 papers · 0 benchmarks
MuSe-CaR (Multimodal Sentiment Analysis in Car Reviews)
The MuSe-CAR database is a large, multimodal (video, audio, and text) dataset which has been gathered in-the-wild with the intention of further understanding Multimodal Sentiment Analysis in-the-wild, e.g., the emotional engagement that…
9 papers · 0 benchmarks
NCLS (Neural Cross-Lingual Summarization Corpora)
Presents two high-quality large-scale CLS datasets based on existing monolingual summarization datasets.
9 papers · 0 benchmarks
NELA-GT-2018 is a dataset for the study of misinformation that consists of 713k articles collected between 02/2018-11/2018.
9 papers · 0 benchmarks
NICO (Non-I.I.D. Image dataset with Contexts)
I.I.D.
9 papers · 2 benchmarks
A popular dataset for node classification on heterogeneous graphs.
9 papers · 1 benchmark
OASIS-1 (Open Access Series of Imaging Studies)
The Open Access Series of Imaging Studies (OASIS) is a project aimed at making neuroimaging data sets of the brain freely available to the scientific community.
MRI
9 papers · 0 benchmarks
Presents half a million samples and structured meta-data to encourage further research and societal engagement.
9 papers · 1 benchmark
OntoGUM is an OntoNotes-like coreference dataset converted from GUM, an English corpus covering 12 genres using deterministic rules.
9 papers · 1 benchmark
PEC (Persona-Based Empathetic Conversational)
A novel large-scale multi-domain dataset for persona-based empathetic conversations.
9 papers · 0 benchmarks
PPM is a portrait matting benchmark with the following characteristics: - Fine Annotation - All images are labeled and checked carefully.
9 papers · 1 benchmark
PSI-AVA is a dataset designed for holistic surgical scene understanding.
9 papers · 0 benchmarks
PTB-TIR is a Thermal InfraRed (TIR) pedestrian tracking benchmark, which provides 60 TIR sequences with mannuly annoations.
9 papers · 0 benchmarks
The Pascal Panoptic Parts dataset consists of annotations for the part-aware panoptic segmentation task on the PASCAL VOC 2010 dataset.
9 papers · 2 benchmarks
PathTrack is a dataset for person tracking which contains more than 15,000 person trajectories in 720 sequences.
9 papers · 0 benchmarks
PerSeg is a dataset for personalized segmentation.
9 papers · 1 benchmark
PhenoBench (PhenoBench — A Large Dataset and Benchmarks for Semantic Image Interpretation in the Agricultural Domain)
The PhenoBench dataset contains multiple image segmentation challenges from the agricultural domain.
9 papers · 0 benchmarks
Most existing MOT datasets are captured using pinhole cameras, which are characterized by a narrow-FoV and linear sensor motion.
9 papers · 1 benchmark
Consists of 330,000 sketches and 204,000 photos spanning across 110 categories.
9 papers · 0 benchmarks
RADDet (Range-Azimuth-Doppler based Radar Dataset)
RADDet is a radar dataset that contains radar data in the form of Range-Azimuth-Doppler tensors along with the bounding boxes on the tensor for dynamic road users, category labels, and 2D bounding boxes on the Cartesian Bird-Eye-View range…
9 papers · 0 benchmarks
RAVEN-FAIR is a modified version of the RAVEN dataset.
9 papers · 0 benchmarks
RU-APC (Rutgers APC)
The RU-APC (Rutgers APC) dataset is a valuable resource for researchers and developers working on robotic perception solutions for warehouse picking challenges.
9 papers · 0 benchmarks
RadQA (A Question Answering Dataset to Improve Comprehension of Radiology Reports)
RadQA is a radiology question answering dataset with 3074 questions posed against radiology reports and annotated with their corresponding answer spans (resulting in a total of 6148 question-answer evidence pairs) by physicians.
9 papers · 1 benchmark
A human-curated ChineseReading Comprehension dataset on Opinion.
9 papers · 0 benchmarks
ReDWeb-S is a large-scale challenging dataset for Salient Object Detection.
9 papers · 0 benchmarks
Roadside Perception 3D (Rope3D) is a dataset for autonomous driving and monocular 3D object detection task consisting of 50k images and over 1.5M 3D objects in various scenes, which are captured under different settings including various…
9 papers · 1 benchmark
The SALMon dataset and benchmark was introduced in the paper "A Suite for Acoustic Language Model Evaluation", with the goal of evaluating the modelling abilities of speech language models with regards to different kinds of acoustic…
9 papers · 1 benchmark
SAMRS is a remote sensing segmentation dataset which provides object category, location, and instance information that can be used for semantic segmentation, instance segmentation, and object detection, either individually or in…
9 papers · 0 benchmarks
Speech encompasses a wealth of information, including but not limited to content, paralinguistic, and environmental information.
9 papers · 0 benchmarks
SHERLOCK is a corpus of 363K commonsense inferences grounded in 103K images.
9 papers · 0 benchmarks
SOBA (Shadow-OBject Association)
A new dataset called SOBA, named after Shadow-OBject Association, with 3,623 pairs of shadow and object instances in 1,000 photos, each with individual labeled masks.
9 papers · 1 benchmark
SPARTQA (SPAtial Reasoning on Textual Question Answering)
SpartQA is a textual question answering benchmark for spatial reasoning on natural language text which contains more realistic spatial phenomena not covered by prior datasets and that is challenging for state-of-the-art language models…
9 papers · 0 benchmarks
SPARTQA - (SPAtial Reasoning on Textual Question Answering.)
We take advantage of the ground truth of NLVR images, design CFGs to generate stories, and use spatial reasoning rules to ask and answer spatial reasoning questions.
9 papers · 0 benchmarks
SPEECH-COCO contains speech captions that are generated using text-to-speech (TTS) synthesis resulting in 616,767 spoken captions (more than 600h) paired with images.
9 papers · 0 benchmarks
Provides four new test sets for the Stanford Question Answering Dataset (SQuAD) and evaluate the ability of question-answering systems to generalize to new data.
9 papers · 3 benchmarks
SUM is a new benchmark dataset of semantic urban meshes which covers about 4 km2 in Helsinki (Finland), with six classes: Ground, Vegetation, Building, Water, Vehicle, and Boat.
9 papers · 0 benchmarks
SciRepEval is a comprehensive benchmark for training and evaluating scientific document representations.
9 papers · 0 benchmarks
The SciTail dataset is an entailment dataset created from multiple-choice science exams and web sentences.
9 papers · 1 benchmark
SelQA is a dataset that consists of questions generated through crowdsourcing and sentence length answers that are drawn from the ten most prevalent topics in the English Wikipedia.
9 papers · 0 benchmarks
A large-scale dataset for the point cloud completion task on the ShapeNet dataset.
9 papers · 1 benchmark
SketchyScene is a large-scale dataset of scene sketches to advance research on sketch understanding at both the object and scene level.
9 papers · 0 benchmarks
The Standardized Project Gutenberg Corpus (SPGC) is an open science approach to a curated version of the complete PG data containing more than 50,000 books and more than 3×109 word-tokens.
9 papers · 0 benchmarks
SwissDial is an annotated parallel corpus of spoken Swiss German across 8 major dialects, plus a Standard German reference.
9 papers · 0 benchmarks
TEMPO (Localizing Moments in Video with Temporal Language)
TEMPOral reasoning in video and language (TEMPO) is a dataset that consists of two parts: a dataset with real videos and template sentences (TEMPO - Template Language) which allows for controlled studies on temporal language, and a human…
9 papers · 0 benchmarks
TRIP (Tiered Reasoning for Intuitive Physics)
Tiered Reasoning for Intuitive Physics (TRIP) is a novel commonsense reasoning dataset with dense annotations that enable multi-tiered evaluation of machines’ reasoning process.
9 papers · 0 benchmarks
TSAC (Tunisian Sentiment Analysis Corpus)
Tunisian Sentiment Analysis Corpus (TSAC) is a Tunisian Dialect corpus of 17.000 comments from Facebook.
9 papers · 0 benchmarks
TSSB (Time Series Segmentation Benchmark)
The time series segmentation benchmark (TSSB) currently contains 75 annotated time series (TS) with 1-9 segments.
9 papers · 1 benchmark
TUM-GAID (TUM Gait from Audio, Image and Depth) collects 305 subjects performing two walking trajectories in an indoor environment.
9 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.