Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 206 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 9841–9888 of 12,172
ProNCI consists of 22.5K proper noun compounds along with their free-form semantic interpretations.
1 paper · 0 benchmarks
This dataset tests the capabilities of language models to correctly capture the meaning of words denoting probabilities (WEP), e.g.
1 paper · 1 benchmark
Functionally correct (ok) and incorrect (buggy) solutions to five Probleable Problems: http://arxiv.org/abs/2405.15123 The ok solutions correspond to attempts that successfully probed all ambiguities in the given specification; the buggy…
1 paper · 0 benchmarks
Procedural Human Action Videos contains a total of 39,982 videos, with more than 1,000 examples for each action of 35 categories.
1 paper · 0 benchmarks
Processed CMIP5 data used for testing the CNN-LSTM model.
1 paper · 0 benchmarks
Processed Twitter is a dataset that is used for Twitter topic recognition.
1 paper · 0 benchmarks
A novel stance detection dataset covering 419 different controversial issues and their related pros and cons collected by procon.org in nonpartisan format.
1 paper · 0 benchmarks
Product Page is a large-scale and realistic dataset of webpages.
1 paper · 0 benchmarks
The corpus contains review sentences mostly of products in electronics domain, annotated and segregated into 4 comparison categories.
1 paper · 1 benchmark
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
ProofNetVerif is an evaluation benchmark comprising 3,752 entries, each including an informal mathematical statement, its reference formalization, a predicted formalization, and a binary label indicating semantic equivalence.
1 paper · 0 benchmarks
Dataset that can be used to evaluate both general semantic flow techniques and region-based approaches such as proposal flow.
1 paper · 0 benchmarks
OOD split of the Mol-Instructions Dataset about Protein Annotation.
1 paper · 0 benchmarks
This is a large-scale court judgment dataset, where each judgment is a summary of the case description with a patternized style.
1 paper · 0 benchmarks
PsOCR (Pashto OCR Dataset)
PsOCR is a large-scale synthetic dataset for Optical Character Recognition in low-resource Pashto language.
1 paper · 0 benchmarks
Pre-training dataset used in paper "From Artificially Real to Real: Leveraging Pseudo Data from Large Language Models for Low-Resource Molecule Discovery" (AAAI 2024) PseudoMD-1M dataset is the first artificially-real dataset for…
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
A.2.1 AN OPEN, LARGE-SCALE DATASET FOR ZERO-SHOT DRUG DISCOVERY DERIVED FROM PUBCHEM We constructed a large public dataset extracted from PubChem (Kim et al., 2019; Preuer et al., 2018), an open chemistry database, and the largest…
1 paper · 0 benchmarks
he goal of this project is to use high throughput screening approaches to identify and develop novel, highly selective small molecule allosteric modulators of the D1 DAR for use as in vitro and in vivo pharmacological tools and in…
1 paper · 0 benchmarks
This dataset gathers 14,857 entities, 133 relations, and entities corresponding tokenized text from PubMed.
1 paper · 0 benchmarks
This dataset gathers three types of pairs: Title-to-Abstract (Training: 22,811/Development: 2095/Test: 2095), Abstract-to-Conclusion and Future work (Training: 22,811/Development: 2095/Test: 2095), Conclusion and Future work-to-Title…
1 paper · 0 benchmarks
This dataset contains a probabilistic sample of ~2.4 million PubMed abstracts, enriched with precomputed dense embeddings (title + abstract), from the ncbi/MedCPT-Article-Encoder model.
1 paper · 0 benchmarks
PubMedQA-MetaGen: Metadata-Enriched PubMedQA Corpus Dataset Summary PubMedQA-MetaGen is a metadata-enriched version of the PubMedQA biomedical question-answering dataset, created using the MetaGenBlendedRAG enrichment pipeline.
1 paper · 2 benchmarks
Data for novelty and its impact detection in scientific publications from Microsoft Academic Graph (now OpenAlex)
1 paper · 0 benchmarks
PushWorld is an environment with simplistic physics that requires manipulation planning with both movable obstacles and tools.
1 paper · 0 benchmarks
PuzzTe (Puzzle Textual Entailment)
Puzzles dataset: comparison, knight&knaves, and zebra puzzles.
1 paper · 0 benchmarks
Evaluate LLM on Real-world coding tasks.
1 paper · 0 benchmarks
A benchmark for Python library migration.
1 paper · 0 benchmarks
We create a new dataset from GitTables, a data lake of 1.7M tables extracted from CSV files on GitHub.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
Pyxis is a performance dataset for specialized accelerators on sparse data.
1 paper · 0 benchmarks
QASiNa (Question Answering Sirah Nabawiyah)
Question Answering Sirah Nabawiyah (QASiNa) Dataset is a reading comprehension dataset consists of QA from Sirah Nabawiyah literature in Indonesian Language
1 paper · 0 benchmarks
QASports (A Question Answering Dataset about Sports)
Sport is one of the most popular and revenue-generating forms of entertainment.
1 paper · 0 benchmarks
A high-quality large-scale dataset consisting of 49,000+ data samples for the task of Chinese query-based document summarization.
1 paper · 0 benchmarks
QDSD (Quantum Dots Stability Diagrams)
This Quantum Dots Stability Diagrams (QDSD) Dataset aggregates experimental stability diagrams of quantum dots from different research groups.
1 paper · 0 benchmarks
Synthetic datasets have successfully been used to probe visual question-answering datasets for their reasoning abilities.
1 paper · 1 benchmark
QTL (Quantitative Trait Locus)
We present QTL, a real-life DS-NER application in the animal science domain.
1 paper · 0 benchmarks
Equilibrium structures of the tautobase(reference) optimized at the level of theory of popular quantum chemical databases (QM9,PC9 and ANI-E).
1 paper · 0 benchmarks
QoEVAVE (Quality of Experience Evaluation of Interactive Virtual Environments with Audiovisual Scenes)
Quality of Experience Evaluation of Interactive Virtual Environments with Audiovisual Scenes (QoEVAVE) provides an initial audiovisual database consiting of 12 sequences capturing real-life nature and urban scenes.
1 paper · 0 benchmarks
This dataset contains closed-loop position, velocity, and current trajectories from 83 motor drives, each consisting of a motor and a Harmonic Drive gearbox.
1 paper · 0 benchmarks
A mapping of Quasimodo to the relations of ConceptNet.
1 paper · 0 benchmarks
Qubit energy relaxation versus frequency and time
1 paper · 0 benchmarks
Quechua Collao corpus for automatic emotion recognition in speech.
1 paper · 1 benchmark
Used in the development of Topographs: Topological Reconstruction of Particle Physics Processes using Graph Neural Networks The datasets contain 5.8M ttbar events in the all hadronic decay channel, with jets matched to the truth partons in…
1 paper · 0 benchmarks
The R1-Onevision dataset is a meticulously crafted resource designed to empower models with advanced multimodal reasoning capabilities.
1 paper · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
This paper describes the first open dataset for full-scale and high-speed autonomous racing.
1 paper · 0 benchmarks
RAOS (Rethinking Abdominal Organ Segmentation)
Rethinking Abdominal Organ Segmentation (RAOS) in the clinical scenario: A robustness evaluation benchmark with challenging cases.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.