Browse State-of-the-Art › Miscellaneous
Miscellaneous
Benchmarks are leaderboard tables with at least one row whose task is in this area, counted by the task's area and not by the archive's per-table category tag, which tags 1,232 tables with Miscellaneous; datasets are those the archive tags with a task in this area; papers with code are catalogue papers tagged with a task in this area that list at least one repository. Task images are not shown (the archive's image host no longer serves them).
Parent tasks
68 tasks in Miscellaneous have sub-tasks in the archive's task tree, most benchmarks first, then most papers with code. Each section shows up to 5 sub-tasks; the task page lists them all.
- Question Answering (19)
- Image Generation (23)
- Anomaly Detection (28)
- Classification (24)
- Language Modelling (6)
- Recommendation Systems (11)
- Unsupervised Anomaly Detection (5)
- Molecular Property Prediction (4)
- Stochastic Optimization (2)
- Logical Reasoning (22)
- Fairness (1)
- Emotion Recognition (12)
- Image Reconstruction (5)
- DeepFake Detection (5)
- Image/Document Clustering (1)
- Fact Checking (5)
- Transfer Learning (4)
- regression (2)
- Intrusion Detection (1)
- Weather Forecasting (1)
- Deep Clustering (3)
- Physical Simulations (3)
- Autonomous Driving (6)
- Robotic Grasping (1)
- Protein Structure Prediction (2)
- Causal Inference (3)
- Malware Classification (4)
- 10-shot image generation (21)
- Image Retrieval with Multi-Modal Query (4)
- Semi Supervised Learning for Image Captioning (1)
- Model Compression (1)
- Synthetic Data Generation (2)
- Offline RL (1)
- Ethics (4)
- Brain Decoding (1)
- Multi-modal Classification (1)
- Management (1)
- Gaussian Processes (1)
- Feature Engineering (1)
- Multi-Armed Bandits (1)
- General Knowledge (7)
- 16k (6)
- Product Recommendation (1)
- Intent Recognition (1)
- Computer Security (1)
- Medical Genetics (1)
- Data Visualization (1)
- Artificial Life (1)
- Privacy Preserving Deep Learning (2)
- Mathematical Proofs (1)
- Table Extraction (1)
- Deception Detection (1)
- Chemical Process (1)
- Intelligent Communication (2)
- Cross-Modal Information Retrieval (1)
- Automatic Machine Learning Model Selection (1)
- Image-text Classification (1)
- 1 Image, 2*2 Stitching (2)
- Seismic Interpretation (2)
- Neural Network Security (1)
- 3D (38)
- Advertising (1)
- Atomistic Description (3)
- Ecommerce (4)
- Facial Recognition and Modelling (26)
- Non-Linear Elasticity (2)
- Pcl Detection (2)
- Remote Sensing (8)
Question Answering
142 benchmarks · 4,171 papers with codeMultiple Choice Question Answering (MCQA)
31 benchmarks · 37 papers with code
Zero-Shot Video Question Answer
17 benchmarks · 73 papers with code
Open-Domain Question Answering
15 benchmarks · 238 papers with code
Knowledge Base Question Answering
10 benchmarks · 67 papers with code
Answer Selection
6 benchmarks · 52 papers with code
5 shown of 19 sub-tasks (4 filed under another area). All sub-tasks of Question Answering →
Image Generation
93 benchmarks · 3,102 papers with codeImage-to-Image Translation
38 benchmarks · 550 papers with code
Text-to-Image Generation
17 benchmarks · 546 papers with code
Image Inpainting
12 benchmarks · 331 papers with code
Conditional Image Generation
11 benchmarks · 167 papers with code
Layout-to-Image Generation
11 benchmarks · 24 papers with code
5 shown of 23 sub-tasks (7 filed under another area). All sub-tasks of Image Generation →
Anomaly Detection
76 benchmarks · 1,727 papers with codeImage Manipulation Detection
21 benchmarks · 38 papers with code
Unsupervised Anomaly Detection
18 benchmarks · 226 papers with code
Anomaly Detection In Surveillance Videos
7 benchmarks · 44 papers with code
Anomaly Classification
5 benchmarks · 33 papers with code
3D Anomaly Detection
3 benchmarks · 18 papers with code
5 shown of 28 sub-tasks (4 filed under another area). All sub-tasks of Anomaly Detection →
Classification
58 benchmarks · 3,778 papers with codeGraph Classification
73 benchmarks · 483 papers with code
Text Classification
68 benchmarks · 1,308 papers with code
Audio Classification
22 benchmarks · 183 papers with code
Medical Image Classification
11 benchmarks · 183 papers with code
Multi-class Classification
5 benchmarks · 289 papers with code
5 shown of 24 sub-tasks (5 filed under another area). All sub-tasks of Classification →
Language Modelling
55 benchmarks · 7,012 papers with codeLong-range modeling
2 benchmarks · 69 papers with code
Cross-Document Language Modeling
2 benchmarks · 1 paper with code
Protein Language Model
1 benchmark · 47 papers with code
XLM-R
0 benchmarks · 99 papers with code
Sentence Pair Modeling
0 benchmarks · 6 papers with code
5 shown of 6 sub-tasks (1 filed under another area). All sub-tasks of Language Modelling →
Recommendation Systems
55 benchmarks · 1,997 papers with codeSequential Recommendation
13 benchmarks · 300 papers with code
Session-Based Recommendations
7 benchmarks · 84 papers with code
Multimodal Recommendation
5 benchmarks · 33 papers with code
Multi-Media Recommendation
4 benchmarks · 2 papers with code
Multi-modal Recommendation
3 benchmarks · 18 papers with code
5 shown of 11 sub-tasks (1 filed under another area). All sub-tasks of Recommendation Systems →
Unsupervised Anomaly Detection
18 benchmarks · 226 papers with codeUnsupervised Anomaly Detection with Specified Settings -- 30% anomaly
5 benchmarks · 4 papers with code
Root Cause Ranking
0 benchmarks · 1 paper with code
Anomaly Detection at 30% anomaly
0 benchmarks · 0 papers with code
Anomaly Detection at Various Anomaly Percentages
0 benchmarks · 0 papers with code
Unsupervised Contextual Anomaly Detection
0 benchmarks · 0 papers with code
5 shown of 5 sub-tasks (1 filed under another area).
Molecular Property Prediction
18 benchmarks · 199 papers with code3D Geometry Prediction
2 benchmarks · 6 papers with code
NMR J-coupling
1 benchmark · 1 paper with code
Odor Descriptor Prediction
1 benchmark · 1 paper with code
mixture property prediction
0 benchmarks · 1 paper with code
4 shown of 4 sub-tasks.
Stochastic Optimization
12 benchmarks · 337 papers with codeDistributed Optimization
1 benchmark · 86 papers with code
Evolutionary Algorithms
0 benchmarks · 241 papers with code
2 shown of 2 sub-tasks.
Logical Reasoning
10 benchmarks · 330 papers with codeTemporal Sequences
1 benchmark · 75 papers with code
Physical Intuition
1 benchmark · 15 papers with code
Elementary Mathematics
1 benchmark · 5 papers with code
Epistemic Reasoning
1 benchmark · 5 papers with code
Metaphor Boolean
1 benchmark · 3 papers with code
5 shown of 22 sub-tasks (1 filed under another area). All sub-tasks of Logical Reasoning →
Fairness
9 benchmarks · 1,714 papers with codeExposure Fairness
0 benchmarks · 8 papers with code
1 shown of 1 sub-task.
Emotion Recognition
9 benchmarks · 614 papers with codeSpeech Emotion Recognition
16 benchmarks · 139 papers with code
Emotion Recognition in Conversation
16 benchmarks · 83 papers with code
Multimodal Emotion Recognition
7 benchmarks · 80 papers with code
Emotion Recognition in Context
4 benchmarks · 5 papers with code
EEG Emotion Recognition
3 benchmarks · 14 papers with code
5 shown of 12 sub-tasks (2 filed under another area). All sub-tasks of Emotion Recognition →
Image Reconstruction
8 benchmarks · 712 papers with codeBlind Super-Resolution
18 benchmarks · 44 papers with code
MRI Reconstruction
6 benchmarks · 198 papers with code
CT Reconstruction
0 benchmarks · 67 papers with code
Film Removal
0 benchmarks · 1 paper with code
WiFi CSI-based Image Reconstruction
0 benchmarks · 1 paper with code
5 shown of 5 sub-tasks (2 filed under another area).
DeepFake Detection
8 benchmarks · 237 papers with codeAudio Deepfake Detection
2 benchmarks · 35 papers with code
Multimodal Forgery Detection
1 benchmark · 1 paper with code
Synthetic Speech Detection
0 benchmarks · 12 papers with code
diffusion-generated faces detection
0 benchmarks · 1 paper with code
Human Detection of Deepfakes
0 benchmarks · 1 paper with code
5 shown of 5 sub-tasks.
Image/Document Clustering
8 benchmarks · 6 papers with codeSelf-Organized Clustering
0 benchmarks · 4 papers with code
1 shown of 1 sub-task.
Fact Checking
7 benchmarks · 297 papers with codeMisconceptions
1 benchmark · 54 papers with code
Sentence Ambiguity
1 benchmark · 3 papers with code
FEVER (2-way)
1 benchmark · 1 paper with code
FEVER (3-way)
1 benchmark · 1 paper with code
Known Unknowns
0 benchmarks · 9 papers with code
5 shown of 5 sub-tasks.
Transfer Learning
6 benchmarks · 3,502 papers with codeMulti-Task Learning
8 benchmarks · 1,306 papers with code
Unsupervised Domain Expansion
2 benchmarks · 2 papers with code
Auxiliary Learning
0 benchmarks · 30 papers with code
Transfer Reinforcement Learning
0 benchmarks · 14 papers with code
4 shown of 4 sub-tasks (2 filed under another area).
regression
6 benchmarks · 2,445 papers with codeTravel Time Estimation
1 benchmark · 19 papers with code
quantile regression
0 benchmarks · 102 papers with code
2 shown of 2 sub-tasks.
Intrusion Detection
6 benchmarks · 151 papers with codeNetwork Intrusion Detection
6 benchmarks · 67 papers with code
1 shown of 1 sub-task.
Weather Forecasting
5 benchmarks · 188 papers with codeSolar Irradiance Forecasting
1 benchmark · 10 papers with code
1 shown of 1 sub-task (1 filed under another area).
Deep Clustering
5 benchmarks · 135 papers with codeTrajectory Clustering
0 benchmarks · 6 papers with code
Deep Nonparametric Clustering
0 benchmarks · 1 paper with code
NONPARAMETRIC DEEP CLUSTERING
0 benchmarks · 1 paper with code
3 shown of 3 sub-tasks.
Physical Simulations
5 benchmarks · 52 papers with codeNeural Network simulation
0 benchmarks · 7 papers with code
Optical Tweezers Simulations
0 benchmarks · 1 paper with code
Liquid Simulation
0 benchmarks · 0 papers with code
3 shown of 3 sub-tasks.
Autonomous Driving
4 benchmarks · 2,091 papers with codeMotion Forecasting
1 benchmark · 88 papers with code
Bench2Drive
1 benchmark · 19 papers with code
NavSim
1 benchmark · 14 papers with code
CARLA MAP Leaderboard
1 benchmark · 6 papers with code
3D Pedestrian Tracking
1 benchmark · 4 papers with code
5 shown of 6 sub-tasks (1 filed under another area). All sub-tasks of Autonomous Driving →
Robotic Grasping
4 benchmarks · 94 papers with codeGrasp Contact Prediction
2 benchmarks · 4 papers with code
1 shown of 1 sub-task.
Protein Structure Prediction
4 benchmarks · 67 papers with codeProtein Interface Prediction
0 benchmarks · 4 papers with code
Protein complex prediction
0 benchmarks · 0 papers with code
2 shown of 2 sub-tasks.
Causal Inference
3 benchmarks · 575 papers with codeHeterogeneous Treatment Effect Estimation
1 benchmark · 19 papers with code
Counterfactual Inference
0 benchmarks · 64 papers with code
IHDP-CATE
0 benchmarks · 0 papers with code
3 shown of 3 sub-tasks.
Malware Classification
3 benchmarks · 47 papers with codeMalware Detection
2 benchmarks · 108 papers with code
Android Malware Detection
0 benchmarks · 19 papers with code
Behavioral Malware Classification
0 benchmarks · 0 papers with code
Behavioral Malware Detection
0 benchmarks · 0 papers with code
4 shown of 4 sub-tasks.
10-shot image generation
3 benchmarks · 21 papers with codeSemantic Segmentation
150 benchmarks · 6,644 papers with code
Text-to-Image Generation
17 benchmarks · 546 papers with code
Deblurring
17 benchmarks · 424 papers with code
Motion Synthesis
13 benchmarks · 126 papers with code
Image Deblurring
9 benchmarks · 167 papers with code
5 shown of 21 sub-tasks (14 filed under another area). All sub-tasks of 10-shot image generation →
Image Retrieval with Multi-Modal Query
3 benchmarks · 9 papers with codeCross-Modal Retrieval
13 benchmarks · 244 papers with code
Zero-Shot Cross-Modal Retrieval
3 benchmarks · 22 papers with code
Cross-Modal Information Retrieval
0 benchmarks · 8 papers with code
Multi-Modal Person Identification
0 benchmarks · 0 papers with code
4 shown of 4 sub-tasks (1 filed under another area).
Semi Supervised Learning for Image Captioning
3 benchmarks · 2 papers with codePseudo Label
0 benchmarks · 438 papers with code
1 shown of 1 sub-task.
Model Compression
2 benchmarks · 440 papers with codeNeural Network Compression
1 benchmark · 77 papers with code
1 shown of 1 sub-task.
Synthetic Data Generation
2 benchmarks · 325 papers with codeSynthetic Outliers Evaluation
3 benchmarks · 0 papers with code
Synthetic Data Evaluation
1 benchmark · 1 paper with code
2 shown of 2 sub-tasks.
Offline RL
2 benchmarks · 310 papers with codeDQN Replay Dataset
0 benchmarks · 6 papers with code
1 shown of 1 sub-task.
Ethics
2 benchmarks · 100 papers with codeMoral Scenarios
1 benchmark · 9 papers with code
Moral Permissibility
1 benchmark · 3 papers with code
Business Ethics
1 benchmark · 1 paper with code
Moral Disputes
1 benchmark · 1 paper with code
4 shown of 4 sub-tasks.
Brain Decoding
2 benchmarks · 42 papers with codeBrain Computer Interface
0 benchmarks · 112 papers with code
1 shown of 1 sub-task (1 filed under another area).
Multi-modal Classification
2 benchmarks · 12 papers with codeImage-text Classification
0 benchmarks · 7 papers with code
1 shown of 1 sub-task.
Management
1 benchmark · 1,214 papers with codeAsset Management
0 benchmarks · 10 papers with code
1 shown of 1 sub-task.
Gaussian Processes
1 benchmark · 685 papers with codeGPR
0 benchmarks · 62 papers with code
1 shown of 1 sub-task (1 filed under another area).
Feature Engineering
1 benchmark · 472 papers with codeImputation
4 benchmarks · 479 papers with code
1 shown of 1 sub-task (1 filed under another area).
Multi-Armed Bandits
1 benchmark · 253 papers with codeThompson Sampling
0 benchmarks · 135 papers with code
1 shown of 1 sub-task (1 filed under another area).
General Knowledge
1 benchmark · 173 papers with codeNatural Questions
2 benchmarks · 89 papers with code
TriviaQA
1 benchmark · 60 papers with code
Miscellaneous
1 benchmark · 29 papers with code
Movie Recommendation
1 benchmark · 29 papers with code
Similarities Abstraction
1 benchmark · 2 papers with code
5 shown of 7 sub-tasks. All sub-tasks of General Knowledge →
16k
1 benchmark · 87 papers with codeObject Detection
123 benchmarks · 4,657 papers with code
Image Super-Resolution
69 benchmarks · 783 papers with code
Image Deblurring
9 benchmarks · 167 papers with code
Scene Generation
6 benchmarks · 122 papers with code
Shadow Removal
5 benchmarks · 71 papers with code
5 shown of 6 sub-tasks (6 filed under another area). All sub-tasks of 16k →
Product Recommendation
1 benchmark · 40 papers with codeContext Aware Product Recommendation
0 benchmarks · 2 papers with code
1 shown of 1 sub-task.
Intent Recognition
1 benchmark · 22 papers with codeMultimodal Intent Recognition
3 benchmarks · 10 papers with code
1 shown of 1 sub-task.
Computer Security
1 benchmark · 18 papers with codeFile Type Identification
0 benchmarks · 1 paper with code
1 shown of 1 sub-task.
Medical Genetics
1 benchmark · 3 papers with codeGenetic Risk Prediction
0 benchmarks · 1 paper with code
1 shown of 1 sub-task.
Data Visualization
0 benchmarks · 117 papers with codeTree Map Layout
0 benchmarks · 1 paper with code
1 shown of 1 sub-task.
Artificial Life
0 benchmarks · 33 papers with codeDevelopmental Learning
0 benchmarks · 7 papers with code
1 shown of 1 sub-task.
Privacy Preserving Deep Learning
0 benchmarks · 30 papers with codeMembership Inference Attack
0 benchmarks · 78 papers with code
Homomorphic Encryption for Deep Learning
0 benchmarks · 1 paper with code
2 shown of 2 sub-tasks.
Mathematical Proofs
0 benchmarks · 29 papers with codeAutomated Theorem Proving
9 benchmarks · 110 papers with code
1 shown of 1 sub-task.
Table Extraction
0 benchmarks · 15 papers with codeTable Functional Analysis
0 benchmarks · 1 paper with code
1 shown of 1 sub-task.
Deception Detection
0 benchmarks · 12 papers with codeDeception Detection In Videos
0 benchmarks · 0 papers with code
1 shown of 1 sub-task.
Chemical Process
0 benchmarks · 11 papers with codeGeochemistry
0 benchmarks · 2 papers with code
1 shown of 1 sub-task.
Intelligent Communication
0 benchmarks · 9 papers with codeSemantic Communication
1 benchmark · 44 papers with code
Beam Prediction
1 benchmark · 11 papers with code
2 shown of 2 sub-tasks.
Cross-Modal Information Retrieval
0 benchmarks · 8 papers with codeCross-Modal Retrieval
13 benchmarks · 244 papers with code
1 shown of 1 sub-task (1 filed under another area).
Automatic Machine Learning Model Selection
0 benchmarks · 7 papers with codeSmart Grid Prediction
0 benchmarks · 1 paper with code
1 shown of 1 sub-task.
Image-text Classification
0 benchmarks · 7 papers with codeMultilingual Image-Text Classification
1 benchmark · 1 paper with code
1 shown of 1 sub-task.
1 Image, 2*2 Stitching
0 benchmarks · 6 papers with codeImage-to-Image Translation
38 benchmarks · 550 papers with code
Fake Image Detection
0 benchmarks · 17 papers with code
2 shown of 2 sub-tasks (2 filed under another area).
Seismic Interpretation
0 benchmarks · 3 papers with codeSeismic Detection
1 benchmark · 2 papers with code
Facies Classification
0 benchmarks · 3 papers with code
2 shown of 2 sub-tasks.
Neural Network Security
0 benchmarks · 2 papers with codeWebsite Fingerprinting Defense
1 benchmark · 0 papers with code
1 shown of 1 sub-task (1 filed under another area).
3D
0 benchmarks · 0 papers with codeObject Detection
123 benchmarks · 4,657 papers with code
Pose Estimation
31 benchmarks · 1,679 papers with code
Depth Estimation
14 benchmarks · 1,029 papers with code
3D Reconstruction
10 benchmarks · 793 papers with code
3D Face Reconstruction
9 benchmarks · 83 papers with code
5 shown of 38 sub-tasks (20 filed under another area). All sub-tasks of 3D →
Advertising
0 benchmarks · 0 papers with codeDetecting Adverts
0 benchmarks · 0 papers with code
1 shown of 1 sub-task.
Atomistic Description
0 benchmarks · 0 papers with codeMolecular Property Prediction
18 benchmarks · 199 papers with code
Formation Energy
14 benchmarks · 31 papers with code
Atomic Forces
0 benchmarks · 11 papers with code
3 shown of 3 sub-tasks.
Ecommerce
0 benchmarks · 0 papers with codeProduct Recommendation
1 benchmark · 40 papers with code
Product Categorization
1 benchmark · 6 papers with code
Online Ranker Evaluation
0 benchmarks · 2 papers with code
Online Review Rating
0 benchmarks · 1 paper with code
4 shown of 4 sub-tasks.
Facial Recognition and Modelling
0 benchmarks · 0 papers with codeFace Alignment
26 benchmarks · 106 papers with code
Face Recognition
25 benchmarks · 639 papers with code
Facial Expression Recognition (FER)
25 benchmarks · 155 papers with code
Face Verification
21 benchmarks · 134 papers with code
Age Estimation
16 benchmarks · 85 papers with code
5 shown of 26 sub-tasks (9 filed under another area). All sub-tasks of Facial Recognition and Modelling →
Non-Linear Elasticity
0 benchmarks · 0 papers with codeStress-Strain Relation
1 benchmark · 1 paper with code
Cantilever Beam
0 benchmarks · 4 papers with code
2 shown of 2 sub-tasks.
Pcl Detection
0 benchmarks · 0 papers with codeSemEval-2022 Task 4-1 (Binary PCL Detection)
1 benchmark · 0 papers with code
SemEval-2022 Task 4-2 (Multi-label PCL Detection)
1 benchmark · 0 papers with code
2 shown of 2 sub-tasks (1 filed under another area).
Remote Sensing
0 benchmarks · 0 papers with codeExtracting Buildings In Remote Sensing Images
3 benchmarks · 10 papers with code
Change detection for remote sensing images
2 benchmarks · 24 papers with code
Building change detection for remote sensing images
2 benchmarks · 19 papers with code
The Semantic Segmentation Of Remote Sensing Imagery
2 benchmarks · 11 papers with code
Remote Sensing Image Classification
1 benchmark · 39 papers with code
5 shown of 8 sub-tasks (3 filed under another area). All sub-tasks of Remote Sensing →
Tasks with no parent task
145 tasks in Miscellaneous sit at the top of the archive's task tree with no sub-tasks of their own, most benchmarks first, then most papers with code.
Click-Through Rate Prediction
20 benchmarks · 165 papers with code
Tabular Data Generation
6 benchmarks · 46 papers with code
Two-sample testing
5 benchmarks · 84 papers with code
Crop Classification
5 benchmarks · 21 papers with code
Recipe Generation
5 benchmarks · 15 papers with code
Collaborative Filtering
4 benchmarks · 461 papers with code
Fine-Grained Urban Flow Inference
4 benchmarks · 4 papers with code
Table Detection
3 benchmarks · 25 papers with code
Knowledge Tracing
2 benchmarks · 110 papers with code
Vulnerability Detection
2 benchmarks · 72 papers with code
Interpretability Techniques for Deep Learning
2 benchmarks · 22 papers with code
Crop Yield Prediction
2 benchmarks · 19 papers with code
Next-basket recommendation
2 benchmarks · 11 papers with code
Twitter Bot Detection
2 benchmarks · 10 papers with code
Unconditional Molecule Generation
2 benchmarks · 6 papers with code
GPS Embeddings
2 benchmarks · 1 paper with code
Text-to-3D-Human Generation
2 benchmarks · 1 paper with code
Computational Efficiency
1 benchmark · 1,644 papers with code
Misinformation
1 benchmark · 457 papers with code
Anatomy
1 benchmark · 375 papers with code
Philosophy
1 benchmark · 174 papers with code
Marketing
1 benchmark · 161 papers with code
Astronomy
1 benchmark · 136 papers with code
Clinical Knowledge
1 benchmark · 54 papers with code
Econometrics
1 benchmark · 36 papers with code
Sociology
1 benchmark · 35 papers with code
Nutrition
1 benchmark · 33 papers with code
Parameter Prediction
1 benchmark · 16 papers with code
Logical Fallacies
1 benchmark · 14 papers with code
Multi-target regression
1 benchmark · 14 papers with code
Food recommendation
1 benchmark · 10 papers with code
Label Error Detection
1 benchmark · 8 papers with code
Virology
1 benchmark · 7 papers with code
Human Aging
1 benchmark · 6 papers with code
Jurisprudence
1 benchmark · 6 papers with code
Ancient Text Restoration
1 benchmark · 5 papers with code
Security Studies
1 benchmark · 5 papers with code
Social Media Popularity Prediction
1 benchmark · 4 papers with code
De novo molecule generation from MS/MS spectrum (bonus chemical formulae)
1 benchmark · 3 papers with code
Unconditional Crystal Generation
1 benchmark · 3 papers with code
Auto Debugging
1 benchmark · 2 papers with code
College Medicine
1 benchmark · 2 papers with code
De novo molecule generation from MS/MS spectrum
1 benchmark · 2 papers with code
Flood Inundation Mapping
1 benchmark · 2 papers with code
Human Organs Senses Multiple Choice
1 benchmark · 2 papers with code
Penn Machine Learning Benchmark
1 benchmark · 2 papers with code
Prehistory
1 benchmark · 2 papers with code
Professional Medicine
1 benchmark · 2 papers with code
Professional Psychology
1 benchmark · 2 papers with code
High School European History
1 benchmark · 1 paper with code
High School Geography
1 benchmark · 1 paper with code
High School Government and Politics
1 benchmark · 1 paper with code
High School Macroeconomics
1 benchmark · 1 paper with code
High School Microeconomics
1 benchmark · 1 paper with code
High School Psychology
1 benchmark · 1 paper with code
High School US History
1 benchmark · 1 paper with code
High School World History
1 benchmark · 1 paper with code
Human Sexuality
1 benchmark · 1 paper with code
International Law
1 benchmark · 1 paper with code
Log Solubility
1 benchmark · 1 paper with code
Molecule retrieval from MS/MS spectrum
1 benchmark · 1 paper with code
Molecule retrieval from MS/MS spectrum (bonus chemical formulae)
1 benchmark · 1 paper with code
MS/MS spectrum simulation
1 benchmark · 1 paper with code
MS/MS spectrum simulation (bonus chemical formulae)
1 benchmark · 1 paper with code
Professional Law
1 benchmark · 1 paper with code
Public Relations
1 benchmark · 1 paper with code
US Foreign Policy
1 benchmark · 1 paper with code
World Religions
1 benchmark · 1 paper with code
Making Hiring Decisions
1 benchmark · 0 papers with code
Diversity
0 benchmarks · 3,166 papers with code
Uncertainty Quantification
0 benchmarks · 832 papers with code
Learning-To-Rank
0 benchmarks · 213 papers with code
Survival Analysis
0 benchmarks · 197 papers with code
scientific discovery
0 benchmarks · 192 papers with code
Learning Theory
0 benchmarks · 140 papers with code
Prediction Intervals
0 benchmarks · 131 papers with code
Operator learning
0 benchmarks · 125 papers with code
Open Set Learning
0 benchmarks · 110 papers with code
Counterfactual Explanation
0 benchmarks · 86 papers with code
Fault Detection
0 benchmarks · 80 papers with code
imbalanced classification
0 benchmarks · 80 papers with code
Numerical Integration
0 benchmarks · 70 papers with code
Load Forecasting
0 benchmarks · 50 papers with code
Data Summarization
0 benchmarks · 34 papers with code
Model Discovery
0 benchmarks · 33 papers with code
Traffic Classification
0 benchmarks · 31 papers with code
X-ray Classification
0 benchmarks · 27 papers with code
Geophysics
0 benchmarks · 26 papers with code
Variational Monte Carlo
0 benchmarks · 24 papers with code
Problem Decomposition
0 benchmarks · 19 papers with code
Non-Intrusive Load Monitoring
0 benchmarks · 18 papers with code
Multilingual text classification
0 benchmarks · 16 papers with code
Seismic Imaging
0 benchmarks · 16 papers with code
Robust Design
0 benchmarks · 14 papers with code
Crime Prediction
0 benchmarks · 12 papers with code
Gender Bias Detection
0 benchmarks · 10 papers with code
Cryptanalysis
0 benchmarks · 8 papers with code
PDE Surrogate Modeling
0 benchmarks · 8 papers with code
Air Pollution Prediction
0 benchmarks · 7 papers with code
Photometric Redshift Estimation
0 benchmarks · 7 papers with code
Seismic Inversion
0 benchmarks · 7 papers with code
Radio Interferometry
0 benchmarks · 6 papers with code
Service Composition
0 benchmarks · 6 papers with code
Vector Quantization (k-means problem)
0 benchmarks · 5 papers with code
3D Bin Packing
0 benchmarks · 4 papers with code
Classification Of Variable Stars
0 benchmarks · 4 papers with code
Gravitational Wave Detection
0 benchmarks · 4 papers with code
Mobile Security
0 benchmarks · 4 papers with code
Music Genre Transfer
0 benchmarks · 4 papers with code
X-Ray Diffraction (XRD)
0 benchmarks · 4 papers with code
Air Quality Inference
0 benchmarks · 3 papers with code
Continued fraction
0 benchmarks · 3 papers with code
Dielectric Constant
0 benchmarks · 3 papers with code
Metadata quality
0 benchmarks · 3 papers with code
Sequential Correlation Estimation
0 benchmarks · 3 papers with code
Structured Output Generation
0 benchmarks · 3 papers with code
Zero-Shot Visual Question Answring
0 benchmarks · 3 papers with code
Automatic Cell Counting
0 benchmarks · 2 papers with code
Geometry-based operator learning
0 benchmarks · 2 papers with code
Hindu Knowledge
0 benchmarks · 2 papers with code
Local intrinsic dimension estimation
0 benchmarks · 2 papers with code
Self Adaptive System
0 benchmarks · 2 papers with code
Business Taxonomy Construction
0 benchmarks · 1 paper with code
Crowd Flows Prediction
0 benchmarks · 1 paper with code
Cyber Attack Investigation
0 benchmarks · 1 paper with code
Detect Ground Reflections
0 benchmarks · 1 paper with code
DFT Z isomer pi-pi* wavelength
0 benchmarks · 1 paper with code
Equilibrium Reaction Energy (ev/atom)
0 benchmarks · 1 paper with code
Modeling Local Geometric Structure
0 benchmarks · 1 paper with code
Network Congestion Control
0 benchmarks · 1 paper with code
Oceanic Eddy Classification
0 benchmarks · 1 paper with code
Outdoor Positioning
0 benchmarks · 1 paper with code
Pulsar Prediction
0 benchmarks · 1 paper with code
Sequential Distribution Function Estimation
0 benchmarks · 1 paper with code
Sequential Quantile Estimation
0 benchmarks · 1 paper with code
Surrogate Hydrodynamic Modeling
0 benchmarks · 1 paper with code
Time Offset Calibration
0 benchmarks · 1 paper with code
When should a hot water tank be replaced?
0 benchmarks · 1 paper with code
Wrong PDF attached
0 benchmarks · 1 paper with code
Home Activity Monitoring
0 benchmarks · 0 papers with code
JSONiq Query Execution
0 benchmarks · 0 papers with code
Link Quality Estimation
0 benchmarks · 0 papers with code
Multi-Modal Learning
0 benchmarks · 0 papers with code
Penn Machine Learning Benchmark (Real-World)
0 benchmarks · 0 papers with code
Zero-day intrusion detection
0 benchmarks · 0 papers with code
8 tasks in Miscellaneous are filed only under a parent task from another area and are not listed on this page; the parent's task page carries them.
Task tree and counts are the archive's, frozen 2025-07-28 archive 2025-07-28. Nothing here is re-ranked.