Browse State-of-the-Art › Benchmarking › Papers, page 47
Benchmarking
Papers archive 2025-07-28
archive papers tagged: 5,548 · with a code link: 2,658 · where Syntology ran a sample: 749 (624 with a run with no instrument failure, 125 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (749 of 5,548 tagged: 624 with a run with no instrument failure, 125 where every run was a failure of Syntology's instrument)
Page 47 of 56: papers 4,601 to 4,700 of 5,548, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Beyond Uniform Lipschitz Condition in Differentially Private Optimization21 Jun 2022 0 repositories listed
-
Design of Supervision-Scalable Learning Systems: Methodology and Performance Benchmarking18 Jun 2022 0 repositories listed
-
Colonoscopy 3D Video Dataset with Paired Depth from 2D-3D Registration17 Jun 2022 0 repositories listed
-
Benchmarking Heterogeneous Treatment Effect Models through the Lens of Interpretability16 Jun 2022 0 repositories listed
-
Characteristics of Harmful Text: Towards Rigorous Benchmarking of Language Models16 Jun 2022 0 repositories listed
-
BEHAVIOR in Habitat 2.0: Simulator-Independent Logical Task Description for Benchmarking Embodied AI Agents13 Jun 2022 0 repositories listed
-
SAIBench: Benchmarking AI for Science11 Jun 2022 0 repositories listed
-
Functional Code Building Genetic Programming9 Jun 2022 0 repositories listed
-
Benchmarking Bayesian neural networks and evaluation metrics for regression tasks8 Jun 2022 0 repositories listed
-
Scaling laws in global corporations as a benchmarking approach to assess environmental performance7 Jun 2022 0 repositories listed
-
MorisienMT: A Dataset for Mauritian Creole Machine Translation6 Jun 2022 0 repositories listed
-
Which models are innately best at uncertainty estimation?5 Jun 2022 0 repositories listed
-
A Semi-Automated Live Interlingual Communication Workflow Featuring Intralingual Respeaking: Evaluation and Benchmarking1 Jun 2022 0 repositories listed
-
Benchmarking Language Models for Cyberbullying Identification and Classification from Social-media Texts1 Jun 2022 0 repositories listed
-
Deep One-Class Hate Speech Detection Model1 Jun 2022 0 repositories listed
-
Evaluation of Three Welsh Language POS Taggers1 Jun 2022 0 repositories listed
-
Introducing RezoJDM16k: a French KnowledgeGraph DataSet for Link Prediction1 Jun 2022 0 repositories listed
-
Low-resource Neural Machine Translation: Benchmarking State-of-the-art Transformer for Wolof<->French1 Jun 2022 0 repositories listed
-
MTLens: Machine Translation Output Debugging1 Jun 2022 0 repositories listed
-
Hide and Seek: on the Stealthiness of Attacks against Deep Learning Systems31 May 2022 0 repositories listed
-
NEWTS: A Corpus for News Topic-Focused Summarization31 May 2022 0 repositories listed
-
Benchmarking Unsupervised Anomaly Detection and Localization30 May 2022 0 repositories listed
-
Benchmarking of Deep Learning models on 2D Laminar Flow behind Cylinder26 May 2022 0 repositories listed
-
Large Language Models are Few-Shot Clinical Information Extractors25 May 2022 0 repositories listed
-
Advanced Manufacturing Configuration by Sample-efficient Batch Bayesian Optimization24 May 2022 0 repositories listed
-
RCC-GAN: Regularized Compound Conditional GAN for Large-Scale Tabular Data Synthesis24 May 2022 0 repositories listed
-
Generalization, Mayhems and Limits in Recurrent Proximal Policy Optimization23 May 2022 0 repositories listed
-
Paddy Doctor: A Visual Image Dataset for Automated Paddy Disease Classification and Benchmarking23 May 2022 0 repositories listed
-
Deep Learning vs. Gradient Boosting: Benchmarking state-of-the-art machine learning algorithms for credit scoring21 May 2022 0 repositories listed
-
Self-Supervised Speech Representation Learning: A Review21 May 2022 0 repositories listed
-
Entity Alignment For Knowledge Graphs: Progress, Challenges, and Empirical Studies18 May 2022 0 repositories listed
-
Accented Speech Recognition: Benchmarking, Pre-training, and Diverse Data16 May 2022 0 repositories listed
-
Uncertainty estimation for Cross-dataset performance in Trajectory prediction15 May 2022 0 repositories listed
-
Provably Safe Reinforcement Learning: Conceptual Analysis, Survey, and Benchmarking13 May 2022 0 repositories listed
-
Beyond Static Models and Test Sets: Benchmarking the Potential of Pre-trained Models Across Tasks and Languages12 May 2022 0 repositories listed
-
Subspace Learning Machine (SLM): Methodology and Performance11 May 2022 0 repositories listed
-
LayoutXLM vs. GNN: An Empirical Evaluation of Relation Extraction for Documents9 May 2022 0 repositories listed
-
Design Target Achievement Index: A Differentiable Metric to Enhance Deep Generative Models in Multi-Objective Inverse Design6 May 2022 0 repositories listed
-
VFHQ: A High-Quality Dataset and Benchmark for Video Face Super-Resolution6 May 2022 0 repositories listed
-
Learn-to-Race Challenge 2022: Benchmarking Safe Learning and Cross-domain Generalisation in Autonomous Racing5 May 2022 0 repositories listed
-
Surface Reconstruction from Point Clouds: A Survey and a Benchmark5 May 2022 0 repositories listed
-
On Continual Model Refinement in Out-of-Distribution Data Streams4 May 2022 0 repositories listed
-
Training Mixed-Domain Translation Models via Federated Learning3 May 2022 0 repositories listed
-
Fantastic Questions and Where to Find Them: FairytaleQA – An Authentic Dataset for Narrative Comprehension1 May 2022 0 repositories listed
-
Foundations for learning from noisy quantum experiments28 Apr 2022 0 repositories listed
-
Causal Reasoning Meets Visual Representation Learning: A Prospective Study26 Apr 2022 0 repositories listed
-
Deeper Insights into the Robustness of ViTs towards Common Corruptions26 Apr 2022 0 repositories listed
-
Label Anchored Contrastive Learning for Language Understanding26 Apr 2022 0 repositories listed
-
Benchmarking Answer Verification Methods for Question Answering-Based Summarization Evaluation Metrics21 Apr 2022 0 repositories listed
-
Learning to Fold Real Garments with One Arm: A Case Study in Cloud-Based Robotics Research21 Apr 2022 0 repositories listed
-
Analyzing the Impact of Undersampling on the Benchmarking and Configuration of Evolutionary Algorithms20 Apr 2022 0 repositories listed
-
Multi-label classification for biomedical literature: an overview of the BioCreative VII LitCovid Track for COVID-19 literature topic annotations20 Apr 2022 0 repositories listed
-
Label Efficient Regularization and Propagation for Graph Node Classification19 Apr 2022 0 repositories listed
-
Benchmarking Domain Generalization on EEG-based Emotion Recognition18 Apr 2022 0 repositories listed
-
From Environmental Sound Representation to Robustness of 2D CNN Models Against Adversarial Attacks14 Apr 2022 0 repositories listed
-
SoccerNet-Tracking: Multiple Object Tracking Dataset and Benchmark in Soccer Videos14 Apr 2022 0 repositories listed
-
Benchmarking Active Learning Strategies for Materials Optimization and Discovery12 Apr 2022 0 repositories listed
-
EVOPS Benchmark: Evaluation of Plane Segmentation from RGBD and LiDAR Data12 Apr 2022 0 repositories listed
-
Metaethical Perspectives on 'Benchmarking' AI Ethics11 Apr 2022 0 repositories listed
-
Benchmarking for Public Health Surveillance tasks on Social Media with a Domain-Specific Pretrained Language Model9 Apr 2022 0 repositories listed
-
Disability prediction in multiple sclerosis using performance outcome measures and demographic data8 Apr 2022 0 repositories listed
-
tmVar 3.0: an improved variant concept recognition and normalization tool7 Apr 2022 0 repositories listed
-
A Comparison of Deep Learning MOS Predictors for Speech Synthesis Quality5 Apr 2022 0 repositories listed
-
A lightweight and accurate YOLO-like network for small target detection in Aerial Imagery5 Apr 2022 0 repositories listed
-
Intelligence at the Extreme Edge: A Survey on Reformable TinyML2 Apr 2022 0 repositories listed
-
1 Apr 2022 0 repositories listed
-
Assessing the risk of re-identification arising from an attack on anonymised data31 Mar 2022 0 repositories listed
-
Is Word Error Rate a good evaluation metric for Speech Recognition in Indic Languages?30 Mar 2022 0 repositories listed
-
Treatment Learning Causal Transformer for Noisy Image Classification29 Mar 2022 0 repositories listed
-
A Unified Study of Machine Learning Explanation Evaluation Metrics27 Mar 2022 0 repositories listed
-
Benchmarking Algorithms for Automatic License Plate Recognition27 Mar 2022 0 repositories listed
-
Benchmarking Deep AUROC Optimization: Loss Functions and Algorithmic Choices27 Mar 2022 0 repositories listed
-
LAMBDA: Covering the Solution Set of Black-Box Inequality by Search Space Quantization25 Mar 2022 0 repositories listed
-
Comprehensive Benchmark Datasets for Amharic Scene Text Detection and Recognition23 Mar 2022 0 repositories listed
-
A Perspective on Neural Capacity Estimation: Viability and Reliability22 Mar 2022 0 repositories listed
-
Benchmarking Test-Time Unsupervised Deep Neural Network Adaptation on Edge Devices21 Mar 2022 0 repositories listed
-
Policy Gradients using Variational Quantum Circuits20 Mar 2022 0 repositories listed
-
A Statistical Framework to Investigate the Optimality of Signal-Reconstruction Methods18 Mar 2022 0 repositories listed
-
Fiber Bundle Morphisms as a Framework for Modeling Many-to-Many Maps15 Mar 2022 0 repositories listed
-
From 2D to 3D: Re-thinking Benchmarking of Monocular Depth Prediction15 Mar 2022 0 repositories listed
-
DFTR: Depth-supervised Fusion Transformer for Salient Object Detection12 Mar 2022 0 repositories listed
-
A Closer Look at Debiased Temporal Sentence Grounding in Videos: Dataset, Metric, and Approach10 Mar 2022 0 repositories listed
-
IndicNLG Benchmark: Multilingual Datasets for Diverse NLG Tasks in Indic Languages10 Mar 2022 0 repositories listed
-
Mapping global dynamics of benchmark creation and saturation in artificial intelligence9 Mar 2022 0 repositories listed
-
Metastatic Cancer Outcome Prediction with Injective Multiple Instance Pooling9 Mar 2022 0 repositories listed
-
Score-Based Generative Models for Molecule Generation7 Mar 2022 0 repositories listed
-
Systematic Comparison of Path Planning Algorithms using PathBench7 Mar 2022 0 repositories listed
-
Automated Machine Learning: A Case Study on Non-Intrusive Appliance Load Monitoring6 Mar 2022 0 repositories listed
-
Multi-channel deep convolutional neural networks for multi-classifying thyroid disease6 Mar 2022 0 repositories listed
-
Benchmarking real-time algorithms for in-phase auditory stimulation of low amplitude slow waves with wearable EEG devices during sleep4 Mar 2022 0 repositories listed
-
Graph clustering with Boltzmann machines4 Mar 2022 0 repositories listed
-
Towards Benchmarking and Evaluating Deepfake Detection4 Mar 2022 0 repositories listed
-
Adaptive Gradient Methods with Local Guarantees2 Mar 2022 0 repositories listed
-
Benchmarking Robustness of Deep Learning Classifiers Using Two-Factor Perturbation2 Mar 2022 0 repositories listed
-
Reliable validation of Reinforcement Learning Benchmarks2 Mar 2022 0 repositories listed
-
Prepare for Trouble and Make it Double. Supervised and Unsupervised Stacking for AnomalyBased Intrusion Detection28 Feb 2022 0 repositories listed
-
Towards Class-agnostic Tracking Using Feature Decorrelation in Point Clouds28 Feb 2022 0 repositories listed
-
Generalised Gaussian Process Latent Variable Models (GPLVM) with Stochastic Variational Inference25 Feb 2022 0 repositories listed
-
Spatio-Temporal Latent Graph Structure Learning for Traffic Forecasting25 Feb 2022 0 repositories listed
-
Measuring CLEVRness: Blackbox testing of Visual Reasoning Models24 Feb 2022 0 repositories listed