Browse State-of-the-Art › Diagnostic
Diagnostic
1,213 papers with code · 0 benchmarks · 1 dataset archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 1,213 papers with code (4,513 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
11 Sep 2019 14 repositories listedIn this paper, we challenge the necessity of such hard/soft sampling methods for training accurate deep object detectors.
-
28 Mar 2018 13 repositories listed Syntology ran 0 of 4 samples · 4 unverifiedTraining of neural networks for automated diagnosis of pigmented skin lesions is hampered by the small size and lack of diversity of available datasets of dermatoscopic images.
-
20 Apr 2018 11 repositories listed Syntology ran 5 of 14 samples · 9 unverified · 1 pointer-only (licence)For natural language understanding (NLU) technology to be maximally useful, both practically and as a scientific object of study, it must be general: it must be able to process language in a way that is not exclusively…
-
23 Oct 2023 9 repositories listed Syntology ran 2 of 8 samples · 6 unverified · 1 pointer-only (licence)Our comprehensive case studies within HallusionBench shed light on the challenges of hallucination and illusion in LVLMs.
-
6 Feb 2020 9 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)This large scale study focuses on quantifying what X-rays diagnostic prediction tasks generalize well across multiple different datasets.
-
9 Jul 2015 9 repositories listed Syntology ran 2 of 3 samples · 1 unverified · 3 pointer-only (licence)Importance weighting is a general way to adjust Monte Carlo integration to account for draws from the wrong distribution, but the resulting estimate can be highly variable when the importance ratios have a heavy right…
-
2 Jan 2020 6 repositories listed Syntology ran 6 of 20 samples · 14 unverifiedA variety of algorithms search architectures under different search space.
-
21 Nov 2017 6 repositories listedWe achieve state of the art results on the bAbI textual question-answering dataset with the recurrent relational network, consistently solving 20/20 tasks.
-
25 Apr 2020 5 repositories listedFor detecting COVID-19 in particular, the model performs with a sensitivity of 0.
-
30 Aug 2019 5 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedWe provide sample code in Python and R as well as examples of applications to photometric redshift estimation and likelihood-free cosmological inference via CDE.
-
16 Aug 2019 5 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)The recent success of natural language understanding (NLU) systems has been troubled by results highlighting the failure of these models to generalize in a systematic and robust way.
-
20 Dec 2016 5 repositories listed Syntology ran 2 of 5 samples · 3 unverified · 4 pointer-only (licence)When building artificial intelligence systems that can reason and answer questions about visual data, we need diagnostic tests to analyze our progress and discover shortcomings.
-
17 Feb 2023 4 repositories listed Syntology ran 0 of 11 samples · 11 unverifiedThis work proposes ``jointly amortized neural approximation'' (JANA) of intractable likelihood functions and posterior densities arising in Bayesian surrogate modeling and simulation-based inference.
-
9 Sep 2021 4 repositories listed Syntology ran 0 of 7 samples · 7 unverifiedModels that have learned to construct cross-modal representations using both modalities are expected to perform worse when inputs are missing from a modality.
-
28 Aug 2020 4 repositories listed Syntology ran 4 of 4 samples · 0 unverified · 2 pointer-only (licence)In this paper, we propose NATS-Bench, a unified benchmark on searching for both topology and size, for (almost) any up-to-date NAS algorithm.
-
10 Jun 2020 4 repositories listedThis work opens the door to further investigation of how automatically analysed respiratory patterns could be used as pre-screening signals to aid COVID-19 diagnosis.
-
10 Jul 2019 4 repositories listedRetinal image quality assessment (RIQA) is essential for controlling the quality of retinal imaging and guaranteeing the reliability of diagnoses by ophthalmologists or automated analysis systems.
-
19 Sep 2018 4 repositories listedWith modern software it is easy to train even a~complex model that fits the training data and results in high accuracy on the test set.
-
28 Nov 2024 3 repositories listed Syntology ran 2 of 17 samples · 15 unverifiedWe present DiffVox, a self-supervised framework for Cone-Beam Computed Tomography (CBCT) reconstruction by directly optimizing a voxelgrid representation using physics-based differentiable X-ray rendering.
-
24 Jun 2024 3 repositories listedThis paper introduces a novel, entity-aware metric, termed as Radiological Report (Text) Evaluation (RaTEScore), to assess the quality of medical reports generated by AI models.
-
20 Dec 2023 3 repositories listedConsequently, we refine the cognitive states of cold-start students as diagnostic outcomes via virtual data, aligning with the diagnosis-oriented goal.
-
28 Nov 2023 3 repositories listed Syntology ran 7 of 10 samples · 3 unverifiedWith the rapid development of Multi-modal Large Language Models (MLLMs), a number of diagnostic benchmarks have recently emerged to evaluate the comprehension capabilities of these models.
-
BHASA: A Holistic Southeast Asian Linguistic and Cultural Evaluation Suite for Large Language Models12 Sep 2023 3 repositories listedAs GPT-4 is purportedly one of the best-performing multilingual LLMs at the moment, we use it as a yardstick to gauge the capabilities of LLMs in the context of SEA languages.
-
5 Sep 2023 3 repositories listedAnother strategy to increase the size of a dataset is crowdsourcing, a widely adopted practice in general computer vision with some success in medical image analysis.
-
24 Apr 2023 3 repositories listed Syntology ran 6 of 14 samples · 8 unverified · 5 pointer-only (licence)Medical image segmentation is a critical component in clinical practice, facilitating accurate diagnosis, treatment planning, and disease monitoring.
-
19 May 2022 3 repositories listed Syntology ran 5 of 28 samples · 23 unverifiedThese results suggest that REMEDIS can significantly accelerate the life-cycle of medical imaging AI development thereby presenting an important step forward for medical imaging AI to deliver broad impact.
-
2 Apr 2021 3 repositories listedReal-time diagnostics of complex technical systems such as power plants are critical to keep the system in its working state.
-
31 Dec 2020 3 repositories listedDetecting online hate is a difficult task that even state-of-the-art models struggle with.
-
28 Sep 2020 3 repositories listedNovel Coronavirus (COVID-19) has drastically overwhelmed more than 200 countries affecting millions and claiming almost 1 million lives, since its emergence in late 2019.
-
5 Aug 2020 3 repositories listedA common task in computational text analyses is to quantify how two corpora differ according to a measurement like word frequency, sentiment, or information content.
Syntology lines on 16 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections