Browse State-of-the-Art › ARC
ARC
155 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 155 papers with code (554 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
3 Sep 2021 8 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedWe show that instruction tuning -- finetuning language models on a collection of tasks described via instructions -- substantially improves zero-shot performance on unseen tasks.
-
5 Nov 2019 6 repositories listed Syntology ran 0 of 11 samples · 11 unverifiedTo make deliberate progress towards more intelligent and more human-like artificial systems, we need to be following an appropriate feedback signal: we need to be able to define and evaluate intelligence in a way that…
-
13 Apr 2022 3 repositories listedDespite recent improvements in abstractive summarization, most current approaches generate summaries that are not factually consistent with the source document, severely restricting their trust and usage in real-world…
-
21 Mar 2022 3 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedChain-of-thought prompting combined with pre-trained large language models has achieved encouraging results on complex reasoning tasks.
-
17 Feb 2022 3 repositories listed Syntology ran 5 of 5 samples · 0 unverified · 5 pointer-only (licence)But advancing the state-of-the-art across a broad set of natural language tasks has been hindered by training instabilities and uncertain quality during fine-tuning.
-
8 Oct 2018 3 repositories listedOver the years many ellipse detection algorithms spring up and are studied broadly, while the critical issue of detecting ellipses accurately and efficiently in real-world images remains a challenge.
-
28 Aug 2018 3 repositories listedIn this paper we propose a retriever-reader model that learns to attend on essential terms during the question answering process.
-
3 Feb 2025 2 repositories listedOur results reveal that o-[n] series, particularly later iterations like o3 and o4-mini, significantly outperform the GPT-[n] series and show strong scalability in multimodal reasoning.
-
17 Nov 2024 2 repositories listedExcellent progress has been made recently in solving ARC Challenge problems.
-
1 May 2024 2 repositories listed Syntology ran 1 of 2 samples · 1 unverifiedWe introduce an approach aimed at enhancing the reasoning capabilities of Large Language Models (LLMs) through an iterative preference learning process inspired by the successful strategy employed by AlphaZero.
-
5 Feb 2024 2 repositories listedWe present the Perceptual Abstraction and Reasoning Language (PeARL) language, which allows DreamCoder to solve ARC tasks, and propose a new recognition model that allows us to significantly improve on the previous best…
-
15 Jun 2021 2 repositories listedWe present LARC, the \textit{Language-complete ARC}: a collection of natural language descriptions by a group of human participants who instruct each other on how to solve ARC tasks using language alone, which contains…
-
3 May 2020 2 repositories listed Syntology ran 2 of 3 samples · 1 unverifiedExperiments and analysis on 27 datasets from 13 languages clearly show that techniques developed before the DL era, such as structural learning (global TreeCRF loss) and high-order modeling are still useful, and can…
-
25 Sep 2019 2 repositories listedAdversarial training, which minimizes the maximal risk for label-preserving input perturbations, has proved to be effective for improving the generalization of language models.
-
16 Sep 2019 2 repositories listedWe consider a family of structural descriptors for visual data, namely covariance descriptors (CovDs) that lie on a non-linear symmetric positive definite (SPD) manifold, a special type of Riemannian manifolds.
-
23 Mar 2015 2 repositories listedAt its fastest, Yara can parse about 4000 sentences per second when in greedy mode (1 beam).
-
8 Jul 2025 1 repository listedThis paper presents our submission to Task 1, Subjectivity Detection, of the CheckThat!
-
8 Jul 2025 1 repository listedIn this paper, we, as the DS@GT team for CLEF 2025 CheckThat!
-
8 Jul 2025 1 repository listedNumerical claims, statements involving quantities, comparisons, and temporal references, pose unique challenges for automated fact-checking systems.
-
Tile-Based ViT Inference with Visual-Cluster Priors for Zero-Shot Multi-Species Plant Identification8 Jul 2025 1 repository listedWe describe DS@GT's second-place solution to the PlantCLEF 2025 challenge on multi-species plant identification in vegetation quadrat images.
-
31 May 2025 1 repository listedIn contrast, we exploit the capabilities of MLLMs to interpret non-textual instructions, specifically, adversarial images or audio generated by our novel method, Con Instruction.
-
30 May 2025 1 repository listed Syntology ran 0 of 7 samples · 7 unverifiedWe thus introduce HELM, a family of HypErbolic Large Language Models, offering a geometric rethinking of the Transformer-based LLM that addresses the representational inflexibility, missing set of necessary operations,…
-
14 May 2025 1 repository listedPrecise initialization plays a critical role in the performance of localization algorithms, especially in the context of robotics, autonomous driving, and computer vision.
-
13 May 2025 1 repository listedText-to-audio systems, while increasingly performant, are slow at inference time, thus making their latency unpractical for many creative applications.
-
11 May 2025 1 repository listedGiven a specified multidimensional k-space trajectory, the method optimizes traversal speed (and therefore timing) with position along the trajectory.
-
15 Apr 2025 1 repository listed Syntology ran 3 of 14 samples · 11 unverifiedBecause large language models are expensive to pretrain on different datasets, using smaller-scale experiments to decide on data is crucial for reducing costs.
-
14 Apr 2025 1 repository listedLanguage models rely on semantic priors to perform in-context learning, which leads to poor performance on tasks involving inductive reasoning.
-
27 Mar 2025 1 repository listedThe accurate determination of the beginning of each Hijri month is essential for religious, cultural, and administrative purposes.
-
23 Mar 2025 1 repository listedAdditionally, we propose PHT-CAD, a novel 2D PPA framework that harnesses the modality alignment and reasoning capabilities of Vision-Language Models (VLMs) for precise engineering drawing analysis.
-
19 Mar 2025 1 repository listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)We attribute these issues to the global, fully-connected MLP neural network architecture encoding of current INRs, which lack mechanisms for local representation: MLPs are sensitive to absolute image location and…
Syntology lines on 9 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections