Browse State-of-the-Art › Sentence Completion
Sentence Completion
49 papers with code · 1 benchmark · 2 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| HellaSwag (89 rows) | CompassMTL 567M with Tailor | Task Compass: Scaling Multi-task Pre-training with Task Prefix | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
1 subtask in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 49 papers with code (91 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
28 May 2020 67 repositories listed Syntology ran 15 of 65 samples · 50 unverified · 4 pointer-only (licence)By contrast, humans can generally perform a new language task from only a few examples or from simple instructions - something which current NLP systems still largely struggle to do.
-
26 Jul 2019 67 repositories listed Syntology ran 22 of 48 samples · 26 unverified · 23 pointer-only (licence)Language model pretraining has led to significant performance gains but careful comparison between different approaches is challenging.
-
27 Feb 2023 57 repositories listed Syntology ran 26 of 58 samples · 32 unverified · 4 pointer-only (licence)We introduce LLaMA, a collection of foundation language models ranging from 7B to 65B parameters.
-
1 Dec 2023 35 repositories listed Syntology ran 18 of 62 samples · 44 unverified · 28 pointer-only (licence)Foundation models, now powering most of the exciting applications in deep learning, are almost universally based on the Transformer architecture and its core attention module.
-
18 Jul 2023 19 repositories listed Syntology ran 31 of 52 samples · 21 unverified · 16 pointer-only (licence)In this work, we develop and release Llama 2, a collection of pretrained and fine-tuned large language models (LLMs) ranging in scale from 7 billion to 70 billion parameters.
-
5 Jun 2020 14 repositories listed Syntology ran 4 of 13 samples · 9 unverified · 3 pointer-only (licence)Recent progress in pre-trained neural language models has significantly improved the performance of many natural language processing (NLP) tasks.
-
15 Mar 2023 11 repositories listed Syntology ran 2 of 5 samples · 3 unverified · 1 pointer-only (licence)We report the development of GPT-4, a large-scale, multimodal model which can accept image and text inputs and produce text outputs.
-
3 Sep 2021 8 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedWe show that instruction tuning -- finetuning language models on a collection of tasks described via instructions -- substantially improves zero-shot performance on unseen tasks.
-
5 Apr 2022 7 repositories listed Syntology ran 30 of 37 samples · 7 unverifiedTo further our understanding of the impact of scale on few-shot learning, we trained a 540-billion parameter, densely activated, Transformer language model, which we call Pathways Language Model PaLM.
-
10 Oct 2023 6 repositories listed Syntology ran 9 of 11 samples · 2 unverified · 1 pointer-only (licence)We introduce Mistral 7B v0.
-
9 Jun 2022 5 repositories listed Syntology ran 2 of 11 samples · 9 unverified · 4 pointer-only (licence)In this work, we measure and improve the factual accuracy of large-scale LMs for open-ended text generation.
-
8 Dec 2021 3 repositories listedLanguage modelling provides a step towards intelligent communication systems by harnessing large repositories of written human knowledge to better predict and understand the world.
-
22 Apr 2024 2 repositories listed Syntology ran 6 of 11 samples · 5 unverifiedWe also propose a new high-throughput framework to alleviate the computation and memory bottlenecks during the training and inference of MOE models.
-
5 Jan 2024 2 repositories listedUsing PESC during instruction tuning, our best sparse model outperforms other sparse and dense models and exhibits superior general capabilities compared to GPT-3.
-
10 Oct 2023 2 repositories listed Syntology ran 3 of 3 samples · 0 unverifiedIn this work, we study structured pruning as an effective means to develop smaller LLMs from pre-trained, larger models.
-
16 Sep 2023 2 repositories listedLLMs are increasingly powerful and widely used to assist users in a variety of tasks.
-
23 May 2023 2 repositories listedFurthermore, we show that instruction tuning with CoT Collection allows LMs to possess stronger few-shot learning capabilities on 4 domain-specific tasks, resulting in an improvement of +2.
-
30 Mar 2023 2 repositories listedThe use of NLP in the realm of financial technology is broad and complex, with applications ranging from sentiment analysis and named entity recognition to question answering.
-
7 Feb 2023 2 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)Recently, Language Models (LMs) instruction-tuned on multiple tasks, also known as multitask-prompted fine-tuning (MT), have shown the capability to generalize to unseen tasks.
-
29 Mar 2022 2 repositories listed Syntology ran 8 of 11 samples · 3 unverified · 4 pointer-only (licence)We investigate the optimal model size and number of tokens for training a transformer language model under a given compute budget.
-
28 Jan 2022 2 repositories listedNext, we detail the training process, the design of our training corpus, and our data curation techniques, which we believe is a key ingredient to the success of the model.
-
26 Jan 2021 2 repositories listedWe propose pre-finetuning, an additional large-scale learning stage between language model pre-training and fine-tuning.
-
19 May 2019 2 repositories listed Syntology ran 0 of 6 samples · 6 unverified · 4 pointer-only (licence)In this paper, we show that commonsense inference still proves difficult for even state-of-the-art models, by presenting HellaSwag, a new challenge dataset.
-
8 Apr 2019 2 repositories listedTo produce a more difficult dataset, we introduce a novel procedure for question acquisition in which workers author questions designed to target weaknesses of state-of-the-art neural question answering systems.
-
6 Jan 2016 2 repositories listedIn this paper, we propose Recurrent Memory Network (RMN), a novel RNN architecture, that not only amplifies the power of RNN but also facilitates our understanding of its internal functioning and allows us to discover…
-
21 Oct 2024 1 repository listedEffective communication within universities is crucial for addressing the diverse information needs of students, alumni, and external stakeholders.
-
16 Jun 2024 1 repository listed Syntology ran 4 of 6 samples · 2 unverified · 6 pointer-only (licence)In this paper, we introduce a subspace-inspired Low-Rank Adaptation (LoRA) method, which is computationally efficient, easy to implement, and readily applicable to large language, multimodal, and diffusion models.
-
9 Feb 2024 1 repository listed Syntology ran 3 of 5 samples · 2 unverified · 5 pointer-only (licence)Controlled text generation (CTG) seeks to guide large language model (LLM) output to produce text that conforms to desired criteria.
-
5 Nov 2023 1 repository listedWe present mahaNLP, an open-source natural language processing (NLP) library specifically built for the Marathi language.
-
30 Oct 2023 1 repository listedAn essential task for tourists having a pleasant holiday is to have a well-planned itinerary with relevant recommendations, especially when visiting unfamiliar cities.
Syntology lines on 18 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections