Browse State-of-the-Art › General Knowledge
General Knowledge
173 papers with code · 1 benchmark · 2 datasets archive 2025-07-28
This task aims to evaluate the ability of a model to answer general-knowledge questions.
Source: BIG-bench
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| BIG-bench (2 rows) | Chinchilla-70B (few-shot, k=5) | Training Compute-Optimal Large Language Models | code | Syntology ran 8 of 11 samples · 3 unverified | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
7 subtasks in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 173 papers with code (399 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
29 Jul 2019 8 repositories listed Syntology ran 3 of 17 samples · 14 unverified · 1 pointer-only (licence)We present Joey NMT, a minimalist neural machine translation toolkit based on PyTorch that is specifically designed for novices.
-
12 Dec 2016 6 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)It is designed to represent the general knowledge involved in understanding language, improving natural language applications by allowing the application to better understand the meanings behind the words people use.
-
15 Feb 2017 4 repositories listedAs one of the fundamental tasks in text analysis, phrase mining aims at extracting quality phrases from a text corpus.
-
23 Dec 2023 3 repositories listedChange detection, a prominent research area in remote sensing, is pivotal in observing and analyzing surface transformations.
-
31 May 2023 3 repositories listed Syntology ran 4 of 7 samples · 3 unverified · 1 pointer-only (licence)With the advent of powerful large language models (LLMs) such as GPT or Llama2, which demonstrate an ability to reason and to utilize general knowledge, there is a growing need for techniques which combine the textual…
-
8 Dec 2021 3 repositories listedLanguage modelling provides a step towards intelligent communication systems by harnessing large repositories of written human knowledge to better predict and understand the world.
-
2 Apr 2015 3 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedDistributional models that learn rich semantic word representations are a success story of recent NLP research.
-
22 Dec 2024 2 repositories listedThis survey examines the state of the art in text summarization models, with a specific focus on the abstractive summarization approach.
-
30 Oct 2024 2 repositories listedRecent advances in LLM have been instrumental in autonomous robot control and human-robot interaction by leveraging their vast general knowledge and capabilities to understand and reason across a wide range of tasks and…
-
26 Sep 2024 2 repositories listedPrompt learning has surfaced as an effective approach to enhance the performance of Vision-Language Models (VLMs) like CLIP when applied to downstream tasks.
-
9 Jun 2024 2 repositories listedTo address this issue, we present F-LMM -- grounding frozen off-the-shelf LMMs in human-AI conversations -- a straightforward yet effective design based on the fact that word-pixel correspondences conducive to visual…
-
9 Jun 2024 2 repositories listed Syntology ran 12 of 12 samples · 0 unverified · 12 pointer-only (licence)We evaluated popular LLMs such as Llama, Baichuan, ChatGLM, and GPT models.
-
2 Dec 2023 2 repositories listedChange detection (CD) is a critical task to observe and analyze dynamic processes of land cover.
-
7 Jul 2023 2 repositories listedThe most popular pipeline for learning on graphs with textual node attributes primarily relies on Graph Neural Networks (GNNs), and utilizes shallow text embedding as initial node representations, which has limitations…
-
24 May 2023 2 repositories listedTo mitigate potential negative transfer, we separate the item representations into market embeddings and item embeddings.
-
7 Feb 2023 2 repositories listedA novel proxy is also proposed to preserve the general knowledge in the original LM.
-
21 Jan 2023 2 repositories listedThis paper shows that the existing methods are suboptimal and proposes a novel method to perform a more informed adaptation of the knowledge in the LM by (1) soft-masking the attention heads based on their importance to…
-
15 Nov 2022 2 repositories listedIn this paper, we focus on the compression of DETR with knowledge distillation.
-
28 Jun 2022 2 repositories listedSolving Chinese character riddles is a challenging task that demands understanding of character glyph, general knowledge, and a grasp of figurative language.
-
29 Mar 2022 2 repositories listed Syntology ran 8 of 11 samples · 3 unverified · 4 pointer-only (licence)We investigate the optimal model size and number of tokens for training a transformer language model under a given compute budget.
-
18 Oct 2021 2 repositories listedThere is currently no simple, unified way to compare, analyse or evaluate metrics across a representative set of tasks.
-
18 May 2021 2 repositories listedBased on our previous MetaAdapter that implicitly leverages adapters, we propose a novel algorithms called SimAdapter for explicitly learning knowledge from adapters.
-
14 Feb 2020 2 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedHowever, expressing the knowledge in a formal (logical or probabilistic) representation has been a major obstacle to this research.
-
31 Dec 2019 2 repositories listedOpen-domain question answering (QA) is known to involve several underlying knowledge and reasoning challenges, but are models actually learning such knowledge when trained on benchmark tasks?
-
22 Nov 2019 2 repositories listedThe key challenge of multi-domain translation lies in simultaneously encoding both the general knowledge shared across domains and the particular knowledge distinctive to each domain in a unified model.
-
29 Mar 2019 2 repositories listed Syntology ran 2 of 3 samples · 1 unverified · 3 pointer-only (licence)Insufficient or even unavailable training data of emerging classes is a big challenge of many classification tasks, including text classification.
-
1 Mar 2018 2 repositories listedThis paper describes our system for SemEval-2018 Task 11: Machine Comprehension using Commonsense Knowledge.
-
11 Apr 2017 2 repositories listedThis paper describes Luminoso's participation in SemEval 2017 Task 2, "Multilingual and Cross-lingual Semantic Word Similarity", with a system based on ConceptNet.
-
26 Jun 2025 1 repository listed Syntology ran 3 of 12 samples · 9 unverified · 12 pointer-only (licence)Causal reasoning capability is critical in advancing large language models (LLMs) toward strong artificial intelligence.
-
12 Jun 2025 1 repository listedTo address these gaps, we propose TaxoAdapt, a framework that dynamically adapts an LLM-generated taxonomy to a given corpus across multiple dimensions.
Syntology lines on 9 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections