Browse State-of-the-Art › World Knowledge
World Knowledge
358 papers with code · 0 benchmarks · 2 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 358 papers with code (818 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
29 May 2023 29 repositories listed Syntology ran 6 of 31 samples · 25 unverified · 2 pointer-only (licence)Existing methods for gaining such steerability collect human labels of the relative quality of model generations and fine-tune the unsupervised LM to align with these preferences, often with reinforcement learning from…
-
7 Sep 2020 18 repositories listed Syntology ran 5 of 26 samples · 21 unverified · 1 pointer-only (licence)By comprehensively evaluating the breadth and depth of a model's academic and professional understanding, our test can be used to analyze models across many tasks and to identify important shortcomings.
-
22 May 2020 18 repositories listed Syntology ran 4 of 6 samples · 2 unverifiedLarge pre-trained language models have been shown to store factual knowledge in their parameters, and achieve state-of-the-art results when fine-tuned on downstream NLP tasks.
-
10 Oct 2023 6 repositories listed Syntology ran 9 of 11 samples · 2 unverified · 1 pointer-only (licence)We introduce Mistral 7B v0.
-
10 Feb 2020 6 repositories listed Syntology ran 4 of 4 samples · 0 unverifiedLanguage model pre-training has been shown to capture a surprising amount of world knowledge, crucial for NLP tasks such as question answering.
-
10 Apr 2018 5 repositories listedImagining a scene described in natural language with realistic layout and appearance of entities is the ultimate test of spatial, visual, and semantic world knowledge.
-
28 Jun 2024 4 repositories listed Syntology ran 1 of 4 samples · 3 unverified · 1 pointer-only (licence)We propose a novel persona-driven data synthesis methodology that leverages various perspectives within a large language model (LLM) to create diverse synthetic data.
-
30 Sep 2022 4 repositories listed Syntology ran 6 of 7 samples · 1 unverified · 7 pointer-only (licence)Knowledge graph embedding aims to predict the missing relations between entities in knowledge graphs.
-
2 Nov 2018 4 repositories listed Syntology ran 0 of 3 samples · 3 unverifiedTo investigate question answering with prior knowledge, we present CommonsenseQA: a challenging new dataset for commonsense question answering.
-
12 Dec 2023 3 repositories listedVisual language models (VLMs) rapidly progressed with the recent success of large language models.
-
11 Dec 2023 3 repositories listedWe discover that the retrieval unit choice significantly impacts the performance of both retrieval and downstream tasks.
-
15 Nov 2023 3 repositories listedAlthough Large Language Models (LLMs) demonstrate remarkable ability in processing and generating human-like text, they do have limitations when it comes to comprehending and expressing world knowledge that extends…
-
5 Aug 2020 3 repositories listed Syntology ran 1 of 7 samples · 6 unverifiedWe show how to assess a language model's knowledge of basic concepts of morality.
-
1 May 2019 3 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedUnderstanding human's language requires complex world knowledge.
-
10 Mar 2025 2 repositories listed Syntology ran 5 of 5 samples · 0 unverified · 5 pointer-only (licence)Text-to-Image (T2I) models are capable of generating high-quality artistic creations and visual content.
-
22 Feb 2025 2 repositories listedHumor is prevalent in online communications and it often relies on more than one modality (e.
-
SeaLLMs 3: Open Foundation and Chat Multilingual Large Language Models for Southeast Asian Languages29 Jul 2024 2 repositories listed Syntology ran 1 of 2 samples · 1 unverified · 2 pointer-only (licence)Large Language Models (LLMs) have shown remarkable abilities across various tasks, yet their development has predominantly centered on high-resource languages like English and Chinese, leaving low-resource languages…
-
16 Jul 2024 2 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)In this paper, we introduce a new task, Reasoning Video Object Segmentation (ReasonVOS).
-
11 Jun 2024 2 repositories listed Syntology ran 5 of 6 samples · 1 unverifiedRecently, powerful Large Language Models (LLMs) have become easily accessible to hundreds of millions of users world-wide.
-
15 May 2024 2 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)We present Elements of World Knowledge (EWOK), a framework for evaluating world modeling in language models by testing their ability to use knowledge of a concept to match a target text with a plausible/implausible…
-
15 Apr 2024 2 repositories listedTo further augment the multi-modal representations, MyGO incorporates fine-grained contrastive learning to highlight the specificity of the entity representations.
-
11 Mar 2024 2 repositories listed Syntology ran 6 of 7 samples · 1 unverified · 7 pointer-only (licence)In this paper, we delve into the fine-tuning methods of LLMs and conduct extensive experiments to investigate the impact of fine-tuning methods for large models on the existing multimodal model in the medical domain…
-
12 Feb 2024 2 repositories listedWe tackle the challenge of building real-world multimodal assistants for complex real-world tasks.
-
14 Nov 2023 2 repositories listedWe present BYOKG, a universal question-answering (QA) system that can operate on any knowledge graph (KG), requires no human-annotated training data, and can be ready to use within a day -- attributes that are…
-
5 Oct 2023 2 repositories listedSpecifically, we introduce FreshQA, a novel dynamic QA benchmark encompassing a diverse range of question and answer types, including questions that require fast-changing world knowledge as well as questions with false…
-
23 Aug 2023 2 repositories listedWe introduce Topical-Chat, a knowledge-grounded human-human conversation dataset where the underlying knowledge spans 8 broad topics and conversation partners don't have explicitly defined roles, to help further…
-
20 Aug 2023 2 repositories listed Syntology ran 6 of 8 samples · 2 unverifiedThe recent surge in research interest in applying large language models (LLMs) to decision-making tasks has flourished by leveraging the extensive world knowledge embedded in LLMs.
-
1 Aug 2023 2 repositories listedIn this work, we propose a new segmentation task -- reasoning segmentation.
-
2 Apr 2023 2 repositories listedIn the research of end-to-end dialogue systems, using real-world knowledge to generate natural, fluent, and human-like utterances with correct answers is crucial.
-
22 Jun 2022 2 repositories listed Syntology ran 3 of 9 samples · 6 unverified · 3 pointer-only (licence)We present the Pathways Autoregressive Text-to-Image (Parti) model, which generates high-fidelity photorealistic images and supports content-rich synthesis involving complex compositions and world knowledge.
Syntology lines on 18 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections