Papers › Neural Attentive Bag-of-Entities Model for Text Classification

Neural Attentive Bag-of-Entities Model for Text Classification

3 Sep 2019CONLL 2019 11arXiv:1909.01259archive 2025-07-28

Ikuya Yamada, Hiroyuki Shindo

This study proposes a Neural Attentive Bag-of-Entities model, which is a neural network model that performs text classification using entities in a knowledge base. Entities provide unambiguous and relevant semantic signals that are beneficial for capturing semantics in texts. We combine simple high-recall entity detection based on a dictionary, to detect entities in a document, with a novel neural attention mechanism that enables the model to focus on a small number of unambiguous and relevant entities. We tested the effectiveness of our model using two standard text classification datasets (i.e., the 20 Newsgroups and R8 datasets) and a popular factoid question answering dataset based on a trivia quiz game. As a result, our model achieved state-of-the-art results on all datasets. The source code of the proposed model is available online at https://github.com/wikipedia2vec/wikipedia2vec.

PaperPDFConference PDFCode

In Syntology View this paper on Syntology: its repositories, every harvested function with whether it ran, its licence and the call to fetch it.

Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

wikipedia2vec/wikipedia2vec officialmentioned in papermentioned on GitHubNOASSERTION report
studio-ousia/wikipedia2vec mentioned on GitHub report
wikipedia2vec/wikipedia2vec mentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

ClassificationGeneral ClassificationQuestion AnsweringText Classification

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Text Classification 20NEWS NABoE-full Accuracy 86.8 #9 of 16 Archive leaderboard report
Text Classification 20NEWS NABoE-full F-measure 86.2 #9 of 16 Archive leaderboard report
Text Classification R8 NABoE-full Accuracy 97.1 #16 of 21 Archive leaderboard report
Text Classification R8 NABoE-full F-measure 91.7 #16 of 21 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections