Papers › Named Entity Recognition in Tweets: An Experimental Study

Named Entity Recognition in Tweets: An Experimental Study

1 Jul 2011Conference on Empirical Methods in Natural Language Processing 2011 7archive 2025-07-28

Alan Ritter, Sam Clark, Mausam Etzioni, Oren Etzioni

People tweet more than 100 Million times daily, yielding a noisy, informal, but sometimes informative corpus of 140-character messages that mirrors the zeitgeist in an unprecedented manner. The performance of standard NLP tools is severely degraded on tweets. This paper addresses this issue by re-building the NLP pipeline beginning with part-of-speech tagging, through chunking, to named-entity recognition. Our novel T-NER system doubles F1 score compared with the Stanford NER system. T-NER leverages the redundancy inherent in tweets to achieve this performance, using LabeledLDA to exploit Freebase dictionaries as a source of distant supervision. LabeledLDA outperforms cotraining, increasing F1 by 25% over ten common entity types. Our NLP tools are available at: http:// github.com/aritter/twitter_nlp

PaperPDFCode

Code

aritter/twitter_nlp mentioned in paper report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

ChunkingNERNamed Entity RecognitionNamed Entity Recognition (NER)Part-Of-Speech Taggingnamed-entity-recognition

Datasets

Introduced by this paper, per the archive.

Ritter PoS

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections