Browse State-of-the-Art › Lemmatization › Papers, page 3
Lemmatization
Papers archive 2025-07-28
archive papers tagged: 351 · with a code link: 68 · where Syntology ran a sample: 3 (2 with a run with no instrument failure, 1 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3 of 351 tagged: 2 with a run with no instrument failure, 1 where every run was a failure of Syntology's instrument)
Page 3 of 4: papers 201 to 300 of 351, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
KeyXtract Twitter Model - An Essential Keywords Extraction Model for Twitter Designed using NLP Tools9 Aug 2017 0 repositories listed
-
DT_Team at SemEval-2017 Task 1: Semantic Similarity Using Alignments, Sentence-Level Embeddings and Gaussian Mixture Model Output1 Aug 2017 0 repositories listed
-
ECNU at SemEval-2017 Task 4: Evaluating Effective Features on Machine Learning Methods for Twitter Message Polarity Classification1 Aug 2017 0 repositories listed
-
LABDA at SemEval-2017 Task 10: Relation Classification between keyphrases via Convolutional Neural Network1 Aug 2017 0 repositories listed
-
Lexical Correction of Polish Twitter Political Data1 Aug 2017 0 repositories listed
-
1 Aug 2017 0 repositories listed
-
QLUT at SemEval-2017 Task 1: Semantic Textual Similarity Based on Word Embeddings1 Aug 2017 0 repositories listed
-
RACAI's Natural Language Processing pipeline for Universal Dependencies1 Aug 2017 0 repositories listed
-
Tokenizing, POS Tagging, Lemmatizing and Parsing UD 2.0 with UDPipe1 Aug 2017 0 repositories listed
-
Context Sensitive Lemmatization Using Two Successive Bidirectional Gated Recurrent Networks1 Jul 2017 0 repositories listed
-
SU-RUG at the CoNLL-SIGMORPHON 2017 shared task: Morphological Inflection with Attentional Sequence-to-Sequence Models12 Jun 2017 0 repositories listed
-
Synergistic Union of Word2Vec and Lexicon for Domain Specific Semantic Similarity6 Jun 2017 0 repositories listed
-
Exploring Properties of Intralingual and Interlingual Association Measures Visually1 May 2017 0 repositories listed
-
Multilingwis\mbox² -- Explore Your Parallel Corpus1 May 2017 0 repositories listed
-
Services for text simplification and analysis1 May 2017 0 repositories listed
-
A data-driven approach to verbal multiword expression detection. PARSEME Shared Task system description paper1 Apr 2017 0 repositories listed
-
Adapting a State-of-the-Art Tagger for South Slavic Languages to Non-Standard Text1 Apr 2017 0 repositories listed
-
Gender Profiling for Slovene Twitter communication: the Influence of Gender Marking, Content and Style1 Apr 2017 0 repositories listed
-
Spelling Correction for Morphologically Rich Language: a Case Study of Russian1 Apr 2017 0 repositories listed
-
The First Cross-Lingual Challenge on Recognition, Normalization, and Matching of Named Entities in Slavic Languages1 Apr 2017 0 repositories listed
-
Distributional regularities of verbs and verbal adjectives: Treebank evidence and broader implications1 Jan 2017 0 repositories listed
-
Acquisition of semantic relations between terms: how far can we get with standard NLP tools?1 Dec 2016 0 repositories listed
-
Automatic Translation of English Text to Indian Sign Language Synthetic Animations1 Dec 2016 0 repositories listed
-
ENIAM: Categorial Syntactic-Semantic Parser for Polish1 Dec 2016 0 repositories listed
-
Improving Neural Translation Models with Linguistic Factors1 Dec 2016 0 repositories listed
-
Improving the Morphological Analysis of Classical Sanskrit1 Dec 2016 0 repositories listed
-
The impact of simple feature engineering in multilingual medical NER1 Dec 2016 0 repositories listed
-
The Power of Language Music: Arabic Lemmatization through Patterns1 Dec 2016 0 repositories listed
-
YAMAMA: Yet Another Multi-Dialect Arabic Morphological Analyzer1 Dec 2016 0 repositories listed
-
LAMB: A Good Shepherd of Morphologically Rich Languages1 Nov 2016 0 repositories listed
-
Towards error annotation in a learner corpus of Portuguese1 Nov 2016 0 repositories listed
-
Still not there? Comparing Traditional Sequence-to-Sequence Models to Encoder-Decoder Neural Networks on Monotone String Translation Tasks25 Oct 2016 0 repositories listed
-
Authorship Attribution Based on Life-Like Network Automata20 Oct 2016 0 repositories listed
-
An Analysis of Lemmatization on Topic Models of Morphologically Rich Language13 Aug 2016 0 repositories listed
-
An NLP Pipeline for Coptic1 Aug 2016 0 repositories listed
-
Dealing with word-internal modification and spelling variation in data-driven lemmatization1 Aug 2016 0 repositories listed
-
English-French Document Alignment Based on Keywords and Statistical Translation1 Aug 2016 0 repositories listed
-
Leveraging Inflection Tables for Stemming and Lemmatization.1 Aug 2016 0 repositories listed
-
Morphological Reinflection via Discriminative String Transduction1 Aug 2016 0 repositories listed
-
NRC Russian-English Machine Translation System for WMT 20161 Aug 2016 0 repositories listed
-
Predicting the Compositionality of Nominal Compounds: Giving Word Embeddings a Hard Time1 Aug 2016 0 repositories listed
-
The Kyoto University Cross-Lingual Pronoun Translation System1 Aug 2016 0 repositories listed
-
UdS-(retrain|distributional|surface): Improving POS Tagging for OOV Words in German CMC and Web Data1 Aug 2016 0 repositories listed
-
Using longest common subsequence and character models to predict word forms1 Aug 2016 0 repositories listed
-
ASOBEK at SemEval-2016 Task 1: Sentence Representation with Character N-gram Embeddings for Semantic Textual Similarity1 Jun 2016 0 repositories listed
-
HHU at SemEval-2016 Task 1: Multiple Approaches to Measuring Semantic Textual Similarity1 Jun 2016 0 repositories listed
-
Leveraging Data-Driven Methods in Word-Level Language Identification for a Multilingual Alpine Heritage Corpus1 Jun 2016 0 repositories listed
-
The GW/UMD CLPsych 2016 Shared Task System1 Jun 2016 0 repositories listed
-
Weighting Finite-State Transductions With Neural Context1 Jun 2016 0 repositories listed
-
A Neural Lemmatizer for Bengali1 May 2016 0 repositories listed
-
CEPLEXicon ― A Lexicon of Child European Portuguese1 May 2016 0 repositories listed
-
FOLK-Gold ― A Gold Standard for Part-of-Speech-Tagging of Spoken German1 May 2016 0 repositories listed
-
Lemmatization and Morphological Tagging in German and Latin: A Comparison and a Survey of the State-of-the-art1 May 2016 0 repositories listed
-
Merging Data Resources for Inflectional and Derivational Morphology in Czech1 May 2016 0 repositories listed
-
Rule-based Automatic Multi-word Term Extraction and Lemmatization1 May 2016 0 repositories listed
-
The COPLE2 corpus: a learner corpus for Portuguese1 May 2016 0 repositories listed
-
The IPR-cleared Corpus of Contemporary Written and Spoken Romanian Language1 May 2016 0 repositories listed
-
UDPipe: Trainable Pipeline for Processing CoNLL-U Files Performing Tokenization, Morphological Analysis, POS Tagging and Parsing1 May 2016 0 repositories listed
-
N-Gramas de Caractere como Técnica de Normalização Morfológica para Língua Portuguesa: Um Estudo em Categorização de Textos (Character N-grams as a Morphological Normalization Technique for Portuguese Language: A Study in Text Categorization)1 Nov 2015 0 repositories listed
-
Realignment from Finer-grained Alignment to Coarser-grained Alignment to Enhance Mongolian-Chinese SMT1 Oct 2015 0 repositories listed
-
Do we need bigram alignment models? On the effect of alignment quality on transduction accuracy in G2P1 Sep 2015 0 repositories listed
-
E-law Module Supporting Lawyers in the Process of Knowledge Discovery from Legal Documents1 Sep 2015 0 repositories listed
-
Morphological Analysis for Unsegmented Languages using Recurrent Neural Network Language Model1 Sep 2015 0 repositories listed
-
Processing and Normalizing Hashtags1 Sep 2015 0 repositories listed
-
Counting What Counts: Decompounding for Keyphrase Extraction1 Jul 2015 0 repositories listed
-
IWNLP: Inverse Wiktionary for Natural Language Processing1 Jul 2015 0 repositories listed
-
Learning Representations for Text-level Discourse Parsing1 Jul 2015 0 repositories listed
-
Lexicon-assisted tagging and lemmatization in Latin: A comparison of six taggers and two lemmatization methods1 Jul 2015 0 repositories listed
-
Multiple Many-to-Many Sequence Alignment for Combining String-Valued Variables: A G2P Experiment1 Jul 2015 0 repositories listed
-
A Publicly Available Cross-Platform Lemmatizer for Bulgarian13 Jun 2015 0 repositories listed
-
Evaluation of the Accuracy of the BGLemmatizer13 Jun 2015 0 repositories listed
-
JAIST: Combining multiple features for Answer Selection in Community Question Answering1 Jun 2015 0 repositories listed
-
Enhancing Sumerian Lemmatization by Unsupervised Named-Entity Recognition1 May 2015 0 repositories listed
-
Persian Sentiment Analyzer: A Framework based on a Novel Feature Selection Method27 Dec 2014 0 repositories listed
-
Non-Standard Words as Features for Text Categorization28 Aug 2014 0 repositories listed
-
AI-KU: Using Co-Occurrence Modeling for Semantic Similarity1 Aug 2014 0 repositories listed
-
Illinois-LH: A Denotational and Distributional Approach to Semantics1 Aug 2014 0 repositories listed
-
Towards Semantic Validation of a Derivational Lexicon1 Aug 2014 0 repositories listed
-
DKPro Keyphrases: Flexible and Reusable Keyphrase Extraction Experiments1 Jun 2014 0 repositories listed
-
Open-Source Tools for Morphology, Lemmatization, POS Tagging and Named Entity Recognition1 Jun 2014 0 repositories listed
-
The CMU Machine Translation Systems at WMT 20141 Jun 2014 0 repositories listed
-
Tolerant BLEU: a Submission to the WMT14 Metrics Task1 Jun 2014 0 repositories listed
-
A set of open source tools for Turkish natural language processing1 May 2014 0 repositories listed
-
An efficient language independent toolkit for complete morphological disambiguation1 May 2014 0 repositories listed
-
Automatic Extraction of Synonyms for German Particle Verbs from Parallel Data with Distributional Similarity as a Re-Ranking Feature1 May 2014 0 repositories listed
-
Compounds and distributional thesauri1 May 2014 0 repositories listed
-
CoRoLa --- The Reference Corpus of Contemporary Romanian Language1 May 2014 0 repositories listed
-
CroDeriV: a new resource for processing Croatian morphology1 May 2014 0 repositories listed
-
DerivBase.hr: A High-Coverage Derivational Morphology Resource for Croatian1 May 2014 0 repositories listed
-
Evaluating Lemmatization Models for Machine-Assisted Corpus-Dictionary Linkage1 May 2014 0 repositories listed
-
MADAMIRA: A Fast, Comprehensive Tool for Morphological Analysis and Disambiguation of Arabic1 May 2014 0 repositories listed
-
Optimizing a Distributional Semantic Model for the Prediction of German Particle Verb Compositionality1 May 2014 0 repositories listed
-
Polish Coreference Corpus in Numbers1 May 2014 0 repositories listed
-
Sharing Cultural Heritage: the Clavius on the Web Project1 May 2014 0 repositories listed
-
The SETimes.HR Linguistically Annotated Corpus of Croatian1 May 2014 0 repositories listed
-
The SYN-series corpora of written Czech1 May 2014 0 repositories listed
-
Two-Step Machine Translation with Lattices1 May 2014 0 repositories listed
-
Using Resource-Rich Languages to Improve Morphological Analysis of Under-Resourced Languages1 May 2014 0 repositories listed
-
Word-Formation Network for Czech1 May 2014 0 repositories listed
-
Facilitating Multi-Lingual Sense Annotation: Human Mediated Lemmatizer1 Jan 2014 0 repositories listed