Browse State-of-the-Art › Language Identification › Papers, page 6
Language Identification
Papers archive 2025-07-28
archive papers tagged: 794 · with a code link: 143 · where Syntology ran a sample: 14 (10 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (14 of 794 tagged: 10 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument)
Page 6 of 8: papers 501 to 600 of 794, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
HeLI-based Experiments in Swiss German Dialect Identification1 Aug 2018 0 repositories listed
-
IIT (BHU) System for Indo-Aryan Language Identification (ILI) at VarDial 20181 Aug 2018 0 repositories listed
-
Iterative Language Model Adaptation for Indo-Aryan Language Identification1 Aug 2018 0 repositories listed
-
Language and the Shifting Sands of Domain, Space and Time (Invited Talk)1 Aug 2018 0 repositories listed
-
Language Identification and Morphosyntactic Tagging: The Second VarDial Evaluation Campaign1 Aug 2018 0 repositories listed
-
Neural Network Architectures for Arabic Dialect Identification1 Aug 2018 0 repositories listed
-
Punctuation as Native Language Interference1 Aug 2018 0 repositories listed
-
Tübingen-Oslo Team at the VarDial 2018 Evaluation Campaign: An Analysis of N-gram Features in Language Variety Identification1 Aug 2018 0 repositories listed
-
Discriminating between Indo-Aryan Languages Using SVM Ensembles9 Jul 2018 0 repositories listed
-
Automatic Detection of Code-switching Style from Acoustics1 Jul 2018 0 repositories listed
-
Automatic Token and Turn Level Language Identification for Code-Switched Text Dialog: An Analysis Across Language Pairs and Corpora1 Jul 2018 0 repositories listed
-
Code-Switched Named Entity Recognition with Embedding Attention1 Jul 2018 0 repositories listed
-
Language Identification and Analysis of Code-Switched Social Media Text1 Jul 2018 0 repositories listed
-
Language Identification and Named Entity Recognition in Hinglish Code Mixed Tweets1 Jul 2018 0 repositories listed
-
Language Modeling for Code-Mixing: The Role of Linguistic Theory based Synthetic Data1 Jul 2018 0 repositories listed
-
Named Entity Recognition on Code-Switched Data Using Conditional Random Fields1 Jul 2018 0 repositories listed
-
Simple Features for Strong Performance on Named Entity Recognition in Code-Switched Twitter Data1 Jul 2018 0 repositories listed
-
Transliteration Better than Translation? Answering Code-mixed Questions over a Knowledge Base1 Jul 2018 0 repositories listed
-
Twitter Universal Dependency Parsing for African-American and Mainstream American English1 Jul 2018 0 repositories listed
-
Automatic Language Identification for Romance Languages using Stop Words and Diacritics14 Jun 2018 0 repositories listed
-
Gender Prediction in English-Hindi Code-Mixed Social Media Content : Corpus and Baseline System14 Jun 2018 0 repositories listed
-
Addition of Code Mixed Features to Enhance the Sentiment Prediction of Song Lyrics11 Jun 2018 0 repositories listed
-
A Comparison of Character Neural Language Model and Bootstrapping for Language Identification in Multilingual Noisy Texts1 Jun 2018 0 repositories listed
-
Cross-corpus Native Language Identification via Statistical Embedding1 Jun 2018 0 repositories listed
-
Linguistic Features of Sarcasm and Metaphor Production Quality1 Jun 2018 0 repositories listed
-
Predicting Foreign Language Usage from English-Only Social Media Posts1 Jun 2018 0 repositories listed
-
Tübingen-Oslo at SemEval-2018 Task 2: SVMs perform better than RNNs in Emoji Prediction1 Jun 2018 0 repositories listed
-
Using Classifier Features to Determine Language Transfer on Morphemes1 Jun 2018 0 repositories listed
-
A Regression Model of Recurrent Deep Neural Networks for Noise Robust Estimation of the Fundamental Frequency Contour of Speech8 May 2018 0 repositories listed
-
Arabic Dialect Identification in the Context of Bivalency and Code-Switching1 May 2018 0 repositories listed
-
Automatic Identification of Maghreb Dialects Using a Dictionary-Based Approach1 May 2018 0 repositories listed
-
Building a TOCFL Learner Corpus for Chinese Grammatical Error Diagnosis1 May 2018 0 repositories listed
-
Building Parallel Monolingual Gan Chinese Dialects Corpus1 May 2018 0 repositories listed
-
Classification of Closely Related Sub-dialects of Arabic Using Support-Vector Machines1 May 2018 0 repositories listed
-
Collecting Code-Switched Data from Social Media1 May 2018 0 repositories listed
-
Coreference Resolution in FreeLing 4.01 May 2018 0 repositories listed
-
Discovering Parallel Language Resources for Training MT Engines1 May 2018 0 repositories listed
-
Discriminating between Similar Languages on Imbalanced Conversational Texts1 May 2018 0 repositories listed
-
From `Solved Problems' to New Challenges: A Report on LDC Activities1 May 2018 0 repositories listed
-
Shami: A Corpus of Levantine Arabic Dialects1 May 2018 0 repositories listed
-
Text Normalization Infrastructure that Scales to Hundreds of Language Varieties1 May 2018 0 repositories listed
-
The French-Algerian Code-Switching Triggered audio corpus (FACST)1 May 2018 0 repositories listed
-
Towards Language Technology for Mi'kmaq1 May 2018 0 repositories listed
-
VAST: A Corpus of Video Annotation for Speech Technologies1 May 2018 0 repositories listed
-
30 Apr 2018 0 repositories listed
-
Staircase Network: structural language identification via hierarchical attentive units30 Apr 2018 0 repositories listed
-
21 Apr 2018 0 repositories listed
-
Automatic Language Identification System for Hindi and Magahi13 Apr 2018 0 repositories listed
-
A Novel Learnable Dictionary Encoding Layer for End-to-End Language Identification2 Apr 2018 0 repositories listed
-
Insights into End-to-End Learning Scheme for Language Identification2 Apr 2018 0 repositories listed
-
Automatic Identification of Closely-related Indian Languages: Resources and Experiments26 Mar 2018 0 repositories listed
-
Language Identification of Bengali-English Code-Mixed data using Character & Phonetic based LSTM Models10 Mar 2018 0 repositories listed
-
Methods for Spoken Language Identification16 Dec 2017 0 repositories listed
-
Curriculum Design for Code-switching: Experiments with Language Identification and Language Modeling with Deep Neural Networks1 Dec 2017 0 repositories listed
-
Using Social Networks to Improve Language Variety Identification with Neural Networks1 Nov 2017 0 repositories listed
-
A Dataset and Classifier for Recognizing Social Media English1 Sep 2017 0 repositories listed
-
A deep-learning based native-language classification by using a latent semantic analysis for the NLI Shared Task 20171 Sep 2017 0 repositories listed
-
A Report on the 2017 Native Language Identification Shared Task1 Sep 2017 0 repositories listed
-
A Shallow Neural Network for Native Language Identification with Character N-grams1 Sep 2017 0 repositories listed
-
All that is English may be Hindi: Enhancing language identification through automatic ranking of the likeliness of word borrowing in social media1 Sep 2017 0 repositories listed
-
Building Dialectal Arabic Corpora1 Sep 2017 0 repositories listed
-
CIC-FBK Approach to Native Language Identification1 Sep 2017 0 repositories listed
-
Classifier Stacking for Native Language Identification1 Sep 2017 0 repositories listed
-
Combining Textual and Speech Features in the NLI Task Using State-of-the-Art Machine Learning Techniques1 Sep 2017 0 repositories listed
-
Ensemble Methods for Native Language Identification1 Sep 2017 0 repositories listed
-
Exploring Optimal Voting in Native Language Identification1 Sep 2017 0 repositories listed
-
1 Sep 2017 0 repositories listed
-
Fusion of Simple Models for Native Language Identification1 Sep 2017 0 repositories listed
-
Native Language Identification Using a Mixture of Character and Word N-grams1 Sep 2017 0 repositories listed
-
Native Language Identification using Phonetic Algorithms1 Sep 2017 0 repositories listed
-
Neural Networks and Spelling Features for Native Language Identification1 Sep 2017 0 repositories listed
-
Stacked Sentence-Document Classifier Approach for Improving Native Language Identification1 Sep 2017 0 repositories listed
-
The Power of Character N-grams in Native Language Identification1 Sep 2017 0 repositories listed
-
Vector Space Model as Cognitive Space for Text Classification21 Aug 2017 0 repositories listed
-
Lump at SemEval-2017 Task 1: Towards an Interlingua Semantic Similarity1 Aug 2017 0 repositories listed
-
Can string kernels pass the test of time in Native Language Identification?26 Jul 2017 0 repositories listed
-
All that is English may be Hindi: Enhancing language identification through automatic ranking of likeliness of word borrowing in social media25 Jul 2017 0 repositories listed
-
Native Language Identification on Text and Speech22 Jul 2017 0 repositories listed
-
Open-Set Language Identification16 Jul 2017 0 repositories listed
-
Feature Hashing for Language and Dialect Identification1 Jul 2017 0 repositories listed
-
Improving Native Language Identification by Using Spelling Errors1 Jul 2017 0 repositories listed
-
Incorporating Dialectal Variability for Socially Equitable Language Identification1 Jul 2017 0 repositories listed
-
Racial Disparity in Natural Language Processing: A Case Study of Social Media African-American English30 Jun 2017 0 repositories listed
-
Phone-aware Neural Language Identification9 May 2017 0 repositories listed
-
Phonetic Temporal Neural Model for Language Identification9 May 2017 0 repositories listed
-
Evaluation of language identification methods using 285 languages1 May 2017 0 repositories listed
-
Learning with learner corpora: Using the TLE for native language identification1 May 2017 0 repositories listed
-
Machine Learning for Rhetorical Figure Detection: More Chiasmus with Less Annotation1 May 2017 0 repositories listed
-
A Code-Switching Corpus of Turkish-German Conversations1 Apr 2017 0 repositories listed
-
A Perplexity-Based Method for Similar Languages Discrimination1 Apr 2017 0 repositories listed
-
CLUZH at VarDial GDI 2017: Testing a Variety of Machine Learning Tools for the Classification of Swiss German Dialects1 Apr 2017 0 repositories listed
-
Discriminating between Similar Languages with Word-level Convolutional Neural Networks1 Apr 2017 0 repositories listed
-
Evaluating HeLI with Non-Linear Mappings1 Apr 2017 0 repositories listed
-
Exploring Lexical and Syntactic Features for Language Variety Identification1 Apr 2017 0 repositories listed
-
Findings of the VarDial Evaluation Campaign 20171 Apr 2017 0 repositories listed
-
Identification of Languages in Algerian Arabic Multilingual Documents1 Apr 2017 0 repositories listed
-
Improving the Character Ngram Model for the DSL Task with BM25 Weighting and Less Frequently Used Feature Sets1 Apr 2017 0 repositories listed
-
Tübingen system in VarDial 2017 shared task: experiments with language identification and cross-lingual parsing1 Apr 2017 0 repositories listed
-
Twitter Language Identification Of Similar Languages And Dialects Without Ground Truth1 Apr 2017 0 repositories listed
-
URIEL and lang2vec: Representing languages as typological, geographical, and phylogenetic vectors1 Apr 2017 0 repositories listed