Browse State-of-the-Art › Language Identification › Papers, page 5
Language Identification
Papers archive 2025-07-28
archive papers tagged: 794 · with a code link: 143 · where Syntology ran a sample: 14 (10 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (14 of 794 tagged: 10 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument)
Page 5 of 8: papers 401 to 500 of 794, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
LT@Helsinki at SemEval-2020 Task 12: Multilingual or language-specific BERT?3 Aug 2020 0 repositories listed
-
SalamNET at SemEval-2020 Task12: Deep Learning Approach for Arabic Offensive Language Detection28 Jul 2020 0 repositories listed
-
Duluth at SemEval-2020 Task 12: Offensive Tweet Identification in English with Logistic Regression25 Jul 2020 0 repositories listed
-
XD at SemEval-2020 Task 12: Ensemble Approach to Offensive Language Identification in Social Media Using Transformer Encoders21 Jul 2020 0 repositories listed
-
Dialect Diversity in Text Summarization on Twitter15 Jul 2020 0 repositories listed
-
Fine-grained Language Identification with Multilingual CapsNet Model12 Jul 2020 0 repositories listed
-
The ASRU 2019 Mandarin-English Code-Switching Speech Recognition Challenge: Open Datasets, Tracks, Methods and Results12 Jul 2020 0 repositories listed
-
Feature Selection on Noisy Twitter Short Text Messages for Language Identification11 Jul 2020 0 repositories listed
-
Streaming End-to-End Bilingual ASR Systems with Joint Language Identification8 Jul 2020 0 repositories listed
-
Cross-lingual Inductive Transfer to Detect Offensive Language7 Jul 2020 0 repositories listed
-
A Report on the 2020 VUA and TOEFL Metaphor Detection Shared Task1 Jul 2020 0 repositories listed
-
An Assessment of Language Identification Methods on Tweets and Wikipedia Articles1 Jul 2020 0 repositories listed
-
GLUECoS: An Evaluation Benchmark for Code-Switched NLP1 Jul 2020 0 repositories listed
-
Investigating the effect of auxiliary objectives for the automated grading of learner English speech transcriptions1 Jul 2020 0 repositories listed
-
OpusFilter: A Configurable Parallel Corpus Filtering Toolbox1 Jul 2020 0 repositories listed
-
Adaptation de domaine non supervisée pour la reconnaissance de la langue par régularisation d'un réseau de neurones (Unsupervised domain adaptation for language identification by regularization of a neural network)1 Jun 2020 0 repositories listed
-
Lexical Normalization for Code-switched Data and its Effect on POS-tagging1 Jun 2020 0 repositories listed
-
Streaming Language Identification using Combination of Acoustic Representations and ASR Hypotheses1 Jun 2020 0 repositories listed
-
Identification/Segmentation of Indian Regional Languages with Singular Value Decomposition based Feature Embedding17 May 2020 0 repositories listed
-
9 May 2020 0 repositories listed
-
LIIR at SemEval-2020 Task 12: A Cross-Lingual Augmentation Approach for Multilingual Offensive Language Identification7 May 2020 0 repositories listed
-
Building Web Corpora for Minority Languages1 May 2020 0 repositories listed
-
Hitachi at SemEval-2020 Task 12: Offensive Language Identification with Noisy Labels using Statistical Sampling and Post-Processing1 May 2020 0 repositories listed
-
On The Performance of Time-Pooling Strategies for End-to-End Spoken Language Identification1 May 2020 0 repositories listed
-
OpusTools and Parallel Corpus Diagnostics1 May 2020 0 repositories listed
-
Search Query Language Identification Using Weak Labeling1 May 2020 0 repositories listed
-
Two LRL & Distractor Corpora from Web Information Retrieval and a Small Case Study in Language Identification without Training Corpora1 May 2020 0 repositories listed
-
SOLID: A Large-Scale Semi-Supervised Dataset for Offensive Language Identification29 Apr 2020 0 repositories listed
-
Detect Language of Transliterated Texts26 Apr 2020 0 repositories listed
-
GLUECoS : An Evaluation Benchmark for Code-Switched NLP26 Apr 2020 0 repositories listed
-
Mapping Languages: The Corpus of Global Language Use2 Apr 2020 0 repositories listed
-
Towards Relevance and Sequence Modeling in Language Recognition2 Apr 2020 0 repositories listed
-
Rnn-transducer with language bias for end-to-end Mandarin-English code-switching speech recognition19 Feb 2020 0 repositories listed
-
Identification of Indian Languages using Ghost-VLAD pooling5 Feb 2020 0 repositories listed
-
Improving Language Identification for Multilingual Speakers29 Jan 2020 0 repositories listed
-
Universal and non-universal text statistics: Clustering coefficient for language identification18 Nov 2019 0 repositories listed
-
Multilingual Grammar Induction with Continuous Language Identification1 Nov 2019 0 repositories listed
-
Normalization of Indonesian-English Code-Mixed Twitter Data1 Nov 2019 0 repositories listed
-
Signal Combination for Language Identification21 Oct 2019 0 repositories listed
-
Language Identification on Massive Datasets of Short Message using an Attention Mechanism CNN15 Oct 2019 0 repositories listed
-
9 Oct 2019 0 repositories listed
-
Overview for the Second Shared Task on Language Identification in Code-Switched Data28 Sep 2019 0 repositories listed
-
Hope Speech Detection: A Computational Analysis of the Voice of Peace11 Sep 2019 0 repositories listed
-
Two-stage Training for Chinese Dialect Recognition6 Aug 2019 0 repositories listed
-
Anglicized Words and Misspelled Cognates in Native Language Identification1 Aug 2019 0 repositories listed
-
Learning Multilingual Meta-Embeddings for Code-Switching Named Entity Recognition1 Aug 2019 0 repositories listed
-
Regression or classification? Automated Essay Scoring for Norwegian1 Aug 2019 0 repositories listed
-
Joint Language Identification of Code-Switching Speech using Attention based E2E Network15 Jul 2019 0 repositories listed
-
Adversarial Training for Multilingual Acoustic Modeling17 Jun 2019 0 repositories listed
-
A Report on the Third VarDial Evaluation Campaign1 Jun 2019 0 repositories listed
-
BNU-HKBU UIC NLP Team 2 at SemEval-2019 Task 6: Detecting Offensive Language Using BERT model1 Jun 2019 0 repositories listed
-
CAMsterdam at SemEval-2019 Task 6: Neural and graph-based feature extraction for the identification of offensive tweets1 Jun 2019 0 repositories listed
-
CN-HIT-MI.T at SemEval-2019 Task 6: Offensive Language Identification Based on BiLSTM with Double Attention1 Jun 2019 0 repositories listed
-
ConvAI at SemEval-2019 Task 6: Offensive Language Identification and Categorization with Perspective and BERT1 Jun 2019 0 repositories listed
-
DeepAnalyzer at SemEval-2019 Task 6: A deep learning-based ensemble method for identifying offensive tweets1 Jun 2019 0 repositories listed
-
Discriminating between Mandarin Chinese and Swiss-German varieties using adaptive language models1 Jun 2019 0 repositories listed
-
Emad at SemEval-2019 Task 6: Offensive Language Identification using Traditional Machine Learning and Deep Learning approaches1 Jun 2019 0 repositories listed
-
HAD-Tübingen at SemEval-2019 Task 6: Deep Learning Analysis of Offensive Language on Twitter: Identification and Categorization1 Jun 2019 0 repositories listed
-
HHU at SemEval-2019 Task 6: Context Does Matter - Tackling Offensive Language Identification and Categorization with ELMo1 Jun 2019 0 repositories listed
-
Improving Cuneiform Language Identification with BERT1 Jun 2019 0 repositories listed
-
Joint Approach to Deromanization of Code-mixed Texts1 Jun 2019 0 repositories listed
-
Language Discrimination and Transfer Learning for Similar Languages: Experiments with Feature Combinations and Adaptation1 Jun 2019 0 repositories listed
-
Naive Bayes and BiLSTM Ensemble for Discriminating between Mainland and Taiwan Variation of Mandarin Chinese1 Jun 2019 0 repositories listed
-
NULI at SemEval-2019 Task 6: Transfer Learning for Offensive Language Detection using Bidirectional Transformers1 Jun 2019 0 repositories listed
-
SSN_NLP at SemEval-2019 Task 6: Offensive Language Identification in Social Media using Traditional and Deep Machine Learning Approaches1 Jun 2019 0 repositories listed
-
The Titans at SemEval-2019 Task 6: Offensive Language Identification, Categorization and Target Identification1 Jun 2019 0 repositories listed
-
TwistBytes - Identification of Cuneiform Languages and German Dialects at VarDial 20191 Jun 2019 0 repositories listed
-
Typological Features for Multilingual Delexicalised Dependency Parsing1 Jun 2019 0 repositories listed
-
Multiclass Language Identification using Deep Learning on Spectral Images of Audio Signals10 May 2019 0 repositories listed
-
Distributional Interaction of Concreteness and Abstractness in Verb--Noun Subcategorisation1 May 2019 0 repositories listed
-
Experiments in Cuneiform Language Identification27 Apr 2019 0 repositories listed
-
Subword-Level Language Identification for Intra-Word Code-Switching3 Apr 2019 0 repositories listed
-
Language Model Adaptation for Language and Dialect Identification of Text26 Mar 2019 0 repositories listed
-
An Exploration of State-of-the-art Methods for Offensive Language Detection15 Mar 2019 0 repositories listed
-
Absit invidia verbo: Comparing Deep Learning methods for offensive language14 Mar 2019 0 repositories listed
-
Language and Dialect Identification of Cuneiform Texts5 Mar 2019 0 repositories listed
-
Towards NLP with Deep Learning: Convolutional Neural Networks and Recurrent Neural Networks for Offensive Language Identification in Social Media2 Mar 2019 0 repositories listed
-
Utterance-level end-to-end language identification using attention-based CNN-BLSTM20 Feb 2019 0 repositories listed
-
Albanian Language Identification in Text Documents14 Jan 2019 0 repositories listed
-
Corpora of social media in minority Uralic languages1 Jan 2019 0 repositories listed
-
Domain Attentive Fusion for End-to-end Dialect Identification with Unknown Target Domain4 Dec 2018 0 repositories listed
-
Transductive Learning with String Kernels for Cross-Domain Text Classification2 Nov 2018 0 repositories listed
-
Towards End-to-End Code-Switching Speech Recognition31 Oct 2018 0 repositories listed
-
Strategies for Language Identification in Code-Mixed Low Resource Languages16 Oct 2018 0 repositories listed
-
A Fast, Compact, Accurate Model for Language Identification of Codemixed Text9 Oct 2018 0 repositories listed
-
Native Language Identification with User Generated Content1 Oct 2018 0 repositories listed
-
Part-of-Speech Tagging for Code-Switched, Transliterated Texts without Explicit Language Identification1 Oct 2018 0 repositories listed
-
Similarity Dependent Chinese Restaurant Process for Cognate Identification in Multilingual Wordlists1 Oct 2018 0 repositories listed
-
The Role of Emotions in Native Language Identification1 Oct 2018 0 repositories listed
-
Hindi-English Code-Switching Speech Corpus24 Sep 2018 0 repositories listed
-
Language Identification with Deep Bottleneck Features18 Sep 2018 0 repositories listed
-
Native Language Identification With Classifier Stacking and Ensembles1 Sep 2018 0 repositories listed
-
Improving the results of string kernels in sentiment analysis and Arabic dialect identification by adapting them to your test set25 Aug 2018 0 repositories listed
-
Language Identification in Code-Mixed Data using Multichannel Neural Networks and Context Capture21 Aug 2018 0 repositories listed
-
Character Level Convolutional Neural Network for Indo-Aryan Language Identification1 Aug 2018 0 repositories listed
-
Computationally efficient discrimination between language varieties with large feature vectors and regularized classifiers1 Aug 2018 0 repositories listed
-
Deep Models for Arabic Dialect Identification on Benchmarked Data1 Aug 2018 0 repositories listed
-
Exploring Classifier Combinations for Language Variety Identification1 Aug 2018 0 repositories listed
-
Fully Connected Neural Network with Advance Preprocessor to Identify Aggression over Facebook and Twitter1 Aug 2018 0 repositories listed
-
HeLI-based Experiments in Discriminating Between Dutch and Flemish Subtitles1 Aug 2018 0 repositories listed