Browse State-of-the-Art › Text Categorization › Papers, page 3
Text Categorization
Papers archive 2025-07-28
archive papers tagged: 247 · with a code link: 44 · where Syntology ran a sample: 0 (0 with a run with no instrument failure, 0 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran: none of the 247 tagged papers has a Syntology run on record
Page 3 of 3: papers 201 to 247 of 247, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
ColLex.en: Automatically Generating and Evaluating a Full-form Lexicon for English1 May 2014 0 repositories listed
-
Data Mining with Shallow vs. Linguistic Features to Study Diversification of Scientific Registers1 May 2014 0 repositories listed
-
Dense Components in the Structure of WordNet1 May 2014 0 repositories listed
-
Mapping WordNet Domains, WordNet Topics and Wikipedia Categories to Generate Multilingual Domain Specific Resources1 May 2014 0 repositories listed
-
SenTube: A Corpus for Sentiment Analysis on YouTube Social Media1 May 2014 0 repositories listed
-
VarClass: An Open-source Language Identification Tool for Language Varieties1 May 2014 0 repositories listed
-
Wikipedia-based Semantic Interpretation for Natural Language Processing15 Jan 2014 0 repositories listed
-
Creation of Lexical Relations for IndoWordNet1 Jan 2014 0 repositories listed
-
Compressive Feature Learning1 Dec 2013 0 repositories listed
-
Automatic Corpora Construction for Text Classification1 Oct 2013 0 repositories listed
-
The Bregman Variational Dual-Tree Framework26 Sep 2013 0 repositories listed
-
Cross-Language Plagiarism Detection Methods1 Sep 2013 0 repositories listed
-
Text segmentation for Language Identification in Greek Forums1 Sep 2013 0 repositories listed
-
Towards Basque Oral Poetry Analysis: A Machine Learning Approach1 Sep 2013 0 repositories listed
-
A Semi-supervised Approach for Natural Language Call Routing1 Aug 2013 0 repositories listed
-
Bridging Languages through Etymology: The case of cross language text categorization1 Aug 2013 0 repositories listed
-
Categorization of Turkish News Documents with Morphological Analysis1 Aug 2013 0 repositories listed
-
GPKEX: Genetically Programmed Keyphrase Extraction from Croatian Texts1 Aug 2013 0 repositories listed
-
Latent Semantic Matching: Application to Cross-language Text Categorization without Alignment Information1 Aug 2013 0 repositories listed
-
Text Classification from Positive and Unlabeled Data using Misclassified Data Correction1 Aug 2013 0 repositories listed
-
TopicSpam: a Topic-Model based approach for spam detection1 Aug 2013 0 repositories listed
-
An Examination of Regret in Bullying Tweets1 Jun 2013 0 repositories listed
-
Cross-lingual and generic text categorization (Apprentissage d'une classification thématique générique et cross-langue à partir des catégories de la Wikipédia) [in French]1 Jun 2013 0 repositories listed
-
DeepPurple: Lexical, String and Affective Feature Fusion for Sentence-Level Semantic Similarity Estimation1 Jun 2013 0 repositories listed
-
ECNUCS: Measuring Short Text Semantic Equivalence Using Multiple Similarity Measurements1 Jun 2013 0 repositories listed
-
From high heels to weed attics: a syntactic investigation of chick lit and literature1 Jun 2013 0 repositories listed
-
Measuring Term Informativeness in Context1 Jun 2013 0 repositories listed
-
Negative Deceptive Opinion Spam1 Jun 2013 0 repositories listed
-
The Story of the Characters, the DNA and the Native Language1 Jun 2013 0 repositories listed
-
Feature Selection Based on Term Frequency and T-Test for Text Categorization3 May 2013 0 repositories listed
-
Semantic Similarity Computation for Abstract and Concrete Nouns Using Network-based Distributional Semantic Models1 Mar 2013 0 repositories listed
-
Baselines and Bigrams: Simple, Good Sentiment and Topic Classification1 Jul 2012 0 repositories listed
-
DeepPurple: Estimating Sentence Semantic Similarity using N-gram Regression Models and Web Snippets1 Jul 2012 0 repositories listed
-
Graph-based Semi-Supervised Learning Algorithms for NLP1 Jul 2012 0 repositories listed
-
langid.py: An Off-the-shelf Language Identification Tool1 Jul 2012 0 repositories listed
-
Modeling Topic Dependencies in Hierarchical Text Categorization1 Jul 2012 0 repositories listed
-
State-of-the-Art Kernels for Natural Language Processing1 Jul 2012 0 repositories listed
-
Towards Building a Multilingual Semantic Network: Identifying Interlingual Links in Wikipedia1 Jul 2012 0 repositories listed
-
Effect of small sample size on text categorization with support vector machines1 Jun 2012 0 repositories listed
-
A good space: Lexical predictors in word space evaluation1 May 2012 0 repositories listed
-
French and German Corpora for Audience-based Text Type Classification1 May 2012 0 repositories listed
-
Improving K-Nearest Neighbor Efficacy for Farsi Text Classification1 May 2012 0 repositories listed
-
Irregularity Detection in Categorized Document Corpora1 May 2012 0 repositories listed
-
Is it Useful to Support Users with Lexical Resources? A User Study.1 May 2012 0 repositories listed
-
SemSim: Resources for Normalized Semantic Similarity Computation Using Lexical Networks1 May 2012 0 repositories listed
-
Learning from Multiple Partially Observed Views - an Application to Multilingual Text Categorization1 Dec 2009 0 repositories listed
-
Semi-supervised Learning with Weakly-Related Unlabeled Data : Towards Better Text Categorization1 Dec 2008 0 repositories listed