Browse State-of-the-Art › Text Classification › Papers, page 19
Text Classification
Papers archive 2025-07-28
archive papers tagged: 3,635 · with a code link: 1,308 · where Syntology ran a sample: 272 (219 with a run with no instrument failure, 53 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (272 of 3,635 tagged: 219 with a run with no instrument failure, 53 where every run was a failure of Syntology's instrument)
Page 19 of 37: papers 1,801 to 1,900 of 3,635, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
3 Jul 2023 0 repositories listed
-
Automatic Counterfactual Augmentation for Robust Text Classification Based on Word-Group Search1 Jul 2023 0 repositories listed
-
Meta-training with Demonstration Retrieval for Efficient Few-shot Learning30 Jun 2023 0 repositories listed
-
Investigating Cross-Domain Behaviors of BERT in Review Understanding27 Jun 2023 0 repositories listed
-
Deconstructing Classifiers: Towards A Data Reconstruction Attack Against Text Classification Models23 Jun 2023 0 repositories listed
-
Learning to Specialize: Joint Gating-Expert Training for Adaptive MoEs in Decentralized Settings14 Jun 2023 0 repositories listed
-
Soft Language Clustering for Multilingual Model Pre-training13 Jun 2023 0 repositories listed
-
Imbalanced Multi-label Classification for Business-related Text with Moderately Large Label Spaces12 Jun 2023 0 repositories listed
-
Textual Augmentation Techniques Applied to Low Resource Machine Translation: Case of Swahili12 Jun 2023 0 repositories listed
-
Assessing Phrase Break of ESL Speech with Pre-trained Language Models and Large Language Models8 Jun 2023 0 repositories listed
-
Interpretable Medical Diagnostics with Structured Data Extraction by Large Language Models8 Jun 2023 0 repositories listed
-
Leveraging Language Identification to Enhance Code-Mixed Text Classification8 Jun 2023 0 repositories listed
-
Analysis of the Fed's communication by using textual entailment model of Zero-Shot classification7 Jun 2023 0 repositories listed
-
5 Jun 2023 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Word Embeddings for Banking Industry2 Jun 2023 0 repositories listed
-
Adversarial Clean Label Backdoor Attacks and Defenses on Text Classification Systems31 May 2023 0 repositories listed
-
Analyzing Text Representations by Measuring Task Alignment31 May 2023 0 repositories listed
-
Cross Encoding as Augmentation: Towards Effective Educational Text Classification30 May 2023 0 repositories listed
-
Machine Learning Approach for Cancer Entities Association and Classification30 May 2023 0 repositories listed
-
A Framework For Refining Text Classification and Object Recognition from Academic Articles27 May 2023 0 repositories listed
-
D-CALM: A Dynamic Clustering-based Active Learning Approach for Mitigating Bias26 May 2023 0 repositories listed
-
Perturbation-based Self-supervised Attention for Attention Bias in Text Classification25 May 2023 0 repositories listed
-
Estimating class separability of text embeddings with persistent homology24 May 2023 0 repositories listed
-
Active Learning for Natural Language Generation24 May 2023 0 repositories listed
-
EXnet: Efficient In-context Learning for Data-less Text classification24 May 2023 0 repositories listed
-
Is Summary Useful or Not? An Extrinsic Human Evaluation of Text Summaries on Downstream Tasks24 May 2023 0 repositories listed
-
PESCO: Prompt-enhanced Self Contrastive Learning for Zero-shot Text Classification24 May 2023 0 repositories listed
-
Enhancing Black-Box Few-Shot Text Classification with Prompt-Based Data Augmentation23 May 2023 0 repositories listed
-
Handling Realistic Label Noise in BERT Text Classification23 May 2023 0 repositories listed
-
Out-of-Distribution Generalization in Text Classification: Past, Present, and Future23 May 2023 0 repositories listed
-
Regex-augmented Domain Transfer Topic Classification based on a Pre-trained Language Model: An application in Financial Domain23 May 2023 0 repositories listed
-
A Comprehensive Survey of Sentence Representations: From the BERT Epoch to the ChatGPT Era and Beyond22 May 2023 0 repositories listed
-
On Bias and Fairness in NLP: Investigating the Impact of Bias and Debiasing in Language Models on the Fairness of Toxicity Detection22 May 2023 0 repositories listed
-
Quantum Text Classifier -- A Synchronistic Approach Towards Classical and Quantum Machine Learning22 May 2023 0 repositories listed
-
Retrieval-augmented Multi-label Text Classification22 May 2023 0 repositories listed
-
Self-Evolution Learning for Mixup: Enhance Data Augmentation on Few-Shot Text Classification Tasks22 May 2023 0 repositories listed
-
The Grammar and Syntax Based Corpus Analysis Tool For The Ukrainian Language22 May 2023 0 repositories listed
-
F-PABEE: Flexible-patience-based Early Exiting for Single-label and Multi-label text Classification Tasks21 May 2023 0 repositories listed
-
Large-Scale Text Analysis Using Generative Language Models: A Case Study in Discovering Public Value Expressions in AI Patents17 May 2023 0 repositories listed
-
UOR: Universal Backdoor Attacks on Pre-trained Language Models16 May 2023 0 repositories listed
-
Sensitivity and Robustness of Large Language Models to Prompt Template in Japanese Text Classification Tasks15 May 2023 0 repositories listed
-
RepCL: Exploring Effective Representation for Continual Text Classification12 May 2023 0 repositories listed
-
Two-in-One: A Model Hijacking Attack Against Text Generation Models12 May 2023 0 repositories listed
-
How Good are Commercial Large Language Models on African Languages?11 May 2023 0 repositories listed
-
10 May 2023 0 repositories listed
-
SPSQL: Step-by-step Parsing Based Framework for Text-to-SQL Generation10 May 2023 0 repositories listed
-
Consistent Text Categorization using Data Augmentation in e-Commerce9 May 2023 0 repositories listed
-
What Do Patients Say About Their Disease Symptoms? Deep Multilabel Text Classification With Human-in-the-Loop Curation for Automatic Labeling of Patient Self Reports of Problems8 May 2023 0 repositories listed
-
Rhetorical Role Labeling of Legal Documents using Transformers and Graph Neural Networks6 May 2023 0 repositories listed
-
Augmenting Low-Resource Text Classification with Graph-Grounded Pre-training and Prompting5 May 2023 0 repositories listed
-
DN at SemEval-2023 Task 12: Low-Resource Language Text Classification via Multilingual Pretrained Language Model Fine-tuning4 May 2023 0 repositories listed
-
Enhancing Pashto Text Classification using Language Processing Techniques for Single And Multi-Label Analysis4 May 2023 0 repositories listed
-
Towards Weakly-Supervised Hate Speech Classification Across Datasets4 May 2023 0 repositories listed
-
Tuning Traditional Language Processing Approaches for Pashto Text Classification4 May 2023 0 repositories listed
-
Backdoor Learning on Sequence to Sequence Models3 May 2023 0 repositories listed
-
Using Language Models on Low-end Hardware3 May 2023 0 repositories listed
-
Privacy-Preserving In-Context Learning for Large Language Models2 May 2023 0 repositories listed
-
Prompt as Triggers for Backdoor Attack: Examining the Vulnerability in Language Models2 May 2023 0 repositories listed
-
An Iterative Algorithm for Rescaled Hyperbolic Functions Regression1 May 2023 0 repositories listed
-
Company classification using zero-shot learning1 May 2023 0 repositories listed
-
NLP-LTU at SemEval-2023 Task 10: The Impact of Data Augmentation and Semi-Supervised Learning Techniques on Text Classification Performance on an Imbalanced Dataset25 Apr 2023 0 repositories listed
-
Graph Neural Networks for Text Classification: A Survey23 Apr 2023 0 repositories listed
-
Downstream Task-Oriented Neural Tokenizer Optimization with Vocabulary Restriction as Post Processing21 Apr 2023 0 repositories listed
-
Text2Time: Transformer-based Article Time Period Prediction21 Apr 2023 0 repositories listed
-
ESimCSE Unsupervised Contrastive Learning Jointly with UDA Semi-Supervised Learning for Large Label System Text Classification Mode19 Apr 2023 0 repositories listed
-
Shuffle & Divide: Contrastive Learning for Long Text19 Apr 2023 0 repositories listed
-
A Two-Stage Framework with Self-Supervised Distillation For Cross-Domain Text Classification18 Apr 2023 0 repositories listed
-
Comparative study on Judgment Text Classification for Transformer Based Models18 Apr 2023 0 repositories listed
-
Label Dependencies-aware Set Prediction Networks for Multi-label Text Classification14 Apr 2023 0 repositories listed
-
Improved Naive Bayes with Mislabeled Data13 Apr 2023 0 repositories listed
-
Classification of news spreading barriers10 Apr 2023 0 repositories listed
-
Continual Graph Convolutional Network for Text Classification9 Apr 2023 0 repositories listed
-
Performance of Data Augmentation Methods for Brazilian Portuguese Text Classification5 Apr 2023 0 repositories listed
-
Multidimensional Perceptron for Efficient and Explainable Long Text Classification4 Apr 2023 0 repositories listed
-
Attention is Not Always What You Need: Towards Efficient Classification of Domain-Specific Text31 Mar 2023 0 repositories listed
-
Comparing Abstractive Summaries Generated by ChatGPT to Real Summaries Through Blinded Reviewers and Text Classification Algorithms30 Mar 2023 0 repositories listed
-
Model and Evaluation: Towards Fairness in Multilingual Text Classification28 Mar 2023 0 repositories listed
-
Beyond Toxic: Toxicity Detection Datasets are Not Enough for Brand Safety27 Mar 2023 0 repositories listed
-
Boosting Few-Shot Text Classification via Distribution Estimation26 Mar 2023 0 repositories listed
-
Improving Transformer Performance for French Clinical Notes Classification Using Mixture of Experts on a Limited Dataset22 Mar 2023 0 repositories listed
-
Analyzing the Generalizability of Deep Contextualized Language Representations For Text Classification22 Mar 2023 0 repositories listed
-
Fine-tuning ClimateBert transformer with ClimaText for the disclosure analysis of climate-related financial risks21 Mar 2023 0 repositories listed
-
Input-length-shortening and text generation via attention values14 Mar 2023 0 repositories listed
-
Large Language Models in the Workplace: A Case Study on Prompt Engineering for Job Type Classification13 Mar 2023 0 repositories listed
-
SA-CNN: Application to text categorization issues using simulated annealing-based convolutional neural network optimization13 Mar 2023 0 repositories listed
-
Transformer-based approaches to Sentiment Detection13 Mar 2023 0 repositories listed
-
ChatGPT: Beginning of an End of Manual Linguistic Data Annotation? Use Case of Automatic Genre Identification7 Mar 2023 0 repositories listed
-
Exploring Data Augmentation Methods on Social Media Corpora3 Mar 2023 0 repositories listed
-
Will Affective Computing Emerge from Foundation Models and General AI? A First Evaluation on ChatGPT3 Mar 2023 0 repositories listed
-
Adopting the Multi-answer Questioning Task with an Auxiliary Metric for Extreme Multi-label Text Classification Utilizing the Label Hierarchy2 Mar 2023 0 repositories listed
-
Document Provenance and Authentication through Authorship Classification2 Mar 2023 0 repositories listed
-
Fairness Evaluation in Text Classification: Machine Learning Practitioner Perspectives of Individual and Group Fairness1 Mar 2023 0 repositories listed
-
Make Every Example Count: On the Stability and Utility of Self-Influence for Learning from Noisy NLP Datasets27 Feb 2023 0 repositories listed
-
Mask-guided BERT for Few Shot Text Classification21 Feb 2023 0 repositories listed
-
Identifying Semantically Difficult Samples to Improve Text Classification13 Feb 2023 0 repositories listed
-
Towards Agile Text Classifiers for Everyone13 Feb 2023 0 repositories listed
-
10 Feb 2023 0 repositories listed
-
CRL+: A Novel Semi-Supervised Deep Active Contrastive Representation Learning-Based Text Classification Model for Insurance Data8 Feb 2023 0 repositories listed
-
AutoWS: Automated Weak Supervision Framework for Text Classification7 Feb 2023 0 repositories listed
-
Natural Language Processing for Policymaking7 Feb 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.