Browse State-of-the-Art › text-classification › Papers, page 14
text-classification
Papers archive 2025-07-28
archive papers tagged: 3,054 · with a code link: 1,083 · where Syntology ran a sample: 204 (161 with a run with no instrument failure, 43 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (204 of 3,054 tagged: 161 with a run with no instrument failure, 43 where every run was a failure of Syntology's instrument)
Page 14 of 31: papers 1,301 to 1,400 of 3,054, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A Novel Two-Step Fine-Tuning Pipeline for Cold-Start Active Learning in Text Classification Tasks24 Jul 2024 0 repositories listed
-
23 Jul 2024 0 repositories listed Syntology 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
AIDE: Antithetical, Intent-based, and Diverse Example-Based Explanations22 Jul 2024 0 repositories listed
-
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts21 Jul 2024 0 repositories listed
-
A Comprehensive Review of Few-shot Action Recognition20 Jul 2024 0 repositories listed
-
Domain-specific or Uncertainty-aware models: Does it really make a difference for biomedical text classification?17 Jul 2024 0 repositories listed
-
Domain-Hierarchy Adaptation via Chain of Iterative Reasoning for Few-shot Hierarchical Text Classification12 Jul 2024 0 repositories listed
-
GPC: Generative and General Pathology Image Classifier12 Jul 2024 0 repositories listed
-
Beyond Text: Leveraging Multi-Task Learning and Cognitive Appraisal Theory for Post-Purchase Intention Analysis11 Jul 2024 0 repositories listed
-
NoisyAG-News: A Benchmark for Addressing Instance-Dependent Noise in Text Classification9 Jul 2024 0 repositories listed
-
New Directions in Text Classification Research: Maximizing The Performance of Sentiment Classification from Limited Data8 Jul 2024 0 repositories listed
-
Bias Correction in Machine Learning-based Classification of Rare Events4 Jul 2024 0 repositories listed
-
Exploring the Role of Transliteration in In-Context Learning for Low-resource Languages Written in Non-Latin Scripts2 Jul 2024 0 repositories listed
-
Calibrated Large Language Models for Binary Question Answering1 Jul 2024 0 repositories listed
-
Protecting Privacy in Classifiers by Token Manipulation1 Jul 2024 0 repositories listed
-
LegalTurk Optimized BERT for Multi-Label Text Classification and NER30 Jun 2024 0 repositories listed
-
A Recipe of Parallel Corpora Exploitation for Multilingual Large Language Models29 Jun 2024 0 repositories listed
-
LLM-Generated Natural Language Meets Scaling Laws: New Explorations and Data Augmentation Methods29 Jun 2024 0 repositories listed
-
Data Generation Using Large Language Models for Text Classification: An Empirical Case Study27 Jun 2024 0 repositories listed
-
Detecting Machine-Generated Texts: Not Just "AI vs Humans" and Explainability is Complicated26 Jun 2024 0 repositories listed
-
Knowledge Distillation in Automated Annotation: Supervised Text Classification with LLM-Generated Training Labels25 Jun 2024 0 repositories listed
-
Understanding Language Model Circuits through Knowledge Editing25 Jun 2024 0 repositories listed
-
Context-augmented Retrieval: A Novel Framework for Fast Information Retrieval based Response Generation using Large Language Model24 Jun 2024 0 repositories listed
-
Hierarchical thematic classification of major conference proceedings21 Jun 2024 0 repositories listed
-
Self-supervised Interpretable Concept-based Models for Text Classification20 Jun 2024 0 repositories listed
-
An Empirical Investigation of Matrix Factorization Methods for Pre-trained Transformers17 Jun 2024 0 repositories listed
-
Leveraging Foundation Models for Multi-modal Federated Learning with Incomplete Modality16 Jun 2024 0 repositories listed
-
Self-Regulated Data-Free Knowledge Amalgamation for Text Classification16 Jun 2024 0 repositories listed
-
Universal Cross-Lingual Text Classification16 Jun 2024 0 repositories listed
-
Adversarial Evasion Attack Efficiency against Large Language Models12 Jun 2024 0 repositories listed
-
Robust Latent Representation Tuning for Image-text Classification10 Jun 2024 0 repositories listed
-
SecureNet: A Comparative Study of DeBERTa and Large Language Models for Phishing Detection10 Jun 2024 0 repositories listed
-
AICoderEval: Improving AI Domain Code Generation of Large Language Models7 Jun 2024 0 repositories listed
-
6 Jun 2024 0 repositories listed
-
IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models5 Jun 2024 0 repositories listed
-
Fuzzy Convolution Neural Networks for Tabular Data Classification4 Jun 2024 0 repositories listed
-
Annotation Guidelines-Based Knowledge Augmentation: Towards Enhancing Large Language Models for Educational Text Classification3 Jun 2024 0 repositories listed
-
Continuous Geometry-Aware Graph Diffusion via Hyperbolic Neural PDE3 Jun 2024 0 repositories listed
-
Progressive Inference: Explaining Decoder-Only Sequence Classification Models Using Intermediate Predictions3 Jun 2024 0 repositories listed
-
Synergizing Unsupervised and Supervised Learning: A Hybrid Approach for Accurate Natural Language Task Modeling3 Jun 2024 0 repositories listed
-
Conformal Transformation of Kernels: A Geometric Perspective on Text Classification1 Jun 2024 0 repositories listed
-
Fast Explanations via Policy Gradient-Optimized Explainer29 May 2024 0 repositories listed
-
Mashee at SemEval-2024 Task 8: The Impact of Samples Quality on the Performance of In-Context Learning for Machine Text Classification28 May 2024 0 repositories listed
-
Bayesian WeakS-to-Strong from Text Classification to Generation24 May 2024 0 repositories listed
-
ChronosLex: Time-aware Incremental Training for Temporal Generalization of Legal Classification Tasks23 May 2024 0 repositories listed
-
Explaining Black-box Model Predictions via Two-level Nested Feature Attributions with Consistency Property23 May 2024 0 repositories listed
-
Exploration of Masked and Causal Language Modelling for Text Generation21 May 2024 0 repositories listed
-
A Constraint-Enforcing Reward for Adversarial Attacks on Text Classifiers20 May 2024 0 repositories listed
-
Exploring Ordinality in Text Classification: A Comparative Study of Explicit and Implicit Techniques20 May 2024 0 repositories listed
-
Simple-Sampling and Hard-Mixup with Prototypes to Rebalance Contrastive Learning for Text Classification19 May 2024 0 repositories listed
-
The Curious Case of Class Accuracy Imbalance in LLMs: Post-hoc Debiasing via Nonlinear Integer Programming13 May 2024 0 repositories listed
-
Fine-tuning the SwissBERT Encoder Model for Embedding Sentences and Documents13 May 2024 0 repositories listed
-
Enhancing Suicide Risk Detection on Social Media through Semi-Supervised Deep Label Smoothing9 May 2024 0 repositories listed
-
Machine learning models for abstract screening task - A systematic literature review application for health economics and outcome research9 May 2024 0 repositories listed
-
Using Machine Translation to Augment Multilingual Classification9 May 2024 0 repositories listed
-
Semi-Supervised Disease Classification based on Limited Medical Image Data7 May 2024 0 repositories listed
-
Liberating Seen Classes: Boosting Few-Shot and Zero-Shot Text Classification via Anchor Generation and Classification Reframing6 May 2024 0 repositories listed
-
On Adversarial Examples for Text Classification by Perturbing Latent Representations6 May 2024 0 repositories listed
-
Assessing Adversarial Robustness of Large Language Models: An Empirical Study4 May 2024 0 repositories listed
-
Exploring Extreme Quantization in Spiking Language Models4 May 2024 0 repositories listed
-
Learning label-label correlations in Extreme Multi-label Classification via Label Features3 May 2024 0 repositories listed
-
The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification3 May 2024 0 repositories listed
-
PVF (Parameter Vulnerability Factor): A Scalable Metric for Understanding AI Vulnerability Against SDCs in Model Parameters2 May 2024 0 repositories listed
-
Lightweight Conceptual Dictionary Learning for Text Classification Using Information Compression28 Apr 2024 0 repositories listed
-
TextGram: Towards a better domain-adaptive pretraining28 Apr 2024 0 repositories listed
-
GuideWalk: A Novel Graph-Based Word Embedding for Enhanced Text Classification25 Apr 2024 0 repositories listed
-
Social Media and Artificial Intelligence for Sustainable Cities and Societies: A Water Quality Analysis Use-case23 Apr 2024 0 repositories listed
-
Transformer-Based Classification Outcome Prediction for Multimodal Stroke Treatment19 Apr 2024 0 repositories listed
-
A Novel ICD Coding Method Based on Associated and Hierarchical Code Description Distillation17 Apr 2024 0 repositories listed
-
Small Language Models are Good Too: An Empirical Study of Zero-Shot Classification17 Apr 2024 0 repositories listed
-
Empowering Interdisciplinary Research with BERT-Based Models: An Approach Through SciBERT-CNN with Topic Modeling16 Apr 2024 0 repositories listed
-
Quantization of Large Language Models with an Overdetermined Basis15 Apr 2024 0 repositories listed
-
Exploring Contrastive Learning for Long-Tailed Multi-Label Text Classification12 Apr 2024 0 repositories listed
-
Semantic Stealth: Adversarial Text Attacks on NLP Using Several Methods8 Apr 2024 0 repositories listed
-
Text clustering applied to data augmentation in legal contexts8 Apr 2024 0 repositories listed
-
Large Language Model (LLM) AI text generation detection based on transformer deep learning algorithm6 Apr 2024 0 repositories listed
-
Adversarial Attacks and Dimensionality in Text Classifiers3 Apr 2024 0 repositories listed
-
Enhancing Low-Resource LLMs Classification with PEFT and Synthetic Data3 Apr 2024 0 repositories listed
-
Ukrainian Texts Classification: Exploration of Cross-lingual Knowledge Transfer Approaches2 Apr 2024 0 repositories listed
-
AISPACE at SemEval-2024 task 8: A Class-balanced Soft-voting System for Detecting Multi-generator Machine-generated Text1 Apr 2024 0 repositories listed
-
Automatic explanation of the classification of Spanish legal judgments in jurisdiction-dependent law categories with tree estimators30 Mar 2024 0 repositories listed
-
Shortcuts Arising from Contrast: Effective and Covert Clean-Label Attacks in Prompt-Based Learning30 Mar 2024 0 repositories listed
-
Identifying Banking Transaction Descriptions via Support Vector Machine Short-Text Classification Based on a Specialized Labelled Corpus29 Mar 2024 0 repositories listed
-
Evaluating Large Language Models for Health-Related Text Classification Tasks with Public Social Media Data27 Mar 2024 0 repositories listed
-
Language Models for Text Classification: Is In-Context Learning Enough?26 Mar 2024 0 repositories listed
-
LARA: Linguistic-Adaptive Retrieval-Augmentation for Multi-Turn Intent Classification25 Mar 2024 0 repositories listed
-
23 Mar 2024 0 repositories listed
-
MasonTigers at SemEval-2024 Task 8: Performance Analysis of Transformer-based Models on Machine-Generated Text Detection22 Mar 2024 0 repositories listed
-
Multi-Level Explanations for Generative Language Models21 Mar 2024 0 repositories listed
-
Visual Analytics for Fine-grained Text Classification Models and Datasets21 Mar 2024 0 repositories listed
-
Vi-Mistral-X: Building a Vietnamese Language Model with Advanced Continual Pre-training20 Mar 2024 0 repositories listed
-
CrossTune: Black-Box Few-Shot Classification with Label Enhancement19 Mar 2024 0 repositories listed
-
Simple Hack for Transformers against Heavy Long-Text Classification on a Time- and Memory-Limited GPU Service19 Mar 2024 0 repositories listed
-
A Disease Labeler for Chinese Chest X-Ray Report Generation18 Mar 2024 0 repositories listed
-
A Modified Word Saliency-Based Adversarial Attack on Text Classification Models17 Mar 2024 0 repositories listed
-
Forging the Forger: An Attempt to Improve Authorship Verification via Data Augmentation17 Mar 2024 0 repositories listed
-
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection12 Mar 2024 0 repositories listed
-
Evolving Knowledge Distillation with Large Language Models and Active Learning11 Mar 2024 0 repositories listed
-
One size doesn't fit all: Predicting the Number of Examples for In-Context Learning11 Mar 2024 0 repositories listed
-
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers8 Mar 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.