Methods › General › Fine-Tuning › Discriminative Fine-Tuning › Papers, page 17
Discriminative Fine-Tuning
Papers archive 2025-07-28
archive papers tagged: 1,990 · with a code link: 794 · where Syntology ran a sample: 271 (223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (271 of 1,990 tagged: 223 with a run with no instrument failure, 48 where every run was a failure of Syntology's instrument)
Page 17 of 20: papers 1,601 to 1,700 of 1,990, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Provably Confidential Language Modelling 16 Jan 2022 · 0 repositories
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 16 Jan 2022 · 0 repositories
-
When a sentence does not introduce a discourse entity, Transformer-based models still often refer to it 16 Jan 2022 · 0 repositories
-
Why Does Surprisal From Smaller GPT-2 Models Provide Better Fit to Human Reading Times? 16 Jan 2022 · 0 repositories
-
Assemble Foundation Models for Automatic Code Summarization 13 Jan 2022 · 1 repository · arXiv:2201.05222
-
Submix: Practical Private Prediction for Large-Scale Language Models 4 Jan 2022 · 0 repositories · arXiv:2201.00971
-
EvoMoE: An Evolutional Mixture-of-Experts Training Framework via Dense-To-Sparse Gate 29 Dec 2021 · 2 repositories · arXiv:2112.14397Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Call for Customized Conversation: Customized Conversation Grounding Persona and Knowledge 16 Dec 2021 · 3 repositories · arXiv:2112.08619
-
Efficient Hierarchical Domain Adaptation for Pretrained Language Models 16 Dec 2021 · 1 repository · arXiv:2112.08786
-
Reconsidering the Past: Optimizing Hidden States in Language Models 16 Dec 2021 · 0 repositories · arXiv:2112.08653
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 13 Dec 2021 · 1 repository · arXiv:2112.06598Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Improving Logical-Level Natural Language Generation with Topic-Conditioned Data Augmentation and Logical Form Generation 12 Dec 2021 · 0 repositories · arXiv:2112.06240
-
Improving Knowledge Graph Representation Learning by Structure Contextual Pre-training 8 Dec 2021 · 0 repositories · arXiv:2112.04087
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 4 Dec 2021 · 0 repositories · arXiv:2112.05787
-
Think Big, Teach Small: Do Language Models Distil Occam’s Razor? 1 Dec 2021 · 1 repository
-
Chemical Identification and Indexing in PubMed Articles via BERT and Text-to-Text Approaches 30 Nov 2021 · 0 repositories · arXiv:2111.15622
-
Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network Models 30 Nov 2021 · 1 repository · arXiv:2112.00029Syntology official: harvested, nothing ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 5 pointer-only (licence)
-
Context Matters in Semantically Controlled Language Generation for Task-oriented Dialogue Systems 28 Nov 2021 · 0 repositories · arXiv:2111.14119
-
Transformer-based Korean Pretrained Language Models: A Survey on Three Years of Progress 25 Nov 2021 · 0 repositories · arXiv:2112.03014
-
ClipCap: CLIP Prefix for Image Captioning 18 Nov 2021 · 4 repositories · arXiv:2111.09734Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Guiding Generative Language Models for Data Augmentation in Few-Shot Text Classification 17 Nov 2021 · 0 repositories · arXiv:2111.09064
-
An Information Theoretic Measurement of Topical Relevance in Learner Essays 16 Nov 2021 · 0 repositories
-
Building Chinese Biomedical Language Models via Multi-Level Text Discrimination 16 Nov 2021 · 1 repository
-
CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation 16 Nov 2021 · 0 repositories
-
ELLE: Efficient Lifelong Pre-training for Emerging Data 16 Nov 2021 · 0 repositories
-
End-to-end Task-oriented Dialog Policy Learning based on Pre-trained Language Model 16 Nov 2021 · 0 repositories
-
Exploring and Adapting Chinese GPT to Pinyin Input Method 16 Nov 2021 · 0 repositories
-
Generative Pre-Trained Transformer for Design Concept Generation: An Exploration 16 Nov 2021 · 0 repositories · arXiv:2111.08489
-
GLM: General Language Model Pretraining with Autoregressive Blank Infilling 16 Nov 2021 · 0 repositories
-
Impact of Tokenization on Language Models: An Analysis for Turkish 16 Nov 2021 · 0 repositories
-
Knowledge Graph is in Rescue: Task Oriented Dialogue System for Response Generation without NLU and DM 16 Nov 2021 · 0 repositories
-
Life after BERT: What do Other Muppets Understand about Language? 16 Nov 2021 · 0 repositories
-
Moving the Eiffel Tower to ROME: Tracing and Editing Facts in GPT 16 Nov 2021 · 0 repositories
-
Representation of Ambiguity in Pre-Trained Sentence Embeddings 16 Nov 2021 · 0 repositories
-
Softmax Bottleneck Makes Language Models Unable to Represent Multi-mode Word Distributions 16 Nov 2021 · 0 repositories
-
Tell me who you are and i'll tell you what to do: A Persona Grounded Task Oriented Dialogue Generation System 16 Nov 2021 · 0 repositories
-
When classifying grammatical role, BERT doesn't care about word order... except when it matters 16 Nov 2021 · 0 repositories
-
Exploring Story Generation with Multi-task Objectives in Variational Autoencoders 15 Nov 2021 · 0 repositories · arXiv:2111.08133
-
IIITT@Dravidian-CodeMix-FIRE2021: Transliterate or translate? Sentiment analysis of code-mixed text in Dravidian languages 15 Nov 2021 · 1 repository · arXiv:2111.07906
-
A Novel Corpus of Discourse Structure in Humans and Computers 10 Nov 2021 · 1 repository · arXiv:2111.05940
-
DistIR: An Intermediate Representation and Simulator for Efficient Neural Network Distribution 9 Nov 2021 · 0 repositories · arXiv:2111.05426
-
FPM: A Collection of Large-scale Foundation Pre-trained Language Models 9 Nov 2021 · 0 repositories · arXiv:2111.04909
-
DSEE: Dually Sparsity-embedded Efficient Tuning of Pre-trained Language Models 30 Oct 2021 · 1 repository · arXiv:2111.00160Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 10 pointer-only (licence)
-
Amendable Generation for Dialogue State Tracking 29 Oct 2021 · 1 repository · arXiv:2110.15659
-
A Sequence to Sequence Model for Extracting Multiple Product Name Entities from Dialog 28 Oct 2021 · 0 repositories · arXiv:2110.14843
-
Colossal-AI: A Unified Deep Learning System For Large-Scale Parallel Training 28 Oct 2021 · 1 repository · arXiv:2110.14883
-
Generating artificial texts as substitution or complement of training data 25 Oct 2021 · 0 repositories · arXiv:2110.13016
-
Fast Model Editing at Scale 21 Oct 2021 · 3 repositories · arXiv:2110.11309Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Illiterate DALL-E Learns to Compose 17 Oct 2021 · 1 repository · arXiv:2110.11405Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Reminding the Incremental Language Model via Data-Free Self-Distillation 17 Oct 2021 · 0 repositories · arXiv:2110.08745
-
A Short Study on Compressing Decoder-Based Language Models 16 Oct 2021 · 0 repositories · arXiv:2110.08460
-
Evaluation of Transfer Learning for Polish with a text-to-text model 16 Oct 2021 · 0 repositories
-
Hydra: A System for Large Multi-Model Deep Learning 16 Oct 2021 · 1 repository · arXiv:2110.08633
-
PAGnol: An Extra-Large French Generative Model 16 Oct 2021 · 0 repositories · arXiv:2110.08554
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 16 Oct 2021 · 0 repositories
-
Kronecker Decomposition for GPT Compression 15 Oct 2021 · 0 repositories · arXiv:2110.08152
-
Building Chinese Biomedical Language Models via Multi-Level Text Discrimination 14 Oct 2021 · 1 repository · arXiv:2110.07244
-
Leveraging Generative Models for Covert Messaging: Challenges and Tradeoffs for "Dead-Drop" Deployments 13 Oct 2021 · 0 repositories · arXiv:2110.07009
-
Language Modelling via Learning to Rank 13 Oct 2021 · 0 repositories · arXiv:2110.06961
-
LightSeq2: Accelerated Training for Transformer-based Models on GPUs 12 Oct 2021 · 1 repository · arXiv:2110.05722
-
Multi-Task Learning for Situated Multi-Domain End-to-End Dialogue Systems 11 Oct 2021 · 0 repositories · arXiv:2110.05221
-
Vector-quantized Image Modeling with Improved VQGAN 9 Oct 2021 · 5 repositories · arXiv:2110.04627Syntology 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Word Acquisition in Neural Language Models 5 Oct 2021 · 1 repository · arXiv:2110.02406
-
Adversarial Examples Generation for Reducing Implicit Gender Bias in Pre-trained Models 3 Oct 2021 · 0 repositories · arXiv:2110.01094
-
Low Frequency Names Exhibit Bias and Overfitting in Contextualizing Language Models 1 Oct 2021 · 0 repositories · arXiv:2110.00672
-
Illiterate DALL·E Learns to Compose 29 Sep 2021 · 0 repositories
-
Language Model Pre-training Improves Generalization in Policy Learning 29 Sep 2021 · 0 repositories
-
Mapping Language Models to Grounded Conceptual Spaces 29 Sep 2021 · 0 repositories
-
Offline Reinforcement Learning for Large Scale Language Action Spaces 29 Sep 2021 · 0 repositories
-
SeqPATE: Differentially Private Text Generation via Knowledge Distillation 29 Sep 2021 · 0 repositories
-
Language Models as Recommender Systems: Evaluations and Limitations 22 Sep 2021 · 0 repositories
-
A Plug-and-Play Method for Controlled Text Generation 20 Sep 2021 · 1 repository · arXiv:2109.09707
-
Model Bias in NLP -- Application to Hate Speech Classification using transfer learning techniques 20 Sep 2021 · 0 repositories · arXiv:2109.09725
-
Learning Low-frequency Patterns with A Pre-trained Document-Grounded Conversation Model 17 Sep 2021 · 0 repositories
-
Relating Neural Text Degeneration to Exposure Bias 17 Sep 2021 · 0 repositories · arXiv:2109.08705
-
Language Models are Few-shot Multilingual Learners 16 Sep 2021 · 1 repository · arXiv:2109.07684Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Improving Text Auto-Completion with Next Phrase Prediction 15 Sep 2021 · 0 repositories · arXiv:2109.07067
-
A Temporal Variational Model for Story Generation 14 Sep 2021 · 3 repositories · arXiv:2109.06807
-
Multilingual Translation via Grafting Pre-trained Language Models 11 Sep 2021 · 1 repository · arXiv:2109.05256
-
TopicRefine: Joint Topic Prediction and Dialogue Response Generation for Multi-turn End-to-End Dialogue System 11 Sep 2021 · 0 repositories · arXiv:2109.05187
-
Enhancing Self-Disclosure In Neural Dialog Models By Candidate Re-ranking 10 Sep 2021 · 0 repositories · arXiv:2109.05090
-
All Bark and No Bite: Rogue Dimensions in Transformer Language Models Obscure Representational Quality 9 Sep 2021 · 1 repository · arXiv:2109.04404
-
Variational Latent-State GPT for Semi-Supervised Task-Oriented Dialog Systems 9 Sep 2021 · 2 repositories · arXiv:2109.04314
-
TruthfulQA: Measuring How Models Mimic Human Falsehoods 8 Sep 2021 · 3 repositories · arXiv:2109.07958Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Empathetic Dialogue Generation with Pre-trained RoBERTa-GPT2 and External Knowledge 7 Sep 2021 · 0 repositories · arXiv:2109.03004
-
NumGPT: Improving Numeracy Ability of Generative Pre-trained Models 7 Sep 2021 · 0 repositories · arXiv:2109.03137
-
Text-Free Prosody-Aware Generative Spoken Language Modeling 7 Sep 2021 · 1 repository · arXiv:2109.03264
-
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation 2 Sep 2021 · 5 repositories · arXiv:2109.00859Syntology official: harvested, nothing ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
ConQX: Semantic Expansion of Spoken Queries for Intent Detection based on Conditioned Text Generation 2 Sep 2021 · 0 repositories · arXiv:2109.00729
-
Fight Fire with Fire: Fine-tuning Hate Detectors using Large Samples of Generated Hate Speech 1 Sep 2021 · 0 repositories · arXiv:2109.00591
-
OptAGAN: Entropy-based finetuning on text VAE-GAN 1 Sep 2021 · 1 repository · arXiv:2109.00239
-
Task-Oriented Dialogue System as Natural Language Generation 31 Aug 2021 · 1 repository · arXiv:2108.13679Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
HeadlineCause: A Dataset of News Headlines for Detecting Causalities 28 Aug 2021 · 1 repository · arXiv:2108.12626
-
Offensive Language Identification in Low-resourced Code-mixed Dravidian languages using Pseudo-labeling 27 Aug 2021 · 1 repository · arXiv:2108.12177
-
Towards Offensive Language Identification for Tamil Code-Mixed YouTube Comments and Posts 24 Aug 2021 · 1 repository · arXiv:2108.10939
-
Table Caption Generation in Scholarly Documents Leveraging Pre-trained Language Models 18 Aug 2021 · 1 repository · arXiv:2108.08111
-
The Stability-Efficiency Dilemma: Investigating Sequence Length Warmup for Training GPT Models 13 Aug 2021 · 1 repository · arXiv:2108.06084
-
AMMUS : A Survey of Transformer-based Pretrained Models in Natural Language Processing 12 Aug 2021 · 1 repository · arXiv:2108.05542
-
PatrickStar: Parallel Training of Pre-trained Models via Chunk-based Memory Management 12 Aug 2021 · 1 repository · arXiv:2108.05818
-
Offensive Language and Hate Speech Detection with Deep Learning and Transfer Learning 6 Aug 2021 · 0 repositories · arXiv:2108.03305