Methods › Natural Language Processing › Autoregressive Transformers › GPT-2 › Papers, page 6
GPT-2
Papers archive 2025-07-28
archive papers tagged: 768 · with a code link: 339 · where Syntology ran a sample: 125 (99 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (125 of 768 tagged: 99 with a run with no instrument failure, 26 where every run was a failure of Syntology's instrument)
Page 6 of 8: papers 501 to 600 of 768, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Efficient Hierarchical Domain Adaptation for Pretrained Language Models 16 Jan 2022 · 0 repositories
-
Elastic Weight Consolidation for Reduction of Catastrophic Forgetting in GPT-2 16 Jan 2022 · 0 repositories
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture 16 Jan 2022 · 0 repositories
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 16 Jan 2022 · 0 repositories
-
When a sentence does not introduce a discourse entity, Transformer-based models still often refer to it 16 Jan 2022 · 0 repositories
-
Why Does Surprisal From Smaller GPT-2 Models Provide Better Fit to Human Reading Times? 16 Jan 2022 · 0 repositories
-
Assemble Foundation Models for Automatic Code Summarization 13 Jan 2022 · 1 repository · arXiv:2201.05222
-
Submix: Practical Private Prediction for Large-Scale Language Models 4 Jan 2022 · 0 repositories · arXiv:2201.00971
-
Call for Customized Conversation: Customized Conversation Grounding Persona and Knowledge 16 Dec 2021 · 3 repositories · arXiv:2112.08619
-
Efficient Hierarchical Domain Adaptation for Pretrained Language Models 16 Dec 2021 · 1 repository · arXiv:2112.08786
-
Reconsidering the Past: Optimizing Hidden States in Language Models 16 Dec 2021 · 0 repositories · arXiv:2112.08653
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 13 Dec 2021 · 1 repository · arXiv:2112.06598Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Improving Logical-Level Natural Language Generation with Topic-Conditioned Data Augmentation and Logical Form Generation 12 Dec 2021 · 0 repositories · arXiv:2112.06240
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 4 Dec 2021 · 0 repositories · arXiv:2112.05787
-
Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network Models 30 Nov 2021 · 1 repository · arXiv:2112.00029Syntology official: harvested, nothing ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 5 pointer-only (licence)
-
Context Matters in Semantically Controlled Language Generation for Task-oriented Dialogue Systems 28 Nov 2021 · 0 repositories · arXiv:2111.14119
-
ClipCap: CLIP Prefix for Image Captioning 18 Nov 2021 · 4 repositories · arXiv:2111.09734Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Guiding Generative Language Models for Data Augmentation in Few-Shot Text Classification 17 Nov 2021 · 0 repositories · arXiv:2111.09064
-
An Information Theoretic Measurement of Topical Relevance in Learner Essays 16 Nov 2021 · 0 repositories
-
End-to-end Task-oriented Dialog Policy Learning based on Pre-trained Language Model 16 Nov 2021 · 0 repositories
-
Generative Pre-Trained Transformer for Design Concept Generation: An Exploration 16 Nov 2021 · 0 repositories · arXiv:2111.08489
-
Knowledge Graph is in Rescue: Task Oriented Dialogue System for Response Generation without NLU and DM 16 Nov 2021 · 0 repositories
-
Moving the Eiffel Tower to ROME: Tracing and Editing Facts in GPT 16 Nov 2021 · 0 repositories
-
Representation of Ambiguity in Pre-Trained Sentence Embeddings 16 Nov 2021 · 0 repositories
-
Softmax Bottleneck Makes Language Models Unable to Represent Multi-mode Word Distributions 16 Nov 2021 · 0 repositories
-
Tell me who you are and i'll tell you what to do: A Persona Grounded Task Oriented Dialogue Generation System 16 Nov 2021 · 0 repositories
-
When classifying grammatical role, BERT doesn't care about word order... except when it matters 16 Nov 2021 · 0 repositories
-
Exploring Story Generation with Multi-task Objectives in Variational Autoencoders 15 Nov 2021 · 0 repositories · arXiv:2111.08133
-
A Novel Corpus of Discourse Structure in Humans and Computers 10 Nov 2021 · 1 repository · arXiv:2111.05940
-
DistIR: An Intermediate Representation and Simulator for Efficient Neural Network Distribution 9 Nov 2021 · 0 repositories · arXiv:2111.05426
-
FPM: A Collection of Large-scale Foundation Pre-trained Language Models 9 Nov 2021 · 0 repositories · arXiv:2111.04909
-
DSEE: Dually Sparsity-embedded Efficient Tuning of Pre-trained Language Models 30 Oct 2021 · 1 repository · arXiv:2111.00160Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 10 pointer-only (licence)
-
Amendable Generation for Dialogue State Tracking 29 Oct 2021 · 1 repository · arXiv:2110.15659
-
A Sequence to Sequence Model for Extracting Multiple Product Name Entities from Dialog 28 Oct 2021 · 0 repositories · arXiv:2110.14843
-
Generating artificial texts as substitution or complement of training data 25 Oct 2021 · 0 repositories · arXiv:2110.13016
-
Reminding the Incremental Language Model via Data-Free Self-Distillation 17 Oct 2021 · 0 repositories · arXiv:2110.08745
-
Taming Visually Guided Sound Generation 17 Oct 2021 · 3 repositories · arXiv:2110.08791Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
A Short Study on Compressing Decoder-Based Language Models 16 Oct 2021 · 0 repositories · arXiv:2110.08460
-
Evaluation of Transfer Learning for Polish with a text-to-text model 16 Oct 2021 · 0 repositories
-
Hydra: A System for Large Multi-Model Deep Learning 16 Oct 2021 · 1 repository · arXiv:2110.08633
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 16 Oct 2021 · 0 repositories
-
Kronecker Decomposition for GPT Compression 15 Oct 2021 · 0 repositories · arXiv:2110.08152
-
Leveraging Generative Models for Covert Messaging: Challenges and Tradeoffs for "Dead-Drop" Deployments 13 Oct 2021 · 0 repositories · arXiv:2110.07009
-
Language Modelling via Learning to Rank 13 Oct 2021 · 0 repositories · arXiv:2110.06961
-
Multi-Task Learning for Situated Multi-Domain End-to-End Dialogue Systems 11 Oct 2021 · 0 repositories · arXiv:2110.05221
-
Word Acquisition in Neural Language Models 5 Oct 2021 · 1 repository · arXiv:2110.02406
-
Low Frequency Names Exhibit Bias and Overfitting in Contextualizing Language Models 1 Oct 2021 · 0 repositories · arXiv:2110.00672
-
Language Model Pre-training Improves Generalization in Policy Learning 29 Sep 2021 · 0 repositories
-
Mapping Language Models to Grounded Conceptual Spaces 29 Sep 2021 · 0 repositories
-
Offline Reinforcement Learning for Large Scale Language Action Spaces 29 Sep 2021 · 0 repositories
-
A Plug-and-Play Method for Controlled Text Generation 20 Sep 2021 · 1 repository · arXiv:2109.09707
-
Model Bias in NLP -- Application to Hate Speech Classification using transfer learning techniques 20 Sep 2021 · 0 repositories · arXiv:2109.09725
-
Learning Low-frequency Patterns with A Pre-trained Document-Grounded Conversation Model 17 Sep 2021 · 0 repositories
-
Relating Neural Text Degeneration to Exposure Bias 17 Sep 2021 · 0 repositories · arXiv:2109.08705
-
Improving Text Auto-Completion with Next Phrase Prediction 15 Sep 2021 · 0 repositories · arXiv:2109.07067
-
A Temporal Variational Model for Story Generation 14 Sep 2021 · 3 repositories · arXiv:2109.06807
-
Enhancing Self-Disclosure In Neural Dialog Models By Candidate Re-ranking 10 Sep 2021 · 0 repositories · arXiv:2109.05090
-
All Bark and No Bite: Rogue Dimensions in Transformer Language Models Obscure Representational Quality 9 Sep 2021 · 1 repository · arXiv:2109.04404
-
Variational Latent-State GPT for Semi-Supervised Task-Oriented Dialog Systems 9 Sep 2021 · 2 repositories · arXiv:2109.04314
-
TruthfulQA: Measuring How Models Mimic Human Falsehoods 8 Sep 2021 · 3 repositories · arXiv:2109.07958Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Empathetic Dialogue Generation with Pre-trained RoBERTa-GPT2 and External Knowledge 7 Sep 2021 · 0 repositories · arXiv:2109.03004
-
Text-Free Prosody-Aware Generative Spoken Language Modeling 7 Sep 2021 · 1 repository · arXiv:2109.03264
-
ConQX: Semantic Expansion of Spoken Queries for Intent Detection based on Conditioned Text Generation 2 Sep 2021 · 0 repositories · arXiv:2109.00729
-
OptAGAN: Entropy-based finetuning on text VAE-GAN 1 Sep 2021 · 1 repository · arXiv:2109.00239
-
Task-Oriented Dialogue System as Natural Language Generation 31 Aug 2021 · 1 repository · arXiv:2108.13679Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
HeadlineCause: A Dataset of News Headlines for Detecting Causalities 28 Aug 2021 · 1 repository · arXiv:2108.12626
-
Table Caption Generation in Scholarly Documents Leveraging Pre-trained Language Models 18 Aug 2021 · 1 repository · arXiv:2108.08111
-
The Stability-Efficiency Dilemma: Investigating Sequence Length Warmup for Training GPT Models 13 Aug 2021 · 1 repository · arXiv:2108.06084
-
Offensive Language and Hate Speech Detection with Deep Learning and Transfer Learning 6 Aug 2021 · 0 repositories · arXiv:2108.03305
-
Q-Pain: A Question Answering Dataset to Measure Social Bias in Pain Management 3 Aug 2021 · 0 repositories · arXiv:2108.01764
-
Best of Both Worlds: Making High Accuracy Non-incremental Transformer-based Disfluency Detection Incremental 1 Aug 2021 · 0 repositories
-
Employing Argumentation Knowledge Graphs for Neural Argument Generation 1 Aug 2021 · 1 repository
-
Explanations for CommonsenseQA: New Dataset and Models 1 Aug 2021 · 0 repositories
-
KuiLeiXi: a Chinese Open-Ended Text Adventure Game 1 Aug 2021 · 0 repositories
-
PRAL: A Tailored Pre-Training Model for Task-Oriented Dialog Generation 1 Aug 2021 · 0 repositories
-
TGEA: An Error-Annotated Dataset and Benchmark Tasks for TextGeneration from Pretrained Language Models 1 Aug 2021 · 0 repositories
-
Unleash GPT-2 Power for Event Detection 1 Aug 2021 · 0 repositories
-
Adapting GPT, GPT-2 and BERT Language Models for Speech Recognition 29 Jul 2021 · 0 repositories · arXiv:2108.07789
-
Chimera: Efficiently Training Large-Scale Neural Networks with Bidirectional Pipelines 14 Jul 2021 · 1 repository · arXiv:2107.06925
-
Is GPT-3 Text Indistinguishable from Human Text? Scarecrow: A Framework for Scrutinizing Machine Text 2 Jul 2021 · 0 repositories · arXiv:2107.01294
-
Toward Less Hidden Cost of Code Completion with Acceptance and Ranking Models 26 Jun 2021 · 0 repositories · arXiv:2106.13928
-
LoRA: Low-Rank Adaptation of Large Language Models 17 Jun 2021 · 74 repositories · arXiv:2106.09685Syntology community repositories only · 51 ran (of which 19 constructed an object rather than computing a result; 44 with no instrument failure: 1 honoured, 0 violated, 43 with no contract checked; 7 where Syntology's instrument failed) · 33 unverified (of 84 harvested samples) · 30 pointer-only (licence)
-
Textual Data Distributions: Kullback Leibler Textual Distributions Contrasts on GPT-2 Generated Texts, with Supervised, Unsupervised Learning on Vaccine & Market Topics & Sentiment 15 Jun 2021 · 0 repositories · arXiv:2107.02025
-
Generate, Annotate, and Learn: NLP with Synthetic Text 11 Jun 2021 · 1 repository · arXiv:2106.06168Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Auto-tagging of Short Conversational Sentences using Transformer Methods 3 Jun 2021 · 0 repositories · arXiv:2106.01735
-
Towards a Comprehensive Understanding and Accurate Evaluation of Societal Biases in Pre-Trained Transformers 1 Jun 2021 · 0 repositories
-
Generative Adversarial Imitation Learning for Empathy-based AI 27 May 2021 · 0 repositories · arXiv:2105.13328
-
Methods for Detoxification of Texts for the Russian Language 19 May 2021 · 3 repositories · arXiv:2105.09052
-
Neural Predictive Text for Grammatical Error Prevention 16 May 2021 · 0 repositories
-
SLGPT: Using Transfer Learning to Directly Generate Simulink Model Files and Find Bugs in the Simulink Toolchain 16 May 2021 · 1 repository · arXiv:2105.07465
-
BERT Busters: Outlier Dimensions that Disrupt Transformers 14 May 2021 · 0 repositories · arXiv:2105.06990
-
BERT is to NLP what AlexNet is to CV: Can Pre-Trained Language Models Identify Analogies? 11 May 2021 · 1 repository · arXiv:2105.04949
-
EL-Attention: Memory Efficient Lossless Attention for Generation 11 May 2021 · 1 repository · arXiv:2105.04779Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
e-ViL: A Dataset and Benchmark for Natural Language Explanations in Vision-Language Tasks 8 May 2021 · 2 repositories · arXiv:2105.03761Syntology official (archive's flag): 5 ran · 7 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Mitigating Political Bias in Language Models Through Reinforced Calibration 30 Apr 2021 · 0 repositories · arXiv:2104.14795
-
Extractive and Abstractive Explanations for Fact-Checking and Evaluation of News 27 Apr 2021 · 0 repositories · arXiv:2104.12918
-
UoT-UWF-PartAI at SemEval-2021 Task 5: Self Attention Based Bi-GRU with Multi-Embedding Representation for Toxicity Highlighter 27 Apr 2021 · 0 repositories · arXiv:2104.13164
-
Accounting for Agreement Phenomena in Sentence Comprehension with Transformer Language Models: Effects of Similarity-based Interference on Surprisal and Attention 26 Apr 2021 · 0 repositories · arXiv:2104.12874
-
Easy and Efficient Transformer : Scalable Inference Solution For large NLP model 26 Apr 2021 · 1 repository · arXiv:2104.12470
-
Efficient pre-training objectives for Transformers 20 Apr 2021 · 0 repositories · arXiv:2104.09694