Methods › Natural Language Processing › Autoregressive Transformers › GPT › Papers, page 12
GPT
Papers archive 2025-07-28
archive papers tagged: 1,212 · with a code link: 453 · where Syntology ran a sample: 152 (130 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (152 of 1,212 tagged: 130 with a run with no instrument failure, 22 where every run was a failure of Syntology's instrument)
Page 12 of 13: papers 1,101 to 1,200 of 1,212, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
NLX-GPT: A Model for Natural Language Explanations in Vision and Vision-Language Tasks 9 Mar 2022 · 1 repository · arXiv:2203.05081
-
Exploring and Adapting Chinese GPT to Pinyin Input Method 1 Mar 2022 · 1 repository · arXiv:2203.00249
-
Consistent Dropout for Policy Gradient Reinforcement Learning 23 Feb 2022 · 0 repositories · arXiv:2202.11818
-
SGPT: GPT Sentence Embeddings for Semantic Search 17 Feb 2022 · 1 repository · arXiv:2202.08904Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
What are the best systems? New perspectives on NLP Benchmarking 8 Feb 2022 · 1 repository · arXiv:2202.03799
-
L3Cube-MahaCorpus and MahaBERT: Marathi Monolingual Corpus, Marathi BERT Language Models, and Resources 2 Feb 2022 · 1 repository · arXiv:2202.01159
-
A Frustratingly Simple Approach for End-to-End Image Captioning 30 Jan 2022 · 0 repositories · arXiv:2201.12723
-
DNNFuser: Generative Pre-Trained Transformer as a Generalized Mapper for Layer Fusion in DNN Accelerators 26 Jan 2022 · 0 repositories · arXiv:2201.11218
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 16 Jan 2022 · 0 repositories
-
Provably Confidential Language Modelling 16 Jan 2022 · 0 repositories
-
EvoMoE: An Evolutional Mixture-of-Experts Training Framework via Dense-To-Sparse Gate 29 Dec 2021 · 2 repositories · arXiv:2112.14397Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Think Big, Teach Small: Do Language Models Distil Occam’s Razor? 1 Dec 2021 · 1 repository
-
Chemical Identification and Indexing in PubMed Articles via BERT and Text-to-Text Approaches 30 Nov 2021 · 0 repositories · arXiv:2111.15622
-
Transformer-based Korean Pretrained Language Models: A Survey on Three Years of Progress 25 Nov 2021 · 0 repositories · arXiv:2112.03014
-
Building Chinese Biomedical Language Models via Multi-Level Text Discrimination 16 Nov 2021 · 1 repository
-
CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation 16 Nov 2021 · 0 repositories
-
ELLE: Efficient Lifelong Pre-training for Emerging Data 16 Nov 2021 · 0 repositories
-
Exploring and Adapting Chinese GPT to Pinyin Input Method 16 Nov 2021 · 0 repositories
-
GLM: General Language Model Pretraining with Autoregressive Blank Infilling 16 Nov 2021 · 0 repositories
-
Impact of Tokenization on Language Models: An Analysis for Turkish 16 Nov 2021 · 0 repositories
-
Life after BERT: What do Other Muppets Understand about Language? 16 Nov 2021 · 0 repositories
-
Colossal-AI: A Unified Deep Learning System For Large-Scale Parallel Training 28 Oct 2021 · 1 repository · arXiv:2110.14883
-
Fast Model Editing at Scale 21 Oct 2021 · 3 repositories · arXiv:2110.11309Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Illiterate DALL-E Learns to Compose 17 Oct 2021 · 1 repository · arXiv:2110.11405Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
PAGnol: An Extra-Large French Generative Model 16 Oct 2021 · 0 repositories · arXiv:2110.08554
-
Kronecker Decomposition for GPT Compression 15 Oct 2021 · 0 repositories · arXiv:2110.08152
-
Building Chinese Biomedical Language Models via Multi-Level Text Discrimination 14 Oct 2021 · 1 repository · arXiv:2110.07244
-
LightSeq2: Accelerated Training for Transformer-based Models on GPUs 12 Oct 2021 · 1 repository · arXiv:2110.05722
-
Vector-quantized Image Modeling with Improved VQGAN 9 Oct 2021 · 5 repositories · arXiv:2110.04627Syntology 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Adversarial Examples Generation for Reducing Implicit Gender Bias in Pre-trained Models 3 Oct 2021 · 0 repositories · arXiv:2110.01094
-
Illiterate DALL·E Learns to Compose 29 Sep 2021 · 0 repositories
-
SeqPATE: Differentially Private Text Generation via Knowledge Distillation 29 Sep 2021 · 0 repositories
-
Language Models as Recommender Systems: Evaluations and Limitations 22 Sep 2021 · 0 repositories
-
Language Models are Few-shot Multilingual Learners 16 Sep 2021 · 1 repository · arXiv:2109.07684Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Multilingual Translation via Grafting Pre-trained Language Models 11 Sep 2021 · 1 repository · arXiv:2109.05256
-
TopicRefine: Joint Topic Prediction and Dialogue Response Generation for Multi-turn End-to-End Dialogue System 11 Sep 2021 · 0 repositories · arXiv:2109.05187
-
Variational Latent-State GPT for Semi-Supervised Task-Oriented Dialog Systems 9 Sep 2021 · 2 repositories · arXiv:2109.04314
-
NumGPT: Improving Numeracy Ability of Generative Pre-trained Models 7 Sep 2021 · 0 repositories · arXiv:2109.03137
-
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation 2 Sep 2021 · 5 repositories · arXiv:2109.00859Syntology official: harvested, nothing ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Fight Fire with Fire: Fine-tuning Hate Detectors using Large Samples of Generated Hate Speech 1 Sep 2021 · 0 repositories · arXiv:2109.00591
-
The Stability-Efficiency Dilemma: Investigating Sequence Length Warmup for Training GPT Models 13 Aug 2021 · 1 repository · arXiv:2108.06084
-
AMMUS : A Survey of Transformer-based Pretrained Models in Natural Language Processing 12 Aug 2021 · 1 repository · arXiv:2108.05542
-
PatrickStar: Parallel Training of Pre-trained Models via Chunk-based Memory Management 12 Aug 2021 · 1 repository · arXiv:2108.05818
-
RockGPT: Reconstructing three-dimensional digital rocks from single two-dimensional slice from the perspective of video generation 5 Aug 2021 · 0 repositories · arXiv:2108.03132
-
Adapting GPT, GPT-2 and BERT Language Models for Speech Recognition 29 Jul 2021 · 0 repositories · arXiv:2108.07789
-
Scalable Memory Protection in the PENGLAI Enclave 14 Jul 2021 · 1 repository
-
Evaluating Large Language Models Trained on Code 7 Jul 2021 · 13 repositories · arXiv:2107.03374Syntology official (archive's flag): 2 ran · 26 ran (of which 0 constructed an object rather than computing a result; 24 with no instrument failure: 1 honoured, 0 violated, 23 with no contract checked; 2 where Syntology's instrument failed) · 13 unverified (of 39 harvested samples) · 4 pointer-only (licence)
-
SymbolicGPT: A Generative Transformer Model for Symbolic Regression 27 Jun 2021 · 2 repositories · arXiv:2106.14131Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
Pre-Trained Models: Past, Present and Future 14 Jun 2021 · 0 repositories · arXiv:2106.07139
-
Engines of Power: Electricity, AI, and General-Purpose Military Transformations 8 Jun 2021 · 0 repositories · arXiv:2106.04338
-
Do Syntactic Probes Probe Syntax? Experiments with Jabberwocky Probing 4 Jun 2021 · 0 repositories · arXiv:2106.02559
-
Data Curation and Quality Assurance for Machine Learning-based Cyber Intrusion Detection 20 May 2021 · 1 repository · arXiv:2105.10041
-
Analyzing COVID-19 Tweets with Transformer-based Language Models 20 Apr 2021 · 0 repositories · arXiv:2104.10259
-
Understanding Transformers for Bot Detection in Twitter 13 Apr 2021 · 1 repository · arXiv:2104.06182
-
KI-BERT: Infusing Knowledge Context for Better Language and Domain Understanding 9 Apr 2021 · 0 repositories · arXiv:2104.08145
-
The NLP Cookbook: Modern Recipes for Transformer based Deep Learning Architectures 23 Mar 2021 · 0 repositories · arXiv:2104.10640
-
GLM: General Language Model Pretraining with Autoregressive Blank Infilling 18 Mar 2021 · 8 repositories · arXiv:2103.10360Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
GPT Understands, Too 18 Mar 2021 · 10 repositories · arXiv:2103.10385Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
From Universal Language Model to Downstream Task: Improving RoBERTa-Based Vietnamese Hate Speech Detection 24 Feb 2021 · 0 repositories · arXiv:2102.12162
-
Exploring Transformers in Natural Language Generation: GPT, BERT, and XLNet 16 Feb 2021 · 1 repository · arXiv:2102.08036
-
Query expansion with artificially generated texts 16 Dec 2020 · 0 repositories · arXiv:2012.08787
-
RPT: Relational Pre-trained Transformer Is Almost All You Need towards Democratizing Data Preparation 4 Dec 2020 · 0 repositories · arXiv:2012.02469
-
Attention Mechanism, Transformers, BERT, and GPT: Tutorial and Survey 17 Nov 2020 · 0 repositories
-
Tabular Transformers for Modeling Multivariate Time Series 3 Nov 2020 · 1 repository · arXiv:2011.01843Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
The Amazing World of Neural Language Generation 1 Nov 2020 · 0 repositories
-
LightSeq: A High Performance Inference Library for Transformers 23 Oct 2020 · 1 repository · arXiv:2010.13887
-
Performance of Transfer Learning Model vs. Traditional Neural Network in Low System Resource Environment 20 Oct 2020 · 0 repositories · arXiv:2011.07962
-
DA-Transformer: Distance-aware Transformer 14 Oct 2020 · 0 repositories · arXiv:2010.06925
-
Prior Art Search and Reranking for Generated Patent Text 19 Sep 2020 · 0 repositories · arXiv:2009.09132
-
Hierarchical GPT with Congruent Transformers for Multi-Sentence Language Models 18 Sep 2020 · 0 repositories · arXiv:2009.08636
-
Comparative Evaluation of Pretrained Transfer Learning Models on Automatic Short Answer Grading 2 Sep 2020 · 1 repository · arXiv:2009.01303
-
Knowledge Efficient Deep Learning for Natural Language Processing 28 Aug 2020 · 0 repositories · arXiv:2008.12878
-
Dynamics of feed forward induced interference training 24 Aug 2020 · 0 repositories · arXiv:2008.11111
-
On-The-Fly Information Retrieval Augmentation for Language Models 3 Jul 2020 · 0 repositories · arXiv:2007.01528
-
Roles and Utilization of Attention Heads in Transformer-based Neural Language Models 1 Jul 2020 · 1 repository
-
Memory-Efficient Pipeline-Parallel DNN Training 16 Jun 2020 · 1 repository · arXiv:2006.09503Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 10 harvested samples)
-
Emergence of Separable Manifolds in Deep Language Representations 1 Jun 2020 · 1 repository · arXiv:2006.01095
-
On the Generation of Medical Dialogues for COVID-19 11 May 2020 · 0 repositories · arXiv:2005.05442
-
Spying on your neighbors: Fine-grained probing of contextual embeddings for information about surrounding words 4 May 2020 · 0 repositories · arXiv:2005.01810Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Multilingual Corpus Creation for Multilingual Semantic Similarity Task 1 May 2020 · 0 repositories
-
ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators 23 Mar 2020 · 19 repositories · arXiv:2003.10555Syntology official: no sample here; runs from other or unrecorded repositories · 31 ran (of which 7 constructed an object rather than computing a result; 18 with no instrument failure: 2 honoured, 2 violated, 14 with no contract checked; 13 where Syntology's instrument failed) · 9 unverified (of 40 harvested samples) · 10 pointer-only (licence)
-
Training Large Neural Networks with Constant Memory using a New Execution Algorithm 13 Feb 2020 · 2 repositories · arXiv:2002.05645Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Joint Contextual Modeling for ASR Correction and Language Understanding 28 Jan 2020 · 0 repositories · arXiv:2002.00750
-
OTEANN: Estimating the Transparency of Orthographies with an Artificial Neural Network 31 Dec 2019 · 2 repositories · arXiv:1912.13321
-
A Comparative Study of Pretrained Language Models on Thai Social Text Categorization 3 Dec 2019 · 0 repositories · arXiv:1912.01580
-
Evaluating Commonsense in Pre-trained Language Models 27 Nov 2019 · 1 repository · arXiv:1911.11931
-
Zero-Shot Paraphrase Generation with Multilingual Language Models 9 Nov 2019 · 0 repositories · arXiv:1911.03597
-
Inspecting Unification of Encoding and Matching with Transformer: A Case Study of Machine Reading Comprehension 1 Nov 2019 · 0 repositories
-
Natural Language Generation for Effective Knowledge Distillation 1 Nov 2019 · 1 repository
-
An Empirical Study of Efficient ASR Rescoring with Transformers 24 Oct 2019 · 0 repositories · arXiv:1910.11450
-
Evolution of transfer learning in natural language processing 16 Oct 2019 · 0 repositories · arXiv:1910.07370
-
Q8BERT: Quantized 8Bit BERT 14 Oct 2019 · 5 repositories · arXiv:1910.06188Syntology official (archive's flag): 5 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
ZeRO: Memory Optimizations Toward Training Trillion Parameter Models 4 Oct 2019 · 10 repositories · arXiv:1910.02054Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Extremely Small BERT Models from Mixed-Vocabulary Training 25 Sep 2019 · 0 repositories · arXiv:1909.11687
-
How Additional Knowledge can Improve Natural Language Commonsense Question Answering? 19 Sep 2019 · 0 repositories · arXiv:1909.08855
-
Reasoning Over Semantic-Level Graph for Fact Checking 9 Sep 2019 · 0 repositories · arXiv:1909.03745
-
Effective Use of Transformer Networks for Entity Tracking 5 Sep 2019 · 1 repository · arXiv:1909.02635Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Semantics-aware BERT for Language Understanding 5 Sep 2019 · 1 repository · arXiv:1909.02209Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 3 pointer-only (licence)
-
Adversarial Learning with Contextual Embeddings for Zero-resource Cross-lingual Classification and NER 31 Aug 2019 · 0 repositories · arXiv:1909.00153
-
Quantity doesn't buy quality syntax with neural language models 31 Aug 2019 · 0 repositories · arXiv:1909.00111