Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 19
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 19 of 20: papers 1,801 to 1,900 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Gaudí: Conversational Interactions with Deep Representations to Generate Image Collections 5 Dec 2021 · 0 repositories · arXiv:2112.04404
-
Searching for Efficient Transformers for Language Modeling 1 Dec 2021 · 0 repositories
-
Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer 1 Dec 2021 · 1 repository
-
Domain Prompt Learning for Efficiently Adapting CLIP to Unseen Domains 25 Nov 2021 · 1 repository · arXiv:2111.12853Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Active Dialogue Simulation in Conversational Systems 16 Nov 2021 · 0 repositories
-
Data Augmentation for Intent Classification with Generic Large Language Models 16 Nov 2021 · 0 repositories
-
ElitePLM: An Empirical Study on General Language Ability Evaluation of Pretrained Language Models 16 Nov 2021 · 0 repositories
-
Generative Pre-Trained Transformer for Design Concept Generation: An Exploration 16 Nov 2021 · 0 repositories · arXiv:2111.08489
-
Improving GPT-3 after deployment with a dynamic memory of feedback 16 Nov 2021 · 0 repositories
-
On the Multilingual Capabilities of Very Large-Scale English Language Models 16 Nov 2021 · 0 repositories
-
The Power of Prompt Tuning for Low-Resource Semantic Parsing 16 Nov 2021 · 0 repositories
-
Towards Coding Social Science Datasets with Language Models 16 Nov 2021 · 0 repositories
-
Scaling Law for Recommendation Models: Towards General-purpose User Representations 15 Nov 2021 · 0 repositories · arXiv:2111.11294
-
Amazon SageMaker Model Parallelism: A General and Flexible Framework for Large Model Training 10 Nov 2021 · 0 repositories · arXiv:2111.05972
-
An Explanation of In-context Learning as Implicit Bayesian Inference 3 Nov 2021 · 1 repository · arXiv:2111.02080Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TheEyeCorpus: Experiments in Reducing NLP Bias and Identifiability for Large LMs 3 Nov 2021 · 0 repositories
-
Risks of AI Foundation Models in Education 19 Oct 2021 · 0 repositories · arXiv:2110.10024
-
Knowledge Inheritance for Pre-trained Language Models 16 Oct 2021 · 0 repositories
-
Sharpness-Aware Minimization Improves Language Model Generalization 16 Oct 2021 · 0 repositories · arXiv:2110.08529
-
The Power of Prompt Tuning for Low-Resource Semantic Parsing 16 Oct 2021 · 0 repositories · arXiv:2110.08525
-
Can Machines Learn Morality? The Delphi Experiment 14 Oct 2021 · 1 repository · arXiv:2110.07574
-
Symbolic Knowledge Distillation: from General Language Models to Commonsense Models 14 Oct 2021 · 1 repository · arXiv:2110.07178
-
Scaling Laws for the Few-Shot Adaptation of Pre-trained Image Classifiers 13 Oct 2021 · 0 repositories · arXiv:2110.06990
-
LiST: Lite Prompted Self-training Makes Parameter-Efficient Few-shot Learners 12 Oct 2021 · 1 repository · arXiv:2110.06274Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 2 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Yuan 1.0: Large-Scale Pre-trained Language Model in Zero-Shot and Few-Shot Learning 10 Oct 2021 · 1 repository · arXiv:2110.04725
-
M6-10T: A Sharing-Delinking Paradigm for Efficient Multi-Trillion Parameter Pretraining 8 Oct 2021 · 0 repositories · arXiv:2110.03888
-
Leveraging the Inductive Bias of Large Language Models for Abstract Textual Reasoning 5 Oct 2021 · 0 repositories · arXiv:2110.02370
-
Collaborative Storytelling with Human Actors and AI Narrators 29 Sep 2021 · 0 repositories · arXiv:2109.14728
-
RAFT: A Real-World Few-Shot Text Classification Benchmark 28 Sep 2021 · 1 repository · arXiv:2109.14076
-
TURINGBENCH: A Benchmark Environment for Turing Test in the Age of Neural Text Generation 27 Sep 2021 · 3 repositories · arXiv:2109.13296Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Recursively Summarizing Books with Human Feedback 22 Sep 2021 · 0 repositories · arXiv:2109.10862
-
Towards Zero-Label Language Learning 19 Sep 2021 · 0 repositories · arXiv:2109.09193
-
Primer: Searching for Efficient Transformers for Language Modeling 17 Sep 2021 · 4 repositories · arXiv:2109.08668Syntology official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 2 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
An Empirical Study of GPT-3 for Few-Shot Knowledge-Based VQA 10 Sep 2021 · 1 repository · arXiv:2109.05014Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
What Changes Can Large-scale Language Models Bring? Intensive Study on HyperCLOVA: Billions-scale Korean Generative Pretrained Transformers 10 Sep 2021 · 2 repositories · arXiv:2109.04650Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Medically Aware GPT-3 as a Data Generator for Medical Dialogue Summarization 9 Sep 2021 · 0 repositories · arXiv:2110.07356
-
TruthfulQA: Measuring How Models Mimic Human Falsehoods 8 Sep 2021 · 3 repositories · arXiv:2109.07958Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
General-Purpose Question-Answering with Macaw 6 Sep 2021 · 2 repositories · arXiv:2109.02593
-
GPT-3 Models are Poor Few-Shot Learners in the Biomedical Domain 6 Sep 2021 · 1 repository · arXiv:2109.02555
-
Finetuned Language Models Are Zero-Shot Learners 3 Sep 2021 · 8 repositories · arXiv:2109.01652Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
So Cloze yet so Far: N400 Amplitude is Better Predicted by Distributional Information than Human Predictability Judgements 2 Sep 2021 · 0 repositories · arXiv:2109.01226
-
MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics 31 Aug 2021 · 4 repositories · arXiv:2109.00110Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
On the Multilingual Capabilities of Very Large-Scale English Language Models 30 Aug 2021 · 2 repositories · arXiv:2108.13349
-
Want To Reduce Labeling Cost? GPT-3 Can Help 30 Aug 2021 · 1 repository · arXiv:2108.13487
-
CGEMs: A Metric Model for Automatic Code Generation using GPT-3 23 Aug 2021 · 0 repositories · arXiv:2108.10168
-
Random Offset Block Embedding Array (ROBE) for CriteoTB Benchmark MLPerf DLRM Model : 1000× Compression and 3.1× Faster Inference 4 Aug 2021 · 0 repositories · arXiv:2108.02191
-
Q-Pain: A Question Answering Dataset to Measure Social Bias in Pain Management 3 Aug 2021 · 0 repositories · arXiv:2108.01764
-
BERTAC: Enhancing Transformer-based Language Models with Adversarially Pretrained Convolutional Neural Networks 1 Aug 2021 · 1 repository
-
KuiLeiXi: a Chinese Open-Ended Text Adventure Game 1 Aug 2021 · 0 repositories
-
Evaluating Large Language Models Trained on Code 7 Jul 2021 · 13 repositories · arXiv:2107.03374Syntology official (archive's flag): 2 ran · 26 ran (of which 0 constructed an object rather than computing a result; 24 with no instrument failure: 1 honoured, 0 violated, 23 with no contract checked; 2 where Syntology's instrument failed) · 13 unverified (of 39 harvested samples) · 4 pointer-only (licence)
-
Not Quite 'Ask a Librarian': AI on the Nature, Value, and Future of LIS 7 Jul 2021 · 0 repositories · arXiv:2107.05383
-
ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation 5 Jul 2021 · 2 repositories · arXiv:2107.02137
-
Is GPT-3 Text Indistinguishable from Human Text? Scarecrow: A Framework for Scrutinizing Machine Text 2 Jul 2021 · 0 repositories · arXiv:2107.01294
-
What's in a Measurement? Using GPT-3 on SemEval 2021 Task 8 -- MeasEval 28 Jun 2021 · 0 repositories · arXiv:2106.14720
-
Process for Adapting Language Models to Society (PALMS) with Values-Targeted Datasets 18 Jun 2021 · 0 repositories · arXiv:2106.10328
-
LoRA: Low-Rank Adaptation of Large Language Models 17 Jun 2021 · 74 repositories · arXiv:2106.09685Syntology community repositories only · 51 ran (of which 19 constructed an object rather than computing a result; 44 with no instrument failure: 1 honoured, 0 violated, 43 with no contract checked; 7 where Syntology's instrument failed) · 33 unverified (of 84 harvested samples) · 30 pointer-only (licence)
-
GPT3-to-plan: Extracting plans from text using GPT-3 14 Jun 2021 · 1 repository · arXiv:2106.07131
-
Programming Puzzles 10 Jun 2021 · 3 repositories · arXiv:2106.05784
-
TIMEDIAL: Temporal Commonsense Reasoning in Dialog 8 Jun 2021 · 1 repository · arXiv:2106.04571
-
GPT Perdetry Test: Generating new meanings for new words 1 Jun 2021 · 0 repositories
-
Knowledge Inheritance for Pre-trained Language Models 28 May 2021 · 2 repositories · arXiv:2105.13880
-
RetGen: A Joint framework for Retrieval and Grounded Text Generation Modeling 14 May 2021 · 1 repository · arXiv:2105.06597
-
DExperts: Decoding-Time Controlled Text Generation with Experts and Anti-Experts 7 May 2021 · 1 repository · arXiv:2105.03023Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
One Model to Rule them All: Towards Zero-Shot Learning for Databases 3 May 2021 · 0 repositories · arXiv:2105.00642
-
Unreasonable Effectiveness of Rule-Based Heuristics in Solving Russian SuperGLUE Tasks 3 May 2021 · 0 repositories · arXiv:2105.01192
-
Entailment as Few-Shot Learner 29 Apr 2021 · 3 repositories · arXiv:2104.14690Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
PanGu-α: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation 26 Apr 2021 · 5 repositories · arXiv:2104.12369
-
A Token-level Reference-free Hallucination Detection Benchmark for Free-form Text Generation 18 Apr 2021 · 2 repositories · arXiv:2104.08704Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity 18 Apr 2021 · 2 repositories · arXiv:2104.08786Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples)
-
GPT3Mix: Leveraging Large-scale Language Models for Text Augmentation 18 Apr 2021 · 1 repository · arXiv:2104.08826Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Cross-Task Generalization via Natural Language Crowdsourcing Instructions 18 Apr 2021 · 3 repositories · arXiv:2104.08773Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
The Power of Scale for Parameter-Efficient Prompt Tuning 18 Apr 2021 · 12 repositories · arXiv:2104.08691Syntology community repositories only · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
An Adversarially-Learned Turing Test for Dialog Generation Models 16 Apr 2021 · 1 repository · arXiv:2104.08231
-
Surface Form Competition: Why the Highest Probability Answer Isn't Always Right 16 Apr 2021 · 2 repositories · arXiv:2104.08315Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Text2App: A Framework for Creating Android Apps from Text Descriptions 16 Apr 2021 · 2 repositories · arXiv:2104.08301
-
Adapting Language Models for Zero-shot Learning by Meta-tuning on Dataset and Prompt Collections 10 Apr 2021 · 1 repository · arXiv:2104.04670
-
Russian Paraphrasers: Paraphrase with Transformers 1 Apr 2021 · 2 repositories
-
Automatic Graph Partitioning for Very Large-scale Deep Learning 30 Mar 2021 · 0 repositories · arXiv:2103.16063
-
Detecting Hate Speech with GPT-3 23 Mar 2021 · 2 repositories · arXiv:2103.12407
-
Calibrate Before Use: Improving Few-Shot Performance of Language Models 19 Feb 2021 · 5 repositories · arXiv:2102.09690Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples)
-
TeraPipe: Token-Level Pipeline Parallelism for Training Large-Scale Language Models 16 Feb 2021 · 1 repository · arXiv:2102.07988Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Prompt Programming for Large Language Models: Beyond the Few-Shot Paradigm 15 Feb 2021 · 0 repositories · arXiv:2102.07350
-
Multiversal views on language models 12 Feb 2021 · 0 repositories · arXiv:2102.06391
-
PipeTransformer: Automated Elastic Pipelining for Distributed Training of Transformers 5 Feb 2021 · 1 repository · arXiv:2102.03161Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Understanding Emails and Drafting Responses -- An Approach Using GPT-3 5 Feb 2021 · 0 repositories · arXiv:2102.03062
-
Understanding the Capabilities, Limitations, and Societal Impact of Large Language Models 4 Feb 2021 · 0 repositories · arXiv:2102.02503
-
"Is depression related to cannabis?": A knowledge-infused model for Entity and Relation Extraction with Limited Supervision 1 Feb 2021 · 0 repositories · arXiv:2102.01222
-
Persistent Anti-Muslim Bias in Large Language Models 14 Jan 2021 · 1 repository · arXiv:2101.05783
-
How Multipurpose Are Language Models? 1 Jan 2021 · 0 repositories
-
WARP: Word-level Adversarial ReProgramming 1 Jan 2021 · 1 repository · arXiv:2101.00121
-
Making Pre-trained Language Models Better Few-shot Learners 31 Dec 2020 · 9 repositories · arXiv:2012.15723Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 9 harvested samples) · 7 pointer-only (licence)
-
The Pile: An 800GB Dataset of Diverse Text for Language Modeling 31 Dec 2020 · 22 repositories · arXiv:2101.00027Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Revisiting Linformer with a modified self-attention with linear complexity 16 Dec 2020 · 0 repositories · arXiv:2101.10277
-
Hardware Beyond Backpropagation: a Photonic Co-Processor for Direct Feedback Alignment 11 Dec 2020 · 0 repositories · arXiv:2012.06373
-
CPM: A Large-scale Generative Chinese Pre-trained Language Model 1 Dec 2020 · 10 repositories · arXiv:2012.00413
-
Increasing Learning Efficiency of Self-Attention Networks through Direct Position Interactions, Learnable Temperature, and Convoluted Attention 1 Dec 2020 · 1 repository
-
Do Fine-tuned Commonsense Language Models Really Generalize? 18 Nov 2020 · 0 repositories · arXiv:2011.09159
-
COMET-ATOMIC 2020: On Symbolic and Neural Commonsense Knowledge Graphs 12 Oct 2020 · 3 repositories · arXiv:2010.05953Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Toward a Thermodynamics of Meaning 24 Sep 2020 · 1 repository · arXiv:2009.11963
-
It's Not Just Size That Matters: Small Language Models Are Also Few-Shot Learners 15 Sep 2020 · 5 repositories · arXiv:2009.07118