Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 16
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 16 of 20: papers 1,501 to 1,600 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Automated Interactive Domain-Specific Conversational Agents that Understand Human Dialogs 15 Mar 2023 · 0 repositories · arXiv:2303.08941
-
SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models 15 Mar 2023 · 1 repository · arXiv:2303.08896Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family 14 Mar 2023 · 2 repositories · arXiv:2303.07992
-
RE-MOVE: An Adaptive Policy Design for Robotic Navigation Tasks in Dynamic Environments via Language-Based Feedback 14 Mar 2023 · 0 repositories · arXiv:2303.07622
-
Large Language Models in the Workplace: A Case Study on Prompt Engineering for Job Type Classification 13 Mar 2023 · 0 repositories · arXiv:2303.07142
-
Large Language Models Know Your Contextual Search Intent: A Prompting Framework for Conversational Search 12 Mar 2023 · 2 repositories · arXiv:2303.06573
-
ChatGPT may Pass the Bar Exam soon, but has a Long Way to Go for the LexGLUE benchmark 9 Mar 2023 · 1 repository · arXiv:2304.12202
-
ICL-D3IE: In-Context Learning with Diverse Demonstrations Updating for Document Information Extraction 9 Mar 2023 · 1 repository · arXiv:2303.05063
-
ChatGPT Participates in a Computer Science Exam 8 Mar 2023 · 1 repository · arXiv:2303.09461
-
Cost-Effective Hyperparameter Optimization for Large Language Model Generation Inference 8 Mar 2023 · 3 repositories · arXiv:2303.04673Syntology community repositories only · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Stealing the Decoding Algorithms of Language Models 8 Mar 2023 · 1 repository · arXiv:2303.04729
-
A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT 7 Mar 2023 · 1 repository · arXiv:2303.04226
-
Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot Learners 3 Mar 2023 · 3 repositories · arXiv:2303.02151Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
Prophet: Prompting Large Language Models with Complementary Answer Heuristics for Knowledge-based Visual Question Answering 3 Mar 2023 · 1 repository · arXiv:2303.01903Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
WiCE: Real-World Entailment for Claims in Wikipedia 2 Mar 2023 · 2 repositories · arXiv:2303.01432Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Framework for Neurosymbolic Robot Action Planning using Large Language Models 1 Mar 2023 · 1 repository · arXiv:2303.00438Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
How Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding Tasks 1 Mar 2023 · 0 repositories · arXiv:2303.00293
-
ToxVis: Enabling Interpretability of Implicit vs. Explicit Toxicity Detection Models with Interactive Visualization 1 Mar 2023 · 0 repositories · arXiv:2303.09402
-
Zero-Shot Cross-Lingual Summarization via Large Language Models 28 Feb 2023 · 0 repositories · arXiv:2302.14229
-
LLaMA: Open and Efficient Foundation Language Models 27 Feb 2023 · 57 repositories · arXiv:2302.13971Syntology official: no sample here; runs from other or unrecorded repositories · 37 ran (of which 9 constructed an object rather than computing a result; 25 with no instrument failure: 3 honoured, 0 violated, 22 with no contract checked; 12 where Syntology's instrument failed) · 21 unverified (of 58 harvested samples) · 4 pointer-only (licence)
-
Reward Design with Language Models 27 Feb 2023 · 1 repository · arXiv:2303.00001
-
Systematic Rectification of Language Models via Dead-end Analysis 27 Feb 2023 · 1 repository · arXiv:2302.14003Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Comparing Sentence-Level Suggestions to Message-Level Suggestions in AI-Mediated Communication 26 Feb 2023 · 0 repositories · arXiv:2302.13382
-
Fast Attention Requires Bounded Entries 26 Feb 2023 · 0 repositories · arXiv:2302.13214
-
Human-in-the-Loop Schema Induction 25 Feb 2023 · 0 repositories · arXiv:2302.13048
-
Spanish Built Factual Freectianary (Spanish-BFF): the first AI-generated free dictionary 24 Feb 2023 · 0 repositories · arXiv:2302.12746
-
Testing AI on language comprehension tasks reveals insensitivity to underlying meaning 23 Feb 2023 · 0 repositories · arXiv:2302.12313
-
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure 23 Feb 2023 · 1 repository · arXiv:2302.12239Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
kNN-Adapter: Efficient Domain Adaptation for Black-Box Language Models 21 Feb 2023 · 0 repositories · arXiv:2302.10879
-
ChatIE: Zero-Shot Information Extraction via Chatting with ChatGPT 20 Feb 2023 · 1 repository · arXiv:2302.10205
-
A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT 18 Feb 2023 · 0 repositories · arXiv:2302.09419
-
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints 17 Feb 2023 · 1 repository · arXiv:2302.09185
-
GPT4MIA: Utilizing Generative Pre-trained Transformer (GPT-3) as A Plug-and-Play Transductive Model for Medical Image Analysis 17 Feb 2023 · 0 repositories · arXiv:2302.08722
-
Prompting Large Language Models With the Socratic Method 17 Feb 2023 · 0 repositories · arXiv:2303.08769
-
For Generated Text, Is NLI-Neutral Text the Best Text? 16 Feb 2023 · 1 repository · arXiv:2302.08577
-
Learning Performance-Improving Code Edits 15 Feb 2023 · 2 repositories · arXiv:2302.07867Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 11 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
ScatterShot: Interactive In-context Example Curation for Text Transformation 14 Feb 2023 · 1 repository · arXiv:2302.07346
-
Can GPT-3 Perform Statutory Reasoning? 13 Feb 2023 · 1 repository · arXiv:2302.06100
-
STREET: A Multi-Task Structured Reasoning and Explanation Benchmark 13 Feb 2023 · 0 repositories · arXiv:2302.06729
-
A Brief Report on LawGPT 1.0: A Virtual Legal Assistant Based on GPT-3 11 Feb 2023 · 0 repositories · arXiv:2302.05729
-
Reliable Natural Language Understanding with Large Language Models and Answer Set Programming 7 Feb 2023 · 0 repositories · arXiv:2302.03780
-
What Matters In The Structured Pruning of Generative Language Models? 7 Feb 2023 · 1 repository · arXiv:2302.03773Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 8 harvested samples)
-
Evaluating Large Language Models in Theory of Mind Tasks 4 Feb 2023 · 0 repositories · arXiv:2302.02083
-
Creating a Large Language Model of a Philosopher 2 Feb 2023 · 0 repositories · arXiv:2302.01339
-
Co-Writing with Opinionated Language Models Affects Users' Views 1 Feb 2023 · 0 repositories · arXiv:2302.00560
-
Improving Few-Shot Generalization by Exploring and Exploiting Auxiliary Data 1 Feb 2023 · 1 repository · arXiv:2302.00674
-
Numeracy from Literacy: Data Science as an Emergent Skill from Large Language Models 31 Jan 2023 · 0 repositories · arXiv:2301.13382
-
Adaptive Machine Translation with Large Language Models 30 Jan 2023 · 1 repository · arXiv:2301.13294
-
REPLUG: Retrieval-Augmented Black-Box Language Models 30 Jan 2023 · 3 repositories · arXiv:2301.12652Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Specializing Smaller Language Models towards Multi-Step Reasoning 30 Jan 2023 · 2 repositories · arXiv:2301.12726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A Discerning Several Thousand Judgments: GPT-3 Rates the Article + Adjective + Numeral + Noun Construction 29 Jan 2023 · 0 repositories · arXiv:2301.12564
-
Towards Equitable Representation in Text-to-Image Synthesis Models with the Cross-Cultural Understanding Benchmark (CCUB) Dataset 28 Jan 2023 · 1 repository · arXiv:2301.12073
-
ThoughtSource: A central hub for large language model reasoning data 27 Jan 2023 · 1 repository · arXiv:2301.11596Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 7 harvested samples)
-
Understanding the Effectiveness of Very Large Language Models on Dialog Evaluation 27 Jan 2023 · 0 repositories · arXiv:2301.12004
-
Causal Reasoning of Entities and Events in Procedural Texts 26 Jan 2023 · 1 repository · arXiv:2301.10896Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 15 harvested samples)
-
ExaRanker: Explanation-Augmented Neural Ranker 25 Jan 2023 · 1 repository · arXiv:2301.10521
-
The Next Chapter: A Study of Large Language Models in Storytelling 24 Jan 2023 · 0 repositories · arXiv:2301.09790
-
Large Language Models as Fiduciaries: A Case Study Toward Robustly Communicating With Artificial Intelligence Through Legal Standards 24 Jan 2023 · 0 repositories · arXiv:2301.10095
-
Large language models can segment narrative events similarly to humans 24 Jan 2023 · 0 repositories · arXiv:2301.10297
-
Multitask Instruction-based Prompting for Fallacy Recognition 24 Jan 2023 · 0 repositories · arXiv:2301.09992
-
AI model GPT-3 (dis)informs us better than humans 23 Jan 2023 · 0 repositories · arXiv:2301.11924
-
SuperScaler: Supporting Flexible DNN Parallelization via a Unified Abstraction 21 Jan 2023 · 0 repositories · arXiv:2301.08984
-
Is ChatGPT A Good Translator? Yes With GPT-4 As The Engine 20 Jan 2023 · 1 repository · arXiv:2301.08745
-
Batch Prompting: Efficient Inference with Large Language Model APIs 19 Jan 2023 · 2 repositories · arXiv:2301.08721
-
GPT as Knowledge Worker: A Zero-Shot Evaluation of (AI)CPA Capabilities 11 Jan 2023 · 1 repository · arXiv:2301.04408
-
Recommending Root-Cause and Mitigation Steps for Cloud Incidents using Large Language Models 10 Jan 2023 · 0 repositories · arXiv:2301.03797
-
Critical Perspectives: A Benchmark Revealing Pitfalls in PerspectiveAPI 5 Jan 2023 · 1 repository · arXiv:2301.01874
-
InPars-v2: Large Language Models as Efficient Dataset Generators for Information Retrieval 4 Jan 2023 · 1 repository · arXiv:2301.01820
-
UniHD at TSAR-2022 Shared Task: Is Compute All We Need for Lexical Simplification? 4 Jan 2023 · 1 repository · arXiv:2301.01764
-
Large Language Models as Corporate Lobbyists 3 Jan 2023 · 1 repository · arXiv:2301.01181
-
Fusing Pre-Trained Language Models With Multimodal Prompts Through Reinforcement Learning 1 Jan 2023 · 1 repository
-
PromptCap: Prompt-Guided Image Captioning for VQA with GPT-3 1 Jan 2023 · 0 repositories
-
Rethinking with Retrieval: Faithful Large Language Model Inference 31 Dec 2022 · 1 repository · arXiv:2301.00303
-
Targeted Phishing Campaigns using Large Scale Language Models 30 Dec 2022 · 0 repositories · arXiv:2301.00665
-
GPT Takes the Bar Exam 29 Dec 2022 · 5 repositories · arXiv:2212.14402
-
Maximizing Use-Case Specificity through Precision Model Tuning 29 Dec 2022 · 0 repositories · arXiv:2212.14206
-
DeepCuts: Single-Shot Interpretability based Pruning for BERT 27 Dec 2022 · 1 repository · arXiv:2212.13392
-
Using Large Language Models to Generate Engaging Captions for Data Visualizations 27 Dec 2022 · 0 repositories · arXiv:2212.14047
-
Biologically Inspired Design Concept Generation Using Generative Pre-Trained Transformers 26 Dec 2022 · 0 repositories · arXiv:2212.13196
-
JASMINE: Arabic GPT Models for Few-Shot Learning 21 Dec 2022 · 0 repositories · arXiv:2212.10755
-
Controllable Text Generation with Language Constraints 20 Dec 2022 · 0 repositories · arXiv:2212.10466
-
Do language models have coherent mental models of everyday things? 20 Dec 2022 · 1 repository · arXiv:2212.10029Syntology official: harvested, nothing ran · 0 ran · 5 unverified (of 5 harvested samples)
-
DocAsRef: An Empirical Study on Repurposing Reference-Based Summary Quality Metrics Reference-Freely 20 Dec 2022 · 1 repository · arXiv:2212.10013
-
Generic Temporal Reasoning with Differential Analysis and Explanation 20 Dec 2022 · 0 repositories · arXiv:2212.10467
-
Go-tuning: Improving Zero-shot Learning Abilities of Smaller Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10461
-
Is GPT-3 a Good Data Annotator? 20 Dec 2022 · 1 repository · arXiv:2212.10450
-
Evaluating Psychological Safety of Large Language Models 20 Dec 2022 · 0 repositories · arXiv:2212.10529
-
Large Language Models Are Reasoning Teachers 20 Dec 2022 · 1 repository · arXiv:2212.10071Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
PairReranker: Pairwise Reranking for Natural Language Generation 20 Dec 2022 · 0 repositories · arXiv:2212.10555
-
Pay Attention to Your Tone: Introducing a New Dataset for Polite Language Rewrite 20 Dec 2022 · 1 repository · arXiv:2212.10190
-
True Detective: A Deep Abductive Reasoning Benchmark Undoable for GPT-3 and Challenging for GPT-4 20 Dec 2022 · 0 repositories · arXiv:2212.10114
-
Emergent Analogical Reasoning in Large Language Models 19 Dec 2022 · 2 repositories · arXiv:2212.09196Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Evaluating Human-Language Model Interaction 19 Dec 2022 · 1 repository · arXiv:2212.09746
-
Large Language Models are Better Reasoners with Self-Verification 19 Dec 2022 · 1 repository · arXiv:2212.09561Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
LENS: A Learnable Evaluation Metric for Text Simplification 19 Dec 2022 · 1 repository · arXiv:2212.09739
-
Reasoning with Language Model Prompting: A Survey 19 Dec 2022 · 2 repositories · arXiv:2212.09597
-
Can Retriever-Augmented Language Models Reason? The Blame Game Between the Retriever and the Language Model 18 Dec 2022 · 1 repository · arXiv:2212.09146
-
Self-Prompting Large Language Models for Zero-Shot Open-Domain QA 16 Dec 2022 · 1 repository · arXiv:2212.08635
-
Revisiting the Gold Standard: Grounding Summarization Evaluation with Robust Human Evaluation 15 Dec 2022 · 2 repositories · arXiv:2212.07981Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
CREPE: Can Vision-Language Foundation Models Reason Compositionally? 13 Dec 2022 · 1 repository · arXiv:2212.07796Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)