Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 11
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 11 of 20: papers 1,001 to 1,100 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Entity Matching using Large Language Models 17 Oct 2023 · 1 repository · arXiv:2310.11244
-
LLMs as Hackers: Autonomous Linux Privilege Escalation Attacks 17 Oct 2023 · 1 repository · arXiv:2310.11409
-
Intent Detection and Slot Filling for Home Assistants: Dataset and Analysis for Bangla and Sylheti 17 Oct 2023 · 1 repository · arXiv:2310.10935
-
Probing the Creativity of Large Language Models: Can models produce divergent semantic association? 17 Oct 2023 · 1 repository · arXiv:2310.11158Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Utilising a Large Language Model to Annotate Subject Metadata: A Case Study in an Australian National Research Data Catalogue 17 Oct 2023 · 0 repositories · arXiv:2310.11318
-
Battle of the Large Language Models: Dolly vs LLaMA vs Vicuna vs Guanaco vs Bard vs ChatGPT -- A Text-to-SQL Parsing Comparison 16 Oct 2023 · 0 repositories · arXiv:2310.10190
-
BioPlanner: Automatic Evaluation of LLMs on Protocol Planning in Biology 16 Oct 2023 · 1 repository · arXiv:2310.10632Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Fine-tuning ChatGPT for Automatic Scoring 16 Oct 2023 · 0 repositories · arXiv:2310.10072
-
Prediction of Arabic Legal Rulings using Large Language Models 16 Oct 2023 · 0 repositories · arXiv:2310.10260
-
TRANSOM: An Efficient Fault-Tolerant System for Training LLMs 16 Oct 2023 · 1 repository · arXiv:2310.10046Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
Large Language Model-Aware In-Context Learning for Code Generation 15 Oct 2023 · 0 repositories · arXiv:2310.09748
-
Large Language Models for In-Context Student Modeling: Synthesizing Student's Behavior in Visual Programming 15 Oct 2023 · 1 repository · arXiv:2310.10690
-
Assessing and Enhancing the Robustness of Large Language Models with Task Structure Variations for Logical Reasoning 13 Oct 2023 · 1 repository · arXiv:2310.09430Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Human-in-the-loop Machine Translation with Large Language Model 13 Oct 2023 · 1 repository · arXiv:2310.08908
-
Table-GPT: Table-tuned GPT for Diverse Table Tasks 13 Oct 2023 · 0 repositories · arXiv:2310.09263
-
Jailbreaking Black Box Large Language Models in Twenty Queries 12 Oct 2023 · 1 repository · arXiv:2310.08419Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 7 harvested samples)
-
Large language models can replicate cross-cultural differences in personality 12 Oct 2023 · 0 repositories · arXiv:2310.10679
-
Promptor: A Conversational and Autonomous Prompt Generation Agent for Intelligent Text Entry Techniques 12 Oct 2023 · 0 repositories · arXiv:2310.08101
-
QASiNa: Religious Domain Question Answering using Sirah Nabawiyah 12 Oct 2023 · 1 repository · arXiv:2310.08102
-
Diversity of Thought Improves Reasoning Abilities of LLMs 11 Oct 2023 · 0 repositories · arXiv:2310.07088
-
Do Large Language Models have Shared Weaknesses in Medical Question Answering? 11 Oct 2023 · 0 repositories · arXiv:2310.07225
-
Found in the Middle: Permutation Self-Consistency Improves Listwise Ranking in Large Language Models 11 Oct 2023 · 1 repository · arXiv:2310.07712Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Large Language Models Are Zero-Shot Time Series Forecasters 11 Oct 2023 · 2 repositories · arXiv:2310.07820Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Diffusion Models for Wireless Communications 11 Oct 2023 · 0 repositories · arXiv:2310.07312
-
Automated clinical coding using off-the-shelf large language models 10 Oct 2023 · 0 repositories · arXiv:2310.06552
-
GeoLLM: Extracting Geospatial Knowledge from Large Language Models 10 Oct 2023 · 1 repository · arXiv:2310.06213Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Large Language Models for Propaganda Detection 10 Oct 2023 · 2 repositories · arXiv:2310.06422
-
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression 10 Oct 2023 · 3 repositories · arXiv:2310.06839
-
Cabbage Sweeter than Cake? Analysing the Potential of Large Language Models for Learning Conceptual Spaces 9 Oct 2023 · 0 repositories · arXiv:2310.05481
-
Exploring the Maze of Multilingual Modeling 9 Oct 2023 · 0 repositories · arXiv:2310.05404
-
SC-Safety: A Multi-round Open-ended Question Adversarial Safety Benchmark for Large Language Models in Chinese 9 Oct 2023 · 0 repositories · arXiv:2310.05818
-
The Program Testing Ability of Large Language Models for Code 9 Oct 2023 · 0 repositories · arXiv:2310.05727
-
Are Emily and Greg Still More Employable than Lakisha and Jamal? Investigating Algorithmic Hiring Bias in the Era of ChatGPT 8 Oct 2023 · 0 repositories · arXiv:2310.05135
-
LLM4VV: Developing LLM-Driven Testsuite for Compiler Validation 8 Oct 2023 · 1 repository · arXiv:2310.04963
-
Zero-Shot Detection of Machine-Generated Codes 8 Oct 2023 · 1 repository · arXiv:2310.05103
-
Large Language Models Only Pass Primary School Exams in Indonesia: A Comprehensive Test on IndoMMLU 7 Oct 2023 · 1 repository · arXiv:2310.04928Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models 6 Oct 2023 · 2 repositories · arXiv:2310.04406Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Agent Instructs Large Language Models to be General Zero-Shot Reasoners 5 Oct 2023 · 1 repository · arXiv:2310.03710Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples)
-
Automating Human Tutor-Style Programming Feedback: Leveraging GPT-4 Tutor Model for Hint Generation and GPT-3.5 Student Model for Hint Validation 5 Oct 2023 · 2 repositories · arXiv:2310.03780Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines 5 Oct 2023 · 3 repositories · arXiv:2310.03714Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To! 5 Oct 2023 · 1 repository · arXiv:2310.03693Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
A Survey of GPT-3 Family Large Language Models Including ChatGPT and GPT-4 4 Oct 2023 · 0 repositories · arXiv:2310.12321
-
Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning 4 Oct 2023 · 1 repository · arXiv:2310.03094
-
Retrieval meets Long Context Large Language Models 4 Oct 2023 · 0 repositories · arXiv:2310.03025
-
Instances Need More Care: Rewriting Prompts for Instances with LLMs in the Loop Yields Better Zero-Shot Performance 3 Oct 2023 · 1 repository · arXiv:2310.02107Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
GPT-Driver: Learning to Drive with GPT 2 Oct 2023 · 1 repository · arXiv:2310.01415Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
LLM Lies: Hallucinations are not Bugs, but Features as Adversarial Examples 2 Oct 2023 · 1 repository · arXiv:2310.01469Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
BooookScore: A systematic exploration of book-length summarization in the era of LLMs 1 Oct 2023 · 2 repositories · arXiv:2310.00785Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Gaze-Driven Sentence Simplification for Language Learners: Enhancing Comprehension and Readability 30 Sep 2023 · 0 repositories · arXiv:2310.00355
-
A Large Language Model Approach to Educational Survey Feedback Analysis 29 Sep 2023 · 0 repositories · arXiv:2309.17447
-
Benchmarking the Abilities of Large Language Models for RDF Knowledge Graph Creation and Comprehension: How Well Do LLMs Speak Turtle? 29 Sep 2023 · 3 repositories · arXiv:2309.17122
-
DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks 29 Sep 2023 · 1 repository · arXiv:2309.17167
-
Cooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation 29 Sep 2023 · 2 repositories · arXiv:2309.17234Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Split and Merge: Aligning Position Biases in LLM-based Evaluators 29 Sep 2023 · 0 repositories · arXiv:2310.01432
-
AE-GPT: Using Large Language Models to Extract Adverse Events from Surveillance Reports-A Use Case with Influenza Vaccine Adverse Events 28 Sep 2023 · 0 repositories · arXiv:2309.16150
-
GPT-Fathom: Benchmarking Large Language Models to Decipher the Evolutionary Path towards GPT-4 and Beyond 28 Sep 2023 · 1 repository · arXiv:2309.16583Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Stress Testing Chain-of-Thought Prompting for Large Language Models 28 Sep 2023 · 0 repositories · arXiv:2309.16621
-
NLPBench: Evaluating Large Language Models on Solving NLP Problems 27 Sep 2023 · 1 repository · arXiv:2309.15630
-
Exploring Small Language Models with Prompt-Learning Paradigm for Efficient Domain-Specific Text Classification 26 Sep 2023 · 0 repositories · arXiv:2309.14779
-
How to Catch an AI Liar: Lie Detection in Black-Box LLMs by Asking Unrelated Questions 26 Sep 2023 · 1 repository · arXiv:2309.15840Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
RankVicuna: Zero-Shot Listwise Document Reranking with Open-Source Large Language Models 26 Sep 2023 · 3 repositories · arXiv:2309.15088
-
Supersonic: Learning to Generate Source Code Optimizations in C/C++ 26 Sep 2023 · 1 repository · arXiv:2309.14846
-
Evaluating Cognitive Maps and Planning in Large Language Models with CogEval 25 Sep 2023 · 0 repositories · arXiv:2309.15129
-
Watch Your Language: Investigating Content Moderation with Large Language Models 25 Sep 2023 · 0 repositories · arXiv:2309.14517
-
Does the "most sinfully decadent cake ever" taste good? Answering Yes/No Questions from Figurative Contexts 24 Sep 2023 · 0 repositories · arXiv:2309.13748
-
A Chat About Boring Problems: Studying GPT-based text normalization 23 Sep 2023 · 0 repositories · arXiv:2309.13426
-
Probing the Moral Development of Large Language Models through Defining Issues Test 23 Sep 2023 · 0 repositories · arXiv:2309.13356
-
BenLLMEval: A Comprehensive Evaluation into the Potentials and Pitfalls of Large Language Models on Bengali NLP 22 Sep 2023 · 0 repositories · arXiv:2309.13173
-
Contextual Emotion Estimation from Image Captions 22 Sep 2023 · 0 repositories · arXiv:2309.13136
-
Large Language Models Are Also Good Prototypical Commonsense Reasoners 22 Sep 2023 · 0 repositories · arXiv:2309.13165
-
Goal-Oriented Prompt Attack and Safety Evaluation for LLMs 21 Sep 2023 · 2 repositories · arXiv:2309.11830
-
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models 21 Sep 2023 · 1 repository · arXiv:2309.12284Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 14 where Syntology's instrument failed) · 7 unverified (of 22 harvested samples)
-
TART: A plug-and-play Transformer module for task-agnostic reasoning 21 Sep 2023 · 1 repository
-
The Cambridge Law Corpus: A Dataset for Legal AI Research 21 Sep 2023 · 0 repositories · arXiv:2309.12269
-
The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A" 21 Sep 2023 · 2 repositories · arXiv:2309.12288Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
TOA: Task-oriented Active VQA 21 Sep 2023 · 0 repositories
-
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11674Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Controlled Generation with Prompt Insertion for Natural Language Explanations in Grammatical Error Correction 20 Sep 2023 · 1 repository · arXiv:2309.11439
-
Design of Chain-of-Thought in Math Problem Solving 20 Sep 2023 · 1 repository · arXiv:2309.11054
-
Fictional Worlds, Real Connections: Developing Community Storytelling Social Chatbots through LLMs 20 Sep 2023 · 0 repositories · arXiv:2309.11478
-
Generative AI in Mafia-like Game Simulation 20 Sep 2023 · 0 repositories · arXiv:2309.11672
-
Safurai 001: New Qualitative Approach for Code LLM Evaluation 20 Sep 2023 · 1 repository · arXiv:2309.11385
-
Language as the Medium: Multimodal Video Classification through text only 19 Sep 2023 · 0 repositories · arXiv:2309.10783
-
Writer-Defined AI Personas for On-Demand Feedback Generation 19 Sep 2023 · 0 repositories · arXiv:2309.10433
-
Evaluation of GPT-3 for Anti-Cancer Drug Sensitivity Prediction 18 Sep 2023 · 0 repositories · arXiv:2309.10016
-
Contrastive Decoding Improves Reasoning in Large Language Models 17 Sep 2023 · 0 repositories · arXiv:2309.09117
-
Do Large GPT Models Discover Moral Dimensions in Language Representations? A Topological Study Of Sentence Embeddings 17 Sep 2023 · 0 repositories · arXiv:2309.09397
-
From Cooking Recipes to Robot Task Trees -- Improving Planning Correctness and Task Efficiency by Leveraging LLMs with a Knowledge Network 17 Sep 2023 · 0 repositories · arXiv:2309.09181
-
Decoder-only Architecture for Speech Recognition with CTC Prompts and Text Data Augmentation 16 Sep 2023 · 0 repositories · arXiv:2309.08876
-
Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 16 Sep 2023 · 1 repository · arXiv:2309.08963
-
Advancing the Evaluation of Traditional Chinese Language Models: Towards a Comprehensive Benchmark Suite 15 Sep 2023 · 1 repository · arXiv:2309.08448
-
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models 15 Sep 2023 · 1 repository · arXiv:2309.08573
-
EvoPrompt: Connecting LLMs with Evolutionary Algorithms Yields Powerful Prompt Optimizers 15 Sep 2023 · 2 repositories · arXiv:2309.08532Syntology 11 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer 15 Sep 2023 · 1 repository · arXiv:2309.08583
-
Large Language Models for Failure Mode Classification: An Investigation 15 Sep 2023 · 1 repository · arXiv:2309.08181
-
An Empirical Evaluation of Prompting Strategies for Large Language Models in Zero-Shot Clinical Natural Language Processing 14 Sep 2023 · 0 repositories · arXiv:2309.08008
-
Two Timin': Repairing Smart Contracts With A Two-Layered Approach 14 Sep 2023 · 0 repositories · arXiv:2309.07841
-
Large Language Models Can Infer Psychological Dispositions of Social Media Users 13 Sep 2023 · 0 repositories · arXiv:2309.08631
-
Comparing Llama-2 and GPT-3 LLMs for HPC kernels generation 12 Sep 2023 · 0 repositories · arXiv:2309.07103
-
Exploring Large Language Models for Ontology Alignment 12 Sep 2023 · 1 repository · arXiv:2309.07172