Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 10
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 10 of 20: papers 901 to 1,000 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Evil Geniuses: Delving into the Safety of LLM-based Agents 20 Nov 2023 · 1 repository · arXiv:2311.11855Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Towards Human-Level Text Coding with LLMs: The Case of Fatherhood Roles in Public Policy Documents 20 Nov 2023 · 1 repository · arXiv:2311.11844
-
Refactoring Programs Using Large Language Models with Few-Shot Examples 20 Nov 2023 · 0 repositories · arXiv:2311.11690
-
Spot the Bot: Distinguishing Human-Written and Bot-Generated Texts Using Clustering and Information Theory Techniques 19 Nov 2023 · 0 repositories · arXiv:2311.11441
-
Behavior Optimized Image Generation 18 Nov 2023 · 0 repositories · arXiv:2311.10995
-
Advancements in Generative AI: A Comprehensive Review of GANs, GPT, Autoencoders, Diffusion Model, and Transformers 17 Nov 2023 · 0 repositories · arXiv:2311.10242
-
Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2 17 Nov 2023 · 3 repositories · arXiv:2311.10702
-
Fumbling in Babel: An Investigation into ChatGPT's Language Identification Ability 16 Nov 2023 · 0 repositories · arXiv:2311.09696
-
Generative AI for Hate Speech Detection: Evaluation and Findings 16 Nov 2023 · 0 repositories · arXiv:2311.09993
-
Human Still Wins over LLM: An Empirical Study of Active Learning on Domain-Specific Annotation Tasks 16 Nov 2023 · 0 repositories · arXiv:2311.09825
-
INTERVENOR: Prompting the Coding Ability of Large Language Models with the Interactive Chain of Repair 16 Nov 2023 · 1 repository · arXiv:2311.09868
-
FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains 16 Nov 2023 · 1 repository · arXiv:2311.09797Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
On Retrieval Augmentation and the Limitations of Language Model Training 16 Nov 2023 · 0 repositories · arXiv:2311.09615
-
Reducing Privacy Risks in Online Self-Disclosures with Language Models 16 Nov 2023 · 0 repositories · arXiv:2311.09538
-
Can Large Language Models Follow Concept Annotation Guidelines? A Case Study on Scientific and Financial Domains 15 Nov 2023 · 1 repository · arXiv:2311.08704
-
Evaluating Gender Bias in the Translation of Gender-Neutral Languages into English 15 Nov 2023 · 0 repositories · arXiv:2311.08836
-
ToolTalk: Evaluating Tool-Usage in a Conversational Setting 15 Nov 2023 · 1 repository · arXiv:2311.10775Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
"We Demand Justice!": Towards Social Context Grounding of Political Texts 15 Nov 2023 · 1 repository · arXiv:2311.09106
-
CPopQA: Ranking Cultural Concept Popularity by LLMs 14 Nov 2023 · 0 repositories · arXiv:2311.07897
-
Evaluating LLMs on Document-Based QA: Exact Answer Selection and Numerical Extraction using Cogtale dataset 14 Nov 2023 · 0 repositories · arXiv:2311.07878
-
Language Models are Better Bug Detector Through Code-Pair Classification 14 Nov 2023 · 1 repository · arXiv:2311.07957
-
Do large language models and humans have similar behaviors in causal inference with script knowledge? 13 Nov 2023 · 1 repository · arXiv:2311.07311
-
It's Not Easy Being Wrong: Large Language Models Struggle with Process of Elimination Reasoning 13 Nov 2023 · 1 repository · arXiv:2311.07532
-
MEGAVERSE: Benchmarking Large Language Models Across Languages, Modalities, Models and Tasks 13 Nov 2023 · 0 repositories · arXiv:2311.07463
-
Speech-based Slot Filling using Large Language Models 13 Nov 2023 · 0 repositories · arXiv:2311.07418
-
STEER: Unified Style Transfer with Expert Reinforcement 13 Nov 2023 · 1 repository · arXiv:2311.07167
-
From Complex to Simple: Unraveling the Cognitive Tree for Reasoning with Small Language Models 12 Nov 2023 · 0 repositories · arXiv:2311.06754
-
GIELLM: Japanese General Information Extraction Large Language Model Utilizing Mutual Reinforcement Effect 12 Nov 2023 · 0 repositories · arXiv:2311.06838
-
Large Language Models are In-context Teachers for Knowledge Reasoning 12 Nov 2023 · 0 repositories · arXiv:2311.06985
-
Data Contamination Quiz: A Tool to Detect and Estimate Contamination in Large Language Models 10 Nov 2023 · 2 repositories · arXiv:2311.06233Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Large Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document Summarisation 9 Nov 2023 · 0 repositories · arXiv:2311.05169
-
Rethinking Benchmark and Contamination for Language Models with Rephrased Samples 8 Nov 2023 · 1 repository · arXiv:2311.04850Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Evaluating Large Language Models in Ophthalmology 7 Nov 2023 · 0 repositories · arXiv:2311.04933
-
Identifying and Mitigating Vulnerabilities in LLM-Integrated Applications 7 Nov 2023 · 0 repositories · arXiv:2311.16153
-
DeepInception: Hypnotize Large Language Model to Be Jailbreaker 6 Nov 2023 · 1 repository · arXiv:2311.03191Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
In-Context Learning for Knowledge Base Question Answering for Unmanned Systems based on Large Language Models 6 Nov 2023 · 0 repositories · arXiv:2311.02956
-
Unraveling Downstream Gender Bias from Large Language Models: A Study on AI Educational Writing Assistance 6 Nov 2023 · 1 repository · arXiv:2311.03311
-
Evaluating the Potential of Leading Large Language Models in Reasoning Biology Questions 5 Nov 2023 · 0 repositories · arXiv:2311.07582
-
Extraction of Atypical Aspects from Customer Reviews: Datasets and Experiments with Language Models 5 Nov 2023 · 2 repositories · arXiv:2311.02702
-
UID as a Guiding Metric for Automated Authorship Obfuscation 5 Nov 2023 · 0 repositories · arXiv:2312.03709
-
Automating Governing Knowledge Commons and Contextual Integrity (GKC-CI) Privacy Policy Annotations with Large Language Models 3 Nov 2023 · 1 repository · arXiv:2311.02192
-
COSMIC: Data Efficient Instruction-tuning For Speech In-Context Learning 3 Nov 2023 · 0 repositories · arXiv:2311.02248
-
Efficient Black-Box Adversarial Attacks on Neural Text Detectors 3 Nov 2023 · 1 repository · arXiv:2311.01873
-
Exploring the Numerical Reasoning Capabilities of Language Models: A Comprehensive Analysis on Tabular Data 3 Nov 2023 · 0 repositories · arXiv:2311.02216
-
Long Story Short: a Summarize-then-Search Method for Long Video Question Answering 2 Nov 2023 · 1 repository · arXiv:2311.01233
-
Measuring Five Accountable Talk Moves to Improve Instruction at Scale 2 Nov 2023 · 0 repositories · arXiv:2311.10749
-
Server-side Rescoring of Spoken Entity-centric Knowledge Queries for Virtual Assistants 2 Nov 2023 · 0 repositories · arXiv:2311.01398
-
Are Large Language Models Reliable Judges? A Study on the Factuality Evaluation Capabilities of LLMs 1 Nov 2023 · 0 repositories · arXiv:2311.00681
-
Continuous Training and Fine-tuning for Domain-Specific Language Models in Medical Question Answering 1 Nov 2023 · 0 repositories · arXiv:2311.00204
-
Is GPT Powerful Enough to Analyze the Emotions of Memes? 1 Nov 2023 · 0 repositories · arXiv:2311.00223
-
Unsupervised Lexical Simplification with Context Augmentation 1 Nov 2023 · 1 repository · arXiv:2311.00310
-
Do large language models solve verbal analogies like children do? 31 Oct 2023 · 0 repositories · arXiv:2310.20384
-
Does GPT-4 pass the Turing test? 31 Oct 2023 · 0 repositories · arXiv:2310.20216
-
Efficient Classification of Student Help Requests in Programming Courses Using Large Language Models 31 Oct 2023 · 0 repositories · arXiv:2310.20105
-
Interactive Multi-fidelity Learning for Cost-effective Adaptation of Language Model with Sparse Human Supervision 31 Oct 2023 · 0 repositories · arXiv:2310.20153
-
PsyCoT: Psychological Questionnaire as Powerful Chain-of-Thought for Personality Detection 31 Oct 2023 · 1 repository · arXiv:2310.20256
-
Interpretable-by-Design Text Understanding with Iteratively Generated Concept Bottleneck 30 Oct 2023 · 1 repository · arXiv:2310.19660
-
Remember what you did so you know what to do next 30 Oct 2023 · 0 repositories · arXiv:2311.01468
-
EtiCor: Corpus for Analyzing LLMs for Etiquettes 29 Oct 2023 · 1 repository · arXiv:2310.18974
-
Large language models for aspect-based sentiment analysis 27 Oct 2023 · 1 repository · arXiv:2310.18025Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
OffMix-3L: A Novel Code-Mixed Dataset in Bangla-English-Hindi for Offensive Language Identification 27 Oct 2023 · 1 repository · arXiv:2310.18387
-
SentMix-3L: A Bangla-English-Hindi Code-Mixed Dataset for Sentiment Analysis 27 Oct 2023 · 1 repository · arXiv:2310.18023
-
FedPEAT: Convergence of Federated Learning, Parameter-Efficient Fine Tuning, and Emulator Assisted Tuning for Artificial Intelligence Foundation Models with Mobile Edge Computing 26 Oct 2023 · 0 repositories · arXiv:2310.17491
-
In-Context Learning Dynamics with Random Binary Sequences 26 Oct 2023 · 1 repository · arXiv:2310.17639Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
"You Are An Expert Linguistic Annotator": Limits of LLMs as Analyzers of Abstract Meaning Representation 26 Oct 2023 · 0 repositories · arXiv:2310.17793
-
BOOST: Harnessing Black-Box Control to Boost Commonsense in LMs' Generation 25 Oct 2023 · 0 repositories · arXiv:2310.17054
-
Decoding Stumpers: Large Language Models vs. Human Problem-Solvers 25 Oct 2023 · 0 repositories · arXiv:2310.16411
-
How well can machine-generated texts be identified and can language models be trained to avoid identification? 25 Oct 2023 · 0 repositories · arXiv:2310.16992
-
Muslim-Violence Bias Persists in Debiased GPT Models 25 Oct 2023 · 0 repositories · arXiv:2310.18368
-
R³ Prompting: Review, Rephrase and Resolve for Chain-of-Thought Reasoning in Large Language Models under Noisy Context 25 Oct 2023 · 0 repositories · arXiv:2310.16535
-
A Communication Theory Perspective on Prompting Engineering Methods for Large Language Models 24 Oct 2023 · 0 repositories · arXiv:2310.18358
-
AI-enhanced Auto-correction of Programming Exercises: How Effective is GPT-3.5? 24 Oct 2023 · 0 repositories · arXiv:2311.10737
-
Background Summarization of Event Timelines 24 Oct 2023 · 1 repository · arXiv:2310.16197
-
Dissecting In-Context Learning of Translations in GPTs 24 Oct 2023 · 0 repositories · arXiv:2310.15987
-
Fighting Fire with Fire: The Dual Role of LLMs in Crafting and Detecting Elusive Disinformation 24 Oct 2023 · 1 repository · arXiv:2310.15515
-
The Janus Interface: How Fine-Tuning in Large Language Models Amplifies the Privacy Risks 24 Oct 2023 · 1 repository · arXiv:2310.15469
-
What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations 24 Oct 2023 · 0 repositories · arXiv:2310.15431
-
Causal Inference Using LLM-Guided Discovery 23 Oct 2023 · 0 repositories · arXiv:2310.15117
-
Evaluating Spatial Understanding of Large Language Models 23 Oct 2023 · 1 repository · arXiv:2310.14540Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Evaluating the Knowledge Base Completion Potential of GPT 23 Oct 2023 · 0 repositories · arXiv:2310.14771
-
GPT-4 as an Effective Zero-Shot Evaluator for Scientific Figure Captions 23 Oct 2023 · 0 repositories · arXiv:2310.15405
-
InstructExcel: A Benchmark for Natural Language Instruction in Excel 23 Oct 2023 · 0 repositories · arXiv:2310.14495
-
Language Models Hallucinate, but May Excel at Fact Verification 23 Oct 2023 · 1 repository · arXiv:2310.14564Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
LINC: A Neurosymbolic Approach for Logical Reasoning by Combining Language Models with First-Order Logic Provers 23 Oct 2023 · 1 repository · arXiv:2310.15164Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
LLM-in-the-loop: Leveraging Large Language Model for Thematic Analysis 23 Oct 2023 · 1 repository · arXiv:2310.15100Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
TeleQnA: A Benchmark Dataset to Assess Large Language Models Telecommunications Knowledge 23 Oct 2023 · 1 repository · arXiv:2310.15051
-
Can Language Models Laugh at YouTube Short-form Videos? 22 Oct 2023 · 1 repository · arXiv:2310.14159
-
Is ChatGPT a game changer for geocoding -- a benchmark for geocoding address parsing techniques 22 Oct 2023 · 1 repository · arXiv:2310.14360
-
Text generation for dataset augmentation in security classification tasks 22 Oct 2023 · 1 repository · arXiv:2310.14429
-
HateRephrase: Zero- and Few-Shot Reduction of Hate Intensity in Online Posts using Large Language Models 21 Oct 2023 · 0 repositories · arXiv:2310.13985
-
A Simple Baseline for Knowledge-Based Visual Question Answering 20 Oct 2023 · 0 repositories · arXiv:2310.13570Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Cache me if you Can: an Online Cost-aware Teacher-Student framework to Reduce the Calls to Large Language Models 20 Oct 2023 · 1 repository · arXiv:2310.13395Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
She had Cobalt Blue Eyes: Prompt Testing to Create Aligned and Sustainable Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.18333
-
The Perils & Promises of Fact-checking with Large Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.13549
-
WordArt Designer: User-Driven Artistic Typography Synthesis using Large Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.18332
-
AgentTuning: Enabling Generalized Agent Abilities for LLMs 19 Oct 2023 · 1 repository · arXiv:2310.12823Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring In-Context Learning of Textless Speech Language Model for Speech Classification Tasks 19 Oct 2023 · 0 repositories · arXiv:2310.12477
-
Experimental Narratives: A Comparison of Human Crowdsourced Storytelling and AI Storytelling 19 Oct 2023 · 0 repositories · arXiv:2310.12902
-
ExtractGPT: Exploring the Potential of Large Language Models for Product Attribute Value Extraction 19 Oct 2023 · 1 repository · arXiv:2310.12537
-
Evaluating the Symbol Binding Ability of Large Language Models for Multiple-Choice Questions in Vietnamese General Education 18 Oct 2023 · 0 repositories · arXiv:2310.12059