Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 9
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 9 of 20: papers 801 to 900 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Improving Classification Performance With Human Feedback: Label a few, we label the rest 17 Jan 2024 · 0 repositories · arXiv:2401.09555
-
Application of LLM Agents in Recruitment: A Novel Framework for Resume Screening 16 Jan 2024 · 0 repositories · arXiv:2401.08315
-
RAG vs Fine-tuning: Pipelines, Tradeoffs, and a Case Study on Agriculture 16 Jan 2024 · 0 repositories · arXiv:2401.08406
-
Tuning Language Models by Proxy 16 Jan 2024 · 2 repositories · arXiv:2401.08565Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Harnessing Large Language Models Over Transformer Models for Detecting Bengali Depressive Social Media Text: A Comprehensive Study 14 Jan 2024 · 1 repository · arXiv:2401.07310
-
Assessing Large Language Models in Mechanical Engineering Education: A Study on Mechanics-Focused Conceptual Understanding 13 Jan 2024 · 0 repositories · arXiv:2401.12983
-
Comparing GPT-4 and Open-Source Language Models in Misinformation Mitigation 12 Jan 2024 · 0 repositories · arXiv:2401.06920
-
Human-AI Collaborative Essay Scoring: A Dual-Process Framework with LLMs 12 Jan 2024 · 1 repository · arXiv:2401.06431
-
How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs 12 Jan 2024 · 2 repositories · arXiv:2401.06373Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Intention Analysis Makes LLMs A Good Jailbreak Defender 12 Jan 2024 · 1 repository · arXiv:2401.06561Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
PersianMind: A Cross-Lingual Persian-English Large Language Model 12 Jan 2024 · 0 repositories · arXiv:2401.06466
-
PizzaCommonSense: Learning to Model Commonsense Reasoning about Intermediate Steps in Cooking Recipes 12 Jan 2024 · 1 repository · arXiv:2401.06930
-
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs 11 Jan 2024 · 0 repositories · arXiv:2401.05940
-
The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models 11 Jan 2024 · 1 repository · arXiv:2401.05618Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning 10 Jan 2024 · 1 repository · arXiv:2401.05268Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks 10 Jan 2024 · 1 repository · arXiv:2401.05507Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Can Active Label Correction Improve LLM-based Modular AI Systems? 10 Jan 2024 · 0 repositories · arXiv:2401.05467
-
Distortions in Judged Spatial Relations in Large Language Models 8 Jan 2024 · 0 repositories · arXiv:2401.04218
-
LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems 8 Jan 2024 · 1 repository · arXiv:2401.05443
-
Mixtral of Experts 8 Jan 2024 · 6 repositories · arXiv:2401.04088Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 5 harvested samples)
-
Exploring Defeasibility in Causal Reasoning 6 Jan 2024 · 0 repositories · arXiv:2401.03183
-
Using Large Language Models to Assess Tutors' Performance in Reacting to Students Making Math Errors 6 Jan 2024 · 0 repositories · arXiv:2401.03238
-
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism 5 Jan 2024 · 1 repository · arXiv:2401.02954
-
Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks 5 Jan 2024 · 2 repositories · arXiv:2401.02731
-
Re-evaluating the Memory-balanced Pipeline Parallelism: BPipe 4 Jan 2024 · 0 repositories · arXiv:2401.02088
-
Vietnamese Poem Generation & The Prospect Of Cross-Language Poem-To-Poem Translation 2 Jan 2024 · 1 repository · arXiv:2401.01078
-
LogicAsker: Evaluating and Improving the Logical Reasoning Ability of Large Language Models 1 Jan 2024 · 1 repository · arXiv:2401.00757Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Large Language Models aren't all that you need 1 Jan 2024 · 0 repositories · arXiv:2401.00698
-
Advancing TTP Analysis: Harnessing the Power of Large Language Models with Retrieval Augmented Generation 30 Dec 2023 · 1 repository · arXiv:2401.00280
-
Jatmo: Prompt Injection Defense by Task-Specific Finetuning 29 Dec 2023 · 1 repository · arXiv:2312.17673Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Evaluating the Performance of Large Language Models for Spanish Language in Undergraduate Admissions Exams 28 Dec 2023 · 0 repositories · arXiv:2312.16845
-
Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4 26 Dec 2023 · 2 repositories · arXiv:2312.16171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
SecQA: A Concise Question-Answering Dataset for Evaluating Large Language Models in Computer Security 26 Dec 2023 · 1 repository · arXiv:2312.15838Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Task Contamination: Language Models May Not Be Few-Shot Anymore 26 Dec 2023 · 0 repositories · arXiv:2312.16337
-
Efficacy of Machine-Generated Instructions 22 Dec 2023 · 0 repositories · arXiv:2312.14423
-
FM-OV3D: Foundation Model-based Cross-modal Knowledge Blending for Open-Vocabulary 3D Detection 22 Dec 2023 · 0 repositories · arXiv:2312.14465
-
Refining GPT-3 Embeddings with a Siamese Structure for Technical Post Duplicate Detection 22 Dec 2023 · 1 repository · arXiv:2312.15068
-
Argue with Me Tersely: Towards Sentence-Level Counter-Argument Generation 21 Dec 2023 · 1 repository · arXiv:2312.13608
-
ChatGPT as a commenter to the news: can LLMs generate human-like opinions? 21 Dec 2023 · 1 repository · arXiv:2312.13961
-
InfoVisDial: An Informative Visual Dialogue Dataset by Bridging Large Multimodal and Language Models 21 Dec 2023 · 0 repositories · arXiv:2312.13503
-
Team Irisapu Project Description for DRC2023 21 Dec 2023 · 0 repositories · arXiv:2312.13765
-
Typhoon: Thai Large Language Models 21 Dec 2023 · 0 repositories · arXiv:2312.13951
-
AgentCoder: Multi-Agent-based Code Generation with Iterative Testing and Optimisation 20 Dec 2023 · 1 repository · arXiv:2312.13010Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Benchmarking and Analyzing In-context Learning, Fine-tuning and Supervised Learning for Biomedical Knowledge Curation: a focused study on chemical entities of biological interest 20 Dec 2023 · 0 repositories · arXiv:2312.12989
-
Can ChatGPT be Your Personal Medical Assistant? 19 Dec 2023 · 0 repositories · arXiv:2312.12006
-
Large Language Models in Medical Term Classification and Unexpected Misalignment Between Response and Reasoning 19 Dec 2023 · 0 repositories · arXiv:2312.14184
-
Evaluating AI Vocational Skills Through Professional Testing 17 Dec 2023 · 0 repositories · arXiv:2312.10603
-
HyperPIE: Hyperparameter Information Extraction from Scientific Publications 17 Dec 2023 · 1 repository · arXiv:2312.10638
-
Mixed Distillation Helps Smaller Language Model Better Reasoning 17 Dec 2023 · 0 repositories · arXiv:2312.10730
-
Multi-Label Classification of COVID-Tweets Using Large Language Models 17 Dec 2023 · 1 repository · arXiv:2312.10748
-
A Comparative Analysis of Large Language Models for Code Documentation Generation 16 Dec 2023 · 0 repositories · arXiv:2312.10349
-
A Novel Dataset for Financial Education Text Simplification in Spanish 15 Dec 2023 · 0 repositories · arXiv:2312.09897
-
Distilling Large Language Models for Matching Patients to Clinical Trials 15 Dec 2023 · 0 repositories · arXiv:2312.09958
-
Fewer is More: Boosting LLM Reasoning with Reinforced Context Pruning 14 Dec 2023 · 0 repositories · arXiv:2312.08901
-
Dynamic Retrieval-Augmented Generation 14 Dec 2023 · 0 repositories · arXiv:2312.08976
-
Self-Evaluation Improves Selective Generation in Large Language Models 14 Dec 2023 · 0 repositories · arXiv:2312.09300
-
TinyGSM: achieving >80% on GSM8k with small language models 14 Dec 2023 · 0 repositories · arXiv:2312.09241
-
Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision 14 Dec 2023 · 0 repositories · arXiv:2312.09390
-
Weaving Pathways for Justice with GPT: LLM-driven automated drafting of interactive legal applications 14 Dec 2023 · 1 repository · arXiv:2312.09198
-
Large Language Models are Complex Table Parsers 13 Dec 2023 · 0 repositories · arXiv:2312.11521
-
AI Control: Improving Safety Despite Intentional Subversion 12 Dec 2023 · 1 repository · arXiv:2312.06942Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Image Content Generation with Causal Reasoning 12 Dec 2023 · 1 repository · arXiv:2312.07132Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Multilingual large language models leak human stereotypes across language boundaries 12 Dec 2023 · 1 repository · arXiv:2312.07141
-
Reducing Energy Bloat in Large Model Training 12 Dec 2023 · 2 repositories · arXiv:2312.06902
-
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions 11 Dec 2023 · 1 repository · arXiv:2312.12450Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Evaluating ChatGPT as a Question Answering System: A Comprehensive Analysis and Comparison with Existing Models 11 Dec 2023 · 0 repositories · arXiv:2312.07592
-
Generative Large Language Models Are All-purpose Text Analytics Engines: Text-to-text Learning Is All Your Need 11 Dec 2023 · 0 repositories · arXiv:2312.06099
-
Exploring the Limits of ChatGPT in Software Security Applications 8 Dec 2023 · 0 repositories · arXiv:2312.05275
-
LLM Interactive Optimization of Open Source Python Libraries -- Case Studies and Generalization 8 Dec 2023 · 0 repositories · arXiv:2312.14949
-
On Sarcasm Detection with OpenAI GPT-based Models 7 Dec 2023 · 0 repositories · arXiv:2312.04642
-
Holmes: Towards Distributed Training Across Clusters with Heterogeneous NIC Environment 6 Dec 2023 · 0 repositories · arXiv:2312.03549
-
A Hardware Evaluation Framework for Large Language Model Inference 5 Dec 2023 · 0 repositories · arXiv:2312.03134
-
DRAFT: Dense Retrieval Augmented Few-shot Topic classifier Framework 5 Dec 2023 · 1 repository · arXiv:2312.02532
-
GPT vs Human for Scientific Reviews: A Dual Source Review on Applications of ChatGPT in Science 5 Dec 2023 · 0 repositories · arXiv:2312.03769
-
Rank-without-GPT: Building GPT-Independent Listwise Rerankers on Open-Source Large Language Models 5 Dec 2023 · 0 repositories · arXiv:2312.02969
-
A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly 4 Dec 2023 · 0 repositories · arXiv:2312.02003
-
Jellyfish: A Large Language Model for Data Preprocessing 4 Dec 2023 · 0 repositories · arXiv:2312.01678
-
Tree of Attacks: Jailbreaking Black-Box LLMs Automatically 4 Dec 2023 · 2 repositories · arXiv:2312.02119Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
NLEBench+NorGLM: A Comprehensive Empirical Analysis and Benchmark Dataset for Generative Language Models in Norwegian 3 Dec 2023 · 1 repository · arXiv:2312.01314Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Harnessing the Power of Prompt-based Techniques for Generating School-Level Questions using Large Language Models 2 Dec 2023 · 1 repository · arXiv:2312.01032
-
Applying Large Language Models and Chain-of-Thought for Automatic Scoring 30 Nov 2023 · 0 repositories · arXiv:2312.03748
-
IAG: Induction-Augmented Generation Framework for Answering Reasoning Questions 30 Nov 2023 · 0 repositories · arXiv:2311.18397
-
Robust Concept Erasure via Kernelized Rate-Distortion Maximization 30 Nov 2023 · 1 repository · arXiv:2312.00194Syntology official (archive's flag): 20 ran · 20 ran (of which 0 constructed an object rather than computing a result; 18 with no instrument failure: 0 honoured, 0 violated, 18 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 21 harvested samples) · 2 pointer-only (licence)
-
Biomedical knowledge graph-optimized prompt generation for large language models 29 Nov 2023 · 1 repository · arXiv:2311.17330Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
BERT Goes Off-Topic: Investigating the Domain Transfer Challenge using Genre Classification 27 Nov 2023 · 1 repository · arXiv:2311.16083
-
Decoding Logic Errors: A Comparative Study on Bug Detection by Students and Large Language Models 27 Nov 2023 · 0 repositories · arXiv:2311.16017
-
Machine-Generated Text Detection using Deep Learning 26 Nov 2023 · 1 repository · arXiv:2311.15425
-
GPT Struct Me: Probing GPT Models on Narrative Entity Extraction 24 Nov 2023 · 1 repository · arXiv:2311.14583
-
Large Language Models as Automated Aligners for benchmarking Vision-Language Models 24 Nov 2023 · 0 repositories · arXiv:2311.14580
-
Machine Translation for Ge'ez Language 24 Nov 2023 · 0 repositories · arXiv:2311.14530
-
A Cross Attention Approach to Diagnostic Explainability using Clinical Practice Guidelines for Depression 23 Nov 2023 · 1 repository · arXiv:2311.13852
-
Hardware Resilience Properties of Text-Guided Image Classifiers 23 Nov 2023 · 1 repository · arXiv:2311.14062Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Minimizing Factual Inconsistency and Hallucination in Large Language Models 23 Nov 2023 · 0 repositories · arXiv:2311.13878
-
Generation of Explanations for Logic Reasoning 22 Nov 2023 · 0 repositories · arXiv:2311.13455
-
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning 22 Nov 2023 · 0 repositories · arXiv:2311.13721
-
PG-Video-LLaVA: Pixel Grounding Large Video-Language Models 22 Nov 2023 · 1 repository · arXiv:2311.13435Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
AlignedCoT: Prompting Large Language Models via Native-Speaking Demonstrations 22 Nov 2023 · 1 repository · arXiv:2311.13538
-
@ve: A Chatbot for Latin 22 Nov 2023 · 0 repositories · arXiv:2311.14741
-
A Survey on Large Language Models for Personalized and Explainable Recommendations 21 Nov 2023 · 0 repositories · arXiv:2311.12338
-
InterPrompt: Interpretable Prompting for Interrelated Interpersonal Risk Factors in Reddit Posts 21 Nov 2023 · 0 repositories · arXiv:2311.12404