Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 7
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 7 of 20: papers 601 to 700 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
ELLEN: Extremely Lightly Supervised Learning For Efficient Named Entity Recognition 26 Mar 2024 · 1 repository · arXiv:2403.17385
-
Enhancing Legal Document Retrieval: A Multi-Phase Approach with Large Language Models 26 Mar 2024 · 0 repositories · arXiv:2403.18093
-
MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution 26 Mar 2024 · 0 repositories · arXiv:2403.17927
-
OmniVid: A Generative Framework for Universal Video Understanding 26 Mar 2024 · 1 repository · arXiv:2403.17935Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Verbing Weirds Language (Models): Evaluation of English Zero-Derivation in Five LLMs 26 Mar 2024 · 0 repositories · arXiv:2403.17856
-
A comparison of Human, GPT-3.5, and GPT-4 Performance in a University-Level Coding Course 25 Mar 2024 · 1 repository · arXiv:2403.16977
-
Iterative Refinement of Project-Level Code Context for Precise Code Generation with Compiler Feedback 25 Mar 2024 · 1 repository · arXiv:2403.16792
-
RepairAgent: An Autonomous, LLM-Based Agent for Program Repair 25 Mar 2024 · 1 repository · arXiv:2403.17134Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
LinkPrompt: Natural and Universal Adversarial Attacks on Prompt-based Language Models 25 Mar 2024 · 1 repository · arXiv:2403.16432Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards Algorithmic Fidelity: Mental Health Representation across Demographics in Synthetic vs. Human-generated Data 25 Mar 2024 · 1 repository · arXiv:2403.16909
-
SQL-Encoder: Improving NL2SQL In-Context Learning Through a Context-Aware Encoder 24 Mar 2024 · 0 repositories · arXiv:2403.16204
-
Using Large Language Models for OntoClean-based Ontology Refinement 23 Mar 2024 · 0 repositories · arXiv:2403.15864
-
Can large language models explore in-context? 22 Mar 2024 · 0 repositories · arXiv:2403.15371
-
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation 22 Mar 2024 · 1 repository · arXiv:2403.14965
-
Text Clustering with Large Language Model Embeddings 22 Mar 2024 · 0 repositories · arXiv:2403.15112
-
VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding 21 Mar 2024 · 1 repository · arXiv:2403.14743
-
Motion Generation from Fine-grained Textual Descriptions 20 Mar 2024 · 1 repository · arXiv:2403.13518
-
Natural Language as Policies: Reasoning for Coordinate-Level Embodied Control with LLMs 20 Mar 2024 · 0 repositories · arXiv:2403.13801
-
PARAMANU-AYN: Pretrain from scratch or Continual Pretraining of LLMs for Legal Domain Adaptation? 20 Mar 2024 · 0 repositories · arXiv:2403.13681
-
Automated Data Curation for Robust Language Model Fine-Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.12776
-
Fine-Tuning Pre-trained Language Models to Detect In-Game Trash Talks 19 Mar 2024 · 0 repositories · arXiv:2403.15458
-
Instructing Large Language Models to Identify and Ignore Irrelevant Conditions 19 Mar 2024 · 1 repository · arXiv:2403.12744
-
Construction of Hyper-Relational Knowledge Graphs Using Pre-Trained Large Language Models 18 Mar 2024 · 0 repositories · arXiv:2403.11786
-
EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models 18 Mar 2024 · 1 repository · arXiv:2403.12171Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Ensuring Safe and High-Quality Outputs: A Guideline Library Approach for Language Models 18 Mar 2024 · 1 repository · arXiv:2403.11838
-
GPT-4 as Evaluator: Evaluating Large Language Models on Pest Management in Agriculture 18 Mar 2024 · 0 repositories · arXiv:2403.11858
-
HateCOT: An Explanation-Enhanced Dataset for Generalizable Offensive Speech Detection via Large Language Models 18 Mar 2024 · 1 repository · arXiv:2403.11456
-
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments 18 Mar 2024 · 1 repository · arXiv:2403.11807Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 2 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Metaphor Understanding Challenge Dataset for LLMs 18 Mar 2024 · 0 repositories · arXiv:2403.11810
-
Leveraging Large Language Models to Detect npm Malicious Packages 18 Mar 2024 · 0 repositories · arXiv:2403.12196
-
Data is all you need: Finetuning LLMs for Chip Design via an Automated design-data augmentation framework 17 Mar 2024 · 1 repository · arXiv:2403.11202Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Large language model-powered chatbots for internationalizing student support in higher education 16 Mar 2024 · 0 repositories · arXiv:2403.14702
-
ExeGPT: Constraint-Aware Resource Scheduling for LLM Inference 15 Mar 2024 · 0 repositories · arXiv:2404.07947
-
Knowledge Condensation and Reasoning for Knowledge-based VQA 15 Mar 2024 · 0 repositories · arXiv:2403.10037
-
CodeUltraFeedback: An LLM-as-a-Judge Dataset for Aligning Large Language Models to Coding Preferences 14 Mar 2024 · 2 repositories · arXiv:2403.09032Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluating LLMs for Gender Disparities in Notable Persons 14 Mar 2024 · 0 repositories · arXiv:2403.09148
-
Komodo: A Linguistic Expedition into Indonesia's Regional Languages 14 Mar 2024 · 0 repositories · arXiv:2403.09362
-
Sabiá-2: A New Generation of Portuguese Large Language Models 14 Mar 2024 · 0 repositories · arXiv:2403.09887
-
Rethinking Generative Large Language Model Evaluation for Semantic Comprehension 12 Mar 2024 · 0 repositories · arXiv:2403.07872
-
SIFiD: Reassess Summary Factual Inconsistency Detection with LLM 12 Mar 2024 · 0 repositories · arXiv:2403.07557
-
The future of document indexing: GPT and Donut revolutionize table of content processing 12 Mar 2024 · 0 repositories · arXiv:2403.07553
-
Development of a Reliable and Accessible Caregiving Language Model (CaLM) 11 Mar 2024 · 0 repositories · arXiv:2403.06857
-
Hybrid Human-LLM Corpus Construction and LLM Evaluation for Rare Linguistic Phenomena 11 Mar 2024 · 0 repositories · arXiv:2403.06965
-
Narrating Causal Graphs with Large Language Models 11 Mar 2024 · 0 repositories · arXiv:2403.07118
-
Bias-Augmented Consistency Training Reduces Biased Reasoning in Chain-of-Thought 8 Mar 2024 · 1 repository · arXiv:2403.05518Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
RAT: Retrieval Augmented Thoughts Elicit Context-Aware Reasoning in Long-Horizon Generation 8 Mar 2024 · 1 repository · arXiv:2403.05313
-
Feedback-Generation for Programming Exercises With GPT-4 7 Mar 2024 · 0 repositories · arXiv:2403.04449
-
Telecom Language Models: Must They Be Large? 7 Mar 2024 · 0 repositories · arXiv:2403.04666
-
Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem 6 Mar 2024 · 1 repository · arXiv:2403.03558
-
Can Large Language Models do Analytical Reasoning? 6 Mar 2024 · 0 repositories · arXiv:2403.04031
-
General2Specialized LLMs Translation for E-commerce 6 Mar 2024 · 0 repositories · arXiv:2403.03689
-
Guiding Enumerative Program Synthesis with Large Language Models 6 Mar 2024 · 0 repositories · arXiv:2403.03997
-
Rapidly Developing High-quality Instruction Data and Evaluation Benchmark for Large Language Models with Minimal Human Effort: A Case Study on Japanese 6 Mar 2024 · 2 repositories · arXiv:2403.03690
-
Evaluating and Optimizing Educational Content with Large Language Model Judgments 5 Mar 2024 · 1 repository · arXiv:2403.02795Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Improving Event Definition Following For Zero-Shot Event Detection 5 Mar 2024 · 0 repositories · arXiv:2403.02586
-
MathScale: Scaling Instruction Tuning for Mathematical Reasoning 5 Mar 2024 · 1 repository · arXiv:2403.02884
-
MeanCache: User-Centric Semantic Caching for LLM Web Services 5 Mar 2024 · 0 repositories · arXiv:2403.02694
-
Zero-Shot Cross-Lingual Document-Level Event Causality Identification with Heterogeneous Graph Contrastive Transfer Learning 5 Mar 2024 · 0 repositories · arXiv:2403.02893
-
Can LLMs Generate Architectural Design Decisions? -An Exploratory Empirical study 4 Mar 2024 · 0 repositories · arXiv:2403.01709
-
Differentially Private Synthetic Data via Foundation Model APIs 2: Text 4 Mar 2024 · 2 repositories · arXiv:2403.01749Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Hypertext Entity Extraction in Webpage 4 Mar 2024 · 0 repositories · arXiv:2403.01698
-
PHAnToM: Persona-based Prompting Has An Effect on Theory-of-Mind Reasoning in Large Language Models 4 Mar 2024 · 0 repositories · arXiv:2403.02246
-
ProTrix: Building Models for Planning and Reasoning over Tables with Sentence Context 4 Mar 2024 · 1 repository · arXiv:2403.02177
-
SciAssess: Benchmarking LLM Proficiency in Scientific Literature Analysis 4 Mar 2024 · 1 repository · arXiv:2403.01976Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Using LLMs for the Extraction and Normalization of Product Attribute Values 4 Mar 2024 · 1 repository · arXiv:2403.02130
-
SERVAL: Synergy Learning between Vertical Models and LLMs towards Oracle-Level Zero-shot Medical Prediction 3 Mar 2024 · 0 repositories · arXiv:2403.01570Syntology 6 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
AutoDefense: Multi-Agent LLM Defense against Jailbreak Attacks 2 Mar 2024 · 1 repository · arXiv:2403.04783Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
LM4OPT: Unveiling the Potential of Large Language Models in Formulating Mathematical Optimization Problems 2 Mar 2024 · 0 repositories · arXiv:2403.01342
-
SoftTiger: A Clinical Foundation Model for Healthcare Workflows 1 Mar 2024 · 1 repository · arXiv:2403.00868
-
ARTiST: Automated Text Simplification for Task Guidance in Augmented Reality 29 Feb 2024 · 1 repository · arXiv:2402.18797
-
LLM-Ensemble: Optimal Large Language Model Ensemble Method for E-commerce Product Attribute Value Extraction 29 Feb 2024 · 0 repositories · arXiv:2403.00863
-
PROC2PDDL: Open-Domain Planning Representations from Texts 29 Feb 2024 · 0 repositories · arXiv:2403.00092
-
VIXEN: Visual Text Comparison Network for Image Difference Captioning 29 Feb 2024 · 0 repositories · arXiv:2402.19119
-
Decomposed Prompting: Unveiling Multilingual Linguistic Structure Knowledge in English-Centric Large Language Models 28 Feb 2024 · 0 repositories · arXiv:2402.18397
-
Keeping LLMs Aligned After Fine-tuning: The Crucial Role of Prompt Templates 28 Feb 2024 · 1 repository · arXiv:2402.18540Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Deep Learning Detection Method for Large Language Models-Generated Scientific Content 27 Feb 2024 · 0 repositories · arXiv:2403.00828
-
Measuring Vision-Language STEM Skills of Neural Models 27 Feb 2024 · 1 repository · arXiv:2402.17205Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ChatMusician: Understanding and Generating Music Intrinsically with LLM 25 Feb 2024 · 1 repository · arXiv:2402.16153Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Deep Learning Approaches for Improving Question Answering Systems in Hepatocellular Carcinoma Research 25 Feb 2024 · 0 repositories · arXiv:2402.16038
-
Knowledge Fusion of Chat LLMs: A Preliminary Technical Report 25 Feb 2024 · 2 repositories · arXiv:2402.16107
-
LLMs Can Defend Themselves Against Jailbreaking in a Practical Manner: A Vision Paper 24 Feb 2024 · 0 repositories · arXiv:2402.15727
-
Look Before You Leap: Problem Elaboration Prompting Improves Mathematical Reasoning in Large Language Models 24 Feb 2024 · 0 repositories · arXiv:2402.15764
-
AttributionBench: How Hard is Automatic Attribution Evaluation? 23 Feb 2024 · 1 repository · arXiv:2402.15089
-
Fine-tuning Large Language Models for Domain-specific Machine Translation 23 Feb 2024 · 0 repositories · arXiv:2402.15061
-
LLMs as Meta-Reviewers' Assistants: A Case Study 23 Feb 2024 · 1 repository · arXiv:2402.15589
-
Can Large Language Models Detect Misinformation in Scientific News Reporting? 22 Feb 2024 · 0 repositories · arXiv:2402.14268
-
Copilot Evaluation Harness: Evaluating LLM-Guided Software Programming 22 Feb 2024 · 0 repositories · arXiv:2402.14261
-
Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge 22 Feb 2024 · 1 repository · arXiv:2402.14310
-
KoCoSa: Korean Context-aware Sarcasm Detection Dataset 22 Feb 2024 · 1 repository · arXiv:2402.14428
-
RoboScript: Code Generation for Free-Form Manipulation Tasks across Real and Simulation 22 Feb 2024 · 0 repositories · arXiv:2402.14623
-
Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs 22 Feb 2024 · 1 repository · arXiv:2402.14903Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Beyond Hate Speech: NLP's Challenges and Opportunities in Uncovering Dehumanizing Language 21 Feb 2024 · 0 repositories · arXiv:2402.13818
-
A Comprehensive Study of Jailbreak Attack versus Defense for Large Language Models 21 Feb 2024 · 1 repository · arXiv:2402.13457Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SYNFAC-EDIT: Synthetic Imitation Edit Feedback for Factual Alignment in Clinical Summarization 21 Feb 2024 · 1 repository · arXiv:2402.13919
-
Benchmarking Retrieval-Augmented Generation for Medicine 20 Feb 2024 · 2 repositories · arXiv:2402.13178Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
EvoGrad: A Dynamic Take on the Winograd Schema Challenge with Human Adversaries 20 Feb 2024 · 0 repositories · arXiv:2402.13372
-
MoELoRA: Contrastive Learning Guided Mixture of Experts on Parameter-Efficient Fine-Tuning for Large Language Models 20 Feb 2024 · 1 repository · arXiv:2402.12851
-
NL2Formula: Generating Spreadsheet Formulas from Natural Language Queries 20 Feb 2024 · 0 repositories · arXiv:2402.14853
-
R³: "This is My SQL, Are You With Me?" A Consensus-Based Multi-Agent System for Text-to-SQL Tasks 20 Feb 2024 · 0 repositories · arXiv:2402.14851
-
The Impact of Demonstrations on Multilingual In-Context Learning: A Multidimensional Analysis 20 Feb 2024 · 1 repository · arXiv:2402.12976