Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 3
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 3 of 20: papers 201 to 300 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SBI-RAG: Enhancing Math Word Problem Solving for Students through Schema-Based Instruction and Retrieval-Augmented Generation 17 Oct 2024 · 1 repository · arXiv:2410.13293
-
Agent Skill Acquisition for Large Language Models via CycleQD 16 Oct 2024 · 1 repository · arXiv:2410.14735Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Table-LLM-Specialist: Language Model Specialists for Tables using Iterative Generator-Validator Fine-tuning 16 Oct 2024 · 0 repositories · arXiv:2410.12164
-
Code-Mixer Ya Nahi: Novel Approaches to Measuring Multilingual LLMs' Code-Mixing Capabilities 14 Oct 2024 · 0 repositories · arXiv:2410.11079
-
Rethinking Legal Judgement Prediction in a Realistic Scenario in the Era of Large Language Models 14 Oct 2024 · 1 repository · arXiv:2410.10542
-
Evaluating Gender Bias of LLMs in Making Morality Judgements 13 Oct 2024 · 0 repositories · arXiv:2410.09992
-
\llinstruct: An Instruction-tuned model for English Language Proficiency Assessments 12 Oct 2024 · 0 repositories · arXiv:2410.09314
-
AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation 11 Oct 2024 · 1 repository · arXiv:2410.09040
-
Observing the Southern US Culture of Honor Using Large-Scale Social Media Analysis 11 Oct 2024 · 0 repositories · arXiv:2410.13887
-
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models 11 Oct 2024 · 1 repository · arXiv:2410.08698
-
FLIER: Few-shot Language Image Models Embedded with Latent Representations 10 Oct 2024 · 0 repositories · arXiv:2410.07648
-
The Rise of AI-Generated Content in Wikipedia 10 Oct 2024 · 1 repository · arXiv:2410.08044Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
AutoFeedback: An LLM-based Framework for Efficient and Accurate API Request Generation 9 Oct 2024 · 0 repositories · arXiv:2410.06943
-
Generative Model for Less-Resourced Language with 1 billion parameters 9 Oct 2024 · 0 repositories · arXiv:2410.06898
-
Large Language Models as Code Executors: An Exploratory Study 9 Oct 2024 · 0 repositories · arXiv:2410.06667
-
MentalArena: Self-play Training of Language Models for Diagnosis and Treatment of Mental Health Disorders 9 Oct 2024 · 1 repository · arXiv:2410.06845
-
Narrative-of-Thought: Improving Temporal Reasoning of Large Language Models via Recounted Narratives 7 Oct 2024 · 1 repository · arXiv:2410.05558
-
On Instruction-Finetuning Neural Machine Translation Models 7 Oct 2024 · 0 repositories · arXiv:2410.05553
-
Large Language Models for Knowledge-Free Network Management: Feasibility Study and Opportunities 6 Oct 2024 · 0 repositories · arXiv:2410.17259
-
Take It Easy: Label-Adaptive Self-Rationalization for Fact Verification and Explanation Generation 5 Oct 2024 · 1 repository · arXiv:2410.04002Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Crafting Narrative Closures: Zero-Shot Learning with SSM Mamba for Short Story Ending Generation 4 Oct 2024 · 0 repositories · arXiv:2410.10848
-
Cross-lingual Transfer for Automatic Question Generation by Learning Interrogative Structures in Target Languages 4 Oct 2024 · 0 repositories · arXiv:2410.03197
-
Structured List-Grounded Question Answering 4 Oct 2024 · 0 repositories · arXiv:2410.03950
-
Towards Linguistically-Aware and Language-Independent Tokenization for Large Language Models (LLMs) 4 Oct 2024 · 0 repositories · arXiv:2410.03568
-
CodeJudge: Evaluating Code Generation with Large Language Models 3 Oct 2024 · 1 repository · arXiv:2410.02184Syntology official (archive's flag): 13 ran · 15 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 2 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 23 harvested samples) · 2 pointer-only (licence)
-
Visual Editing with LLM-based Tool Chaining: An Efficient Distillation Approach for Real-Time Applications 3 Oct 2024 · 1 repository · arXiv:2410.02952
-
Emotion-Aware Embedding Fusion in LLMs (Flan-T5, LLAMA 2, DeepSeek-R1, and ChatGPT 4) for Intelligent Response Generation 2 Oct 2024 · 0 repositories · arXiv:2410.01306
-
AlignSum: Data Pyramid Hierarchical Fine-tuning for Aligning with Human Summarization Preference 1 Oct 2024 · 1 repository · arXiv:2410.00409Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Decoding Hate: Exploring Language Models' Reactions to Hate Speech 1 Oct 2024 · 0 repositories · arXiv:2410.00775
-
Language Enhanced Model for Eye (LEME): An Open-Source Ophthalmology-Specific Large Language Model 1 Oct 2024 · 0 repositories · arXiv:2410.03740
-
A Looming Replication Crisis in Evaluating Behavior in Language Models? Evidence and Solutions 30 Sep 2024 · 0 repositories · arXiv:2409.20303
-
Adapting LLMs for the Medical Domain in Portuguese: A Study on Fine-Tuning and Model Evaluation 30 Sep 2024 · 0 repositories · arXiv:2410.00163
-
Charting the Future: Using Chart Question-Answering for Scalable Evaluation of LLM-Driven Data Visualizations 27 Sep 2024 · 0 repositories · arXiv:2409.18764
-
Efficient In-Domain Question Answering for Resource-Constrained Environments 26 Sep 2024 · 0 repositories · arXiv:2409.17648
-
MaskLLM: Learnable Semi-Structured Sparsity for Large Language Models 26 Sep 2024 · 1 repository · arXiv:2409.17481Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
A Prompting-Based Representation Learning Method for Recommendation with Large Language Models 25 Sep 2024 · 0 repositories · arXiv:2409.16674
-
Using LLM for Real-Time Transcription and Summarization of Doctor-Patient Interactions into ePuskesmas in Indonesia 25 Sep 2024 · 0 repositories · arXiv:2409.17054
-
AI Can Be Cognitively Biased: An Exploratory Study on Threshold Priming in LLM-Based Batch Relevance Assessment 24 Sep 2024 · 0 repositories · arXiv:2409.16022
-
Effectiveness of Cross-linguistic Extraction of Genetic Information using Generative Large Language Models 24 Sep 2024 · 1 repository
-
Selection of Prompt Engineering Techniques for Code Generation through Predicting Code Complexity 24 Sep 2024 · 0 repositories · arXiv:2409.16416
-
Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale 24 Sep 2024 · 0 repositories · arXiv:2409.15637
-
Task-oriented Prompt Enhancement via Script Generation 24 Sep 2024 · 0 repositories · arXiv:2409.16418
-
PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs 23 Sep 2024 · 1 repository · arXiv:2409.14866Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
GEM-RAG: Graphical Eigen Memories For Retrieval Augmented Generation 23 Sep 2024 · 0 repositories · arXiv:2409.15566
-
Location is Key: Leveraging Large Language Model for Functional Bug Localization in Verilog 23 Sep 2024 · 0 repositories · arXiv:2409.15186
-
Can pre-trained language models generate titles for research papers? 22 Sep 2024 · 1 repository · arXiv:2409.14602
-
Evaluating the Quality of Code Comments Generated by Large Language Models for Novice Programmers 22 Sep 2024 · 0 repositories · arXiv:2409.14368
-
Proof Automation with Large Language Models 22 Sep 2024 · 0 repositories · arXiv:2409.14274
-
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling 21 Sep 2024 · 1 repository · arXiv:2409.14175
-
Knowledge in Triples for LLMs: Enhancing Table QA Accuracy with Semantic Extraction 21 Sep 2024 · 0 repositories · arXiv:2409.14192
-
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks 20 Sep 2024 · 1 repository · arXiv:2409.13203Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Enhancing TinyBERT for Financial Sentiment Analysis Using GPT-Augmented FinBERT Distillation 19 Sep 2024 · 1 repository · arXiv:2409.18999
-
On the Effectiveness of LLMs for Manual Test Verifications 19 Sep 2024 · 0 repositories · arXiv:2409.12405
-
Retrieval-Augmented Test Generation: How Far Are We? 19 Sep 2024 · 0 repositories · arXiv:2409.12682
-
MAgICoRe: Multi-Agent, Iterative, Coarse-to-Fine Refinement for Reasoning 18 Sep 2024 · 1 repository · arXiv:2409.12147Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
TART: An Open-Source Tool-Augmented Framework for Explainable Table-based Reasoning 18 Sep 2024 · 1 repository · arXiv:2409.11724
-
Small Language Models can Outperform Humans in Short Creative Writing: A Study Comparing SLMs with Humans and LLMs 17 Sep 2024 · 1 repository · arXiv:2409.11547
-
Benchmarking Large Language Model Uncertainty for Prompt Optimization 16 Sep 2024 · 1 repository · arXiv:2409.10044
-
LLM-DER:A Named Entity Recognition Method Based on Large Language Models for Chinese Coal Chemical Domain 16 Sep 2024 · 0 repositories · arXiv:2409.10077
-
SelECT-SQL: Self-correcting ensemble Chain-of-Thought for Text-to-SQL 16 Sep 2024 · 1 repository · arXiv:2409.10007
-
Detection Made Easy: Potentials of Large Language Models for Solidity Vulnerabilities 15 Sep 2024 · 0 repositories · arXiv:2409.10574
-
RethinkMCTS: Refining Erroneous Thoughts in Monte Carlo Tree Search for Code Generation 15 Sep 2024 · 0 repositories · arXiv:2409.09584
-
An empirical evaluation of using ChatGPT to summarize disputes for recommending similar labor and employment cases in Chinese 14 Sep 2024 · 0 repositories · arXiv:2409.09280
-
Optimizing Ingredient Substitution Using Large Language Models to Enhance Phytochemical Content in Recipes 13 Sep 2024 · 0 repositories · arXiv:2409.08792
-
Can Large Language Models Unlock Novel Scientific Research Ideas? 10 Sep 2024 · 1 repository · arXiv:2409.06185Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 11 harvested samples)
-
Classification performance and reproducibility of GPT-4 omni for information extraction from veterinary electronic health records 9 Sep 2024 · 1 repository · arXiv:2409.13727
-
Elsevier Arena: Human Evaluation of Chemistry/Biology/Health Foundational Large Language Models 9 Sep 2024 · 0 repositories · arXiv:2409.05486
-
FairHome: A Fair Housing and Fair Lending Dataset 9 Sep 2024 · 0 repositories · arXiv:2409.05990
-
Harmonic Reasoning in Large Language Models 9 Sep 2024 · 0 repositories · arXiv:2409.05521
-
Identifying the sources of ideological bias in GPT models through linguistic variation in output 9 Sep 2024 · 0 repositories · arXiv:2409.06043
-
Regression with Large Language Models for Materials and Molecular Property Prediction 9 Sep 2024 · 0 repositories · arXiv:2409.06080
-
Vision-fused Attack: Advancing Aggressive and Stealthy Adversarial Text against Neural Machine Translation 8 Sep 2024 · 1 repository · arXiv:2409.05021Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Towards Safer Online Spaces: Simulating and Assessing Intervention Strategies for Eating Disorder Discussions 6 Sep 2024 · 0 repositories · arXiv:2409.04043
-
MaterialBENCH: Evaluating College-Level Materials Science Problem-Solving Abilities of Large Language Models 5 Sep 2024 · 0 repositories · arXiv:2409.03161
-
A Comparative Study on Large Language Models for Log Parsing 4 Sep 2024 · 0 repositories · arXiv:2409.02474
-
How Privacy-Savvy Are Large Language Models? A Case Study on Compliance and Privacy Technical Review 4 Sep 2024 · 0 repositories · arXiv:2409.02375
-
Irrelevant Alternatives Bias Large Language Model Hiring Decisions 4 Sep 2024 · 0 repositories · arXiv:2409.15299
-
Large Language Models as Efficient Reward Function Searchers for Custom-Environment Multi-Objective Reinforcement Learning 4 Sep 2024 · 0 repositories · arXiv:2409.02428
-
More is More: Addition Bias in Large Language Models 4 Sep 2024 · 1 repository · arXiv:2409.02569
-
Self-Judge: Selective Instruction Following with Alignment Self-Evaluation 2 Sep 2024 · 1 repository · arXiv:2409.00935Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Deep Knowledge-Infusion For Explainable Depression Detection 1 Sep 2024 · 0 repositories · arXiv:2409.02122
-
Can Large Language Models Address Open-Target Stance Detection? 30 Aug 2024 · 0 repositories · arXiv:2409.00222
-
ProGRes: Prompted Generative Rescoring on ASR n-Best 30 Aug 2024 · 1 repository · arXiv:2409.00217
-
FRACTURED-SORRY-Bench: Framework for Revealing Attacks in Conversational Turns Undermining Refusal Efficacy and Defenses over SORRY-Bench (Automated Multi-shot Jailbreaks) 28 Aug 2024 · 0 repositories · arXiv:2408.16163
-
Leveraging Large Language Models for Wireless Symbol Detection via In-Context Learning 28 Aug 2024 · 0 repositories · arXiv:2409.00124
-
Strategic Optimization and Challenges of Large Language Models in Object-Oriented Programming 27 Aug 2024 · 0 repositories · arXiv:2408.14834
-
Zero-Shot Visual Reasoning by Vision-Language Models: Benchmarking and Analysis 27 Aug 2024 · 0 repositories · arXiv:2409.00106
-
CodeGraph: Enhancing Graph Reasoning of LLMs with Code 25 Aug 2024 · 1 repository · arXiv:2408.13863
-
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models 25 Aug 2024 · 1 repository · arXiv:2409.00084
-
Enhancing Automated Program Repair with Solution Design 22 Aug 2024 · 0 repositories · arXiv:2408.12056
-
Optimizing Performance: How Compact Models Match or Exceed GPT's Classification Capabilities through Fine-Tuning 22 Aug 2024 · 0 repositories · arXiv:2409.11408
-
Unlocking Adversarial Suffix Optimization Without Affirmative Phrases: Efficient Black-box Jailbreaking via LLM as Optimizer 21 Aug 2024 · 1 repository · arXiv:2408.11313Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
CTP-LLM: Clinical Trial Phase Transition Prediction Using Large Language Models 20 Aug 2024 · 0 repositories · arXiv:2408.10995
-
How Well Do Large Language Models Serve as End-to-End Secure Code Agents for Python? 20 Aug 2024 · 0 repositories · arXiv:2408.10495
-
Soda-Eval: Open-Domain Dialogue Evaluation in the age of LLMs 20 Aug 2024 · 1 repository · arXiv:2408.10902
-
Towardseffective teaching assistants: From intent-based chatbots to LLM-poweredteachingassistants 20 Aug 2024 · 0 repositories
-
TableBench: A Comprehensive and Complex Benchmark for Table Question Answering 17 Aug 2024 · 0 repositories · arXiv:2408.09174
-
A Mean Field Ansatz for Zero-Shot Weight Transfer 16 Aug 2024 · 0 repositories · arXiv:2408.08681
-
Fine-tuning LLMs for Autonomous Spacecraft Control: A Case Study Using Kerbal Space Program 16 Aug 2024 · 1 repository · arXiv:2408.08676
-
FuseChat: Knowledge Fusion of Chat Models 15 Aug 2024 · 3 repositories · arXiv:2408.07990