Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 4
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 4 of 20: papers 301 to 400 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Leveraging Web-Crawled Data for High-Quality Fine-Tuning 15 Aug 2024 · 1 repository · arXiv:2408.08003
-
Predicting Lung Cancer Patient Prognosis with Large Language Models 15 Aug 2024 · 0 repositories · arXiv:2408.07971
-
CodeMirage: Hallucinations in Code Generated by Large Language Models 14 Aug 2024 · 0 repositories · arXiv:2408.08333
-
SAGE-RT: Synthetic Alignment data Generation for Safety Evaluation and Red Teaming 14 Aug 2024 · 0 repositories · arXiv:2408.11851
-
Evaluating Cultural Adaptability of a Large Language Model via Simulation of Synthetic Personas 13 Aug 2024 · 1 repository · arXiv:2408.06929
-
Kov: Transferable and Naturalistic Black-Box LLM Attacks using Markov Decision Processes and Tree Search 11 Aug 2024 · 1 repository · arXiv:2408.08899
-
PhishLang: A Real-Time, Fully Client-Side Phishing Detection Framework Using MobileBERT 11 Aug 2024 · 2 repositories · arXiv:2408.05667
-
Chain of Condition: Construct, Verify and Solve Conditions for Conditional Question Answering 10 Aug 2024 · 0 repositories · arXiv:2408.05442
-
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis 9 Aug 2024 · 1 repository · arXiv:2408.05006
-
Examining the Behavior of LLM Architectures Within the Framework of Standardized National Exams in Brazil 9 Aug 2024 · 0 repositories · arXiv:2408.05035
-
Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants 7 Aug 2024 · 0 repositories · arXiv:2408.11841
-
Query3D: LLM-Powered Open-Vocabulary Scene Segmentation with Language Embedded 3D Gaussian 7 Aug 2024 · 1 repository · arXiv:2408.03516
-
The Use of Large Language Models (LLM) for Cyber Threat Intelligence (CTI) in Cybercrime Forums 6 Aug 2024 · 0 repositories · arXiv:2408.03354
-
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02416
-
XMainframe: A Large Language Model for Mainframe Modernization 5 Aug 2024 · 1 repository · arXiv:2408.04660
-
Effective Demonstration Annotation for In-Context Learning via Language Model-Based Determinantal Point Process 4 Aug 2024 · 0 repositories · arXiv:2408.02103
-
Leveraging Large Language Models with Chain-of-Thought and Prompt Engineering for Traffic Crash Severity Analysis and Inference 4 Aug 2024 · 0 repositories · arXiv:2408.04652
-
Building Trust in Mental Health Chatbots: Safety Metrics and LLM-Based Evaluation Tools 3 Aug 2024 · 0 repositories · arXiv:2408.04650
-
LLM as Runtime Error Handler: A Promising Pathway to Adaptive Self-Healing of Software Systems 2 Aug 2024 · 0 repositories · arXiv:2408.01055
-
High-Throughput Phenotyping of Clinical Text Using Large Language Models 2 Aug 2024 · 0 repositories · arXiv:2408.01214
-
Improving Retrieval-Augmented Generation in Medicine with Iterative Follow-up Questions 1 Aug 2024 · 1 repository · arXiv:2408.00727
-
AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation 1 Aug 2024 · 1 repository · arXiv:2408.00764Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Improving Faithfulness of Large Language Models in Summarization via Sliding Generation and Self-Consistency 31 Jul 2024 · 0 repositories · arXiv:2407.21443
-
Decomposed Prompting to Answer Questions on a Course Discussion Board 30 Jul 2024 · 1 repository · arXiv:2407.21170
-
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification 30 Jul 2024 · 0 repositories · arXiv:2407.20859
-
Comparison of Large Language Models for Generating Contextually Relevant Questions 30 Jul 2024 · 1 repository · arXiv:2407.20578
-
Enhancing Code Translation in Language Models with Few-Shot Learning via Retrieval-Augmented Generation 29 Jul 2024 · 0 repositories · arXiv:2407.19619
-
Are LLMs Good Annotators for Discourse-level Event Relation Extraction? 28 Jul 2024 · 1 repository · arXiv:2407.19568
-
TAGIFY: LLM-powered Tagging Interface for Improved Data Findability on OGD portals 26 Jul 2024 · 0 repositories · arXiv:2407.18764
-
Using Large Language Models for the Interpretation of Building Regulations 26 Jul 2024 · 0 repositories · arXiv:2407.21060
-
Closing the gap between open-source and commercial large language models for medical evidence summarization 25 Jul 2024 · 0 repositories · arXiv:2408.00588
-
Cost-effective Instruction Learning for Pathology Vision and Language Analysis 25 Jul 2024 · 1 repository · arXiv:2407.17734
-
PEFT-U: Parameter-Efficient Fine-Tuning for User Personalization 25 Jul 2024 · 1 repository · arXiv:2407.18078Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Bailicai: A Domain-Optimized Retrieval-Augmented Generation Framework for Medical Applications 24 Jul 2024 · 0 repositories · arXiv:2407.21055
-
Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles 24 Jul 2024 · 0 repositories · arXiv:2407.17211
-
Data Mixture Inference: What do BPE Tokenizers Reveal about their Training Data? 23 Jul 2024 · 1 repository · arXiv:2407.16607Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing LLM's Cognition via Structurization 23 Jul 2024 · 1 repository · arXiv:2407.16434Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Patched RTC: evaluating LLMs for diverse software development tasks 23 Jul 2024 · 1 repository · arXiv:2407.16557
-
Robust Privacy Amidst Innovation with Large Language Models Through a Critical Assessment of the Risks 23 Jul 2024 · 1 repository · arXiv:2407.16166
-
Imposter.AI: Adversarial Attacks with Hidden Intentions towards Aligned Large Language Models 22 Jul 2024 · 0 repositories · arXiv:2407.15399
-
MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Diversity 22 Jul 2024 · 1 repository · arXiv:2407.15838Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
RadioRAG: Factual large language models for enhanced diagnostics in radiology using online retrieval augmented generation 22 Jul 2024 · 1 repository · arXiv:2407.15621
-
Unlocking the Potential: Benchmarking Large Language Models in Water Engineering and Research 22 Jul 2024 · 0 repositories · arXiv:2407.21045
-
SQLfuse: Enhancing Text-to-SQL Performance through Comprehensive LLM Synergy 19 Jul 2024 · 0 repositories · arXiv:2407.14568
-
Can Open-Source LLMs Compete with Commercial Models? Exploring the Few-Shot Performance of Current GPT Models in Biomedical Tasks 18 Jul 2024 · 1 repository · arXiv:2407.13511
-
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency 18 Jul 2024 · 0 repositories · arXiv:2407.13578
-
Learning-From-Mistakes Prompting for Indigenous Language Translation 18 Jul 2024 · 0 repositories · arXiv:2407.13343
-
PRAGyan -- Connecting the Dots in Tweets 18 Jul 2024 · 0 repositories · arXiv:2407.13909
-
Does Refusal Training in LLMs Generalize to the Past Tense? 16 Jul 2024 · 1 repository · arXiv:2407.11969Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text 16 Jul 2024 · 0 repositories · arXiv:2407.11827
-
Large Language Models as Misleading Assistants in Conversation 16 Jul 2024 · 0 repositories · arXiv:2407.11789
-
Large Visual-Language Models Are Also Good Classifiers: A Study of In-Context Multimodal Fake News Detection 16 Jul 2024 · 0 repositories · arXiv:2407.12879
-
Representation Bias in Political Sample Simulations with Large Language Models 16 Jul 2024 · 0 repositories · arXiv:2407.11409
-
ReFeR: Improving Evaluation and Reasoning through Hierarchy of Models 16 Jul 2024 · 0 repositories · arXiv:2407.12877
-
CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization 15 Jul 2024 · 0 repositories · arXiv:2407.10424
-
Leveraging LLM-Respondents for Item Evaluation: a Psychometric Analysis 15 Jul 2024 · 0 repositories · arXiv:2407.10899
-
Think-on-Graph 2.0: Deep and Faithful Large Language Model Reasoning with Knowledge-guided Retrieval Augmented Generation 15 Jul 2024 · 1 repository · arXiv:2407.10805Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Self-Evolving GPT: A Lifelong Autonomous Experiential Learner 12 Jul 2024 · 0 repositories · arXiv:2407.08937
-
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay 12 Jul 2024 · 1 repository · arXiv:2407.11068
-
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs 12 Jul 2024 · 0 repositories · arXiv:2407.09152
-
GPT-4 is judged more human than humans in displaced and inverted Turing tests 11 Jul 2024 · 0 repositories · arXiv:2407.08853
-
LLMs' morphological analyses of complex FST-generated Finnish words 11 Jul 2024 · 1 repository · arXiv:2407.08269
-
Vox Populi, Vox AI? Using Language Models to Estimate German Public Opinion 11 Jul 2024 · 1 repository · arXiv:2407.08563
-
FsPONER: Few-shot Prompt Optimization for Named Entity Recognition in Domain-specific Scenarios 10 Jul 2024 · 1 repository · arXiv:2407.08035
-
A Guide To Effectively Leveraging LLMs for Low-Resource Text Summarization: Data Augmentation and Semi-supervised Approaches 10 Jul 2024 · 0 repositories · arXiv:2407.07341
-
Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture 10 Jul 2024 · 0 repositories · arXiv:2407.07342
-
AI AI Bias: Large Language Models Favor Their Own Generated Content 9 Jul 2024 · 1 repository · arXiv:2407.12856
-
ChatGPT Doesn't Trust Chargers Fans: Guardrail Sensitivity in Context 9 Jul 2024 · 1 repository · arXiv:2407.06866Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Identification of emotions on Twitter during the 2022 electoral process in Colombia 9 Jul 2024 · 0 repositories · arXiv:2407.07258
-
Prompting Techniques for Secure Code Generation: A Systematic Investigation 9 Jul 2024 · 0 repositories · arXiv:2407.07064
-
Using Large Language Models for Generating Smart Contracts for Health Insurance from Textual Policies 9 Jul 2024 · 0 repositories · arXiv:2407.07019
-
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct 8 Jul 2024 · 1 repository · arXiv:2407.05700Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Large Language Models Understand Layout 8 Jul 2024 · 1 repository · arXiv:2407.05750
-
SHINE: Saliency-aware HIerarchical NEgative Ranking for Compositional Temporal Grounding 6 Jul 2024 · 1 repository · arXiv:2407.05118Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Are Large Language Models Strategic Decision Makers? A Study of Performance and Bias in Two-Player Non-Zero-Sum Games 5 Jul 2024 · 0 repositories · arXiv:2407.04467
-
From Data to Commonsense Reasoning: The Use of Large Language Models for Explainable AI 4 Jul 2024 · 0 repositories · arXiv:2407.03778
-
NutriBench: A Dataset for Evaluating Large Language Models on Nutrition Estimation from Meal Descriptions 4 Jul 2024 · 0 repositories · arXiv:2407.12843
-
Controllable Conversations: Planning-Based Dialogue Agent with Large Language Models 4 Jul 2024 · 1 repository · arXiv:2407.03884
-
Towards Automating Text Annotation: A Case Study on Semantic Proximity Annotation using GPT-4 4 Jul 2024 · 0 repositories · arXiv:2407.04130
-
AgentInstruct: Toward Generative Teaching with Agentic Flows 3 Jul 2024 · 0 repositories · arXiv:2407.03502
-
Regurgitative Training: The Value of Real Data in Training Large Language Models 3 Jul 2024 · 0 repositories · arXiv:2407.12835
-
Assessing the Code Clone Detection Capability of Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02402
-
Beyond Numeric Awards: In-Context Dueling Bandits with LLM Agents 2 Jul 2024 · 0 repositories · arXiv:2407.01887
-
GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning 2 Jul 2024 · 0 repositories · arXiv:2407.01892
-
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation 2 Jul 2024 · 1 repository · arXiv:2407.02056Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Language Model Alignment in Multilingual Trolley Problems 2 Jul 2024 · 2 repositories · arXiv:2407.02273Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters 2 Jul 2024 · 1 repository · arXiv:2407.01902Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Increasing Model Capacity for Free: A Simple Strategy for Parameter Efficient Fine-tuning 1 Jul 2024 · 1 repository · arXiv:2407.01320Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 7 pointer-only (licence)
-
Predicting DC-Link Capacitor Current Ripple in AC-DC Rectifier Circuits Using Fine-Tuned Large Language Models 1 Jul 2024 · 0 repositories · arXiv:2407.01724
-
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs 28 Jun 2024 · 0 repositories · arXiv:2406.20060
-
FRED: Flexible REduction-Distribution Interconnect and Communication Implementation for Wafer-Scale Distributed Training of DNN Models 28 Jun 2024 · 0 repositories · arXiv:2406.19580
-
ShortcutsBench: A Large-Scale Real-world Benchmark for API-based Agents 28 Jun 2024 · 1 repository · arXiv:2407.00132Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data 27 Jun 2024 · 1 repository · arXiv:2406.19292
-
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models 27 Jun 2024 · 0 repositories · arXiv:2406.19358
-
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets 26 Jun 2024 · 0 repositories · arXiv:2406.18518
-
Improving Entity Recognition Using Ensembles of Deep Learning and Fine-tuned Large Language Models: A Case Study on Adverse Event Extraction from Multiple Sources 26 Jun 2024 · 0 repositories · arXiv:2406.18049
-
Evaluation of Instruction-Following Ability for Large Language Models on Story-Ending Generation 24 Jun 2024 · 0 repositories · arXiv:2406.16356
-
PlagBench: Exploring the Duality of Large Language Models in Plagiarism Generation and Detection 24 Jun 2024 · 0 repositories · arXiv:2406.16288
-
The GPT-WritingPrompts Dataset: A Comparative Analysis of Character Portrayal in Short Stories 24 Jun 2024 · 1 repository · arXiv:2406.16767
-
Towards Better Graph-based Cross-document Relation Extraction via Non-bridge Entity Enhancement and Prediction Debiasing 24 Jun 2024 · 1 repository · arXiv:2406.16529