Methods › Natural Language Processing › Transformers › GPT-3 › Papers, page 6
GPT-3
Papers archive 2025-07-28
archive papers tagged: 1,906 · with a code link: 866 · where Syntology ran a sample: 319 (259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (319 of 1,906 tagged: 259 with a run with no instrument failure, 60 where every run was a failure of Syntology's instrument)
Page 6 of 20: papers 501 to 600 of 1,906, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Cross-Language Assessment of Mathematical Capability of ChatGPT 18 May 2024 · 0 repositories · arXiv:2405.11264
-
Evaluation of large language model performance on the Biomedical Language Understanding and Reasoning Benchmark 17 May 2024 · 0 repositories
-
FinTextQA: A Dataset for Long-form Financial Question Answering 16 May 2024 · 0 repositories · arXiv:2405.09980
-
Optimization Techniques for Sentiment Analysis Based on LLM (GPT-3) 16 May 2024 · 0 repositories · arXiv:2405.09770
-
Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis 14 May 2024 · 0 repositories · arXiv:2405.08944
-
GPT-3.5 for Grammatical Error Correction 14 May 2024 · 0 repositories · arXiv:2405.08469
-
Can Language Models Explain Their Own Classification Behavior? 13 May 2024 · 1 repository · arXiv:2405.07436
-
Coding historical causes of death data with Large Language Models 13 May 2024 · 1 repository · arXiv:2405.07560
-
Many-Shot Regurgitation (MSR) Prompting 13 May 2024 · 0 repositories · arXiv:2405.08134
-
Limited Ability of LLMs to Simulate Human Psychological Behaviours: a Psychometric Analysis 12 May 2024 · 1 repository · arXiv:2405.07248Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
An Assessment of Model-On-Model Deception 10 May 2024 · 0 repositories · arXiv:2405.12999
-
Can large language models understand uncommon meanings of common words? 9 May 2024 · 0 repositories · arXiv:2405.05741
-
Digital Diagnostics: The Potential Of Large Language Models In Recognizing Symptoms Of Common Illnesses 9 May 2024 · 0 repositories · arXiv:2405.06712
-
Iris: An AI-Driven Virtual Tutor For Computer Science Education 9 May 2024 · 0 repositories · arXiv:2405.08008
-
People cannot distinguish GPT-4 from a human in a Turing test 9 May 2024 · 0 repositories · arXiv:2405.08007
-
Reddit-Impacts: A Named Entity Recognition Dataset for Analyzing Clinical and Social Effects of Substance Use Derived from Social Media 9 May 2024 · 0 repositories · arXiv:2405.06145
-
LLMs Can Patch Up Missing Relevance Judgments in Evaluation 8 May 2024 · 0 repositories · arXiv:2405.04727
-
Long Context Alignment with Short Instructions and Synthesized Positions 7 May 2024 · 0 repositories · arXiv:2405.03939
-
SUTRA: Scalable Multilingual Language Model Architecture 7 May 2024 · 0 repositories · arXiv:2405.06694
-
The Silicon Ceiling: Auditing GPT's Race and Gender Biases in Hiring 7 May 2024 · 0 repositories · arXiv:2405.04412
-
Hire Me or Not? Examining Language Model's Behavior with Occupation Attributes 6 May 2024 · 1 repository · arXiv:2405.06687Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Large Language Models Reveal Information Operation Goals, Tactics, and Narrative Frames 6 May 2024 · 1 repository · arXiv:2405.03688
-
Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models 5 May 2024 · 0 repositories · arXiv:2405.02917
-
Unraveling the Dominance of Large Language Models Over Transformer Models for Bangla Natural Language Inference: A Comprehensive Study 5 May 2024 · 1 repository · arXiv:2405.02937
-
Evaluating Large Language Models for Structured Science Summarization in the Open Research Knowledge Graph 3 May 2024 · 0 repositories · arXiv:2405.02105
-
Exploring Combinatorial Problem Solving with Large Language Models: A Case Study on the Travelling Salesman Problem Using GPT-3.5 Turbo 3 May 2024 · 0 repositories · arXiv:2405.01997
-
Attribution in Scientific Literature: New Benchmark and Methods 3 May 2024 · 0 repositories · arXiv:2405.02228
-
A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law 2 May 2024 · 1 repository · arXiv:2405.01769
-
Bayesian Optimization with LLM-Based Acquisition Functions for Natural Language Preference Elicitation 2 May 2024 · 0 repositories · arXiv:2405.00981
-
Investigating Wit, Creativity, and Detectability of Large Language Models in Domain-Specific Writing Style Adaptation of Reddit's Showerthoughts 2 May 2024 · 1 repository · arXiv:2405.01660
-
CourseAssist: Pedagogically Appropriate AI Tutor for Computer Science Education 1 May 2024 · 0 repositories · arXiv:2407.10246
-
How Can I Improve? Using GPT to Highlight the Desired and Undesired Parts of Open-ended Responses 1 May 2024 · 0 repositories · arXiv:2405.00291
-
Do Large Language Models Understand Conversational Implicature -- A case study with a chinese sitcom 30 Apr 2024 · 1 repository · arXiv:2404.19509Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Graphical Reasoning: LLM-based Semi-Open Relation Extraction 30 Apr 2024 · 1 repository · arXiv:2405.00216
-
TuBA: Cross-Lingual Transferability of Backdoor Attacks in LLMs with Instruction Tuning 30 Apr 2024 · 0 repositories · arXiv:2404.19597
-
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models 29 Apr 2024 · 1 repository · arXiv:2404.18353
-
Evaluating and Mitigating Linguistic Discrimination in Large Language Models 29 Apr 2024 · 0 repositories · arXiv:2404.18534
-
GPT-4 passes most of the 297 written Polish Board Certification Examinations 29 Apr 2024 · 0 repositories · arXiv:2405.01589
-
PECC: Problem Extraction and Coding Challenges 29 Apr 2024 · 1 repository · arXiv:2404.18766
-
Detection of Conspiracy Theories Beyond Keyword Bias in German-Language Telegram Using Large Language Models 27 Apr 2024 · 0 repositories · arXiv:2404.17985
-
Evaluation of Few-Shot Learning for Classification Tasks in the Polish Language 27 Apr 2024 · 0 repositories · arXiv:2404.17832
-
Automated Data Visualization from Natural Language via Large Language Models: An Exploratory Study 26 Apr 2024 · 1 repository · arXiv:2404.17136
-
IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages 25 Apr 2024 · 1 repository · arXiv:2404.16816
-
Influence of Solution Efficiency and Valence of Instruction on Additive and Subtractive Solution Strategies in Humans and GPT-4 25 Apr 2024 · 0 repositories · arXiv:2404.16692
-
WorldValuesBench: A Large-Scale Benchmark Dataset for Multi-Cultural Value Awareness of Language Models 25 Apr 2024 · 1 repository · arXiv:2404.16308Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
The Promise and Challenges of Using LLMs to Accelerate the Screening Process of Systematic Reviews 24 Apr 2024 · 0 repositories · arXiv:2404.15667
-
Evaluating the Efficacy of Large Language Models in Identifying Phishing Attempts 23 Apr 2024 · 0 repositories · arXiv:2404.15485
-
PRISM: Patient Records Interpretation for Semantic Clinical Trial Matching using Large Language Models 23 Apr 2024 · 0 repositories · arXiv:2404.15549
-
Watch Out for Your Guidance on Generation! Exploring Conditional Backdoor Attacks against Large Language Models 23 Apr 2024 · 0 repositories · arXiv:2404.14795
-
Unsupervised End-to-End Task-Oriented Dialogue with LLMs: The Power of the Noisy Channel 23 Apr 2024 · 1 repository · arXiv:2404.15219
-
Generating Attractive and Authentic Copywriting from Customer Reviews 22 Apr 2024 · 0 repositories · arXiv:2404.13906
-
How Well Can LLMs Echo Us? Evaluating AI Chatbots' Role-Play Ability with ECHO 22 Apr 2024 · 1 repository · arXiv:2404.13957Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Information Re-Organization Improves Reasoning in Large Language Models 22 Apr 2024 · 0 repositories · arXiv:2404.13985
-
Navigating the Path of Writing: Outline-guided Text Generation with Large Language Models 22 Apr 2024 · 0 repositories · arXiv:2404.13919
-
Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone 22 Apr 2024 · 0 repositories · arXiv:2404.14219
-
SVGEditBench: A Benchmark Dataset for Quantitative Assessment of LLM's SVG Editing Capabilities 21 Apr 2024 · 1 repository · arXiv:2404.13710Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Data Alignment for Zero-Shot Concept Generation in Dermatology AI 19 Apr 2024 · 0 repositories · arXiv:2404.13043
-
Dubo-SQL: Diverse Retrieval-Augmented Generation and Fine Tuning for Text-to-SQL 19 Apr 2024 · 1 repository · arXiv:2404.12560
-
The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions 19 Apr 2024 · 1 repository · arXiv:2404.13208
-
From Form(s) to Meaning: Probing the Semantic Depths of Language Models Using Multisense Consistency 18 Apr 2024 · 1 repository · arXiv:2404.12145Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Comparative Analysis of Deep Natural Networks and Large Language Models for Aspect-Based Sentiment Analysis 17 Apr 2024 · 1 repository
-
Reinforcement Learning from Multi-role Debates as Feedback for Bias Mitigation in LLMs 15 Apr 2024 · 0 repositories · arXiv:2404.10160
-
The Fault in our Stars: Quality Assessment of Code Generation Benchmarks 15 Apr 2024 · 0 repositories · arXiv:2404.10155
-
CreativEval: Evaluating Creativity of LLM-Based Hardware Code Generation 12 Apr 2024 · 0 repositories · arXiv:2404.08806
-
Is ChatGPT Transforming Academics' Writing Style? 12 Apr 2024 · 0 repositories · arXiv:2404.08627
-
Small Models Are (Still) Effective Cross-Domain Argument Extractors 12 Apr 2024 · 1 repository · arXiv:2404.08579
-
AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs 11 Apr 2024 · 1 repository · arXiv:2404.07921
-
Characterizing Multimodal Long-form Summarization: A Case Study on Financial Reports 9 Apr 2024 · 0 repositories · arXiv:2404.06162
-
Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs 9 Apr 2024 · 0 repositories · arXiv:2404.07242
-
LTNER: Large Language Model Tagging for Named Entity Recognition with Contextualized Entity Marking 8 Apr 2024 · 0 repositories · arXiv:2404.05624
-
PetKaz at SemEval-2024 Task 3: Advancing Emotion Classification with an LLM for Emotion-Cause Pair Extraction in Conversations 8 Apr 2024 · 1 repository · arXiv:2404.05502
-
Relation Extraction Using Large Language Models: A Case Study on Acupuncture Point Locations 8 Apr 2024 · 0 repositories · arXiv:2404.05415
-
IITK at SemEval-2024 Task 2: Exploring the Capabilities of LLMs for Safe Biomedical Natural Language Inference for Clinical Trials 6 Apr 2024 · 1 repository · arXiv:2404.04510
-
AI-Tutoring in Software Engineering Education 3 Apr 2024 · 0 repositories · arXiv:2404.02548
-
Benchmarking Large Language Models for Persian: A Preliminary Study Focusing on ChatGPT 3 Apr 2024 · 1 repository · arXiv:2404.02403
-
GPT-DETOX: An In-Context Learning-Based Paraphraser for Text Detoxification 3 Apr 2024 · 0 repositories · arXiv:2404.03052
-
uTeBC-NLP at SemEval-2024 Task 9: Can LLMs be Lateral Thinkers? 3 Apr 2024 · 1 repository · arXiv:2404.02474Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Advancing LLM Reasoning Generalists with Preference Trees 2 Apr 2024 · 1 repository · arXiv:2404.02078Syntology official (archive's flag): 18 ran · 18 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 1 violated, 12 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 20 harvested samples) · 2 pointer-only (licence)
-
CMAT: A Multi-Agent Collaboration Tuning Framework for Enhancing Small Language Models 2 Apr 2024 · 1 repository · arXiv:2404.01663
-
Comparative Study of Domain Driven Terms Extraction Using Large Language Models 2 Apr 2024 · 0 repositories · arXiv:2404.02330
-
Deconstructing In-Context Learning: Understanding Prompts via Corruption 2 Apr 2024 · 1 repository · arXiv:2404.02054
-
Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks 2 Apr 2024 · 1 repository · arXiv:2404.02151Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
METAL: Towards Multilingual Meta-Evaluation 2 Apr 2024 · 0 repositories · arXiv:2404.01667
-
SGSH: Stimulate Large Language Models with Skeleton Heuristics for Knowledge Base Question Generation 2 Apr 2024 · 1 repository · arXiv:2404.01923
-
Toward Informal Language Processing: Knowledge of Slang in Large Language Models 2 Apr 2024 · 1 repository · arXiv:2404.02323
-
BERT-Enhanced Retrieval Tool for Homework Plagiarism Detection System 1 Apr 2024 · 0 repositories · arXiv:2404.01582
-
FABLES: Evaluating faithfulness and content selection in book-length summarization 1 Apr 2024 · 3 repositories · arXiv:2404.01261Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Unveiling Divergent Inductive Biases of LLMs on Temporal Data 1 Apr 2024 · 1 repository · arXiv:2404.01453
-
CHOPS: CHat with custOmer Profile Systems for Customer Service with LLMs 31 Mar 2024 · 1 repository · arXiv:2404.01343
-
CoUDA: Coherence Evaluation via Unified Data Augmentation 31 Mar 2024 · 1 repository · arXiv:2404.00681
-
Training-Free Semantic Segmentation via LLM-Supervision 31 Mar 2024 · 0 repositories · arXiv:2404.00701
-
A Comprehensive Study on NLP Data Augmentation for Hate Speech Detection: Legacy Methods, BERT, and LLMs 30 Mar 2024 · 0 repositories · arXiv:2404.00303
-
Small Language Models Learn Enhanced Reasoning Skills from Medical Textbooks 30 Mar 2024 · 0 repositories · arXiv:2404.00376
-
DataAgent: Evaluating Large Language Models' Ability to Answer Zero-Shot, Natural Language Queries 29 Mar 2024 · 0 repositories · arXiv:2404.00188
-
ReALM: Reference Resolution As Language Modeling 29 Mar 2024 · 0 repositories · arXiv:2403.20329
-
FACTOID: FACtual enTailment fOr hallucInation Detection 28 Mar 2024 · 0 repositories · arXiv:2403.19113
-
Evaluating Large Language Models for Health-Related Text Classification Tasks with Public Social Media Data 27 Mar 2024 · 0 repositories · arXiv:2403.19031
-
LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research Practices 27 Mar 2024 · 0 repositories · arXiv:2403.18173
-
Reshaping Free-Text Radiology Notes Into Structured Reports With Generative Transformers 27 Mar 2024 · 1 repository · arXiv:2403.18938
-
Vulnerability Detection with Code Language Models: How Far Are We? 27 Mar 2024 · 1 repository · arXiv:2403.18624Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)