Methods › Natural Language Processing › Autoregressive Transformers › GPT
GPT
Introduced by Alec Radford et al. in Improving Language Understanding by Generative Pre-Training
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
GPT is a Transformer-based architecture and training procedure for natural language processing tasks. Training follows a two-stage procedure. First, a language modeling objective is used on the unlabeled data to learn the initial parameters of a neural network model. Subsequently, these parameters are adapted to a target task using the corresponding supervised objective.
Source in the archive: Improving Language Understanding by Generative Pre-Training, a link on s3-us-west-2.amazonaws.com (archive link, not checked and not linked: not a paper host this site links to).
Papers archive 2025-07-28
30 shown of 1,212, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
Making Language Model a Hierarchical Classifier and Generator 17 Jul 2025 · 1 repository · arXiv:2507.12930
-
Generative Click-through Rate Prediction with Applications to Search Advertising 15 Jul 2025 · 0 repositories · arXiv:2507.11246
-
Behaviour Space Analysis of LLM-driven Meta-heuristic Discovery 4 Jul 2025 · 0 repositories · arXiv:2507.03605
-
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models 28 Jun 2025 · 0 repositories · arXiv:2506.22957
-
Cat and Mouse -- Can Fake Text Generation Outpace Detector Systems? 26 Jun 2025 · 0 repositories · arXiv:2506.21274
-
Large Language Models Acing Chartered Accountancy 26 Jun 2025 · 0 repositories · arXiv:2506.21031
-
Large Language Model-Driven Code Compliance Checking in Building Information Modeling 25 Jun 2025 · 0 repositories · arXiv:2506.20551
-
InsertRank: LLMs can reason over BM25 scores to Improve Listwise Reranking 17 Jun 2025 · 0 repositories · arXiv:2506.14086
-
Toward a Graph Foundation Model: Pre-Training Transformers With Random Walks 17 Jun 2025 · 0 repositories · arXiv:2506.14098
-
NeuralNexus at BEA 2025 Shared Task: Retrieval-Augmented Prompting for Mistake Identification in AI Tutors 12 Jun 2025 · 1 repository · arXiv:2506.10627
-
Latent Multi-Head Attention for Small Language Models 11 Jun 2025 · 0 repositories · arXiv:2506.09342
-
AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP 10 Jun 2025 · 0 repositories · arXiv:2506.08768
-
Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving 10 Jun 2025 · 1 repository · arXiv:2506.08349Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)
-
Generative Voice Bursts during Phone Call 9 Jun 2025 · 0 repositories · arXiv:2506.07526
-
LLM-driven Indoor Scene Layout Generation via Scaled Human-aligned Data Synthesis and Multi-Stage Preference Optimization 9 Jun 2025 · 0 repositories · arXiv:2506.07570
-
RoboCerebra: A Large-scale Benchmark for Long-horizon Robotic Manipulation Evaluation 7 Jun 2025 · 0 repositories · arXiv:2506.06677
-
Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey 6 Jun 2025 · 0 repositories · arXiv:2506.11102
-
The Lock-in Hypothesis: Stagnation by Algorithm 6 Jun 2025 · 0 repositories · arXiv:2506.06166
-
Mathematical Reasoning for Unmanned Aerial Vehicles: A RAG-Based Approach for Complex Arithmetic Reasoning 5 Jun 2025 · 1 repository · arXiv:2506.04998
-
Privacy and Security Threat for OpenAI GPTs 4 Jun 2025 · 0 repositories · arXiv:2506.04036
-
Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks 30 May 2025 · 1 repository · arXiv:2505.24876
-
MOFGPT: Generative Design of Metal-Organic Frameworks using Language Models 30 May 2025 · 1 repository · arXiv:2506.00198
-
When GPT Spills the Tea: Comprehensive Assessment of Knowledge File Leakage in GPTs 30 May 2025 · 0 repositories · arXiv:2506.00197
-
Daunce: Data Attribution through Uncertainty Estimation 29 May 2025 · 0 repositories · arXiv:2505.23223
-
Reducing Latency in LLM-Based Natural Language Commands Processing for Robot Navigation 29 May 2025 · 0 repositories · arXiv:2506.00075
-
Multi-MLLM Knowledge Distillation for Out-of-Context News Detection 28 May 2025 · 0 repositories · arXiv:2505.22517
-
Explainability of Large Language Models using SMILE: Statistical Model-agnostic Interpretability with Local Explanations 27 May 2025 · 1 repository · arXiv:2505.21657
-
From prosthetic memory to prosthetic denial: Auditing whether large language models are prone to mass atrocity denialism 27 May 2025 · 0 repositories · arXiv:2505.21753
-
Automated evaluation of children's speech fluency for low-resource languages 26 May 2025 · 0 repositories · arXiv:2505.19671
-
Beyond Specialization: Benchmarking LLMs for Transliteration of Indian Languages 26 May 2025 · 0 repositories · arXiv:2505.19851
Tasks archive 2025-07-28
20 shown of 583 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections