Methods › Natural Language Processing › Transformers › GPT-Neo
GPT-Neo
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
An implementation of model & data parallel GPT3-like models using the mesh-tensorflow library.
Source: EleutherAI/GPT-Neo
Papers archive 2025-07-28
30 shown of 38, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
IRepair: An Intent-Aware Approach to Repair Data-Driven Errors in Large Language Models 10 Feb 2025 · 0 repositories · arXiv:2502.07072
-
Robust Hybrid Classical-Quantum Transfer Learning Model for Text Classification Using GPT-Neo 125M with LoRA & SMOTE Enhancement 12 Jan 2025 · 1 repository · arXiv:2501.10435
-
LLM Vocabulary Compression for Low-Compute Environments 10 Nov 2024 · 0 repositories · arXiv:2411.06371
-
BERTtime Stories: Investigating the Role of Synthetic Story Data in Language pre-training 20 Oct 2024 · 1 repository · arXiv:2410.15365
-
Reconstruction of Differentially Private Text Sanitization via Large Language Models 16 Oct 2024 · 0 repositories · arXiv:2410.12443
-
The Unreasonable Ineffectiveness of Nucleus Sampling on Mitigating Text Memorization 29 Aug 2024 · 1 repository · arXiv:2408.16345Syntology ran 8 of 11 samples · 3 unverified · 11 pointer-only (licence)
-
WPN: An Unlearning Method Based on N-pair Contrastive Learning in Language Models 18 Aug 2024 · 0 repositories · arXiv:2408.09459
-
Towards Robust and Parameter-Efficient Knowledge Unlearning for LLMs 13 Aug 2024 · 1 repository · arXiv:2408.06621Syntology ran 2 of 2 samples · 0 unverified
-
Semantic Membership Inference Attack against Large Language Models 14 Jun 2024 · 0 repositories · arXiv:2406.10218
-
Investigating Wit, Creativity, and Detectability of Large Language Models in Domain-Specific Writing Style Adaptation of Reddit's Showerthoughts 2 May 2024 · 1 repository · arXiv:2405.01660
-
More than Correlation: Do Large Language Models Learn Causal Representations of Space? 26 Dec 2023 · 0 repositories · arXiv:2312.16257
-
Fairness-Aware Structured Pruning in Transformers 24 Dec 2023 · 1 repository · arXiv:2312.15398Syntology ran 0 of 2 samples · 2 unverified
-
Scalable Extraction of Training Data from (Production) Language Models 28 Nov 2023 · 0 repositories · arXiv:2311.17035
-
Heaps' Law in GPT-Neo Large Language Model Emulated Corpora 10 Nov 2023 · 1 repository · arXiv:2311.06377
-
Watermarking LLMs with Weight Quantization 17 Oct 2023 · 1 repository · arXiv:2310.11237Syntology ran 1 of 10 samples · 9 unverified · 10 pointer-only (licence)
-
TART: A plug-and-play Transformer module for task-agnostic reasoning 21 Sep 2023 · 1 repository
-
Fine-Tuning Large Language Models for Answering Programming Questions with Code Snippets 26 Jun 2023 · 0 repositories
-
Exposing Bias in Online Communities through Large-Scale Language Models 4 Jun 2023 · 0 repositories · arXiv:2306.02294
-
Test-Time Training on Nearest Neighbors for Large Language Models 29 May 2023 · 1 repository · arXiv:2305.18466Syntology ran 0 of 5 samples · 5 unverified
-
Controlling the Extraction of Memorized Data from Large Language Models via Prompt-Tuning 19 May 2023 · 1 repository · arXiv:2305.11759
-
TinyStories: How Small Can Language Models Be and Still Speak Coherent English? 12 May 2023 · 8 repositories · arXiv:2305.07759Syntology ran 3 of 18 samples · 15 unverified
-
Stealing the Decoding Algorithms of Language Models 8 Mar 2023 · 1 repository · arXiv:2303.04729
-
Bag of Tricks for Training Data Extraction from Language Models 9 Feb 2023 · 1 repository · arXiv:2302.04460Syntology ran 0 of 1 samples · 1 unverified · 1 pointer-only (licence)
-
Why Does Surprisal From Larger Transformer-Based Language Models Provide a Poorer Fit to Human Reading Times? 23 Dec 2022 · 0 repositories · arXiv:2212.12131
-
Explicit Knowledge Transfer for Weakly-Supervised Code Generation 30 Nov 2022 · 0 repositories · arXiv:2211.16740
-
GPT-Neo for commonsense reasoning -- a theoretical and practical lens 28 Nov 2022 · 1 repository · arXiv:2211.15593
-
Understanding BLOOM: An empirical study on diverse NLP tasks 27 Nov 2022 · 0 repositories · arXiv:2211.14865
-
Collateral facilitation in humans and language models 9 Nov 2022 · 1 repository · arXiv:2211.05198Syntology ran 1 of 1 samples · 0 unverified
-
Chain of Explanation: New Prompting Method to Generate Higher Quality Natural Language Explanation for Implicit Hate Speech 11 Sep 2022 · 0 repositories · arXiv:2209.04889
-
Materials Transformers Language Models for Generative Materials Design: a benchmark study 27 Jun 2022 · 1 repository · arXiv:2206.13578
Tasks archive 2025-07-28
20 shown of 42 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections