Methods › Natural Language Processing › Language Models › Pythia › Papers where code ran, page 1
Pythia
Papers archive 2025-07-28
archive papers tagged: 60 · with a code link: 36 · where Syntology ran a sample: 23 (21 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (23 of 60 tagged: 21 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 1 of 1: papers 1 to 23 of the 23 tagged papers where Syntology ran at least one harvested sample (21 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
What Happens During the Loss Plateau? Understanding Abrupt Learning in Transformers 16 Jun 2025 · 1 repository · arXiv:2506.13688Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Pretraining Language Models to Ponder in Continuous Space 27 May 2025 · 1 repository · arXiv:2505.20674Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 4 pointer-only (licence)
-
RoSTE: An Efficient Quantization-Aware Supervised Fine-Tuning Approach for Large Language Models 13 Feb 2025 · 0 repositories · arXiv:2502.09003Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Local and Global Decoding in Text Generation 14 Oct 2024 · 1 repository · arXiv:2410.10810Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Demystifying Verbatim Memorization in Large Language Models 25 Jul 2024 · 1 repository · arXiv:2407.17817Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Generalization v.s. Memorization: Tracing Language Models' Capabilities Back to Pretraining Data 20 Jul 2024 · 0 repositories · arXiv:2407.14985Syntology 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 15 harvested samples)
-
MathCAMPS: Fine-grained Synthesis of Mathematical Problems From Human Curricula 1 Jul 2024 · 1 repository · arXiv:2407.00900Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models 10 Jun 2024 · 1 repository · arXiv:2406.06046Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 2 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Causal Estimation of Memorisation Profiles 6 Jun 2024 · 1 repository · arXiv:2406.04327Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
MAP-Neo: Highly Capable and Transparent Bilingual Large Language Model Series 29 May 2024 · 1 repository · arXiv:2405.19327Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations 27 Mar 2024 · 1 repository · arXiv:2403.18167Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
The N+ Implementation Details of RLHF with PPO: A Case Study on TL;DR Summarization 24 Mar 2024 · 1 repository · arXiv:2403.17031Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples)
-
Pandora's White-Box: Precise Training Data Detection and Extraction in Large Language Models 26 Feb 2024 · 1 repository · arXiv:2402.17012Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
CausalGym: Benchmarking causal interpretability methods on linguistic tasks 19 Feb 2024 · 2 repositories · arXiv:2402.12560Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Model Editing with Canonical Examples 9 Feb 2024 · 1 repository · arXiv:2402.06155Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
DYAD: A Descriptive Yet Abjuring Density efficient approximation to linear neural network layers 11 Dec 2023 · 1 repository · arXiv:2312.06881Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
SparQ Attention: Bandwidth-Efficient LLM Inference 8 Dec 2023 · 1 repository · arXiv:2312.04985Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
How do Language Models Bind Entities in Context? 26 Oct 2023 · 1 repository · arXiv:2310.17191Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning 10 Oct 2023 · 2 repositories · arXiv:2310.06694Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Efficient Streaming Language Models with Attention Sinks 29 Sep 2023 · 6 repositories · arXiv:2309.17453Syntology official (archive's flag): 6 ran · 8 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Universal and Transferable Adversarial Attacks on Aligned Language Models 27 Jul 2023 · 25 repositories · arXiv:2307.15043Syntology official: no sample here; runs from other or unrecorded repositories · 39 ran (of which 7 constructed an object rather than computing a result; 19 with no instrument failure: 4 honoured, 1 violated, 14 with no contract checked; 20 where Syntology's instrument failed) · 22 unverified (of 61 harvested samples) · 4 pointer-only (licence)
-
Early Weight Averaging meets High Learning Rates for LLM Pre-training 5 Jun 2023 · 1 repository · arXiv:2306.03241Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Emergent and Predictable Memorization in Large Language Models 21 Apr 2023 · 2 repositories · arXiv:2304.11158Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)