Methods › General › Position Embeddings › ALiBi › Papers where code ran, page 1
Attention with Linear Biases
ALiBi
Papers archive 2025-07-28
archive papers tagged: 19 · with a code link: 10 · where Syntology ran a sample: 8 (8 with a run with no instrument failure, 0 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (8 of 19 tagged: 8 with a run with no instrument failure, 0 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 1 of 1: papers 1 to 8 of the 8 tagged papers where Syntology ran at least one harvested sample (8 with a run with no instrument failure, 0 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
SeqPE: Transformer with Sequential Position Encoding 16 Jun 2025 · 1 repository · arXiv:2506.13277Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 2 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Context-aware Biases for Length Extrapolation 11 Mar 2025 · 1 repository · arXiv:2503.08067Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Mitigate Position Bias in Large Language Models via Scaling a Single Dimension 4 Jun 2024 · 1 repository · arXiv:2406.02536Syntology official (archive's flag): 17 ran · 17 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 6 where Syntology's instrument failed) · 5 unverified (of 22 harvested samples) · 5 pointer-only (licence)
-
BTLM-3B-8K: 7B Parameter Performance in a 3B Parameter Model 20 Sep 2023 · 1 repository · arXiv:2309.11568Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale 23 Jun 2023 · 1 repository · arXiv:2306.15687Syntology 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 3 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
The Impact of Positional Encoding on Length Generalization in Transformers 31 May 2023 · 2 repositories · arXiv:2305.19466Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
A Vector Quantized Approach for Text to Speech Synthesis on Real-World Spontaneous Speech 8 Feb 2023 · 1 repository · arXiv:2302.04215Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation 27 Aug 2021 · 10 repositories · arXiv:2108.12409Syntology official (archive's flag): 4 ran · 16 ran (of which 6 constructed an object rather than computing a result; 14 with no instrument failure: 4 honoured, 0 violated, 10 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 16 harvested samples) · 5 pointer-only (licence)