Methods › Natural Language Processing › Inference Extrapolation › Patching › Papers where code ran, page 1
Activation Patching
Patching
Papers archive 2025-07-28
archive papers tagged: 102 · with a code link: 43 · where Syntology ran a sample: 19 (14 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (19 of 102 tagged: 14 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 1 of 1: papers 1 to 19 of the 19 tagged papers where Syntology ran at least one harvested sample (14 with a run with no instrument failure, 5 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models 25 Jun 2025 · 1 repository · arXiv:2506.20251Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models 9 Jun 2025 · 1 repository · arXiv:2506.07468Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 11 unverified (of 17 harvested samples) · 1 pointer-only (licence)
-
MIB: A Mechanistic Interpretability Benchmark 17 Apr 2025 · 1 repository · arXiv:2504.13151Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Multi-Grained Patch Training for Efficient LLM-based Recommendation 25 Jan 2025 · 1 repository · arXiv:2501.15087Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
WPMixer: Efficient Multi-Resolution Mixing for Long-Term Time Series Forecasting 22 Dec 2024 · 1 repository · arXiv:2412.17176Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Benchmarking and Understanding Compositional Relational Reasoning of LLMs 17 Dec 2024 · 1 repository · arXiv:2412.12841Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Efficient Large-Scale Traffic Forecasting with Transformers: A Spatial Data Management Perspective 13 Dec 2024 · 3 repositories · arXiv:2412.09972Syntology official (archive's flag): 3 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Separating Tongue from Thought: Activation Patching Reveals Language-Agnostic Concept Representations in Transformers 13 Nov 2024 · 1 repository · arXiv:2411.08745Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Beyond Toxic Neurons: A Mechanistic Analysis of DPO for Toxicity Reduction 10 Nov 2024 · 1 repository · arXiv:2411.06424Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
Answer, Assemble, Ace: Understanding How Transformers Answer Multiple Choice Questions 21 Jul 2024 · 0 repositories · arXiv:2407.15018Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Modular Pluralism: Pluralistic Alignment via Multi-LLM Collaboration 22 Jun 2024 · 1 repository · arXiv:2406.15951Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Are Language Models Actually Useful for Time Series Forecasting? 22 Jun 2024 · 3 repositories · arXiv:2406.16964Syntology official (archive's flag): 2 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Finding Safety Neurons in Large Language Models 20 Jun 2024 · 0 repositories · arXiv:2406.14144Syntology 14 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 1 pointer-only (licence)
-
Continuum Attention for Neural Operators 10 Jun 2024 · 1 repository · arXiv:2406.06486Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
AROMA: Preserving Spatial Structure for Latent PDE Modeling with Local Neural Fields 4 Jun 2024 · 1 repository · arXiv:2406.02176Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 2 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Medformer: A Multi-Granularity Patching Transformer for Medical Time-Series Classification 24 May 2024 · 2 repositories · arXiv:2405.19363Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 2 pointer-only (licence)
-
Unveiling LLMs: The Evolution of Latent Representations in a Dynamic Knowledge Graph 4 Apr 2024 · 1 repository · arXiv:2404.03623Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms 26 Mar 2024 · 3 repositories · arXiv:2403.17806Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Monotonic Representation of Numeric Properties in Language Models 15 Mar 2024 · 1 repository · arXiv:2403.10381Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)