Browse State-of-the-Art › Memorization › Papers, page 8
Memorization
Papers archive 2025-07-28
archive papers tagged: 1,088 · with a code link: 438 · where Syntology ran a sample: 177 (148 with a run with no instrument failure, 29 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (177 of 1,088 tagged: 148 with a run with no instrument failure, 29 where every run was a failure of Syntology's instrument)
Page 8 of 11: papers 701 to 800 of 1,088, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Learnable Privacy Neurons Localization in Language Models16 May 2024 0 repositories listed
-
Dynamic Loss Decay based Robust Oriented Object Detection on Remote Sensing Images with Noisy Labels15 May 2024 0 repositories listed
-
Generalized Holographic Reduced Representations15 May 2024 0 repositories listed
-
Beyond Scaling Laws: Understanding Transformer Performance with Associative Memory14 May 2024 0 repositories listed
-
Thinking Tokens for Language Modeling14 May 2024 0 repositories listed
-
To Each (Textual Sequence) Its Own: Improving Memorized-Data Unlearning in Large Language Models6 May 2024 0 repositories listed
-
Exploring prompts to elicit memorization in masked language model-based named entity recognition5 May 2024 0 repositories listed
-
Mothman at SemEval-2024 Task 9: An Iterative System for Chain-of-Thought Prompt Optimization3 May 2024 0 repositories listed
-
Report on the AAPM Grand Challenge on deep generative modeling for learning medical image statistics3 May 2024 0 repositories listed
-
Quantifying Memorization and Detecting Training Data of Pre-trained Language Models using Japanese Newspaper26 Apr 2024 0 repositories listed
-
Rethinking LLM Memorization through the Lens of Adversarial Compression23 Apr 2024 0 repositories listed
-
Reliable Model Watermarking: Defending Against Theft without Compromising on Evasion21 Apr 2024 0 repositories listed
-
The Positivity of the Neural Tangent Kernel19 Apr 2024 0 repositories listed
-
The Fault in our Stars: Quality Assessment of Code Generation Benchmarks15 Apr 2024 0 repositories listed
-
AI Knowledge and Reasoning: Emulating Expert Creativity in Scientific Research5 Apr 2024 0 repositories listed
-
GP-MoLFormer: A Foundation Model For Molecular Generation4 Apr 2024 0 repositories listed
-
Towards Better Generalization in Open-Domain Question Answering by Mitigating Context Memorization2 Apr 2024 0 repositories listed
-
What Can Transformer Learn with Varying Depth? Case Studies on Sequence Learning Tasks2 Apr 2024 0 repositories listed
-
SoK: A Review of Differentially Private Linear Models For High-Dimensional Data1 Apr 2024 0 repositories listed
-
Towards Memorization-Free Diffusion Models1 Apr 2024 0 repositories listed
-
Soften to Defend: Towards Adversarial Robustness via Self-Guided Label Refinement14 Mar 2024 0 repositories listed
-
Ethos: Rectifying Language Models in Orthogonal Parameter Space13 Mar 2024 0 repositories listed
-
LLM-Oriented Retrieval Tuner4 Mar 2024 0 repositories listed
-
ROME: Memorization Insights from Text, Logits and Representation1 Mar 2024 0 repositories listed
-
Learning Associative Memories with Gradient Descent28 Feb 2024 0 repositories listed
-
Unveiling Privacy, Memorization, and Input Curvature Links28 Feb 2024 0 repositories listed
-
Unified View of Grokking, Double Descent and Emergent Abilities: A Perspective from Circuits Competition23 Feb 2024 0 repositories listed
-
Towards Uncovering How Large Language Model Works: An Explainability Perspective16 Feb 2024 0 repositories listed
-
Neural Information Organizing and Processing -- Neural Machines15 Feb 2024 0 repositories listed
-
Information Complexity of Stochastic Convex Optimization: Applications to Generalization and Memorization14 Feb 2024 0 repositories listed
-
Future Prediction Can be a Strong Evidence of Good History Representation in Partially Observable Environments11 Feb 2024 0 repositories listed
-
Social Evolution of Published Text and The Emergence of Artificial Intelligence Through Large Language Models and The Problem of Toxicity and Bias11 Feb 2024 0 repositories listed
-
Wasserstein proximal operators describe score-based generative models and resolve memorization9 Feb 2024 0 repositories listed
-
Revisiting Early-Learning Regularization When Federated Learning Meets Noisy Labels8 Feb 2024 0 repositories listed
-
Selective Forgetting: Advancing Machine Unlearning Techniques and Evaluation in Language Models8 Feb 2024 0 repositories listed
-
Analyzing the Neural Tangent Kernel of Periodically Activated Coordinate Networks7 Feb 2024 0 repositories listed
-
EMN: Brain-inspired Elastic Memory Network for Quick Domain Adaptive Feature Mapping4 Feb 2024 0 repositories listed
-
Déjà Vu Memorization in Vision-Language Models3 Feb 2024 0 repositories listed
-
Human-Centered Privacy Research in the Age of Large Language Models3 Feb 2024 0 repositories listed
-
Expressive Power of ReLU and Step Networks under Floating-Point Operations26 Jan 2024 0 repositories listed
-
Critical Data Size of Language Models from a Grokking Perspective19 Jan 2024 0 repositories listed
-
Understanding Learning through the Lens of Dynamical Invariants19 Jan 2024 0 repositories listed
-
Learning with Structural Labels for Learning with Noisy Labels1 Jan 2024 0 repositories listed
-
27 Dec 2023 0 repositories listed
-
BloomVQA: Assessing Hierarchical Multi-modal Comprehension20 Dec 2023 0 repositories listed
-
Social Learning: Towards Collaborative Learning with Large Language Models18 Dec 2023 0 repositories listed
-
Lifted RDT based capacity analysis of the 1-hidden layer treelike sign perceptrons neural networks13 Dec 2023 0 repositories listed
-
Memory Triggers: Unveiling Memorization in Text-To-Image Generative Models through Word-Level Duplication6 Dec 2023 0 repositories listed
-
Understanding (Un)Intended Memorization in Text-to-Image Generative Models6 Dec 2023 0 repositories listed
-
Scalable Extraction of Training Data from (Production) Language Models28 Nov 2023 0 repositories listed
-
Positional Description Matters for Transformers Arithmetic22 Nov 2023 0 repositories listed
-
CSGNN: Conquering Noisy Node labels via Dynamic Class-wise Selection20 Nov 2023 0 repositories listed
-
On Retrieval Augmentation and the Limitations of Language Model Training16 Nov 2023 0 repositories listed
-
Does Pre-trained Language Model Actually Infer Unseen Links in Knowledge Graph Completion?15 Nov 2023 0 repositories listed
-
Preserving Privacy in GANs Against Membership Inference Attack6 Nov 2023 0 repositories listed
-
The statistical thermodynamics of generative diffusion models: Phase transitions, symmetry breaking and critical instability26 Oct 2023 0 repositories listed
-
Grokking in Linear Estimators -- A Solvable Model that Groks without Understanding25 Oct 2023 0 repositories listed
-
SoK: Memorization in General-Purpose Large Language Models24 Oct 2023 0 repositories listed
-
MoPe: Model Perturbation-based Privacy Attacks on Language Models22 Oct 2023 0 repositories listed
-
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks19 Oct 2023 0 repositories listed
-
Training Dynamics of Deep Network Linear Regions19 Oct 2023 0 repositories listed
-
Unintended Memorization in Large ASR Models, and How to Mitigate It18 Oct 2023 0 repositories listed
-
Combating Label Noise With A General Surrogate Model For Sample Selection16 Oct 2023 0 repositories listed
-
Generation or Replication: Auscultating Audio Latent Diffusion Models16 Oct 2023 0 repositories listed
-
Why Train More? Effective and Efficient Membership Inference via Memorization12 Oct 2023 0 repositories listed
-
Exploring Memorization in Fine-tuned Language Models10 Oct 2023 0 repositories listed
-
Grokking as Compression: A Nonlinear Complexity Perspective9 Oct 2023 0 repositories listed
-
What do larger image classifiers memorise?9 Oct 2023 0 repositories listed
-
Probing Large Language Models from A Human Behavioral Perspective8 Oct 2023 0 repositories listed
-
How Much Training Data is Memorized in Overparameterized Autoencoders? An Inverse Problem Perspective on Memorization Evaluation4 Oct 2023 0 repositories listed
-
Scaling Laws for Associative Memories4 Oct 2023 0 repositories listed
-
30 Sep 2023 0 repositories listed Syntology 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
Identifying and Mitigating Privacy Risks Stemming from Language Models: A Survey27 Sep 2023 0 repositories listed
-
22 Sep 2023 0 repositories listed
-
Extreme Image Transformations Facilitate Robust Latent Object Representations19 Sep 2023 0 repositories listed
-
Analysis of the Memorization and Generalization Capabilities of AI Agents: Are Continual Learners Robust?18 Sep 2023 0 repositories listed
-
Collectionless Artificial Intelligence13 Sep 2023 0 repositories listed
-
Text Encoders Lack Knowledge: Leveraging Generative LLMs for Domain-Specific Semantic Textual Similarity12 Sep 2023 0 repositories listed
-
Quantifying and Attributing the Hallucination of Large Language Models via Association Analysis11 Sep 2023 0 repositories listed
-
When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale8 Sep 2023 0 repositories listed
-
FLM-101B: An Open LLM and How to Train It with $100K Budget7 Sep 2023 0 repositories listed
-
On the Planning, Search, and Memorization Capabilities of Large Language Models5 Sep 2023 0 repositories listed
-
Least Squares Maximum and Weighted Generalization-Memorization Machines31 Aug 2023 0 repositories listed
-
Quantifying and Analyzing Entity-level Memorization in Large Language Models30 Aug 2023 0 repositories listed
-
Large language models converge toward human-like concept organization29 Aug 2023 0 repositories listed
-
Continuous Reinforcement Learning-based Dynamic Difficulty Adjustment in a Visual Working Memory Game24 Aug 2023 0 repositories listed
-
Smoothness Similarity Regularization for Few-Shot GAN Adaptation18 Aug 2023 0 repositories listed
-
U-Turn Diffusion14 Aug 2023 0 repositories listed
-
LLaMA-E: Empowering E-commerce Authoring with Object-Interleaved Instruction Following9 Aug 2023 0 repositories listed
-
Arithmetic with Language Models: from Memorization to Computation2 Aug 2023 0 repositories listed
-
Excitatory/Inhibitory Balance Emerges as a Key Factor for RBN Performance, Overriding Attractor Dynamics2 Aug 2023 0 repositories listed
-
Training Data Protection with Compositional Diffusion Models2 Aug 2023 0 repositories listed
-
Understanding Activation Patterns in Artificial Neural Networks by Exploring Stochastic Processes1 Aug 2023 0 repositories listed
-
Are Transformers with One Layer Self-Attention Using Low-Rank Weight Matrices Universal Approximators?26 Jul 2023 0 repositories listed
-
Gradient-Based Word Substitution for Obstinate Adversarial Examples Generation in Language Models24 Jul 2023 0 repositories listed
-
Distribution Shift Matters for Knowledge Distillation with Webly Collected Images21 Jul 2023 0 repositories listed
-
What can we learn from Data Leakage and Unlearning for Law?19 Jul 2023 0 repositories listed
-
Towards Model-Size Agnostic, Compute-Free, Memorization-based Inference of Deep Learning14 Jul 2023 0 repositories listed
-
Memorization Through the Lens of Curvature of Loss Function Around Samples11 Jul 2023 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.