Browse State-of-the-Art › Memorization › Papers, page 6
Memorization
Papers archive 2025-07-28
archive papers tagged: 1,088 · with a code link: 438 · where Syntology ran a sample: 177 (148 with a run with no instrument failure, 29 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (177 of 1,088 tagged: 148 with a run with no instrument failure, 29 where every run was a failure of Syntology's instrument)
Page 6 of 11: papers 501 to 600 of 1,088, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Few-Shot Generation of Brain Tumors for Secure and Fair Data Sharing31 Mar 2025 0 repositories listed
-
WinoWhat: A Parallel Corpus of Paraphrased WinoGrande Sentences with Common Sense Categorization31 Mar 2025 0 repositories listed
-
Factored Agents: Decoupling In-Context Learning and Memorization for Robust Tool Use29 Mar 2025 0 repositories listed
-
SUV: Scalable Large Language Model Copyright Compliance with Regularized Selective Unlearning29 Mar 2025 0 repositories listed
-
The Reasoning-Memorization Interplay in Language Models Is Mediated by a Single Direction29 Mar 2025 0 repositories listed
-
Malicious and Unintentional Disclosure Risks in Large Language Models for Code Generation27 Mar 2025 0 repositories listed
-
Quantifying the Ease of Reproducing Training Data in Unconditional Diffusion Models25 Mar 2025 0 repositories listed
-
Exploring the Hidden Reasoning Process of Large Language Models by Misleading Them20 Mar 2025 0 repositories listed
-
BLIA: Detect model memorization in binary classification model through passive Label Inference attack17 Mar 2025 0 repositories listed
-
Empirical Privacy Variance16 Mar 2025 0 repositories listed
-
PrivacyScalpel: Enhancing LLM Privacy via Interpretable Feature Intervention with Sparse Autoencoders14 Mar 2025 0 repositories listed
-
DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation13 Mar 2025 0 repositories listed
-
Trustworthy Machine Learning via Memorization and the Granular Long-Tail: A Survey on Interactions, Tradeoffs, and Beyond10 Mar 2025 0 repositories listed
-
Privacy Auditing of Large Language Models9 Mar 2025 0 repositories listed
-
Mitigating Memorization in LLMs using Activation Steering8 Mar 2025 0 repositories listed
-
CMMCoT: Enhancing Complex Multi-Image Comprehension via Multi-Modal Chain-of-Thought and Memory Augmentation7 Mar 2025 0 repositories listed
-
Dynamic-KGQA: A Scalable Framework for Generating Adaptive Question Answering Datasets6 Mar 2025 0 repositories listed
-
Memorize or Generalize? Evaluating LLM Code Generation with Evolved Questions4 Mar 2025 0 repositories listed
-
Privacy-Preserving Fair Synthetic Tabular Data4 Mar 2025 0 repositories listed
-
Superficial Self-Improved Reasoners Benefit from Model Merging3 Mar 2025 0 repositories listed
-
Asynchronous Personalized Federated Learning through Global Memorization1 Mar 2025 0 repositories listed
-
Holistic Audit Dataset Generation for LLM Unlearning via Knowledge Graph Traversal and Redundancy Removal26 Feb 2025 0 repositories listed
-
On the Interpolation Effect of Score Smoothing26 Feb 2025 0 repositories listed
-
IGDA: Interactive Graph Discovery through Large Language Model Agents24 Feb 2025 0 repositories listed
-
On the Dichotomy Between Privacy and Traceability in ℓₚ Stochastic Convex Optimization24 Feb 2025 0 repositories listed
-
Reasoning with Latent Thoughts: On the Power of Looped Transformers24 Feb 2025 0 repositories listed
-
Swallowing the Poison Pills: Insights from Vulnerability Disparity Among LLMs23 Feb 2025 0 repositories listed
-
Interrogating LLM design under a fair learning doctrine22 Feb 2025 0 repositories listed
-
CopyJudge: Automated Copyright Infringement Identification and Mitigation in Text-to-Image Diffusion Models21 Feb 2025 0 repositories listed
-
Generative AI Training and Copyright Law21 Feb 2025 0 repositories listed
-
LIFT: Improving Long Context Understanding of Large Language Models through Long Input Fine-Tuning20 Feb 2025 0 repositories listed
-
Obliviate: Efficient Unmemorization for Protecting Intellectual Property in Large Language Models20 Feb 2025 0 repositories listed
-
Quantifying Memorization and Retriever Performance in Retrieval-Augmented Vision-Language Models19 Feb 2025 0 repositories listed
-
None of the Others: a General Technique to Distinguish Reasoning from Memorization in Multiple-Choice LLM Evaluation Benchmarks18 Feb 2025 0 repositories listed
-
Pruning as a Defense: Reducing Memorization in Large Language Models18 Feb 2025 0 repositories listed
-
Continual Learning Should Move Beyond Incremental Classification17 Feb 2025 0 repositories listed
-
Rethinking Benign Overfitting in Two-Layer Neural Networks17 Feb 2025 0 repositories listed
-
Logarithmic Width Suffices for Robust Memorization16 Feb 2025 0 repositories listed
-
Retrieval-augmented Encoders for Extreme Multi-label Text Classification15 Feb 2025 0 repositories listed
-
The Vendiscope: An Algorithmic Microscope For Data Collections15 Feb 2025 0 repositories listed
-
Diffusing DeBias: a Recipe for Turning a Bug into a Feature13 Feb 2025 0 repositories listed
-
Democratizing AI: Open-source Scalable LLM Training on GPU-based Supercomputers12 Feb 2025 0 repositories listed
-
Captured by Captions: On Memorization and its Mitigation in CLIP Models11 Feb 2025 0 repositories listed
-
MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations10 Feb 2025 0 repositories listed
-
Mitigating Sensitive Information Leakage in LLMs4Code through Machine Unlearning9 Feb 2025 0 repositories listed
-
A Lightweight Method to Disrupt Memorized Sequences in LLM7 Feb 2025 0 repositories listed
-
An Analysis for Reasoning Bias of Language Models with Small Initialization5 Feb 2025 0 repositories listed
-
Taking a Big Step: Large Learning Rates in Denoising Score Matching Prevent Memorization5 Feb 2025 0 repositories listed
-
TReMu: Towards Neuro-Symbolic Temporal Reasoning for LLM-Agents with Memory in Multi-Session Dialogues3 Feb 2025 0 repositories listed
-
Compositional Generalization Requires More Than Disentangled Representations30 Jan 2025 0 repositories listed
-
Integrating LMM Planners and 3D Skill Policies for Generalizable Manipulation30 Jan 2025 0 repositories listed
-
FUNU: Boosting Machine Unlearning Efficiency by Filtering Unnecessary Unlearning28 Jan 2025 0 repositories listed
-
Memorize and Rank: Elevating Large Language Models for Clinical Diagnosis Prediction28 Jan 2025 0 repositories listed
-
SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training28 Jan 2025 0 repositories listed
-
Decoding Generalization from Memorization in Deep Neural Networks24 Jan 2025 0 repositories listed
-
On the Reasoning Capacity of AI Models and How to Quantify It23 Jan 2025 0 repositories listed
-
RPO: Retrieval Preference Optimization for Robust Retrieval-Augmented Generation23 Jan 2025 0 repositories listed
-
Test-time regression: a unifying framework for designing sequence models with associative memory21 Jan 2025 0 repositories listed
-
Synthetic Data Can Mislead Evaluations: Membership Inference as Machine Text Detection20 Jan 2025 0 repositories listed
-
Enhancing Generalization in Chain of Thought Reasoning for Smaller Models16 Jan 2025 0 repositories listed
-
Modeling Neural Networks with Privacy Using Neural Stochastic Differential Equations12 Jan 2025 0 repositories listed
-
ChronoSense: Exploring Temporal Understanding in Large Language Models with Time Intervals of Events6 Jan 2025 0 repositories listed
-
Knowledge Memorization and Rumination for Pre-trained Model-based Class-Incremental Learning1 Jan 2025 0 repositories listed
-
Representation in large language models1 Jan 2025 0 repositories listed
-
Uncovering Memorization Effect in the Presence of Spurious Correlations1 Jan 2025 0 repositories listed
-
Variance-Based Membership Inference Attacks Against Large-Scale Image Captioning Models1 Jan 2025 0 repositories listed
-
Elucidating Flow Matching ODE Dynamics with Respect to Data Geometries25 Dec 2024 0 repositories listed
-
The Impact of Input Order Bias on Large Language Models for Software Fault Localization25 Dec 2024 0 repositories listed
-
Think or Remember? Detecting and Directing LLMs Towards Memorization or Generalization24 Dec 2024 0 repositories listed
-
Accessing the topological properties of human brain functional sub-circuits in Echo State Networks19 Dec 2024 0 repositories listed
-
Memorization Over Reasoning? Exposing and Mitigating Verbatim Memorization in Large Language Models' Character Understanding Evaluation18 Dec 2024 0 repositories listed
-
Knowledge Boundary of Large Language Models: A Survey17 Dec 2024 0 repositories listed
-
The Impact of Generalization Techniques on the Interplay Among Privacy, Utility, and Fairness in Image Classification16 Dec 2024 0 repositories listed
-
Understanding and Mitigating Memorization in Diffusion Models for Tabular Data15 Dec 2024 0 repositories listed
-
Too Big to Fool: Resisting Deception in Language Models13 Dec 2024 0 repositories listed
-
When Can Memorization Improve Fairness?12 Dec 2024 0 repositories listed
-
Underestimated Privacy Risks for Minority Populations in Large Language Model Unlearning11 Dec 2024 0 repositories listed
-
MemHunter: Automated and Verifiable Memorization Detection at Dataset-scale in LLMs10 Dec 2024 0 repositories listed
-
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit9 Dec 2024 0 repositories listed
-
Robust Noisy Correspondence Learning via Self-Drop and Dual-Weight9 Dec 2024 0 repositories listed
-
Sometimes I am a Tree: Data Drives Unstable Hierarchical Generalization5 Dec 2024 0 repositories listed
-
T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts5 Dec 2024 0 repositories listed
-
Understanding Memorization in Generative Models via Sharpness in Probability Landscapes5 Dec 2024 0 repositories listed
-
Improved Localized Machine Unlearning Through the Lens of Memorization3 Dec 2024 0 repositories listed
-
CopyrightShield: Spatial Similarity Guided Backdoor Defense against Copyright Infringement in Diffusion Models2 Dec 2024 0 repositories listed
-
Detecting Memorization in Large Language Models2 Dec 2024 0 repositories listed
-
LoyalDiffusion: A Diffusion Model Guarding Against Data Replication2 Dec 2024 0 repositories listed
-
Learned Random Label Predictions as a Neural Network Complexity Metric29 Nov 2024 0 repositories listed
-
Integrating Functionalities To A System Via Autoencoder Hippocampus Network28 Nov 2024 0 repositories listed
-
Differential learning kinetics govern the transition from memorization to generalization during in-context learning27 Nov 2024 0 repositories listed
-
A solvable generative model with a linear, one-step denoiser26 Nov 2024 0 repositories listed
-
Are Large Language Models Memorizing Bug Benchmarks?20 Nov 2024 0 repositories listed
-
Branches, Assemble! Multi-Branch Cooperation Network for Large-Scale Click-Through Rate Prediction at Taobao20 Nov 2024 0 repositories listed
-
Vertical Validation: Evaluating Implicit Generative Models for Graphs on Thin Support Regions20 Nov 2024 0 repositories listed
-
Education in the Era of Neurosymbolic AI16 Nov 2024 0 repositories listed
-
Measuring Non-Adversarial Reproduction of Training Data in Large Language Models15 Nov 2024 0 repositories listed
-
Unlearning in- vs. out-of-distribution data in LLMs under gradient-based method7 Nov 2024 0 repositories listed
-
LSHBloom: Memory-efficient, Extreme-scale Document Deduplication6 Nov 2024 0 repositories listed
-
Extracting Unlearned Information from LLMs with Activation Steering4 Nov 2024 0 repositories listed
-
Generalizability of Memorization Neural Networks1 Nov 2024 0 repositories listed