Methods › General › Distributed Methods › Tofu
Tofu
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
Tofu is an intra-layer model parallel system that partitions very large DNN models across multiple GPU devices to reduce per-GPU memory footprint. Tofu is designed to partition a dataflow graph of fine-grained tensor operators used by platforms like MXNet and TensorFlow. To optimally partition different operators in a dataflow graph, Tofu uses a recursive search algorithm that minimizes the total communication cost.
Papers archive 2025-07-28
19 shown of 19, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.
-
GUARD: Guided Unlearning and Retention via Data Attribution for Large Language Models 12 Jun 2025 · 0 repositories · arXiv:2506.10946
-
Constrained Entropic Unlearning: A Primal-Dual Framework for Large Language Models 5 Jun 2025 · 0 repositories · arXiv:2506.05314
-
UniErase: Unlearning Token as a Universal Erasure Primitive for Language Models 21 May 2025 · 1 repository · arXiv:2505.15674
-
GUARD: Generation-time LLM Unlearning via Adaptive Restriction and Detection 19 May 2025 · 0 repositories · arXiv:2505.13312
-
UIPE: Enhancing LLM Unlearning by Removing Knowledge Related to Forgetting Targets 6 Mar 2025 · 0 repositories · arXiv:2503.04693
-
CE-U: Cross Entropy Unlearning 3 Mar 2025 · 0 repositories · arXiv:2503.01224
-
Towards Robust Evaluation of Unlearning in LLMs via Data Transformations 23 Nov 2024 · 1 repository · arXiv:2411.15477
-
Unlearning as multi-task optimization: A normalized gradient difference approach with an adaptive learning rate 29 Oct 2024 · 0 repositories · arXiv:2410.22086
-
LLM Unlearning via Loss Adjustment with Only Forget Data 14 Oct 2024 · 0 repositories · arXiv:2410.11143
-
Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning 9 Oct 2024 · 2 repositories · arXiv:2410.07163Syntology ran 8 of 10 samples · 2 unverified
-
Answer When Needed, Forget When Not: Language Models Pretend to Forget via In-Context Knowledge Unlearning 1 Oct 2024 · 0 repositories · arXiv:2410.00382
-
Towards Robust and Parameter-Efficient Knowledge Unlearning for LLMs 13 Aug 2024 · 1 repository · arXiv:2408.06621Syntology ran 2 of 2 samples · 0 unverified
-
Reversing the Forget-Retain Objectives: An Efficient LLM Unlearning Framework from Logit Difference 12 Jun 2024 · 1 repository · arXiv:2406.08607
-
Negative Preference Optimization: From Catastrophic Collapse to Effective Unlearning 8 Apr 2024 · 1 repository · arXiv:2404.05868
-
Token Fusion: Bridging the Gap between Token Pruning and Token Merging 2 Dec 2023 · 0 repositories · arXiv:2312.01026
-
On High-dimensional and Low-rank Tensor Bandits 6 May 2023 · 0 repositories · arXiv:2305.03884
-
TOFU: Towards Obfuscated Federated Updates by Encoding Weight Updates into Gradients from Proxy Data 21 Jan 2022 · 0 repositories · arXiv:2201.08494
-
Topologically Consistent Multi-View Face Inference Using Volumetric Sampling 6 Oct 2021 · 0 repositories · arXiv:2110.02948
-
Supporting Very Large Models using Automatic Dataflow Graph Partitioning 24 Jul 2018 · 0 repositories · arXiv:1807.08887
Tasks archive 2025-07-28
18 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.
Usage over time archive 2025-07-28
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections