Methods › Natural Language Processing › Language Models › Diffusion › Papers, page 20
Diffusion
Papers archive 2025-07-28
archive papers tagged: 13,848 · with a code link: 5,365 · where Syntology ran a sample: 2,249 (1,969 with a run with no instrument failure, 280 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,249 of 13,848 tagged: 1,969 with a run with no instrument failure, 280 where every run was a failure of Syntology's instrument)
Page 20 of 139: papers 1,901 to 2,000 of 13,848, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Perceive, Understand and Restore: Real-World Image Super-Resolution with Autoregressive Multimodal Generative Models 14 Mar 2025 · 1 repository · arXiv:2503.11073
-
PSF-4D: A Progressive Sampling Framework for View Consistent 4D Editing 14 Mar 2025 · 0 repositories · arXiv:2503.11044
-
Safe-VAR: Safe Visual Autoregressive Model for Text-to-Image Generative Watermarking 14 Mar 2025 · 0 repositories · arXiv:2503.11324
-
TASTE-Rob: Advancing Video Generation of Task-Oriented Hand-Object Interaction for Generalizable Robotic Manipulation 14 Mar 2025 · 0 repositories · arXiv:2503.11423
-
Towards A Correct Usage of Cryptography in Semantic Watermarks for Diffusion Models 14 Mar 2025 · 0 repositories · arXiv:2503.11404
-
Towards Better Alignment: Training Diffusion Models with Reinforcement Learning Against Sparse Rewards 14 Mar 2025 · 1 repository · arXiv:2503.11240Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Understanding Flatness in Generative Models: Its Role and Benefits 14 Mar 2025 · 0 repositories · arXiv:2503.11078
-
AdvPaint: Protecting Images from Inpainting Manipulation via Adversarial Attention Disruption 13 Mar 2025 · 1 repository · arXiv:2503.10081Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
AudioX: Diffusion Transformer for Anything-to-Audio Generation 13 Mar 2025 · 0 repositories · arXiv:2503.10522
-
CameraCtrl II: Dynamic Scene Exploration via Camera-controlled Video Diffusion Models 13 Mar 2025 · 0 repositories · arXiv:2503.10592
-
Channel-wise Noise Scheduled Diffusion for Inverse Rendering in Indoor Scenes 13 Mar 2025 · 0 repositories · arXiv:2503.09993
-
CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance 13 Mar 2025 · 0 repositories · arXiv:2503.10391
-
CoDiPhy: A General Framework for Applying Denoising Diffusion Models to the Physical Layer of Wireless Communication Systems 13 Mar 2025 · 0 repositories · arXiv:2503.10297
-
ConsisLoRA: Enhancing Content and Style Consistency for LoRA-based Style Transfer 13 Mar 2025 · 0 repositories · arXiv:2503.10614
-
Cosh-DiT: Co-Speech Gesture Video Synthesis via Hybrid Audio-Visual Diffusion Transformers 13 Mar 2025 · 0 repositories · arXiv:2503.09942
-
CoSTA∗: Cost-Sensitive Toolpath Agent for Multi-turn Image Editing 13 Mar 2025 · 1 repository · arXiv:2503.10613
-
CoStoDet-DDPM: Collaborative Training of Stochastic and Deterministic Models Improves Surgical Workflow Anticipation and Recognition 13 Mar 2025 · 1 repository · arXiv:2503.10216
-
Data augmentation using diffusion models to enhance inverse Ising inference 13 Mar 2025 · 0 repositories · arXiv:2503.10154
-
Distilling Diversity and Control in Diffusion Models 13 Mar 2025 · 0 repositories · arXiv:2503.10637
-
DiT-Air: Revisiting the Efficiency of Diffusion Model Architecture Design in Text to Image Generation 13 Mar 2025 · 0 repositories · arXiv:2503.10618
-
DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image 13 Mar 2025 · 0 repositories · arXiv:2503.10342
-
Enhancing Facial Privacy Protection via Weakening Diffusion Purification 13 Mar 2025 · 1 repository · arXiv:2503.10350
-
Fine-Tuning Diffusion Generative Models via Rich Preference Optimization 13 Mar 2025 · 0 repositories · arXiv:2503.11720
-
GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing 13 Mar 2025 · 1 repository · arXiv:2503.10639Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model 13 Mar 2025 · 0 repositories · arXiv:2503.10631
-
Improving Diffusion-based Inverse Algorithms under Few-Step Constraint via Learnable Linear Extrapolation 13 Mar 2025 · 1 repository · arXiv:2503.10103Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Investigating and Improving Counter-Stereotypical Action Relation in Text-to-Image Diffusion Models 13 Mar 2025 · 0 repositories · arXiv:2503.10037
-
Long Context Tuning for Video Generation 13 Mar 2025 · 0 repositories · arXiv:2503.10589
-
MoEdit: On Learning Quantity Perception for Multi-object Image Editing 13 Mar 2025 · 1 repository · arXiv:2503.10112
-
MuDG: Taming Multi-modal Diffusion with Gaussian Splatting for Urban Scene Reconstruction 13 Mar 2025 · 0 repositories · arXiv:2503.10604
-
NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models 13 Mar 2025 · 0 repositories · arXiv:2503.10626
-
PanoGen++: Domain-Adapted Text-Guided Panoramic Environment Generation for Vision-and-Language Navigation 13 Mar 2025 · 0 repositories · arXiv:2503.09938
-
Probability-Flow ODE in Infinite-Dimensional Function Spaces 13 Mar 2025 · 0 repositories · arXiv:2503.10219
-
Proxy-Tuning: Tailoring Multimodal Autoregressive Models for Subject-Driven Image Generation 13 Mar 2025 · 0 repositories · arXiv:2503.10125
-
RI3D: Few-Shot Gaussian Splatting With Repair and Inpainting Diffusion Priors 13 Mar 2025 · 1 repository · arXiv:2503.10860
-
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion 13 Mar 2025 · 0 repositories · arXiv:2503.10488
-
Studying Classifier(-Free) Guidance From a Classifier-Centric Perspective 13 Mar 2025 · 0 repositories · arXiv:2503.10638
-
VideoMerge: Towards Training-free Long Video Generation 13 Mar 2025 · 0 repositories · arXiv:2503.09926
-
Accelerating Diffusion Sampling via Exploiting Local Transition Coherence 12 Mar 2025 · 0 repositories · arXiv:2503.09675
-
Active Learning Inspired ControlNet Guidance for Augmenting Semantic Segmentation Datasets 12 Mar 2025 · 0 repositories · arXiv:2503.09221
-
AdvAD: Exploring Non-Parametric Diffusion for Imperceptible Adversarial Attacks 12 Mar 2025 · 1 repository · arXiv:2503.09124
-
Alias-Free Latent Diffusion Models:Improving Fractional Shift Equivariance of Diffusion Latent Space 12 Mar 2025 · 1 repository · arXiv:2503.09419Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Block Diffusion: Interpolating Between Autoregressive and Diffusion Language Models 12 Mar 2025 · 2 repositories · arXiv:2503.09573Syntology official (archive's flag): 10 ran · 14 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 8 pointer-only (licence)
-
CM-Diff: A Single Generative Network for Bidirectional Cross-Modality Translation Diffusion Model Between Infrared and Visible Images 12 Mar 2025 · 0 repositories · arXiv:2503.09514
-
Constrained Language Generation with Discrete Diffusion Models 12 Mar 2025 · 0 repositories · arXiv:2503.09790
-
Context-guided Responsible Data Augmentation with Diffusion Models 12 Mar 2025 · 1 repository · arXiv:2503.10687
-
CoRe^2: Collect, Reflect and Refine to Generate Better and Faster 12 Mar 2025 · 1 repository · arXiv:2503.09662
-
Diff-CL: A Novel Cross Pseudo-Supervision Method for Semi-supervised Medical Image Segmentation 12 Mar 2025 · 0 repositories · arXiv:2503.09408
-
Error Analyses of Auto-Regressive Video Diffusion Models: A Unified Framework 12 Mar 2025 · 0 repositories · arXiv:2503.10704
-
FCaS: Fine-grained Cardiac Image Synthesis based on 3D Template Conditional Diffusion Model 12 Mar 2025 · 0 repositories · arXiv:2503.09560
-
I2V3D: Controllable image-to-video generation with 3D guidance 12 Mar 2025 · 0 repositories · arXiv:2503.09733
-
Incomplete Multi-view Clustering via Diffusion Contrastive Generation 12 Mar 2025 · 0 repositories · arXiv:2503.09185
-
Inductive Spatio-Temporal Kriging with Physics-Guided Increment Training Strategy for Air Quality Inference 12 Mar 2025 · 0 repositories · arXiv:2503.09646
-
Leveraging Semantic Attribute Binding for Free-Lunch Color Control in Diffusion Models 12 Mar 2025 · 0 repositories · arXiv:2503.09864
-
Minimax Optimality of the Probability Flow ODE for Diffusion Models 12 Mar 2025 · 0 repositories · arXiv:2503.09583
-
Monte Carlo Diffusion for Generalizable Learning-Based RANSAC 12 Mar 2025 · 0 repositories · arXiv:2503.09410
-
NAMI: Efficient Image Generation via Progressive Rectified Flow Transformers 12 Mar 2025 · 0 repositories · arXiv:2503.09242
-
Other Vehicle Trajectories Are Also Needed: A Driving World Model Unifies Ego-Other Vehicle Trajectories in Video Latant Space 12 Mar 2025 · 0 repositories · arXiv:2503.09215
-
PerCoV2: Improved Ultra-Low Bit-Rate Perceptual Image Compression with Implicit Hierarchical Masked Image Modeling 12 Mar 2025 · 1 repository · arXiv:2503.09368
-
Reangle-A-Video: 4D Video Generation as Video-to-Video Translation 12 Mar 2025 · 0 repositories · arXiv:2503.09151
-
RewardSDS: Aligning Score Distillation via Reward-Weighted Sampling 12 Mar 2025 · 1 repository · arXiv:2503.09601Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Silent Branding Attack: Trigger-free Data Poisoning Attack on Text-to-Image Diffusion Models 12 Mar 2025 · 0 repositories · arXiv:2503.09669
-
Solving Bayesian inverse problems with diffusion priors and off-policy RL 12 Mar 2025 · 0 repositories · arXiv:2503.09746
-
Sparse Autoencoder as a Zero-Shot Classifier for Concept Erasing in Text-to-Image Diffusion Models 12 Mar 2025 · 1 repository · arXiv:2503.09446
-
SuperCarver: Texture-Consistent 3D Geometry Super-Resolution for High-Fidelity Surface Detail Generation 12 Mar 2025 · 0 repositories · arXiv:2503.09439
-
TA-V2A: Textually Assisted Video-to-Audio Generation 12 Mar 2025 · 0 repositories · arXiv:2503.10700
-
The Pitfalls of Imitation Learning when Actions are Continuous 12 Mar 2025 · 0 repositories · arXiv:2503.09722
-
The R2D2 Deep Neural Network Series for Scalable Non-Cartesian Magnetic Resonance Imaging 12 Mar 2025 · 0 repositories · arXiv:2503.09559
-
Theoretical Guarantees for High Order Trajectory Refinement in Generative Flows 12 Mar 2025 · 0 repositories · arXiv:2503.09069
-
TPDiff: Temporal Pyramid Video Diffusion Model 12 Mar 2025 · 0 repositories · arXiv:2503.09566
-
Training Data Provenance Verification: Did Your Model Use Synthetic Data from My Generative Model for Training? 12 Mar 2025 · 1 repository · arXiv:2503.09122
-
UniCombine: Unified Multi-Conditional Combination with Diffusion Transformer 12 Mar 2025 · 0 repositories · arXiv:2503.09277Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Adaptive Anomaly Recovery for Telemanipulation: A Diffusion Model Approach to Vision-Based Tracking 11 Mar 2025 · 0 repositories · arXiv:2503.09632
-
Aligning Text to Image in Diffusion Models is Easier Than You Think 11 Mar 2025 · 1 repository · arXiv:2503.08250Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AnyMoLe: Any Character Motion In-betweening Leveraging Video Diffusion Models 11 Mar 2025 · 2 repositories · arXiv:2503.08417
-
Bokeh Diffusion: Defocus Blur Control in Text-to-Image Diffusion Models 11 Mar 2025 · 0 repositories · arXiv:2503.08434
-
CDI3D: Cross-guided Dense-view Interpolation for 3D Reconstruction 11 Mar 2025 · 0 repositories · arXiv:2503.08005
-
Controlling Latent Diffusion Using Latent CLIP 11 Mar 2025 · 1 repository · arXiv:2503.08455
-
D3PO: Preference-Based Alignment of Discrete Diffusion Models 11 Mar 2025 · 0 repositories · arXiv:2503.08295
-
FP3: A 3D Foundation Policy for Robotic Manipulation 11 Mar 2025 · 0 repositories · arXiv:2503.08950
-
GarmentCrafter: Progressive Novel View Synthesis for Single-View 3D Garment Reconstruction and Editing 11 Mar 2025 · 0 repositories · arXiv:2503.08678
-
Generalizable AI-Generated Image Detection Based on Fractal Self-Similarity in the Spectrum 11 Mar 2025 · 0 repositories · arXiv:2503.08484
-
High-Quality 3D Head Reconstruction from Any Single Portrait Image 11 Mar 2025 · 0 repositories · arXiv:2503.08516
-
Layton: Latent Consistency Tokenizer for 1024-pixel Image Reconstruction and Generation by 256 Tokens 11 Mar 2025 · 0 repositories · arXiv:2503.08377
-
MEAT: Multiview Diffusion Model for Human Generation on Megapixels with Mesh Attention 11 Mar 2025 · 1 repository · arXiv:2503.08664
-
MegaSR: Mining Customized Semantics and Expressive Guidance for Image Super-Resolution 11 Mar 2025 · 1 repository · arXiv:2503.08096
-
MF-VITON: High-Fidelity Mask-Free Virtual Try-On with Minimal Input 11 Mar 2025 · 0 repositories · arXiv:2503.08650
-
Modular Customization of Diffusion Models via Blockwise-Parameterized Low-Rank Adaptation 11 Mar 2025 · 0 repositories · arXiv:2503.08575
-
MVD-HuGaS: Human Gaussians from a Single Image via 3D Human Multi-view Diffusion Prior 11 Mar 2025 · 0 repositories · arXiv:2503.08218
-
NullFace: Training-Free Localized Face Anonymization 11 Mar 2025 · 1 repository · arXiv:2503.08478
-
OminiControl2: Efficient Conditioning for Diffusion Transformers 11 Mar 2025 · 1 repository · arXiv:2503.08280
-
OmniPaint: Mastering Object-Oriented Editing via Disentangled Insertion-Removal Inpainting 11 Mar 2025 · 0 repositories · arXiv:2503.08677
-
Partial differential equation system for binarization of degraded document images 11 Mar 2025 · 0 repositories · arXiv:2503.08017
-
Posterior-Mean Denoising Diffusion Model for Realistic PET Image Reconstruction 11 Mar 2025 · 0 repositories · arXiv:2503.08546
-
Preserving Product Fidelity in Large Scale Image Recontextualization with Diffusion Models 11 Mar 2025 · 0 repositories · arXiv:2503.08729
-
"Principal Components" Enable A New Language of Images 11 Mar 2025 · 1 repository · arXiv:2503.08685Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 3 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 17 harvested samples) · 2 pointer-only (licence)
-
Reconstruct Anything Model: a lightweight foundation model for computational imaging 11 Mar 2025 · 0 repositories · arXiv:2503.08915
-
Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models 11 Mar 2025 · 0 repositories · arXiv:2503.08737Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Rethinking Diffusion Model in High Dimension 11 Mar 2025 · 1 repository · arXiv:2503.08643
-
SAS: Segment Any 3D Scene with Integrated 2D Priors 11 Mar 2025 · 1 repository · arXiv:2503.08512