Browse State-of-the-Art › Video Editing
Video Editing
133 papers with code · 0 benchmarks · 4 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
4 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
1 subtask in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 133 papers with code (346 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
6 May 2023 5 repositories listedOur key idea is to represent motion as a sequence of flow maps in the generation process, which inherently isolate motion from appearance.
-
26 Nov 2020 4 repositories listedIn this work, we propose SoccerNet-v2, a novel large-scale corpus of manual annotations for the SoccerNet video dataset, along with open challenges to encourage more research in soccer understanding and broadcast…
-
7 Dec 2023 3 repositories listed Syntology ran 8 of 9 samples · 1 unverifiedRecent advancements in diffusion-based models have demonstrated significant success in generating images from text.
-
5 Apr 2019 3 repositories listedWe introduce point-to-point video generation that controls the generation process with two control points: the targeted start- and end-frames.
-
10 Mar 2025 2 repositories listed Syntology ran 8 of 19 samples · 11 unverifiedFurther pursuing the unification of generation and editing tasks has yielded significant progress in the domain of image content creation.
-
16 Dec 2024 2 repositories listedTherefore, the controllability of video editing remains a formidable challenge.
-
17 Oct 2024 2 repositories listedOur models set a new state-of-the-art on multiple tasks: text-to-video synthesis, video personalization, video editing, video-to-audio generation, and text-to-audio generation.
-
21 Aug 2024 2 repositories listedTo the best of our knowledge, VE-Bench introduces the first quality assessment dataset for video editing and an effective subjective-aligned quantitative metric for this domain.
-
31 Jul 2024 2 repositories listedTo address this gap, this work conducts a systematic review on SAM for videos in the era of foundation models.
-
28 May 2024 2 repositories listed Syntology ran 5 of 11 samples · 6 unverified · 5 pointer-only (licence)(3) RACCooN also plans to imagine new objects in a given video, so users simply prompt the model to receive a detailed video editing plan for complex video editing.
-
3 Jan 2024 2 repositories listed Syntology ran 5 of 7 samples · 2 unverifiedThis work presents Moonshot, a new video generation model that conditions simultaneously on multimodal inputs of image and text.
-
FaceDNeRF: Semantics-Driven Face Reconstruction, Prompt Editing and Relighting with Diffusion Models1 Jun 2023 2 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)The ability to create high-quality 3D faces from a single image has become increasingly important with wide applications in video conferencing, AR/VR, and advanced video editing in movie industries.
-
7 Feb 2023 2 repositories listedWith information consumption via online video streaming becoming increasingly popular, misinformation video poses a new threat to the health of the online information ecosystem.
-
23 Sep 2021 2 repositories listed Syntology ran 2 of 17 samples · 15 unverifiedWe present a method that decomposes, or "unwraps", an input video into a set of layered 2D atlases, each providing a unified representation of the appearance of an object (or background) over the video.
-
28 Jun 2021 2 repositories listedWith rapidly evolving internet technologies and emerging tools, sports related videos generated online are increasing at an unprecedentedly fast pace.
-
22 Dec 2020 2 repositories listedWe show that a single handheld consumer-grade camera is sufficient to synthesize sophisticated renderings of a dynamic scene from novel virtual camera views, e.
-
23 Apr 2019 2 repositories listedFree-form video inpainting is a very challenging task that could be widely used for video editing such as text removal.
-
13 Apr 2018 2 repositories listedHuman shape estimation is an important task for video editing, animation and fashion industry.
-
17 Jun 2025 1 repository listedAdapting text-to-image (T2I) latent diffusion models for video editing has shown strong visual fidelity and controllability, but challenges remain in maintaining causal relationships in video content.
-
6 Jun 2025 1 repository listedTo overcome these limitations, we introduce FADE, a training-free yet highly effective video editing approach that fully leverages the inherent priors from pre-trained video diffusion models via frequency-aware…
-
29 May 2025 1 repository listedAppearance editing according to user needs is a pivotal task in video editing.
-
29 May 2025 1 repository listedVisual dubbing, the synchronization of facial movements with new speech, is crucial for making content accessible across different languages, enabling broader global reach.
-
16 Apr 2025 1 repository listedIn this work, we propose the Dual Consistency SAM (DC-SAM) method based on prompt-tuning to adapt SAM and SAM2 for in-context segmentation of both images and videos.
-
9 Apr 2025 1 repository listedA versatile video depth estimation model should (1) be accurate and consistent across frames, (2) produce high-resolution depth maps, and (3) support real-time streaming.
-
30 Mar 2025 1 repository listedThis task has attracted increasing attention in the field of computer vision due to its promising applications in video editing and human-agent interaction.
-
26 Mar 2025 1 repository listed Syntology ran 1 of 18 samples · 17 unverifiedThe edited first frame is propagated to subsequent frames to produce the edited video, followed by another round of filtering for frame quality and motion evaluation.
-
26 Mar 2025 1 repository listedOpenness: We open-source the entire series of Wan, including source code and all models, with the goal of fostering the growth of the video generation community.
-
Alias-Free Latent Diffusion Models:Improving Fractional Shift Equivariance of Diffusion Latent Space12 Mar 2025 1 repository listed Syntology ran 3 of 4 samples · 1 unverified · 4 pointer-only (licence)Latent Diffusion Models (LDMs) are known to have an unstable generation process, where even small perturbations or shifts in the input noise can lead to significantly different outputs.
-
7 Mar 2025 1 repository listedVideo inpainting, which aims to restore corrupted video content, has experienced substantial progress.
-
21 Jan 2025 1 repository listedPoint tracking in videos is a fundamental task with applications in robotics, video editing, and more.
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections