Browse State-of-the-Art › Video Reconstruction
Video Reconstruction
55 papers with code · 9 benchmarks · 9 datasets archive 2025-07-28
Source: Deep-SloMo
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
9 leaderboard tables shown for this task, 9 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
9 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 55 papers with code (145 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
23 Mar 2022 9 repositories listed Syntology ran 9 of 13 samples · 4 unverified · 12 pointer-only (licence)Pre-training video transformers on extra large-scale datasets is generally required to achieve premier performance on relatively small datasets.
-
14 Feb 2025 3 repositories listed Syntology ran 2 of 9 samples · 7 unverifiedWe present Step-Video-T2V, a state-of-the-art text-to-video pre-trained model with 30B parameters and the ability to generate videos up to 204 frames in length.
-
26 Oct 2021 3 repositories listed Syntology ran 7 of 10 samples · 3 unverified · 10 pointer-only (licence)In contrast, with NeRV, we can use any neural network compression method as a proxy for video compression, and achieve comparable performance to traditional frame-based video compression approaches (H.
-
29 Feb 2020 3 repositories listed Syntology ran 2 of 10 samples · 8 unverifiedTo achieve this, we decouple appearance and motion information using a self-supervised formulation.
-
11 Jun 2024 2 repositories listed Syntology ran 11 of 14 samples · 3 unverifiedThe resulting BSQ-ViT achieves state-of-the-art visual reconstruction quality on image and video reconstruction benchmarks with 2.
-
27 Mar 2023 2 repositories listed Syntology ran 0 of 2 samples · 2 unverifiedMoreover, on the seemingly implausible x16 interpolation task, our method outperforms existing methods by more than 1.
-
23 Sep 2021 2 repositories listed Syntology ran 2 of 17 samples · 15 unverifiedWe present a method that decomposes, or "unwraps", an input video into a set of layered 2D atlases, each providing a unified representation of the appearance of an object (or background) over the video.
-
22 Apr 2021 2 repositories listed Syntology ran 2 of 6 samples · 4 unverified · 5 pointer-only (licence)To facilitate animation and prevent the leakage of the shape of the driving object, we disentangle shape and pose of objects in the region space.
-
22 May 2025 1 repository listedEvent-based cameras offer unique advantages such as high temporal resolution, high dynamic range, and low power consumption.
-
22 May 2025 1 repository listedWe propose AdapTok, an adaptive temporal causal video tokenizer that can flexibly allocate tokens for different frames based on video content.
-
18 Mar 2025 1 repository listed Syntology ran 2 of 4 samples · 2 unverifiedRecent advances in Latent Video Diffusion Models (LVDMs) have revolutionized video generation by leveraging Video Variational Autoencoders (Video VAEs) to compress intricate video data into a compact latent space.
-
14 Mar 2025 1 repository listed Syntology ran 0 of 1 samples · 1 unverified · 1 pointer-only (licence)Decoding visual stimuli from neural activity is essential for understanding the human brain.
-
5 Mar 2025 1 repository listedTrained using only a basic MSE diffusion loss for reconstruction, along with KL term and LPIPS perceptual loss from scratch, extensive experiments demonstrate that CDT achieves state-of-the-art performance in video…
-
5 Feb 2025 1 repository listedWe consider the problem of efficiently representing casually captured monocular videos in a spatially- and temporally-coherent manner.
-
29 Nov 2024 1 repository listedIn this paper, we propose a novel framework for solving high-definition video inverse problems using latent image diffusion models.
-
30 Oct 2024 1 repository listed Syntology ran 0 of 3 samples · 3 unverified · 3 pointer-only (licence)Inspired by recent work on Poisson denoising, we developed an algorithm that creates a dense image sequence from sparse binary photon data by predicting the photon arrival location probability distribution.
-
28 Oct 2024 1 repository listed Syntology ran 4 of 13 samples · 9 unverifiedBy incorporating the prior model during training, LARP learns a latent space that is not only optimized for video reconstruction but is also structured in a way that is more conducive to autoregressive generation.
-
25 Oct 2024 1 repository listed Syntology ran 0 of 4 samples · 4 unverified · 4 pointer-only (licence)We contend that the key to addressing these challenges lies in accurately decoding both high-level semantics and low-level perception flows, as perceived by the brain in response to video stimuli.
-
2 Sep 2024 1 repository listedWith the same reconstruction quality, the more sufficient the VAE's compression for videos is, the more efficient the LVDMs are.
-
26 Aug 2024 1 repository listedHowever, the key components in recurrent-based VSR networks significantly impact model efficiency, e.
-
31 Jul 2024 1 repository listedTo address this challenge, in this paper, we propose a simple low-bit quantization framework (dubbed Q-SCI) for the end-to-end deep learning-based video SCI reconstruction methods which usually consist of a feature…
-
28 Jun 2024 1 repository listed Syntology ran 5 of 5 samples · 0 unverified · 5 pointer-only (licence)FFM is designed for the fusion of contextual information within neighboring event streams, leveraging the coupling relationship between positive and negative events to alleviate the misleading of noises in the…
-
16 May 2024 1 repository listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)In this paper, we propose a bilateral event mining and complementary network (BMCNet) to fully leverage the potential of each event and capture the shared information to complement each other simultaneously.
-
30 Apr 2024 1 repository listedIn this work, to facilitate the development of real-world HDR video reconstruction, we present Real-HDRV, a large-scale real-world benchmark dataset for HDR video reconstruction, featuring various scenes, diverse motion…
-
6 Apr 2024 1 repository listedHowever, inaccurate alignment usually leads to aligned features with significant artifacts, which will be accumulated during propagation and thus affect video restoration.
-
7 Dec 2023 1 repository listedTo overcome these limitations, we leverage Neural Radiance Fields (NeRFs) to represent videos, conducting stylization in the rendered feature space.
-
3 Sep 2023 1 repository listedWe also demonstrate the integration of image convolution with linear spatial kernels Gaussian, Sobel, and Laplacian as an application of our architecture.
-
22 Aug 2023 1 repository listed Syntology ran 12 of 18 samples · 6 unverified · 18 pointer-only (licence)In this paper, we propose an end-to-end HDR video composition framework, which aligns LDR frames in the feature space and then merges aligned features into an HDR frame, without relying on pixel-domain optical flow.
-
22 May 2023 1 repository listedIn this paper, we propose a light, simple model-based deep network for E2V reconstruction, explore the diversity for adjacent pixels in V2E generation, and finally build a video-to-events-to-video (V2E2V) architecture…
-
10 May 2023 1 repository listedEvent-based cameras are becoming increasingly popular for their ability to capture high-speed motion with low latency and high dynamic range.
Syntology lines on 16 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections