Papers › Adversarial Video Generation on Complex Datasets

Adversarial Video Generation on Complex Datasets

15 Jul 2019arXiv:1907.06571archive 2025-07-28

Aidan Clark, Jeff Donahue, Karen Simonyan

Generative models of natural images have progressed towards high fidelity samples by the strong leveraging of scale. We attempt to carry this success to the field of video modeling by showing that large Generative Adversarial Networks trained on the complex Kinetics-600 dataset are able to produce video samples of substantially higher complexity and fidelity than previous work. Our proposed model, Dual Video Discriminator GAN (DVD-GAN), scales to longer and higher resolution videos by leveraging a computationally efficient decomposition of its discriminator. We evaluate on the related tasks of video synthesis and video prediction, and achieve new state-of-the-art Fr\'echet Inception Distance for prediction for Kinetics-600, as well as state-of-the-art Inception Score for synthesis on the UCF-101 dataset, alongside establishing a strong baseline for synthesis on Kinetics-600.

PaperPDFConference PDFCode

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

Harrypotterrrr/DVD-GAN mentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

3D Character Animation From A Single PhotoVideo GenerationVideo Prediction

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Video Generation BAIR Robot Pushing DVD-GAN-FP Cond 1 #11 of 31 Archive leaderboard report
Video Generation BAIR Robot Pushing DVD-GAN-FP FVD score 109.8 #11 of 31 Archive leaderboard report
Video Generation BAIR Robot Pushing DVD-GAN-FP Pred 15 #11 of 31 Archive leaderboard report
Video Generation BAIR Robot Pushing DVD-GAN-FP Train 15 #11 of 31 Archive leaderboard report
Video Generation Kinetics-600 12 frames, 128x128 DVD-GAN FID 2.16 #1 of 1 Archive leaderboard report
Video Generation Kinetics-600 12 frames, 64x64 DVD-GAN FVD 31.1 #4 of 4 Archive leaderboard report
Video Generation Kinetics-600 48 frames, 64x64 DVD-GAN FID 12.92 #1 of 1 Archive leaderboard report
Video Generation Kinetics-600 48 frames, 64x64 DVD-GAN Inception Score 219.05 #1 of 1 Archive leaderboard report
Video Prediction BAIR Robot Pushing DVD-GAN-FP FVD 109.8 #5 of 6 Archive leaderboard report
Video Prediction Kinetics-600 12 frames, 64x64 DVD-GAN-FP Cond 5 #14 of 16 Archive leaderboard report
Video Prediction Kinetics-600 12 frames, 64x64 DVD-GAN-FP FVD 69.15±0.78 #14 of 16 Archive leaderboard report
Video Prediction Kinetics-600 12 frames, 64x64 DVD-GAN-FP Pred 11 #14 of 16 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Methods

Introduced by this paper: DVD-GAN, DVD-GAN DBlock

1x1 Convolution3D ConvolutionAdamAverage PoolingBatch NormalizationBigGANCGRUConditional Batch NormalizationConvolutionDVD-GANDVD-GAN DBlockDVD-GAN GBlockDense ConnectionsEarly StoppingFeedforward NetworkGAN Hinge LossGRULinear LayerNon-Local BlockNon-Local OperationOff-Diagonal Orthogonal RegularizationOrthogonal RegularizationProjection DiscriminatorReLUResidual BlockResidual ConnectionSAGANSigmoid ActivationSoftmaxSpectral NormalizationTTURTruncation Trick

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections