{"url":"/task/video-compression","name":"Video Compression","slug":"video-compression","description_markdown":"**Video Compression** is a process of reducing the size of an image or video file by exploiting spatial and temporal redundancies within an image or video frame and across multiple video frames. The ultimate goal of a successful Video Compression system is to reduce data volume while retaining the perceptual quality of the decompressed data.\r\n\r\n\r\n<span class=\"description-source\">Source: [Adversarial Video Compression Guided by Soft Edge Detection ](https://arxiv.org/abs/1811.10673)</span>","categories":[{"name":"Computer Vision","url":"/area/computer-vision"}],"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","slug_source":"archive_url"},"counts":{"papers_tagged":496,"papers_with_code":146,"benchmarks":0,"benchmark_tables_in_archive":0,"benchmark_tables_shown":0,"benchmark_tables_withheld_as_spam":0,"benchmark_definition":"a leaderboard table with at least one row; benchmark_tables_shown also counts the zero-row tables; benchmark_tables_in_archive adds the tables withheld as spam","datasets":5,"subtasks":0,"parent_tasks":1},"benchmarks":[],"datasets":[{"url":"/dataset/vinoground","name":"Vinoground","full_name":"","num_papers_in_archive":17},{"url":"/dataset/bvi-dvc","name":"BVI-DVC","full_name":"","num_papers_in_archive":15},{"url":"/dataset/yt-ugc","name":"YT-UGC","full_name":"YouTube UGC","num_papers_in_archive":12},{"url":"/dataset/deep-fakes-dataset","name":"Deep Fakes Dataset","full_name":"inamibora","num_papers_in_archive":3},{"url":"/dataset/ali","name":"SEPE 8K","full_name":"SEPE 8K","num_papers_in_archive":2}],"subtasks":[],"parent_tasks":[{"url":"/task/video","name":"Video"}],"papers":{"order":"repositories listed in the archive (desc), then date (desc); the archive holds no stars","population":"papers tagged with this task that list at least one repository in the archive","shown":30,"of":146,"tagged_in_all":496,"items":[{"url":"/paper/compressai-a-pytorch-library-and-evaluation","title":"CompressAI: a PyTorch library and evaluation platform for end-to-end compression research","date":"2020-11-05","arxiv_id":"2011.03029","repositories_listed":4,"syntology":{"n":6,"n_ran":6,"n_unverified":0,"n_pointer_only":6}},{"url":"/paper/opendvc-an-open-source-implementation-of-the","title":"OpenDVC: An Open Source Implementation of the DVC Video Compression Method","date":"2020-06-29","arxiv_id":"2006.15862","repositories_listed":4,"syntology":null},{"url":"/paper/sme-net-sparse-motion-estimation-for","title":"SME-Net: Sparse Motion Estimation for Parametric Video Prediction Through Reinforcement Learning","date":"2019-10-01","arxiv_id":null,"repositories_listed":4,"syntology":null},{"url":"/paper/dvc-an-end-to-end-deep-video-compression","title":"DVC: An End-to-end Deep Video Compression Framework","date":"2018-11-30","arxiv_id":"1812.00101","repositories_listed":4,"syntology":{"n":4,"n_ran":3,"n_unverified":1,"n_pointer_only":3}},{"url":"/paper/mganet-a-robust-model-for-quality-enhancement","title":"MGANet: A Robust Model for Quality Enhancement of Compressed Video","date":"2018-11-22","arxiv_id":"1811.09150","repositories_listed":4,"syntology":null},{"url":"/paper/language-model-beats-diffusion-tokenizer-is","title":"Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation","date":"2023-10-09","arxiv_id":"2310.05737","repositories_listed":3,"syntology":{"n":20,"n_ran":12,"n_unverified":8,"n_pointer_only":0}},{"url":"/paper/nerv-neural-representations-for-videos","title":"NeRV: Neural Representations for Videos","date":"2021-10-26","arxiv_id":"2110.13903","repositories_listed":3,"syntology":{"n":10,"n_ran":7,"n_unverified":3,"n_pointer_only":10}},{"url":"/paper/transformer-based-transform-coding","title":"Transformer-based Transform Coding","date":"2021-09-29","arxiv_id":null,"repositories_listed":3,"syntology":null},{"url":"/paper/perceptual-video-compression-with-recurrent","title":"Perceptual Learned Video Compression with Recurrent Conditional GAN","date":"2021-09-07","arxiv_id":"2109.03082","repositories_listed":3,"syntology":{"n":10,"n_ran":0,"n_unverified":10,"n_pointer_only":0}},{"url":"/paper/hierarchical-autoregressive-modeling-for-1","title":"Hierarchical Autoregressive Modeling for Neural Video Compression","date":"2020-10-19","arxiv_id":"2010.10258","repositories_listed":3,"syntology":{"n":29,"n_ran":16,"n_unverified":13,"n_pointer_only":21}},{"url":"/paper/learning-for-video-compression-with","title":"Learning for Video Compression with Hierarchical Quality and Recurrent Enhancement","date":"2020-03-04","arxiv_id":"2003.01966","repositories_listed":3,"syntology":{"n":2,"n_ran":0,"n_unverified":2,"n_pointer_only":0}},{"url":"/paper/disentangled-sequential-autoencoder","title":"Disentangled Sequential Autoencoder","date":"2018-03-08","arxiv_id":"1803.02991","repositories_listed":3,"syntology":{"n":2,"n_ran":2,"n_unverified":0,"n_pointer_only":2}},{"url":"/paper/semantic-perceptual-image-compression-using","title":"Semantic Perceptual Image Compression using Deep Convolution Networks","date":"2016-12-27","arxiv_id":"1612.08712","repositories_listed":3,"syntology":null},{"url":"/paper/video-compression-commander-plug-and-play","title":"Video Compression Commander: Plug-and-Play Inference Acceleration for Video Large Language Models","date":"2025-05-20","arxiv_id":"2505.14454","repositories_listed":2,"syntology":{"n":10,"n_ran":9,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/image-and-video-tokenization-with-binary","title":"Image and Video Tokenization with Binary Spherical Quantization","date":"2024-06-11","arxiv_id":"2406.07548","repositories_listed":2,"syntology":{"n":14,"n_ran":11,"n_unverified":3,"n_pointer_only":0}},{"url":"/paper/image-coding-for-machines-with-edge","title":"Image Coding for Machines with Edge Information Learning Using Segment Anything","date":"2024-03-07","arxiv_id":"2403.04173","repositories_listed":2,"syntology":null},{"url":"/paper/neural-video-compression-with-diverse","title":"Neural Video Compression with Diverse Contexts","date":"2023-02-28","arxiv_id":"2302.14402","repositories_listed":2,"syntology":null},{"url":"/paper/flexible-rate-learned-hierarchical-bi","title":"Flexible-Rate Learned Hierarchical Bi-Directional Video Compression With Motion Refinement and Frame-Level Bit Allocation","date":"2022-06-27","arxiv_id":"2206.13613","repositories_listed":2,"syntology":null},{"url":"/paper/end-to-end-rate-distortion-optimized-learned","title":"End-to-End Rate-Distortion Optimized Learned Hierarchical Bi-Directional Video Compression","date":"2021-12-17","arxiv_id":"2112.09529","repositories_listed":2,"syntology":null},{"url":"/paper/deep-contextual-video-compression","title":"Deep Contextual Video Compression","date":"2021-09-30","arxiv_id":"2109.15047","repositories_listed":2,"syntology":{"n":21,"n_ran":12,"n_unverified":9,"n_pointer_only":0}},{"url":"/paper/dfpn-deformable-frame-prediction-network","title":"DFPN: Deformable Frame Prediction Network","date":"2021-05-26","arxiv_id":"2105.12794","repositories_listed":2,"syntology":null},{"url":"/paper/learning-for-video-compression-with-recurrent","title":"Learning for Video Compression with Recurrent Auto-Encoder and Recurrent Probability Model","date":"2020-06-24","arxiv_id":"2006.13560","repositories_listed":2,"syntology":{"n":5,"n_ran":0,"n_unverified":5,"n_pointer_only":0}},{"url":"/paper/convolutional-tensor-train-lstm-for-spatio","title":"Convolutional Tensor-Train LSTM for Spatio-temporal Learning","date":"2020-02-21","arxiv_id":"2002.09131","repositories_listed":2,"syntology":{"n":2,"n_ran":2,"n_unverified":0,"n_pointer_only":2}},{"url":"/paper/remote-heart-rate-measurement-from-highly","title":"Remote Heart Rate Measurement from Highly Compressed Facial Videos: an End-to-end Deep Learning Solution with Video Enhancement","date":"2019-07-27","arxiv_id":"1907.11921","repositories_listed":2,"syntology":null},{"url":"/paper/enhancing-quality-for-vvc-compressed-videos","title":"Enhancing Quality for VVC Compressed Videos by Jointly Exploiting Spatial Details and Temporal Structure","date":"2019-01-28","arxiv_id":"1901.09575","repositories_listed":2,"syntology":null},{"url":"/paper/video-compression-for-spatiotemporal-earth","title":"Video Compression for Spatiotemporal Earth System Data","date":"2025-06-24","arxiv_id":"2506.19656","repositories_listed":1,"syntology":null},{"url":"/paper/higher-fidelity-perceptual-image-and-video","title":"Higher fidelity perceptual image and video compression with a latent conditioned residual denoising diffusion model","date":"2025-05-19","arxiv_id":"2505.13152","repositories_listed":1,"syntology":null},{"url":"/paper/biecvc-gated-diversification-of-bidirectional","title":"BiECVC: Gated Diversification of Bidirectional Contexts for Learned Video Compression","date":"2025-05-14","arxiv_id":"2505.09193","repositories_listed":1,"syntology":null},{"url":"/paper/fg-dfpn-flow-guided-deformable-frame","title":"FG-DFPN: Flow Guided Deformable Frame Prediction Network","date":"2025-03-14","arxiv_id":"2503.11343","repositories_listed":1,"syntology":null},{"url":"/paper/rethinking-video-tokenization-a-conditioned","title":"Rethinking Video Tokenization: A Conditioned Diffusion-based Approach","date":"2025-03-05","arxiv_id":"2503.03708","repositories_listed":1,"syntology":null}],"syntology_records":13,"syntology_note":"a paper without a record is not a recorded non-run: it may lack an arXiv id or simply be absent from the graph layer"},"description_links":{"kept":0,"unwrapped_to_text":0,"bare_urls_linked":0,"relative_images_dropped":0,"rule":"internal links are kept only when the target slug exists in the catalog"},"syntology":{"read_at":"2026-09-24T18:15:14+00:00","claim":"Per-sample execution status on synthesized fixtures ('ran N of M samples'); not a correctness claim and not a ranking signal.","status_vocabulary":{"ran_honours":"ran, honoured the contract we drafted","ran_violates":"ran, violated the contract we drafted","ran_draft_wrong":"ran; our contract draft was wrong, not the code","ran_fixture":"ran; our fixture could not drive it","ran":"ran on a synthesized input","unverified":"unverified (harvested, no recorded run)"}},"not_shown":{"libraries":"the archive has no per-task library table","trend_sparklines":"the Trend column of the benchmarks table was a rendered image; it is not in the archive","social_and_latest_sorts":"stars and social signals are not in the archive"}}