{"url":"/method/barlow-twins","slug":"barlow-twins","name":"Barlow Twins","full_name":"Barlow Twins","full_name_withheld":false,"description_markdown":"**Barlow Twins** is a self-supervised learning method that applies redundancy-reduction — a principle first proposed in neuroscience — to self supervised learning. The objective function measures the cross-correlation matrix between the embeddings of two identical networks fed with distorted versions of a batch of samples, and tries to make this matrix close to the identity. This causes the embedding vectors of distorted version of a sample to be similar, while minimizing the redundancy between the components of these vectors. Barlow Twins does not require large batches nor asymmetry between the network twins such as a predictor network, gradient stopping, or a moving average on the weight updates. Intriguingly it benefits from very high-dimensional output vectors.","description_state":"present","introduced_year":null,"introduced_by":{"title":"Barlow Twins: Self-Supervised Learning via Redundancy Reduction","paper":"/paper/barlow-twins-self-supervised-learning-via","first_author":"Jure Zbontar","n_authors":5,"url_abs":null,"archive_paper_url":"https://paperswithcode.com/paper/barlow-twins-self-supervised-learning-via"},"source":{"url":"https://arxiv.org/abs/2103.03230v3","title":"Barlow Twins: Self-Supervised Learning via Redundancy Reduction","url_on_a_paper_host":true},"code_snippet_url":null,"code_snippet_url_on_a_code_host":false,"categories":[{"area":"General","area_id":"general","collection":"Self-Supervised Learning","url":"/methods/category/self-supervised-learning","pwc_aliases":[]}],"n_papers_tagged":71,"archive_num_papers":71,"papers_newest_first":[{"paper":null,"title":"Enhancing User Sequence Modeling through Barlow Twins-based Self-Supervised Learning","date":"2025-05-02","arxiv_id":"2505.00953","n_code_links":0,"syntology":null},{"paper":null,"title":"SinSim: Sinkhorn-Regularized SimCLR","date":"2025-02-13","arxiv_id":"2502.10478","n_code_links":0,"syntology":null},{"paper":null,"title":"Self-supervised Benchmark Lottery on ImageNet: Do Marginal Improvements Translate to Improvements on Similar Datasets?","date":"2025-01-26","arxiv_id":"2501.15431","n_code_links":0,"syntology":null},{"paper":"/paper/warlearn-weather-adaptive-representation","title":"WARLearn: Weather-Adaptive Representation Learning","date":"2024-11-21","arxiv_id":"2411.14095","n_code_links":1,"syntology":null},{"paper":"/paper/unsupervised-homography-estimation-on","title":"Unsupervised Homography Estimation on Multimodal Image Pair via Alternating Optimization","date":"2024-11-20","arxiv_id":"2411.13036","n_code_links":1,"syntology":null},{"paper":null,"title":"Infinite Width Limits of Self Supervised Neural Networks","date":"2024-11-17","arxiv_id":"2411.11176","n_code_links":0,"syntology":null},{"paper":"/paper/efficient-self-supervised-barlow-twins-from","title":"Efficient Self-Supervised Barlow Twins from Limited Tissue Slide Cohorts for Colonic Pathology Diagnostics","date":"2024-11-08","arxiv_id":"2411.05959","n_code_links":1,"syntology":null},{"paper":null,"title":"A Theoretical Characterization of Optimal Data Augmentations in Self-Supervised Learning","date":"2024-11-04","arxiv_id":"2411.01767","n_code_links":0,"syntology":null},{"paper":null,"title":"Uncovering RL Integration in SSL Loss: Objective-Specific Implications for Data-Efficient RL","date":"2024-10-22","arxiv_id":"2410.17428","n_code_links":0,"syntology":null},{"paper":null,"title":"Self-Supervised Anomaly Detection in the Wild: Favor Joint Embeddings Methods","date":"2024-10-05","arxiv_id":"2410.04289","n_code_links":0,"syntology":null},{"paper":"/paper/context-aware-predictive-coding-a","title":"Context-Aware Predictive Coding: A Representation Learning Framework for WiFi Sensing","date":"2024-09-20","arxiv_id":null,"n_code_links":1,"syntology":null},{"paper":"/paper/context-aware-predictive-coding-a-1","title":"Context-Aware Predictive Coding: A Representation Learning Framework for WiFi Sensing","date":"2024-09-16","arxiv_id":"2410.01825","n_code_links":1,"syntology":null},{"paper":null,"title":"Learning Robust Representations for Communications over Noisy Channels","date":"2024-09-02","arxiv_id":"2409.01129","n_code_links":0,"syntology":null},{"paper":null,"title":"PreMix: Addressing Label Scarcity in Whole Slide Image Classification with Pre-trained Multiple Instance Learning Aggregators","date":"2024-08-02","arxiv_id":"2408.01162","n_code_links":0,"syntology":null},{"paper":"/paper/2408-00040","title":"Barlow Twins Deep Neural Network for Advanced 1D Drug-Target Interaction Prediction","date":"2024-07-31","arxiv_id":"2408.00040","n_code_links":2,"syntology":null},{"paper":null,"title":"Dense Self-Supervised Learning for Medical Image Segmentation","date":"2024-07-29","arxiv_id":"2407.20395","n_code_links":0,"syntology":null},{"paper":"/paper/enhanced-self-supervised-learning-for-multi","title":"Enhanced Masked Image Modeling to Avoid Model Collapse on Multi-modal MRI Datasets","date":"2024-07-15","arxiv_id":"2407.10377","n_code_links":2,"syntology":null},{"paper":null,"title":"Maximum Manifold Capacity Representations in State Representation Learning","date":"2024-05-22","arxiv_id":"2405.13848","n_code_links":0,"syntology":null},{"paper":"/paper/from-barlow-twins-to-triplet-training","title":"From Barlow Twins to Triplet Training: Differentiating Dementia with Limited Data","date":"2024-04-09","arxiv_id":"2404.06253","n_code_links":1,"syntology":null},{"paper":"/paper/towards-better-understanding-of-contrastive","title":"Towards Better Understanding of Contrastive Sentence Representation Learning: A Unified Paradigm for Gradient","date":"2024-02-28","arxiv_id":"2402.18281","n_code_links":1,"syntology":null},{"paper":"/paper/the-common-stability-mechanism-behind-most","title":"The Common Stability Mechanism behind most Self-Supervised Learning Approaches","date":"2024-02-22","arxiv_id":"2402.14957","n_code_links":1,"syntology":null},{"paper":null,"title":"BarlowTwins-CXR : Enhancing Chest X-Ray abnormality localization in heterogeneous data with cross-domain self-supervised learning","date":"2024-02-09","arxiv_id":"2402.06499","n_code_links":0,"syntology":null},{"paper":"/paper/twinbooster-synergising-large-language-models","title":"TwinBooster: Synergising Large Language Models with Barlow Twins and Gradient Boosting for Enhanced Molecular Property Prediction","date":"2024-01-09","arxiv_id":"2401.04478","n_code_links":1,"syntology":null},{"paper":null,"title":"Enhancing Context Through Contrast","date":"2024-01-06","arxiv_id":"2401.03314","n_code_links":0,"syntology":null},{"paper":"/paper/self-supervised-learning-for-skin-cancer","title":"Self-supervised learning for skin cancer diagnosis with limited training data","date":"2024-01-01","arxiv_id":"2401.00692","n_code_links":1,"syntology":null},{"paper":"/paper/upper-bounding-barlow-twins-a-novel-filter","title":"Upper Bounding Barlow Twins: A Novel Filter for Multi-Relational Clustering","date":"2023-12-21","arxiv_id":"2312.14066","n_code_links":1,"syntology":{"ran":6,"of":9,"unverified":3,"pointer_only":9}},{"paper":"/paper/guarding-barlow-twins-against-overfitting","title":"Guarding Barlow Twins Against Overfitting with Mixed Samples","date":"2023-12-04","arxiv_id":"2312.02151","n_code_links":1,"syntology":null},{"paper":null,"title":"Contrastive Learning of View-Invariant Representations for Facial Expressions Recognition","date":"2023-11-12","arxiv_id":"2311.06852","n_code_links":0,"syntology":null},{"paper":"/paper/adaptive-multi-head-contrastive-learning","title":"Adaptive Multi-head Contrastive Learning","date":"2023-10-09","arxiv_id":"2310.05615","n_code_links":1,"syntology":null},{"paper":"/paper/information-flow-in-self-supervised-learning","title":"Information Flow in Self-Supervised Learning","date":"2023-09-29","arxiv_id":"2309.17281","n_code_links":2,"syntology":null}],"papers_shown":30,"tasks":[{"task":"/task/self-supervised-learning","name":"Self-Supervised Learning","papers":54},{"task":"/task/representation-learning","name":"Representation Learning","papers":23},{"task":"/task/contrastive-learning","name":"Contrastive Learning","papers":16},{"task":"/task/transfer-learning","name":"Transfer Learning","papers":8},{"task":"/task/image-classification","name":"Image Classification","papers":6},{"task":"/task/semantic-segmentation","name":"Semantic Segmentation","papers":5},{"task":"/task/image-classification","name":"image-classification","papers":5},{"task":"/task/active-learning","name":"Active Learning","papers":3},{"task":"/task/benchmarking","name":"Benchmarking","papers":3},{"task":"/task/disentanglement","name":"Disentanglement","papers":3},{"task":"/task/domain-adaptation","name":"Domain Adaptation","papers":3},{"task":"/task/image-segmentation","name":"Image Segmentation","papers":3},{"task":"/task/language-modelling","name":"Language Modelling","papers":3},{"task":"/task/activity-recognition","name":"Activity Recognition","papers":2},{"task":"/task/autonomous-driving","name":"Autonomous Driving","papers":2},{"task":"/task/autonomous-vehicles","name":"Autonomous Vehicles","papers":2},{"task":"/task/clustering","name":"Clustering","papers":2},{"task":"/task/continual-learning","name":"Continual Learning","papers":2},{"task":"/task/data-augmentation","name":"Data Augmentation","papers":2},{"task":"/task/decoder","name":"Decoder","papers":2}],"tasks_shown":20,"n_tasks":105,"usage_by_year":[{"year":"2021","papers":11},{"year":"2022","papers":16},{"year":"2023","papers":19},{"year":"2024","papers":22},{"year":"2025","papers":3}],"row_source":"methods_table","archive":{"source":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","archive_url":"https://paperswithcode.com/method/barlow-twins"},"syntology_read_at":"2026-09-24T18:15:14+00:00"}