Browse State-of-the-Art › Lip Reading › Papers, page 2
Lip Reading
Papers archive 2025-07-28
archive papers tagged: 153 · with a code link: 49 · where Syntology ran a sample: 17 (14 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (17 of 153 tagged: 14 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument)
Page 2 of 2: papers 101 to 153 of 153, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Leveraging Uni-Modal Self-Supervised Learning for Multimodal Audio-visual Speech Recognition16 Nov 2021 0 repositories listed
-
Advances and Challenges in Deep Lip Reading15 Oct 2021 0 repositories listed
-
14 Oct 2021 0 repositories listed
-
Perception Point: Identifying Critical Learning Periods in Speech for Bilingual Networks13 Oct 2021 0 repositories listed
-
Audio-Visual Speech Recognition is Worth 32×32×8 Voxels20 Sep 2021 0 repositories listed
-
LRWR: Large-Scale Benchmark for Lip Reading in Russian language14 Sep 2021 0 repositories listed
-
SimulLR: Simultaneous Lip Reading Transducer with Attention-Guided Adaptive Memory31 Aug 2021 0 repositories listed
-
16 Aug 2021 0 repositories listed
-
Spatio-Temporal Attention Mechanism and Knowledge Distillation for Lip Reading7 Aug 2021 0 repositories listed
-
Facetron: A Multi-speaker Face-to-Speech Model based on Cross-modal Latent Representations26 Jul 2021 0 repositories listed
-
Learning From the Master: Distilling Cross-Modal Advanced Knowledge for Lip Reading19 Jun 2021 0 repositories listed
-
LiRA: Learning Visual Speech Representations from Audio through Self-supervision16 Jun 2021 0 repositories listed
-
Multi-Perspective LSTM for Joint Visual Representation Learning6 May 2021 0 repositories listed
-
End-to-End Video-To-Speech Synthesis using Generative Adversarial Networks27 Apr 2021 0 repositories listed
-
Fusing information streams in end-to-end audio-visual speech recognition19 Apr 2021 0 repositories listed
-
Lip reading using external viseme decoding10 Apr 2021 0 repositories listed
-
Contrastive Self-Supervised Learning of Global-Local Audio-Visual Representations1 Jan 2021 0 repositories listed
-
Lip-reading with Hierarchical Pyramidal Convolution and Self-Attention28 Dec 2020 0 repositories listed
-
Disentangling Homophemes in Lip Reading using Perplexity Analysis28 Nov 2020 0 repositories listed
-
A Study on Lip Localization Techniques used for Lip reading from a Video28 Sep 2020 0 repositories listed
-
Seeing voices and hearing voices: learning discriminative embeddings using cross-modal self-supervision29 Apr 2020 0 repositories listed
-
9 Mar 2020 0 repositories listed
-
Re-synchronization using the Hand Preceding Model for Multi-modal Fusion in Automatic Continuous Cued Speech Recognition23 Feb 2020 0 repositories listed
-
28 Nov 2019 0 repositories listed
-
Towards Pose-invariant Lip-Reading14 Nov 2019 0 repositories listed
-
1 Oct 2019 0 repositories listed
-
30 Aug 2019 0 repositories listed
-
14 Aug 2019 0 repositories listed
-
Realistic Speech-Driven Facial Animation with GANs14 Jun 2019 0 repositories listed
-
MobiVSR: A Visual Speech Recognition Solution for Mobile Devices10 May 2019 0 repositories listed
-
Synthesising 3D Facial Motion from "In-the-Wild" Speech15 Apr 2019 0 repositories listed
-
Learning from Videos with Deep Convolutional LSTM Networks9 Apr 2019 0 repositories listed
-
An Empirical Analysis of Deep Audio-Visual Models for Speech Recognition21 Dec 2018 0 repositories listed
-
Contextual Audio-Visual Switching For Speech Enhancement in Real-World Environments28 Aug 2018 0 repositories listed
-
Lip-Reading Driven Deep Learning Approach for Speech Enhancement31 Jul 2018 0 repositories listed
-
Deep Lip Reading: a comparison of models and an online application15 Jun 2018 0 repositories listed
-
Lip Reading Using Convolutional Auto Encoders as Feature Extractor31 May 2018 0 repositories listed
-
Resource aware design of a deep convolutional-recurrent neural network for speech recognition through audio-visual sensor fusion13 Mar 2018 0 repositories listed
-
Deep Learning for Lip Reading using Audio-Visual Information for Urdu Language15 Feb 2018 0 repositories listed
-
Decoding visemes: improving machine lipreading3 Oct 2017 0 repositories listed
-
Finding phonemes: improving machine lip-reading3 Oct 2017 0 repositories listed
-
Resolution limits on visual speech recognition3 Oct 2017 0 repositories listed
-
Some observations on computer lip-reading: moving from the dream to the reality3 Oct 2017 0 repositories listed
-
Speaker-independent machine lip-reading with speaker-dependent viseme classifiers3 Oct 2017 0 repositories listed
-
Which phoneme-to-viseme maps best improve visual-only computer lip-reading?3 Oct 2017 0 repositories listed
-
Automatic Viseme Vocabulary Construction to Enhance Continuous Lip-reading26 Apr 2017 0 repositories listed
-
Towards Estimating the Upper Bound of Visual-Speech Recognition: The Visual Lip-Reading Feasibility Database26 Apr 2017 0 repositories listed
-
16 Nov 2016 0 repositories listed
-
Video-Based Action Recognition Using Rate-Invariant Analysis of Covariance Trajectories23 Mar 2015 0 repositories listed
-
Definition of Visual Speech Element and Research on a Method of Extracting Feature Vector for Korean Lip-Reading15 Nov 2014 0 repositories listed
-
Visual Words for Automatic Lip-Reading17 Sep 2014 0 repositories listed
-
Visual Speech Recognition3 Sep 2014 0 repositories listed
-
Visual Passwords Using Automatic Lip Reading2 Sep 2014 0 repositories listed