Browse State-of-the-Art › Speaker Recognition › Papers, page 2
Speaker Recognition
Papers archive 2025-07-28
archive papers tagged: 435 · with a code link: 102 · where Syntology ran a sample: 17 (15 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (17 of 435 tagged: 15 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument)
Page 2 of 5: papers 101 to 200 of 435, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
12 Sep 2018 1 repository listed
-
22 Jul 2018 1 repository listed
-
An Exploration of ECAPA-TDNN and x-vector Speaker Representations in Zero-shot Multi-speaker TTS25 Jun 2025 0 repositories listed
-
A Comparative Evaluation of Deep Learning Models for Speech Enhancement in Real-World Noisy Environments17 Jun 2025 0 repositories listed
-
Learning Speaker-Invariant Visual Features for Lipreading9 Jun 2025 0 repositories listed
-
Rhythm Features for Speaker Identification7 Jun 2025 0 repositories listed
-
Synthetic Speech Source Tracing using Metric Learning3 Jun 2025 0 repositories listed
-
Investigating the Reasonable Effectiveness of Speaker Pre-Trained Models and their Synergistic Power for SingMOS Prediction2 Jun 2025 0 repositories listed
-
LASPA: Language Agnostic Speaker Disentanglement with Prefix-Tuned Cross-Attention2 Jun 2025 0 repositories listed
-
Source Tracing of Synthetic Speech Systems Through Paralinguistic Pre-Trained Representations1 Jun 2025 0 repositories listed
-
Pretraining Multi-Speaker Identification for Neural Speaker Diarization30 May 2025 0 repositories listed
-
Analysis of ABC Frontend Audio Systems for the NIST-SRE2421 May 2025 0 repositories listed
-
SoCov: Semi-Orthogonal Parametric Pooling of Covariance Matrix for Speaker Recognition23 Apr 2025 0 repositories listed
-
From Dialect Gaps to Identity Maps: Tackling Variability in Speaker Verification21 Apr 2025 0 repositories listed
-
Audio-to-Image Encoding for Improved Voice Characteristic Detection Using Deep Convolutional Neural Networks7 Mar 2025 0 repositories listed
-
Language Modelling for Speaker Diarization in Telephonic Interviews28 Jan 2025 0 repositories listed
-
VoxVietnam: a Large-Scale Multi-Genre Dataset for Vietnamese Speaker Recognition31 Dec 2024 0 repositories listed
-
Investigating Prosodic Signatures via Speech Pre-Trained Models for Audio Deepfake Source Attribution23 Dec 2024 0 repositories listed
-
Study on Inter and Intra Speaker Variability in Speaker Recognition12 Nov 2024 0 repositories listed
-
Multi-View Multi-Task Modeling with Speech Foundation Models for Speech Forensic Tasks16 Oct 2024 0 repositories listed
-
Investigation of Speaker Representation for Target-Speaker Speech Processing15 Oct 2024 0 repositories listed
-
The OCON model: an old but green solution for distributable supervised classification for acoustic monitoring in smart cities5 Oct 2024 0 repositories listed
-
Enhancing Open-Set Speaker Identification through Rapid Tuning with Speaker Reciprocal Points and Negative Sample24 Sep 2024 0 repositories listed
-
Avengers Assemble: Amalgamation of Non-Semantic Features for Depression Detection22 Sep 2024 0 repositories listed
-
Are Music Foundation Models Better at Singing Voice Deepfake Detection? Far-Better Fuse them with Speech Foundation Models21 Sep 2024 0 repositories listed
-
oboVox Far Field Speaker Recognition: A Novel Data Augmentation Approach with Pretrained Models16 Sep 2024 0 repositories listed
-
Speaker-IPL: Unsupervised Learning of Speaker Characteristics with i-Vector based Pseudo-Labels16 Sep 2024 0 repositories listed
-
Text-To-Speech Synthesis In The Wild13 Sep 2024 0 repositories listed
-
Recursive Attentive Pooling for Extracting Speaker Embeddings from Multi-Speaker Recordings30 Aug 2024 0 repositories listed
-
The VoxCeleb Speaker Recognition Challenge: A Retrospective27 Aug 2024 0 repositories listed
-
Convexity-based Pruning of Speech Representation Models16 Aug 2024 0 repositories listed
-
Long-Term Conversation Analysis: Privacy-Utility Trade-off under Noise and Reverberation1 Aug 2024 0 repositories listed
-
Overview of Speaker Modeling and Its Applications: From the Lens of Deep Speaker Representation Learning21 Jul 2024 0 repositories listed
-
Team HYU ASML ROBOVOX SP Cup 2024 System Description16 Jul 2024 0 repositories listed
-
Phonetic Richness for Improved Automatic Speaker Verification10 Jul 2024 0 repositories listed
-
Analyzing Speech Unit Selection for Textless Speech-to-Speech Translation8 Jul 2024 0 repositories listed
-
We Need Variations in Speech Generation: Sub-center Modelling for Speaker Embeddings5 Jul 2024 0 repositories listed
-
Open-Source Conversational AI with SpeechBrain 1.029 Jun 2024 0 repositories listed
-
CEC: A Noisy Label Detection Method for Speaker Recognition19 Jun 2024 0 repositories listed
-
Challenging margin-based speaker embedding extractors by using the variational information bottleneck18 Jun 2024 0 repositories listed
-
PERSONA: An Application for Emotion Recognition, Gender Recognition and Age Estimation10 Jun 2024 0 repositories listed
-
The Reasonable Effectiveness of Speaker Embeddings for Violence Detection10 Jun 2024 0 repositories listed
-
Fill in the Gap! Combining Self-supervised Representation Learning with Neural Audio Synthesis for Speech Inpainting30 May 2024 0 repositories listed
-
Speaker Characterization by means of Attention Pooling7 May 2024 0 repositories listed
-
Who is Authentic Speaker30 Apr 2024 0 repositories listed
-
Artificial Neural Networks to Recognize Speakers Division from Continuous Bengali Speech18 Apr 2024 0 repositories listed
-
TIMIT Speaker Profiling: A Comparison of Multi-task learning and Single-task learning Approaches18 Apr 2024 0 repositories listed
-
Voice Conversion Augmentation for Speaker Recognition on Defective Datasets1 Apr 2024 0 repositories listed
-
Asymmetric and trial-dependent modeling: the contribution of LIA to SdSV Challenge Task 228 Mar 2024 0 repositories listed
-
Cosine Scoring with Uncertainty for Neural Speaker Embedding11 Mar 2024 0 repositories listed
-
Post-Training Embedding Alignment for Decoupling Enrollment and Runtime Speaker Recognition Models23 Jan 2024 0 repositories listed
-
Voxceleb-ESP: preliminary experiments detecting Spanish celebrities from their voices20 Dec 2023 0 repositories listed
-
Vulnerability of Automatic Identity Recognition to Audio-Visual Deepfakes29 Nov 2023 0 repositories listed
-
Phonetic-aware speaker embedding for far-field speaker verification27 Nov 2023 0 repositories listed
-
Parrot-Trained Adversarial Examples: Pushing the Practicality of Black-Box Audio Attacks against Speaker Recognition Models13 Nov 2023 0 repositories listed
-
Detecting Agreement in Multi-party Conversational AI6 Nov 2023 0 repositories listed
-
Personalizing Keyword Spotting with Speaker Information6 Nov 2023 0 repositories listed
-
Deep Neural Networks for Automatic Speaker Recognition Do Not Learn Supra-Segmental Temporal Features1 Nov 2023 0 repositories listed
-
UniX-Encoder: A Universal X-Channel Speech Encoder for Ad-Hoc Microphone Array Speech Processing25 Oct 2023 0 repositories listed
-
Privacy-oriented manipulation of speaker representations10 Oct 2023 0 repositories listed
-
Thech. Report: Genuinization of Speech waveform PMF for speaker detection spoofing and countermeasures9 Oct 2023 0 repositories listed
-
Disentangling Voice and Content with Self-Supervision for Speaker Recognition2 Oct 2023 0 repositories listed
-
Voice Morphing: Two Identities in One Voice5 Sep 2023 0 repositories listed
-
UNISOUND System for VoxCeleb Speaker Recognition Challenge 202324 Aug 2023 0 repositories listed
-
Graph Neural Network Backend for Speaker Recognition17 Aug 2023 0 repositories listed
-
The DKU-MSXF Speaker Verification System for the VoxCeleb Speaker Recognition Challenge 202317 Aug 2023 0 repositories listed
-
ChinaTelecom System Description to VoxCeleb Speaker Recognition Challenge 202316 Aug 2023 0 repositories listed
-
The ID R&D VoxCeleb Speaker Recognition Challenge 2023 System Description16 Aug 2023 0 repositories listed
-
GIST-AiTeR Speaker Diarization System for VoxCeleb Speaker Recognition Challenge (VoxSRC) 202315 Aug 2023 0 repositories listed
-
The DKU-MSXF Diarization System for the VoxCeleb Speaker Recognition Challenge 202315 Aug 2023 0 repositories listed
-
VoxBlink: A Large Scale Speaker Verification Dataset on Camera14 Aug 2023 0 repositories listed
-
On-Device Speaker Anonymization of Acoustic Embeddings for ASR based onFlexible Location Gradient Reversal Layer25 Jul 2023 0 repositories listed
-
Exploring the Integration of Speech Separation and Recognition with Self-Supervised Learning Representation23 Jul 2023 0 repositories listed
-
Facial Landmark Detection Evaluation on MOBIO Database6 Jul 2023 0 repositories listed
-
VoxWatch: An open-set speaker recognition benchmark on VoxCeleb30 Jun 2023 0 repositories listed
-
Understanding Contrastive Learning Through the Lens of Margins20 Jun 2023 0 repositories listed
-
Meta-Learning Framework for End-to-End Imposter Identification in Unseen Speaker Recognition1 Jun 2023 0 repositories listed
-
STT4SG-350: A Speech Corpus for All Swiss German Dialect Regions30 May 2023 0 repositories listed
-
Transforming the Embeddings: A Lightweight Technique for Speech Emotion Recognition Tasks29 May 2023 0 repositories listed
-
Ordered and Binary Speaker Embedding25 May 2023 0 repositories listed
-
Generalized domain adaptation framework for parametric back-end in speaker recognition24 May 2023 0 repositories listed
-
QFA2SR: Query-Free Adversarial Transfer Attacks to Speaker Recognition Systems23 May 2023 0 repositories listed
-
A Comparative Study of Pre-trained Speech and Audio Embeddings for Speech Emotion Recognition22 Apr 2023 0 repositories listed
-
The Graph feature fusion technique for speaker recognition based on wav2vec2.0 framework19 Mar 2023 0 repositories listed
-
A Study on Bias and Fairness In Deep Speaker Recognition14 Mar 2023 0 repositories listed
-
Self-FiLM: Conditioning GANs with self-supervised representations for bandwidth extension based speaker recognition7 Mar 2023 0 repositories listed
-
Speaker Recognition in Realistic Scenario Using Multimodal Data25 Feb 2023 0 repositories listed
-
A Reinforcement Learning Framework for Online Speaker Diarization21 Feb 2023 0 repositories listed
-
Interpretable Spectrum Transformation Attacks to Speaker Recognition21 Feb 2023 0 repositories listed
-
Speaker and Language Change Detection using Wav2vec2 and Whisper18 Feb 2023 0 repositories listed
-
Audio Representation Learning by Distilling Video as Privileged Information6 Feb 2023 0 repositories listed
-
Leveraging Speaker Embeddings with Adversarial Multi-task Learning for Age Group Classification22 Jan 2023 0 repositories listed
-
A Multi-Purpose Audio-Visual Corpus for Multi-Modal Persian Speech Recognition: the Arman-AV Dataset21 Jan 2023 0 repositories listed
-
The Newsbridge -Telecom SudParis VoxCeleb Speaker Recognition Challenge 2022 System Description17 Jan 2023 0 repositories listed
-
Introducing Model Inversion Attacks on Automatic Speaker Recognition9 Jan 2023 0 repositories listed
-
SLUE Phase-2: A Benchmark Suite of Diverse Spoken Language Understanding Tasks20 Dec 2022 0 repositories listed
-
Probing Deep Speaker Embeddings for Speaker-related Tasks14 Dec 2022 0 repositories listed
-
A Novel Speech Feature Fusion Algorithm for Text-Independent Speaker Recognition1 Dec 2022 0 repositories listed
-
A new Speech Feature Fusion method with cross gate parallel CNN for Speaker Recognition24 Nov 2022 0 repositories listed
-
Multi-source Domain Adaptation for Text-independent Forensic Speaker Recognition17 Nov 2022 0 repositories listed