Browse State-of-the-Art › Speaker Identification › Papers, page 2
Speaker Identification
Papers archive 2025-07-28
archive papers tagged: 248 · with a code link: 74 · where Syntology ran a sample: 15 (11 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (15 of 248 tagged: 11 with a run with no instrument failure, 4 where every run was a failure of Syntology's instrument)
Page 2 of 3: papers 101 to 200 of 248, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Hearing-Loss Compensation Using Deep Neural Networks: A Framework and Results From a Listening Test15 Mar 2024 0 repositories listed
-
A Closer Look at Wav2Vec2 Embeddings for On-Device Single-Channel Speech Enhancement3 Mar 2024 0 repositories listed
-
Unraveling Adversarial Examples against Speaker Identification -- Techniques for Attack Detection and Victim Model Classification29 Feb 2024 0 repositories listed
-
Effect of utterance duration and phonetic content on speaker identification using second-order statistical methods26 Feb 2024 0 repositories listed
-
Significance of Chirp MFCC as a Feature in Speech and Audio Applications19 Feb 2024 0 repositories listed
-
Probing Self-supervised Learning Models with Target Speech Extraction17 Feb 2024 0 repositories listed
-
Speech Rhythm-Based Speaker Embeddings Extraction from Phonemes and Phoneme Duration for Multi-Speaker Speech Synthesis11 Feb 2024 0 repositories listed
-
Post-Training Embedding Alignment for Decoupling Enrollment and Runtime Speaker Recognition Models23 Jan 2024 0 repositories listed
-
Voxceleb-ESP: preliminary experiments detecting Spanish celebrities from their voices20 Dec 2023 0 repositories listed
-
Efficiency-oriented approaches for self-supervised speech representation learning18 Dec 2023 0 repositories listed
-
Privacy-preserving Representation Learning for Speech Understanding26 Oct 2023 0 repositories listed
-
Advanced accent/dialect identification and accentedness assessment with multi-embedding models and automatic speech recognition17 Oct 2023 0 repositories listed
-
End-to-end Multichannel Speaker-Attributed ASR: Speaker Guided Decoder and Input Feature Analysis16 Oct 2023 0 repositories listed
-
Test-Time Training for Speech19 Sep 2023 0 repositories listed
-
Spiking-LEAF: A Learnable Auditory front-end for Spiking Neural Networks18 Sep 2023 0 repositories listed
-
Understanding Self-Supervised Learning of Speech Representation via Invariance and Redundancy Reduction7 Sep 2023 0 repositories listed
-
Read, Look or Listen? What's Needed for Solving a Multimodal Dataset6 Jul 2023 0 repositories listed
-
VoxWatch: An open-set speaker recognition benchmark on VoxCeleb30 Jun 2023 0 repositories listed
-
Meta-Learning Framework for End-to-End Imposter Identification in Unseen Speaker Recognition1 Jun 2023 0 repositories listed
-
Few-Shot Speaker Identification Using Lightweight Prototypical Network with Feature Grouping and Interaction31 May 2023 0 repositories listed
-
Ordered and Binary Speaker Embedding25 May 2023 0 repositories listed
-
On the Transferability of Whisper-based Representations for "In-the-Wild" Cross-Task Downstream Speech Applications23 May 2023 0 repositories listed
-
Security and Privacy Problems in Voice Assistant Applications: A Survey19 Apr 2023 0 repositories listed
-
HiSSNet: Sound Event Detection and Speaker Identification via Hierarchical Prototypical Networks for Low-Resource Headphones13 Mar 2023 0 repositories listed
-
Ensemble knowledge distillation of self-supervised speech models24 Feb 2023 0 repositories listed
-
ExARN: self-attending RNN for target speaker extraction2 Dec 2022 0 repositories listed
-
Multi-Label Training for Text-Independent Speaker Identification14 Nov 2022 0 repositories listed
-
Privacy-Utility Balanced Voice De-Identification Using Adversarial Examples10 Nov 2022 0 repositories listed
-
Symmetric Saliency-based Adversarial Attack To Speaker Identification30 Oct 2022 0 repositories listed
-
Quantitative Evidence on Overlooked Aspects of Enrollment Speaker Embeddings for Target Speaker Separation23 Oct 2022 0 repositories listed
-
Speaker Identification from emotional and noisy speech data using learned voice segregation and Speech VGG23 Oct 2022 0 repositories listed
-
Text Independent Speaker Identification System for Access Control26 Sep 2022 0 repositories listed
-
Computing with Hypervectors for Efficient Speaker Identification28 Aug 2022 0 repositories listed
-
Deep versus Wide: An Analysis of Student Architectures for Task-Agnostic Knowledge Distillation of Self-Supervised Speech Models14 Jul 2022 0 repositories listed
-
Graph-based Multi-View Fusion and Local Adaptation: Mitigating Within-Household Confusability for Speaker Identification8 Jul 2022 0 repositories listed
-
Speaker Diarization and Identification from Single-Channel Classroom Audio Recording Using Virtual Microphones1 Jul 2022 0 repositories listed
-
iEmoTTS: Toward Robust Cross-Speaker Emotion Transfer and Control for Speech Synthesis based on Disentanglement between Prosody and Timbre29 Jun 2022 0 repositories listed
-
Identifying Source Speakers for Voice Conversion based Spoofing Attacks on Speaker Verification Systems18 Jun 2022 0 repositories listed
-
Strategies to Improve Robustness of Target Speech Extraction to Enrollment Variations16 Jun 2022 0 repositories listed
-
Speaker Identification using Speech Recognition29 May 2022 0 repositories listed
-
Silence is Sweeter Than Speech: Self-Supervised Model Using Silence to Store Speaker Information8 May 2022 0 repositories listed
-
VFHQ: A High-Quality Dataset and Benchmark for Video Face Super-Resolution6 May 2022 0 repositories listed
-
Few-Shot Speaker Identification Using Depthwise Separable Convolutional Network with Channel Attention24 Apr 2022 0 repositories listed
-
WaBERT: A Low-resource End-to-end Model for Spoken Language Understanding and Speech-to-BERT Alignment22 Apr 2022 0 repositories listed
-
Listen only to me! How well can target speech extraction handle false alarms?11 Apr 2022 0 repositories listed
-
AdvEst: Adversarial Perturbation Estimation to Classify and Detect Adversarial Attacks against Speaker Identification8 Apr 2022 0 repositories listed
-
Karaoker: Alignment-free singing voice synthesis with speech training data8 Apr 2022 0 repositories listed
-
Improved Relation Networks for End-to-End Speaker Verification and Identification31 Mar 2022 0 repositories listed
-
NeuraGen-A Low-Resource Neural Network based approach for Gender Classification29 Mar 2022 0 repositories listed
-
Speaker Identification Experiments Under Gender De-Identification9 Mar 2022 0 repositories listed
-
On the relevance of bandwidth extension for speaker identification24 Feb 2022 0 repositories listed
-
openFEAT: Improving Speaker Identification by Open-set Few-shot Embedding Adaptation with Transformer24 Feb 2022 0 repositories listed
-
Speech watermarking: an approach for the forensic analysis of digital telephonic recordings23 Feb 2022 0 repositories listed
-
Tubes Among Us: Analog Attack on Automatic Speaker Identification6 Feb 2022 0 repositories listed
-
Cross-Lingual Speaker Identification from Weak Local Evidence16 Jan 2022 0 repositories listed
-
The exploitation of Multiple Feature Extraction Techniques for Speaker Identification in Emotional States under Disguised Voices15 Dec 2021 0 repositories listed
-
Target Speech Extraction: Independent Vector Extraction Guided by Supervised Speaker Identification5 Nov 2021 0 repositories listed
-
A Study of Acoustic Features in Arabic Speaker Identification under Noisy Environmental Conditions23 Oct 2021 0 repositories listed
-
PEAF: Learnable Power Efficient Analog Acoustic Features for Audio Recognition7 Oct 2021 0 repositories listed
-
Transcribe-to-Diarize: Neural Speaker Diarization for Unlimited Number of Speakers using End-to-End Speaker-Attributed ASR7 Oct 2021 0 repositories listed
-
Improving Speaker Identification for Shared Devices by Adapting Embeddings to Speaker Subsets6 Sep 2021 0 repositories listed
-
QASR: QCRI Aljazeera Speech Resource A Large Scale Annotated Arabic Speech Corpus1 Aug 2021 0 repositories listed
-
A Real-time Speaker Diarization System Based on Spatial Spectrum20 Jul 2021 0 repositories listed
-
Representation Learning to Classify and Detect Adversarial Attacks against Speaker and Speech Recognition Systems9 Jul 2021 0 repositories listed
-
QASR: QCRI Aljazeera Speech Resource -- A Large Scale Annotated Arabic Speech Corpus24 Jun 2021 0 repositories listed
-
Fusion of Embeddings Networks for Robust Combination of Text Dependent and Independent Speaker Recognition18 Jun 2021 0 repositories listed
-
Graph-based Label Propagation for Semi-Supervised Speaker Identification15 Jun 2021 0 repositories listed
-
End-to-End Diarization for Variable Number of Speakers with Local-Global Networks and Discriminative Speaker Embeddings5 May 2021 0 repositories listed
-
End-to-End Speaker-Attributed ASR with Transformer5 Apr 2021 0 repositories listed
-
Streaming Multi-talker Speech Recognition with Joint Speaker Identification5 Apr 2021 0 repositories listed
-
A Survey on Paralinguistics in Tamil Speech Processing1 Apr 2021 0 repositories listed
-
Voice Privacy with Smart Digital Assistants in Educational Settings24 Mar 2021 0 repositories listed
-
Triplet loss based embeddings for forensic speaker identification in Spanish24 Feb 2021 0 repositories listed
-
CASA-Based Speaker Identification Using Cascaded GMM-CNN Classifier in Noisy and Emotional Talking Conditions11 Feb 2021 0 repositories listed
-
Speaker attribution with voice profiles by graph-based semi-supervised learning6 Feb 2021 0 repositories listed
-
Hypothesis Stitcher for End-to-End Speaker-attributed ASR on Long-form Multi-talker Recordings6 Jan 2021 0 repositories listed
-
A Study of Few-Shot Audio Classification2 Dec 2020 0 repositories listed
-
How Far Are We from Robust Voice Conversion: A Survey24 Nov 2020 0 repositories listed
-
Multi-Modal Emotion Detection with Transfer Learning13 Nov 2020 0 repositories listed
-
T-vectors: Weakly Supervised Speaker Identification Using Hierarchical Transformer Model29 Oct 2020 0 repositories listed
-
A Lightweight Speaker Recognition System Using Timbre Properties12 Oct 2020 0 repositories listed
-
Remarks on Optimal Scores for Speaker Recognition10 Oct 2020 0 repositories listed
-
SoK: The Faults in our ASRs: An Overview of Attacks against Automatic Speech Recognition and Speaker Identification Systems13 Jul 2020 0 repositories listed
-
Joint Speaker Counting, Speech Recognition, and Speaker Identification for Overlapped Speech of Any Number of Speakers19 Jun 2020 0 repositories listed
-
Integrated Replay Spoofing-aware Text-independent Speaker Verification10 Jun 2020 0 repositories listed
-
Speaker and Posture Classification using Instantaneous Intraspeech Breathing Features25 May 2020 0 repositories listed
-
Weakly Supervised Training of Hierarchical Attention Networks for Speaker Identification15 May 2020 0 repositories listed
-
Speaker Recognition in Bengali Language from Nonlinear Features15 Apr 2020 0 repositories listed
-
End-to-end Recurrent Denoising Autoencoder Embeddings for Speaker Identification13 Mar 2020 0 repositories listed
-
Deep Neural Networks for Automatic Speech Processing: A Survey from Large Corpora to Limited Data9 Mar 2020 0 repositories listed
-
Speaker Identification using EEG7 Mar 2020 0 repositories listed
-
3 Mar 2020 0 repositories listed
-
Speech Enhancement using Self-Adaptation and Multi-Head Self-Attention14 Feb 2020 0 repositories listed
-
Robust Speaker Recognition Using Speech Enhancement And Attention Model14 Jan 2020 0 repositories listed
-
Supervised Speaker Embedding De-Mixing in Two-Speaker Environment14 Jan 2020 0 repositories listed
-
The Deterministic plus Stochastic Model of the Residual Signal and its Applications29 Dec 2019 0 repositories listed
-
Advances in Online Audio-Visual Meeting Transcription10 Dec 2019 0 repositories listed
-
Privacy-Preserving Adversarial Representation Learning in ASR: Reality or Illusion?12 Nov 2019 0 repositories listed
-
Supervised Initialization of LSTM Networks for Fundamental Frequency Detection in Noisy Speech Signals11 Nov 2019 0 repositories listed
-
Reducing audio membership inference attack accuracy to chance: 4 defenses31 Oct 2019 0 repositories listed