{"about":{"non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","site":"https://codewithpapers.app","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page","syntology":{"site":"https://syntology.ai","developers":"https://syntology.ai/developers","mcp":{"server":"https://syntology.ai/mcp","transport":"streamable-http","server_card":"https://syntology.ai/.well-known/mcp/server-card.json","auth":{"type":"trial token, no account","trial_token":"https://syntology.ai/api/oauth/trial/token","method":"POST","docs":"https://syntology.ai/developers"}},"have":"https://syntology.ai/api/graph/have?x=<method, arXiv id or title> (free, answers coverage only)","paper_base":"https://syntology.ai/paper/","atlas_base":"https://app.syntology.ai/?focus="},"machine_readable":[{"url":"https://codewithpapers.app/llms.txt","what":"the machine catalog: every machine-readable file, counted"},{"url":"https://codewithpapers.app/index/manifest.json","what":"paper-to-code index by arXiv id, with Syntology's counts"},{"url":"https://codewithpapers.app/search/manifest.json","what":"site search index (titles, authors) and its files"},{"url":"https://codewithpapers.app/download","what":"bulk files: Syntology's layer, described there"},{"url":"https://codewithpapers.app/build_manifest.json","what":"the build record: inputs, counts, exclusions, probes"}]},"url":"/task/speaker-verification/papers/2","list_of":"/task/speaker-verification","task":"Speaker Verification","archive":{"snapshot":"2025-07-28"},"key_notes":{"n_ran_checked":"legacy name, kept unchanged so existing readers do not break: it counts the samples that ran with no instrument failure (honoured, violated, and ran with no contract checked); it does not mean a contract was checked, and the pages print it as 'K with no instrument failure', not 'K checked'","n_constructed":"a sub-count of the samples that ran, never subtracted from them and never a failure: an executed sample whose run returned an instance of its own class (fixture_out_type equals the entry name): the run built an object and did not compute a result (Syntology's RAN record, counts.constructed)"},"syntology_read_at":"2026-09-28T10:30:06+00:00","order":"archive","order_definition":"repositories listed in the archive (most first), then date (newest first), then slug","page":2,"pages_in_order":8,"rows_per_page":100,"rows":[101,200],"of":746,"counts":{"archive_papers_tagged":746,"with_a_code_link":200,"where_syntology_ran_a_sample":29,"not_listed_spam_title":0,"listed":746,"listed_where_code_ran":29,"where_syntology_ran_a_sample_split":{"with_a_run_with_no_instrument_failure":23,"every_run_a_failure_of_syntologys_instrument":6,"listed_with_a_run_with_no_instrument_failure":23,"listed_every_run_a_failure_of_syntologys_instrument":6,"filter":{"states":["a run with no instrument failure","any run, instrument failures included"],"default":"a run with no instrument failure","note":"on the 'only where code ran' pages the default hides, in the browser, the rows where every run was a failure of Syntology's instrument; the second state shows them again. Rows are hidden, never re-ordered; these twins list every row"}},"definition":"distinct papers the archive tags; 'where Syntology ran a sample' counts papers with at least one harvested sample that ran, which is not a correctness claim"},"first_page":"/task/speaker-verification","prev":"/task/speaker-verification","next":"/task/speaker-verification/papers/3","papers":[{"url":"/paper/pushing-the-limits-of-self-supervised-speaker","slug":"pushing-the-limits-of-self-supervised-speaker","title":"Pushing the limits of self-supervised speaker verification using regularized distillation framework","date":"2022-11-08","arxiv_id":"2211.04168","repositories_listed":1,"syntology":null},{"url":"/paper/integrated-parameter-efficient-tuning-for","slug":"integrated-parameter-efficient-tuning-for","title":"Integrated Parameter-Efficient Tuning for General-Purpose Audio Models","date":"2022-11-04","arxiv_id":"2211.02227","repositories_listed":1,"syntology":null},{"url":"/paper/psvrf-learning-to-restore-pitch-shifted-voice","slug":"psvrf-learning-to-restore-pitch-shifted-voice","title":"PSVRF: Learning to restore Pitch-Shifted Voice without reference","date":"2022-10-06","arxiv_id":"2210.02731","repositories_listed":1,"syntology":null},{"url":"/paper/an-attention-based-backend-allowing-efficient","slug":"an-attention-based-backend-allowing-efficient","title":"An attention-based backend allowing efficient fine-tuning of transformer models for speaker verification","date":"2022-10-03","arxiv_id":"2210.01273","repositories_listed":1,"syntology":null},{"url":"/paper/voice-spoofing-countermeasures-taxonomy-state","slug":"voice-spoofing-countermeasures-taxonomy-state","title":"Voice Spoofing Countermeasures: Taxonomy, State-of-the-art, experimental analysis of generalizability, open challenges, and the way forward","date":"2022-10-02","arxiv_id":"2210.00417","repositories_listed":1,"syntology":null},{"url":"/paper/the-2022-far-field-speaker-verification","slug":"the-2022-far-field-speaker-verification","title":"The 2022 Far-field Speaker Verification Challenge: Exploring domain mismatch and semi-supervised learning under the far-field scenario","date":"2022-09-12","arxiv_id":"2209.05273","repositories_listed":1,"syntology":null},{"url":"/paper/deid-vc-speaker-de-identification-via-zero","slug":"deid-vc-speaker-de-identification-via-zero","title":"DeID-VC: Speaker De-identification via Zero-shot Pseudo Voice Conversion","date":"2022-09-09","arxiv_id":"2209.04530","repositories_listed":1,"syntology":null},{"url":"/paper/on-the-potential-of-jointly-optimised","slug":"on-the-potential-of-jointly-optimised","title":"On the potential of jointly-optimised solutions to spoofing attack detection and automatic speaker verification","date":"2022-09-01","arxiv_id":"2209.00506","repositories_listed":1,"syntology":null},{"url":"/paper/indicsuperb-a-speech-processing-universal","slug":"indicsuperb-a-speech-processing-universal","title":"IndicSUPERB: A Speech Processing Universal Performance Benchmark for Indian languages","date":"2022-08-24","arxiv_id":"2208.11761","repositories_listed":1,"syntology":{"n":7,"n_ran":4,"n_constructed":0,"n_ran_checked":4,"n_instrument":0,"n_unverified":3,"n_honours":0,"n_violates":0,"n_no_contract":4,"n_pointer_only":0,"phrase":"4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified","sample_list":"/paper/indicsuperb-a-speech-processing-universal#ran","syntology_url":"https://syntology.ai/paper/2208.11761","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2208.11761"}},"official":{"repos":["AI4Bharat/indicSUPERB"],"state":"official (archive's flag): 4 ran","n_ran":4,"n_constructed":0,"n_ran_no_instrument_failure":4,"n_unverified":3,"ran_from_kinds":["official"]}}},{"url":"/paper/analysis-of-impact-of-emotions-on-target","slug":"analysis-of-impact-of-emotions-on-target","title":"Analysis of impact of emotions on target speech extraction and speech separation","date":"2022-08-15","arxiv_id":"2208.07091","repositories_listed":1,"syntology":null},{"url":"/paper/non-contrastive-self-supervised-learning-of","slug":"non-contrastive-self-supervised-learning-of","title":"Non-Contrastive Self-Supervised Learning of Utterance-Level Speech Representations","date":"2022-08-10","arxiv_id":"2208.05413","repositories_listed":1,"syntology":null},{"url":"/paper/a-single-self-supervised-model-for-many","slug":"a-single-self-supervised-model-for-many","title":"u-HuBERT: Unified Mixed-Modal Speech Pretraining And Zero-Shot Transfer to Unlabeled Modality","date":"2022-07-14","arxiv_id":"2207.07036","repositories_listed":1,"syntology":null},{"url":"/paper/cross-age-speaker-verification-learning-age","slug":"cross-age-speaker-verification-learning-age","title":"Cross-Age Speaker Verification: Learning Age-Invariant Speaker Embeddings","date":"2022-07-13","arxiv_id":"2207.05929","repositories_listed":1,"syntology":null},{"url":"/paper/label-efficient-self-supervised-speaker","slug":"label-efficient-self-supervised-speaker","title":"Label-Efficient Self-Supervised Speaker Verification With Information Maximization and Contrastive Learning","date":"2022-07-12","arxiv_id":"2207.05506","repositories_listed":1,"syntology":null},{"url":"/paper/fatnet-cost-effective-approach-towards","slug":"fatnet-cost-effective-approach-towards","title":"FAtNet: Cost-Effective Approach Towards Mitigating the Linguistic Bias in Speaker Verification Systems","date":"2022-07-01","arxiv_id":null,"repositories_listed":1,"syntology":null},{"url":"/paper/extended-u-net-for-speaker-verification-in","slug":"extended-u-net-for-speaker-verification-in","title":"Extended U-Net for Speaker Verification in Noisy Environments","date":"2022-06-27","arxiv_id":"2206.13044","repositories_listed":1,"syntology":null},{"url":"/paper/learning-lip-based-audio-visual-speaker","slug":"learning-lip-based-audio-visual-speaker","title":"Learning Lip-Based Audio-Visual Speaker Embeddings with AV-HuBERT","date":"2022-05-15","arxiv_id":"2205.07180","repositories_listed":1,"syntology":null},{"url":"/paper/evi-multilingual-spoken-dialogue-tasks-and-1","slug":"evi-multilingual-spoken-dialogue-tasks-and-1","title":"EVI: Multilingual Spoken Dialogue Tasks and Dataset for Knowledge-Based Enrolment, Verification, and Identification","date":"2022-04-28","arxiv_id":"2204.13496","repositories_listed":1,"syntology":null},{"url":"/paper/multi-task-learning-improves-synthetic-speech","slug":"multi-task-learning-improves-synthetic-speech","title":"Multi-task learning improves synthetic speech detection","date":"2022-04-27","arxiv_id":null,"repositories_listed":1,"syntology":null},{"url":"/paper/dictionary-attacks-on-speaker-verification","slug":"dictionary-attacks-on-speaker-verification","title":"Dictionary Attacks on Speaker Verification","date":"2022-04-24","arxiv_id":"2204.11304","repositories_listed":1,"syntology":null},{"url":"/paper/is-speech-pathology-a-biomarker-in-automatic","slug":"is-speech-pathology-a-biomarker-in-automatic","title":"The effect of speech pathology on automatic speaker verification -- a large-scale study","date":"2022-04-13","arxiv_id":"2204.06450","repositories_listed":1,"syntology":null},{"url":"/paper/selective-kernel-attention-for-robust-speaker","slug":"selective-kernel-attention-for-robust-speaker","title":"Frequency and Multi-Scale Selective Kernel Attention for Speaker Verification","date":"2022-04-03","arxiv_id":"2204.01005","repositories_listed":1,"syntology":null},{"url":"/paper/robust-disentangled-variational-speech","slug":"robust-disentangled-variational-speech","title":"Robust Disentangled Variational Speech Representation Learning for Zero-shot Voice Conversion","date":"2022-03-30","arxiv_id":"2203.16705","repositories_listed":1,"syntology":null},{"url":"/paper/decomposed-temporal-dynamic-cnn-efficient","slug":"decomposed-temporal-dynamic-cnn-efficient","title":"Decomposed Temporal Dynamic CNN: Efficient Time-Adaptive Network for Text-Independent Speaker Verification Explained with Speaker Activation Map","date":"2022-03-29","arxiv_id":"2203.15277","repositories_listed":1,"syntology":null},{"url":"/paper/lighthubert-lightweight-and-configurable","slug":"lighthubert-lightweight-and-configurable","title":"LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT","date":"2022-03-29","arxiv_id":"2203.15610","repositories_listed":1,"syntology":{"n":8,"n_ran":5,"n_constructed":0,"n_ran_checked":4,"n_instrument":1,"n_unverified":3,"n_honours":0,"n_violates":0,"n_no_contract":4,"n_pointer_only":1,"phrase":"5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified","sample_list":"/paper/lighthubert-lightweight-and-configurable#ran","syntology_url":"https://syntology.ai/paper/2203.15610","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2203.15610"}},"official":{"repos":["mechanicalsea/lighthubert"],"state":"official (archive's flag): 5 ran","n_ran":5,"n_constructed":0,"n_ran_no_instrument_failure":4,"n_unverified":3,"ran_from_kinds":["official"]}}},{"url":"/paper/towards-privacy-preserving-speech","slug":"towards-privacy-preserving-speech","title":"A Speech Representation Anonymization Framework via Selective Noise Perturbation","date":"2022-03-26","arxiv_id":"2203.14171","repositories_listed":1,"syntology":null},{"url":"/paper/ecapa-tdnn-for-multi-speaker-text-to-speech","slug":"ecapa-tdnn-for-multi-speaker-text-to-speech","title":"ECAPA-TDNN for Multi-speaker Text-to-speech Synthesis","date":"2022-03-20","arxiv_id":"2203.10473","repositories_listed":1,"syntology":null},{"url":"/paper/parameter-free-attentive-scoring-for-speaker","slug":"parameter-free-attentive-scoring-for-speaker","title":"Parameter-Free Attentive Scoring for Speaker Verification","date":"2022-03-10","arxiv_id":"2203.05642","repositories_listed":1,"syntology":null},{"url":"/paper/explainable-deepfake-and-spoofing-detection","slug":"explainable-deepfake-and-spoofing-detection","title":"Explainable deepfake and spoofing detection: an attack analysis using SHapley Additive exPlanations","date":"2022-02-28","arxiv_id":"2202.13693","repositories_listed":1,"syntology":null},{"url":"/paper/magnitude-aware-probabilistic-speaker","slug":"magnitude-aware-probabilistic-speaker","title":"Magnitude-aware Probabilistic Speaker Embeddings","date":"2022-02-28","arxiv_id":"2202.13826","repositories_listed":1,"syntology":null},{"url":"/paper/improving-fairness-in-speaker-verification","slug":"improving-fairness-in-speaker-verification","title":"Improving fairness in speaker verification via Group-adapted Fusion Network","date":"2022-02-23","arxiv_id":"2202.11323","repositories_listed":1,"syntology":{"n":1,"n_ran":1,"n_constructed":0,"n_ran_checked":1,"n_instrument":0,"n_unverified":0,"n_honours":0,"n_violates":0,"n_no_contract":1,"n_pointer_only":0,"phrase":"1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified","sample_list":"/paper/improving-fairness-in-speaker-verification#ran","syntology_url":"https://syntology.ai/paper/2202.11323","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2202.11323"}},"official":{"repos":["huashen218/voxceleb-fairness"],"state":"official (archive's flag): 1 ran","n_ran":1,"n_constructed":0,"n_ran_no_instrument_failure":1,"n_unverified":0,"ran_from_kinds":["official"]}}},{"url":"/paper/on-the-detection-of-adaptive-adversarial","slug":"on-the-detection-of-adaptive-adversarial","title":"On the Detection of Adaptive Adversarial Attacks in Speaker Verification Systems","date":"2022-02-11","arxiv_id":"2202.05725","repositories_listed":1,"syntology":null},{"url":"/paper/a-new-fusion-strategy-for-spoofing-aware","slug":"a-new-fusion-strategy-for-spoofing-aware","title":"A Probabilistic Fusion Framework for Spoofing Aware Speaker Verification","date":"2022-02-10","arxiv_id":"2202.05253","repositories_listed":1,"syntology":null},{"url":"/paper/bias-in-automated-speaker-recognition","slug":"bias-in-automated-speaker-recognition","title":"Bias in Automated Speaker Recognition","date":"2022-01-24","arxiv_id":"2201.09486","repositories_listed":1,"syntology":null},{"url":"/paper/a-practical-guide-to-logical-access-voice","slug":"a-practical-guide-to-logical-access-voice","title":"A Practical Guide to Logical Access Voice Presentation Attack Detection","date":"2022-01-10","arxiv_id":"2201.03321","repositories_listed":1,"syntology":null},{"url":"/paper/task-specific-optimization-of-virtual-channel","slug":"task-specific-optimization-of-virtual-channel","title":"Task-specific Optimization of Virtual Channel Linear Prediction-based Speech Dereverberation Front-End for Far-Field Speaker Verification","date":"2021-12-27","arxiv_id":"2112.13569","repositories_listed":1,"syntology":null},{"url":"/paper/rawnext-speaker-verification-system-for","slug":"rawnext-speaker-verification-system-for","title":"RawNeXt: Speaker verification system for variable-duration utterances with deep layer aggregation and extended dynamic scaling policies","date":"2021-12-15","arxiv_id":"2112.07935","repositories_listed":1,"syntology":null},{"url":"/paper/multisv-dataset-for-far-field-multi-channel","slug":"multisv-dataset-for-far-field-multi-channel","title":"MultiSV: Dataset for Far-Field Multi-Channel Speaker Verification","date":"2021-11-11","arxiv_id":"2111.06458","repositories_listed":1,"syntology":null},{"url":"/paper/sig-vc-a-speaker-information-guided-zero-shot","slug":"sig-vc-a-speaker-information-guided-zero-shot","title":"SIG-VC: A Speaker Information Guided Zero-shot Voice Conversion System for Both Human Beings and Machines","date":"2021-11-06","arxiv_id":"2111.03811","repositories_listed":1,"syntology":null},{"url":"/paper/cs-rep-making-speaker-verification-networks","slug":"cs-rep-making-speaker-verification-networks","title":"CS-Rep: Making Speaker Verification Networks Embracing Re-parameterization","date":"2021-10-26","arxiv_id":"2110.13465","repositories_listed":1,"syntology":null},{"url":"/paper/real-additive-margin-softmax-for-speaker","slug":"real-additive-margin-softmax-for-speaker","title":"Real Additive Margin Softmax for Speaker Verification","date":"2021-10-18","arxiv_id":"2110.09116","repositories_listed":1,"syntology":null},{"url":"/paper/a-study-of-the-robustness-of-raw-waveform","slug":"a-study-of-the-robustness-of-raw-waveform","title":"A study of the robustness of raw waveform based speaker embeddings under mismatched conditions","date":"2021-10-08","arxiv_id":"2110.04265","repositories_listed":1,"syntology":null},{"url":"/paper/filteraugment-an-acoustic-environmental-data","slug":"filteraugment-an-acoustic-environmental-data","title":"FilterAugment: An Acoustic Environmental Data Augmentation Method","date":"2021-10-07","arxiv_id":"2110.03282","repositories_listed":1,"syntology":null},{"url":"/paper/temporal-dynamic-convolutional-neural-network","slug":"temporal-dynamic-convolutional-neural-network","title":"Temporal Dynamic Convolutional Neural Network for Text-Independent Speaker Verification and Phonemetic Analysis","date":"2021-10-07","arxiv_id":"2110.03213","repositories_listed":1,"syntology":null},{"url":"/paper/pl-eesr-perceptual-loss-based-end-to-end","slug":"pl-eesr-perceptual-loss-based-end-to-end","title":"PL-EESR: Perceptual Loss Based END-TO-END Robust Speaker Representation Extraction","date":"2021-10-03","arxiv_id":"2110.00940","repositories_listed":1,"syntology":{"n":1,"n_ran":1,"n_constructed":0,"n_ran_checked":0,"n_instrument":1,"n_unverified":0,"n_honours":0,"n_violates":0,"n_no_contract":0,"n_pointer_only":1,"phrase":"1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified","sample_list":"/paper/pl-eesr-perceptual-loss-based-end-to-end#ran","syntology_url":"https://syntology.ai/paper/2110.00940","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2110.00940"}},"official":{"repos":["mmmmayi/pl-eesr"],"state":"official (archive's flag): 1 ran","n_ran":1,"n_constructed":0,"n_ran_no_instrument_failure":0,"n_unverified":0,"ran_from_kinds":["official"]}}},{"url":"/paper/speechnas-towards-better-trade-off-between","slug":"speechnas-towards-better-trade-off-between","title":"SpeechNAS: Towards Better Trade-off between Latency and Accuracy for Large-Scale Speaker Verification","date":"2021-09-18","arxiv_id":"2109.08839","repositories_listed":1,"syntology":null},{"url":"/paper/fastaudio-a-learnable-audio-front-end-for","slug":"fastaudio-a-learnable-audio-front-end-for","title":"FastAudio: A Learnable Audio Front-End for Spoof Speech Detection","date":"2021-09-06","arxiv_id":"2109.02774","repositories_listed":1,"syntology":null},{"url":"/paper/efficient-attention-branch-network-with","slug":"efficient-attention-branch-network-with","title":"Efficient Attention Branch Network with Combined Loss Function for Automatic Speaker Verification Spoof Detection","date":"2021-09-05","arxiv_id":"2109.02051","repositories_listed":1,"syntology":null},{"url":"/paper/asvspoof-2021-automatic-speaker-verification","slug":"asvspoof-2021-automatic-speaker-verification","title":"ASVspoof 2021: Automatic Speaker Verification Spoofing and Countermeasures Challenge Evaluation Plan","date":"2021-09-01","arxiv_id":"2109.00535","repositories_listed":1,"syntology":null},{"url":"/paper/ctal-pre-training-cross-modal-transformer-for","slug":"ctal-pre-training-cross-modal-transformer-for","title":"CTAL: Pre-training Cross-modal Transformer for Audio-and-Language Representations","date":"2021-09-01","arxiv_id":"2109.00181","repositories_listed":1,"syntology":{"n":19,"n_ran":15,"n_constructed":2,"n_ran_checked":15,"n_instrument":0,"n_unverified":4,"n_honours":1,"n_violates":0,"n_no_contract":14,"n_pointer_only":6,"phrase":"15 ran (of which 2 constructed an object rather than computing a result; 15 with no instrument failure: 1 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified","sample_list":"/paper/ctal-pre-training-cross-modal-transformer-for#ran","syntology_url":"https://syntology.ai/paper/2109.00181","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2109.00181"}},"official":{"repos":["ydkwim/ctal"],"state":"official (archive's flag): 13 ran","n_ran":13,"n_constructed":0,"n_ran_no_instrument_failure":13,"n_unverified":1,"ran_from_kinds":["found_in_text","official"]}}},{"url":"/paper/fdn-finite-difference-network-with","slug":"fdn-finite-difference-network-with","title":"FDN: Finite Difference Network with Hierarchical Convolutional Features for Text-independent Speaker Verification","date":"2021-08-18","arxiv_id":"2108.07974","repositories_listed":1,"syntology":null},{"url":"/paper/analyzing-speaker-information-in-self","slug":"analyzing-speaker-information-in-self","title":"Analyzing Speaker Information in Self-Supervised Models to Improve Zero-Resource Speech Processing","date":"2021-08-02","arxiv_id":"2108.00917","repositories_listed":1,"syntology":{"n":3,"n_ran":3,"n_constructed":0,"n_ran_checked":0,"n_instrument":3,"n_unverified":0,"n_honours":0,"n_violates":0,"n_no_contract":0,"n_pointer_only":0,"phrase":"3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified","sample_list":"/paper/analyzing-speaker-information-in-self#ran","syntology_url":"https://syntology.ai/paper/2108.00917","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2108.00917"}},"official":null}},{"url":"/paper/end-to-end-spectro-temporal-graph-attention","slug":"end-to-end-spectro-temporal-graph-attention","title":"End-to-End Spectro-Temporal Graph Attention Networks for Speaker Verification Anti-Spoofing and Speech Deepfake Detection","date":"2021-07-27","arxiv_id":"2107.12710","repositories_listed":1,"syntology":null},{"url":"/paper/use-of-speaker-recognition-approaches-for","slug":"use-of-speaker-recognition-approaches-for","title":"Use of speaker recognition approaches for learning and evaluating embedding representations of musical instrument sounds","date":"2021-07-24","arxiv_id":"2107.11506","repositories_listed":1,"syntology":null},{"url":"/paper/spotting-adversarial-samples-for-speaker","slug":"spotting-adversarial-samples-for-speaker","title":"Adversarial Sample Detection for Speaker Verification by Neural Vocoders","date":"2021-07-01","arxiv_id":"2107.00309","repositories_listed":1,"syntology":null},{"url":"/paper/voting-for-the-right-answer-adversarial","slug":"voting-for-the-right-answer-adversarial","title":"Voting for the right answer: Adversarial defense for speaker verification","date":"2021-06-15","arxiv_id":"2106.07868","repositories_listed":1,"syntology":null},{"url":"/paper/visualizing-classifier-adjacency-relations-a","slug":"visualizing-classifier-adjacency-relations-a","title":"Visualizing Classifier Adjacency Relations: A Case Study in Speaker Verification and Voice Anti-Spoofing","date":"2021-06-11","arxiv_id":"2106.06362","repositories_listed":1,"syntology":null},{"url":"/paper/attack-on-practical-speaker-verification","slug":"attack-on-practical-speaker-verification","title":"Attack on practical speaker verification system using universal adversarial perturbations","date":"2021-05-19","arxiv_id":"2105.09022","repositories_listed":1,"syntology":{"n":2,"n_ran":0,"n_constructed":0,"n_ran_checked":0,"n_instrument":0,"n_unverified":2,"n_honours":0,"n_violates":0,"n_no_contract":0,"n_pointer_only":0,"phrase":"0 ran · 2 unverified","sample_list":"/paper/attack-on-practical-speaker-verification#ran","syntology_url":"https://syntology.ai/paper/2105.09022","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2105.09022"}},"official":{"repos":["zhang-wy15/Attack_practical_asv"],"state":"official: harvested, nothing ran","n_ran":0,"n_constructed":0,"n_ran_no_instrument_failure":0,"n_unverified":2,"ran_from_kinds":[]}}},{"url":"/paper/scaling-to-many-languages-with-a-triaged","slug":"scaling-to-many-languages-with-a-triaged","title":"SpeakerStew: Scaling to Many Languages with a Triaged Multilingual Text-Dependent and Text-Independent Speaker Verification System","date":"2021-04-05","arxiv_id":"2104.02125","repositories_listed":1,"syntology":null},{"url":"/paper/attention-back-end-for-automatic-speaker","slug":"attention-back-end-for-automatic-speaker","title":"Attention Back-end for Automatic Speaker Verification with Multiple Enrollment Utterances","date":"2021-04-04","arxiv_id":"2104.01541","repositories_listed":1,"syntology":{"n":1,"n_ran":1,"n_constructed":0,"n_ran_checked":1,"n_instrument":0,"n_unverified":0,"n_honours":0,"n_violates":0,"n_no_contract":1,"n_pointer_only":0,"phrase":"1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified","sample_list":"/paper/attention-back-end-for-automatic-speaker#ran","syntology_url":"https://syntology.ai/paper/2104.01541","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2104.01541"}},"official":{"repos":["nii-yamagishilab/Attention_Backend_for_ASV"],"state":"official (archive's flag): 1 ran","n_ran":1,"n_constructed":0,"n_ran_no_instrument_failure":1,"n_unverified":0,"ran_from_kinds":["official"]}}},{"url":"/paper/target-speaker-verification-with-selective","slug":"target-speaker-verification-with-selective","title":"Target Speaker Verification with Selective Auditory Attention for Single and Multi-talker Speech","date":"2021-03-30","arxiv_id":"2103.16269","repositories_listed":1,"syntology":null},{"url":"/paper/learning-spectro-temporal-representations-of","slug":"learning-spectro-temporal-representations-of","title":"Learning spectro-temporal representations of complex sounds with parameterized neural networks","date":"2021-03-12","arxiv_id":"2103.07125","repositories_listed":1,"syntology":null},{"url":"/paper/the-npu-system-for-the-2020-personalized","slug":"the-npu-system-for-the-2020-personalized","title":"The NPU System for the 2020 Personalized Voice Trigger Challenge","date":"2021-02-26","arxiv_id":"2102.13552","repositories_listed":1,"syntology":null},{"url":"/paper/a-speaker-verification-backend-with-robust","slug":"a-speaker-verification-backend-with-robust","title":"A Speaker Verification Backend with Robust Performance across Conditions","date":"2021-02-02","arxiv_id":"2102.01760","repositories_listed":1,"syntology":null},{"url":"/paper/self-supervised-text-independent-speaker","slug":"self-supervised-text-independent-speaker","title":"Self-supervised Text-independent Speaker Verification using Prototypical Momentum Contrastive Learning","date":"2020-12-13","arxiv_id":"2012.07178","repositories_listed":1,"syntology":null},{"url":"/paper/adversarial-disentanglement-of-speaker","slug":"adversarial-disentanglement-of-speaker","title":"Adversarial Disentanglement of Speaker Representation for Attribute-Driven Privacy Preservation","date":"2020-12-08","arxiv_id":"2012.04454","repositories_listed":1,"syntology":null},{"url":"/paper/mask-proxy-loss-for-text-independent-speaker","slug":"mask-proxy-loss-for-text-independent-speaker","title":"Masked Proxy Loss For Text-Independent Speaker Verification","date":"2020-11-09","arxiv_id":"2011.04491","repositories_listed":1,"syntology":null},{"url":"/paper/end-to-end-anti-spoofing-with-rawnet2","slug":"end-to-end-anti-spoofing-with-rawnet2","title":"End-to-end anti-spoofing with RawNet2","date":"2020-11-02","arxiv_id":"2011.01108","repositories_listed":1,"syntology":{"n":3,"n_ran":3,"n_constructed":0,"n_ran_checked":0,"n_instrument":3,"n_unverified":0,"n_honours":0,"n_violates":0,"n_no_contract":0,"n_pointer_only":3,"phrase":"3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified","sample_list":"/paper/end-to-end-anti-spoofing-with-rawnet2#ran","syntology_url":"https://syntology.ai/paper/2011.01108","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2011.01108"}},"official":{"repos":["eurecom-asp/rawnet2-antispoofing"],"state":"official (archive's flag): 3 ran","n_ran":3,"n_constructed":0,"n_ran_no_instrument_failure":0,"n_unverified":0,"ran_from_kinds":["official"]}}},{"url":"/paper/leveraging-speaker-attribute-information","slug":"leveraging-speaker-attribute-information","title":"Leveraging speaker attribute information using multi task learning for speaker verification and diarization","date":"2020-10-27","arxiv_id":"2010.14269","repositories_listed":1,"syntology":null},{"url":"/paper/raw-x-vector-multi-scale-time-domain-speaker","slug":"raw-x-vector-multi-scale-time-domain-speaker","title":"Y-Vector: Multiscale Waveform Encoder for Speaker Embedding","date":"2020-10-24","arxiv_id":"2010.12951","repositories_listed":1,"syntology":null},{"url":"/paper/unsupervised-representation-learning-for-1","slug":"unsupervised-representation-learning-for-1","title":"Unsupervised Representation Learning for Speaker Recognition via Contrastive Equilibrium Learning","date":"2020-10-22","arxiv_id":"2010.11433","repositories_listed":1,"syntology":null},{"url":"/paper/learning-speaker-embedding-from-text-to","slug":"learning-speaker-embedding-from-text-to","title":"Learning Speaker Embedding from Text-to-Speech","date":"2020-10-21","arxiv_id":"2010.11221","repositories_listed":1,"syntology":null},{"url":"/paper/a-fully-tensorized-recurrent-neural-network","slug":"a-fully-tensorized-recurrent-neural-network","title":"A Fully Tensorized Recurrent Neural Network","date":"2020-10-08","arxiv_id":"2010.04196","repositories_listed":1,"syntology":null},{"url":"/paper/sum-product-networks-for-robust-automatic","slug":"sum-product-networks-for-robust-automatic","title":"Sum-Product Networks for Robust Automatic Speaker Identification","date":"2020-08-13","arxiv_id":"1910.11969","repositories_listed":1,"syntology":null},{"url":"/paper/neural-plda-modeling-for-end-to-end-speaker","slug":"neural-plda-modeling-for-end-to-end-speaker","title":"Neural PLDA Modeling for End-to-End Speaker Verification","date":"2020-08-11","arxiv_id":"2008.04527","repositories_listed":1,"syntology":null},{"url":"/paper/audio-spoofing-verification-using-deep","slug":"audio-spoofing-verification-using-deep","title":"Audio Spoofing Verification using Deep Convolutional Neural Networks by Transfer Learning","date":"2020-08-08","arxiv_id":"2008.03464","repositories_listed":1,"syntology":null},{"url":"/paper/deep-multi-metric-learning-for-text","slug":"deep-multi-metric-learning-for-text","title":"Deep multi-metric learning for text-independent speaker verification","date":"2020-07-17","arxiv_id":"2007.10479","repositories_listed":1,"syntology":null},{"url":"/paper/dynamically-mitigating-data-discrepancy-with","slug":"dynamically-mitigating-data-discrepancy-with","title":"Dynamically Mitigating Data Discrepancy with Balanced Focal Loss for Replay Attack Detection","date":"2020-06-25","arxiv_id":"2006.14563","repositories_listed":1,"syntology":null},{"url":"/paper/from-speaker-verification-to-multispeaker","slug":"from-speaker-verification-to-multispeaker","title":"From Speaker Verification to Multispeaker Speech Synthesis, Deep Transfer with Feedback Constraint","date":"2020-05-10","arxiv_id":"2005.04587","repositories_listed":1,"syntology":null},{"url":"/paper/crop-aggregating-for-short-utterances-speaker","slug":"crop-aggregating-for-short-utterances-speaker","title":"Segment Aggregation for short utterances speaker verification using raw waveforms","date":"2020-05-07","arxiv_id":"2005.03329","repositories_listed":1,"syntology":null},{"url":"/paper/meta-learning-for-short-utterance-speaker","slug":"meta-learning-for-short-utterance-speaker","title":"Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs","date":"2020-04-06","arxiv_id":"2004.02863","repositories_listed":1,"syntology":null},{"url":"/paper/a-comparison-of-metric-learning-loss","slug":"a-comparison-of-metric-learning-loss","title":"A Comparison of Metric Learning Loss Functions for End-To-End Speaker Verification","date":"2020-03-31","arxiv_id":"2003.14021","repositories_listed":1,"syntology":null},{"url":"/paper/nplda-a-deep-neural-plda-model-for-speaker","slug":"nplda-a-deep-neural-plda-model-for-speaker","title":"NPLDA: A Deep Neural PLDA Model for Speaker Verification","date":"2020-02-10","arxiv_id":"2002.03562","repositories_listed":1,"syntology":null},{"url":"/paper/an-initial-investigation-on-optimizing-tandem","slug":"an-initial-investigation-on-optimizing-tandem","title":"An initial investigation on optimizing tandem speaker verification and countermeasure systems using reinforcement learning","date":"2020-02-06","arxiv_id":"2002.03801","repositories_listed":1,"syntology":null},{"url":"/paper/dropclass-and-dropadapt-dropping-classes-for","slug":"dropclass-and-dropadapt-dropping-classes-for","title":"DropClass and DropAdapt: Dropping classes for deep speaker representation learning","date":"2020-02-02","arxiv_id":"2002.00453","repositories_listed":1,"syntology":null},{"url":"/paper/pairwise-discriminative-neural-plda-for","slug":"pairwise-discriminative-neural-plda-for","title":"Pairwise Discriminative Neural PLDA for Speaker Verification","date":"2020-01-20","arxiv_id":"2001.07034","repositories_listed":1,"syntology":null},{"url":"/paper/learning-speaker-embedding-with-momentum","slug":"learning-speaker-embedding-with-momentum","title":"Learning Speaker Embedding with Momentum Contrast","date":"2020-01-07","arxiv_id":"2001.01986","repositories_listed":1,"syntology":null},{"url":"/paper/what-does-a-network-layer-hear-analyzing","slug":"what-does-a-network-layer-hear-analyzing","title":"What does a network layer hear? Analyzing hidden representations of end-to-end ASR through speech synthesis","date":"2019-11-04","arxiv_id":"1911.01102","repositories_listed":1,"syntology":null},{"url":"/paper/robust-speaker-recognition-using-unsupervised","slug":"robust-speaker-recognition-using-unsupervised","title":"Robust speaker recognition using unsupervised adversarial invariance","date":"2019-11-03","arxiv_id":"1911.00940","repositories_listed":1,"syntology":null},{"url":"/paper/spoofing-speaker-verification-systems-with","slug":"spoofing-speaker-verification-systems-with","title":"Spoofing Speaker Verification Systems with Deep Multi-speaker Text-to-speech Synthesis","date":"2019-10-29","arxiv_id":"1910.13054","repositories_listed":1,"syntology":null},{"url":"/paper/feature-enhancement-with-deep-feature-losses","slug":"feature-enhancement-with-deep-feature-losses","title":"Feature Enhancement with Deep Feature Losses for Speaker Verification","date":"2019-10-25","arxiv_id":"1910.11905","repositories_listed":1,"syntology":null},{"url":"/paper/adversarial-attacks-on-spoofing","slug":"adversarial-attacks-on-spoofing","title":"Adversarial Attacks on Spoofing Countermeasures of automatic speaker verification","date":"2019-10-19","arxiv_id":"1910.08716","repositories_listed":1,"syntology":null},{"url":"/paper/deep-residual-neural-networks-for-audio","slug":"deep-residual-neural-networks-for-audio","title":"Deep Residual Neural Networks for Audio Spoofing Detection","date":"2019-06-30","arxiv_id":"1907.00501","repositories_listed":1,"syntology":null},{"url":"/paper/spoof-detection-using-x-vector-and-feature","slug":"spoof-detection-using-x-vector-and-feature","title":"Spoof detection using time-delay shallow neural network and feature switching","date":"2019-04-16","arxiv_id":"1904.07453","repositories_listed":1,"syntology":null},{"url":"/paper/contrastive-predictive-coding-based-feature","slug":"contrastive-predictive-coding-based-feature","title":"Contrastive Predictive Coding Based Feature for Automatic Speaker Verification","date":"2019-04-01","arxiv_id":"1904.01575","repositories_listed":1,"syntology":null},{"url":"/paper/attentive-filtering-networks-for-audio-replay","slug":"attentive-filtering-networks-for-audio-replay","title":"Attentive Filtering Networks for Audio Replay Attack Detection","date":"2018-10-31","arxiv_id":"1810.13048","repositories_listed":1,"syntology":{"n":3,"n_ran":3,"n_constructed":0,"n_ran_checked":3,"n_instrument":0,"n_unverified":0,"n_honours":0,"n_violates":0,"n_no_contract":3,"n_pointer_only":0,"phrase":"3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified","sample_list":"/paper/attentive-filtering-networks-for-audio-replay#ran","syntology_url":"https://syntology.ai/paper/1810.13048","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"1810.13048"}},"official":{"repos":["jefflai108/Attentive-Filtering-Network"],"state":"official (archive's flag): 3 ran","n_ran":3,"n_constructed":0,"n_ran_no_instrument_failure":3,"n_unverified":0,"ran_from_kinds":["official"]}}},{"url":"/paper/noise-invariant-frame-selection-a-simple","slug":"noise-invariant-frame-selection-a-simple","title":"Noise Invariant Frame Selection: A Simple Method to Address the Background Noise Problem for Text-independent Speaker Verification","date":"2018-05-03","arxiv_id":"1805.01259","repositories_listed":1,"syntology":null},{"url":"/paper/t-dcf-a-detection-cost-function-for-the","slug":"t-dcf-a-detection-cost-function-for-the","title":"t-DCF: a Detection Cost Function for the Tandem Assessment of Spoofing Countermeasures and Automatic Speaker Verification","date":"2018-04-25","arxiv_id":"1804.09618","repositories_listed":1,"syntology":null},{"url":"/paper/exploring-the-encoding-layer-and-loss","slug":"exploring-the-encoding-layer-and-loss","title":"Exploring the Encoding Layer and Loss Function in End-to-End Speaker and Language Recognition System","date":"2018-04-14","arxiv_id":"1804.05160","repositories_listed":1,"syntology":null},{"url":"/paper/aicyber-at-semeval-2016-task-4-i-vector-based","slug":"aicyber-at-semeval-2016-task-4-i-vector-based","title":"Aicyber at SemEval-2016 Task 4: i-vector based sentence representation","date":"2016-06-01","arxiv_id":null,"repositories_listed":1,"syntology":null}],"record_sha256":"b21cc3ef0fa01cdb3ce3e0f3715d2a7144f85cae0458e91b71fdbd1e4725985b","record_changed_at":"2026-09-28","record_changed_at_basis":"first_hashed"}