Papers › Improving Spoken Language Identification with Map-Mix

Improving Spoken Language Identification with Map-Mix

16 Feb 2023arXiv:2302.08229archive 2025-07-28

Shangeth Rajaa, Kriti Anandan, Swaraj Dalmia, Tarun Gupta, Eng Siong Chng

The pre-trained multi-lingual XLSR model generalizes well for language identification after fine-tuning on unseen languages. However, the performance significantly degrades when the languages are not very distinct from each other, for example, in the case of dialects. Low resource dialect classification remains a challenging problem to solve. We present a new data augmentation method that leverages model training dynamics of individual data points to improve sampling for latent mixup. The method works well in low-resource settings where generalization is paramount. Our datamaps-based mixup technique, which we call Map-Mix improves weighted F1 scores by 2% compared to the random mixup baseline and results in a significantly well-calibrated model. The code for our method is open sourced on https://github.com/skit-ai/Map-Mix.

PaperPDFCode

Code

skit-ai/map-mix officialmentioned in papermentioned on GitHub report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Data AugmentationLanguage IdentificationSpoken language identification

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

MixupXLSR

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections