Please use this identifier to cite or link to this item:
http://hdl.handle.net/10397/120230
| DC Field | Value | Language |
|---|---|---|
| dc.contributor | Department of Electrical and Electronic Engineering | - |
| dc.creator | Li, J | - |
| dc.creator | Mak, MW | - |
| dc.creator | Rohdin, J | - |
| dc.creator | Lee, KA | - |
| dc.creator | Hermansky, H | - |
| dc.date.accessioned | 2026-07-28T00:26:31Z | - |
| dc.date.available | 2026-07-28T00:26:31Z | - |
| dc.identifier.uri | http://hdl.handle.net/10397/120230 | - |
| dc.description | 26th edition of the Interspeech Conference, August 17-21, 2025, Rotterdam, The Netherlands | en_US |
| dc.language.iso | en | en_US |
| dc.publisher | International Speech Communication Association | en_US |
| dc.rights | The following publication Li, J., Mak, M.-W., Rohdin, J., Lee, K.A., Hermansky, H. (2025) Bayesian Learning for Domain-Invariant Speaker Verification and Anti-Spoofing. Proc. Interspeech 2025, 1123-1127 is available at https://doi.org/10.21437/Interspeech.2025-655. | en_US |
| dc.subject | Anti-spoofing | en_US |
| dc.subject | Bayesian learning | en_US |
| dc.subject | Domain generalization | en_US |
| dc.subject | Speaker verification | en_US |
| dc.title | Bayesian learning for domain-invariant speaker verification and anti-spoofing | en_US |
| dc.type | Conference Paper | en_US |
| dc.identifier.spage | 1123 | - |
| dc.identifier.epage | 1127 | - |
| dc.identifier.doi | 10.21437/Interspeech.2025-655 | - |
| dcterms.abstract | The performance of automatic speaker verification (ASV) and anti-spoofing drops seriously under real-world domain mismatch conditions. The relaxed instance frequency-wise normalization (RFN), which normalizes the frequency components based on the feature statistics along the time and channel axes, is a promising approach to reducing the domain dependence in the feature maps of a speaker embedding network. We advocate that the different frequencies should receive different weights and that the weights' uncertainty due to domain shift should be accounted for. To these ends, we propose leveraging variational inference to model the posterior distribution of the weights, which results in Bayesian weighted RFN (BWRFN). This approach overcomes the limitations of fixed-weight RFN, making it more effective under domain mismatch conditions. Extensive experiments on cross-dataset ASV, cross-TTS anti-spoofing, and spoofing-robust ASV show that BWRFN is significantly better than WRFN and RFN. | - |
| dcterms.accessRights | open access | en_US |
| dcterms.bibliographicCitation | In 26th edition of the Interspeech Conference, to be held August 17-21, 2025, in Rotterdam, The Netherlands, p. 1123-1127 | - |
| dcterms.issued | 2025 | - |
| dc.identifier.scopus | 2-s2.0-105020044480 | - |
| dc.relation.ispartofbook | 26th edition of the Interspeech Conference, to be held August 17-21, 2025, in Rotterdam, The Netherlands | - |
| dc.relation.conference | Conference of the International Speech Communication Association [INTERSPEECH] | - |
| dc.description.validate | 202607 bcch | - |
| dc.description.oa | Version of Record | en_US |
| dc.identifier.FolderNumber | a4741b | en_US |
| dc.identifier.SubFormID | 53836 | en_US |
| dc.description.fundingSource | RGC | en_US |
| dc.description.pubStatus | Published | en_US |
| dc.description.oaCategory | VoR allowed | en_US |
| Appears in Collections: | Conference Paper | |
Files in This Item:
| File | Description | Size | Format | |
|---|---|---|---|---|
| li25h_interspeech.pdf | 306.35 kB | Adobe PDF | View/Open |
Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.



