Please use this identifier to cite or link to this item:
http://hdl.handle.net/10397/120228
| DC Field | Value | Language |
|---|---|---|
| dc.contributor | Department of Electrical and Electronic Engineering | - |
| dc.creator | Jin, Z | - |
| dc.creator | Tu, Y | - |
| dc.creator | Li, Z | - |
| dc.creator | Huang, Z | - |
| dc.creator | Gan, CX | - |
| dc.creator | Mak, MW | - |
| dc.date.accessioned | 2026-07-28T00:26:28Z | - |
| dc.date.available | 2026-07-28T00:26:28Z | - |
| dc.identifier.isbn | 979-8-3503-6874-1 (Electronic) | - |
| dc.identifier.isbn | 979-8-3503-6875-8 (Print on Demand(PoD)) | - |
| dc.identifier.uri | http://hdl.handle.net/10397/120228 | - |
| dc.description | ICASSP 2025 - 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 6-11 April 2025, Hyderabad, India | en_US |
| dc.language.iso | en | en_US |
| dc.publisher | Institute of Electrical and Electronics Engineers | en_US |
| dc.rights | © 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works. | en_US |
| dc.rights | The following publication Z. Jin, Y. Tu, Z. Li, Z. Huang, C. -X. Gan and M. -W. Mak, "Denoising Student Features with Diffusion Models for Knowledge Distillation in Speaker Verification," ICASSP 2025 - 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Hyderabad, India, 2025, pp. 1-5 is available at https://doi.org/10.1109/ICASSP49660.2025.10889980. | en_US |
| dc.subject | Diffusion models | en_US |
| dc.subject | Knowledge distillation | en_US |
| dc.subject | Pre-trained speech models | en_US |
| dc.subject | Short-utterance | en_US |
| dc.subject | Speaker verification | en_US |
| dc.title | Denoising student features with diffusion models for knowledge distillation in speaker verification | en_US |
| dc.type | Conference Paper | en_US |
| dc.identifier.doi | 10.1109/ICASSP49660.2025.10889980 | - |
| dcterms.abstract | In recent years, there has been a surge in the use of a pre-trained speech model as a feature extractor for speaker verification (SV). To reduce model complexity, researchers transfer knowledge from a pre-trained model to a lightweight student model, enabling the latter to reach a performance level not attainable by conventional methods. However, due to the differences in model capacity, the student features contain more noise. This results in discrepancies between the teacher and student features at the intermediate layers, negatively impacting feature-level knowledge distillation (KD). To address this issue, we employ a diffusion model to denoise the student features for KD (DenoKD). This approach enables more effective feature-level distillation. Our method, trained with a small ECAPA-TDNN, achieved a 13% improvement over the baseline on the VoxCeleb1-O test set. Further more, the DenoKD mechanism is found to be effective for SV on short test utterances. | - |
| dcterms.accessRights | open access | en_US |
| dcterms.bibliographicCitation | In 2025 IEEE International Conference on Acoustics, Speech, and Signal Processing: Conference proceedings, https://doi.org/10.1109/ICASSP49660.2025.10889980 | - |
| dcterms.issued | 2025 | - |
| dc.identifier.scopus | 2-s2.0-105009600645 | - |
| dc.relation.ispartofbook | 2025 IEEE International Conference on Acoustics, Speech, and Signal Processing: Conference proceedings | - |
| dc.relation.conference | International Conference on Acoustics, Speech and Signal Processing [ICASSP] | - |
| dc.publisher.place | Piscataway, NJ | en_US |
| dc.description.validate | 202607 bcch | - |
| dc.description.oa | Accepted Manuscript | en_US |
| dc.identifier.FolderNumber | a4740 | en_US |
| dc.identifier.SubFormID | 53833 | en_US |
| dc.description.fundingSource | RGC | en_US |
| dc.description.pubStatus | Published | en_US |
| dc.description.oaCategory | Green (AAM) | en_US |
| Appears in Collections: | Conference Paper | |
Files in This Item:
| File | Description | Size | Format | |
|---|---|---|---|---|
| Jin_Denoising_Student_Features.pdf | Pre-Published version | 2.74 MB | Adobe PDF | View/Open |
Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.



