Please use this identifier to cite or link to this item:
http://hdl.handle.net/10397/120233
| DC Field | Value | Language |
|---|---|---|
| dc.contributor | Department of Electrical and Electronic Engineering | - |
| dc.creator | Gan, CX | - |
| dc.creator | Tu, Y | - |
| dc.creator | Jin, Z | - |
| dc.creator | Mak, MW | - |
| dc.creator | Lee, KA | - |
| dc.date.accessioned | 2026-07-28T00:26:35Z | - |
| dc.date.available | 2026-07-28T00:26:35Z | - |
| dc.identifier.isbn | 979-8-3503-6874-1 (Electronic) | - |
| dc.identifier.isbn | 979-8-3503-6875-8 (Print on Demand(PoD)) | - |
| dc.identifier.uri | http://hdl.handle.net/10397/120233 | - |
| dc.description | ICASSP 2025 - 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 6-11 April 2025, Hyderabad, India | en_US |
| dc.language.iso | en | en_US |
| dc.publisher | Institute of Electrical and Electronics Engineers | en_US |
| dc.rights | © 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works. | en_US |
| dc.rights | The following publication C. -X. Gan, Y. Tu, Z. Jin, M. -W. Mak and K. A. Lee, "Grouped Knowledge Distillation with Adaptive Logit Softening for Speaker Recognition," ICASSP 2025 - 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Hyderabad, India, 2025, pp. 1-5 is available at https://doi.org/10.1109/ICASSP49660.2025.10887814. | en_US |
| dc.subject | Adaptive logit softening | en_US |
| dc.subject | Grouped knowledge transfer | en_US |
| dc.subject | Knowledge distillation | en_US |
| dc.subject | Speaker recognition | en_US |
| dc.title | Grouped knowledge distillation with adaptive logit softening for speaker recognition | en_US |
| dc.type | Conference Paper | en_US |
| dc.identifier.doi | 10.1109/ICASSP49660.2025.10887814 | - |
| dcterms.abstract | Recent works suggest that decoupling the information of non-target speakers from that of the target speaker in knowledge distillation (KD) and subsequently emphasizing the former can lead to significant performance improvement. However, a well-trained teacher model typically produces almost zero non-target speaker posteriors with limited contribution to knowledge transfer, resulting in a less effective KD. To address this problem, we advocate a dual-group knowledge distillation framework, wherein the primary group with top-k speaker posteriors captures most of the speaker discrimination knowledge in an utterance. The non-primary group contributes to the KD through a binary classification (distillation) between the primary and non-primary groups. In addition, adaptive logit softening is proposed to adjust the teacher’s and student’s logits in the binary distillation, further facilitating effective knowledge transfer. The proposed method trained with a simple x-vector pipeline obtains an impressive equal error rate of 1.46%, 1.47%, and 2.70% on three VoxCeleb1 test sets, outperforming the state-of-the-art methods with a noticeable margin. | - |
| dcterms.accessRights | open access | en_US |
| dcterms.bibliographicCitation | In 2025 IEEE International Conference on Acoustics, Speech, and Signal Processing: Conference proceedings, https://doi.org/10.1109/ICASSP49660.2025.10887814 | - |
| dcterms.issued | 2025 | - |
| dc.identifier.scopus | 2-s2.0-105003871932 | - |
| dc.relation.ispartofbook | 2025 IEEE International Conference on Acoustics, Speech, and Signal Processing: Conference proceedings | - |
| dc.relation.conference | International Conference on Acoustics, Speech and Signal Processing [ICASSP] | - |
| dc.publisher.place | Piscataway, NJ | en_US |
| dc.description.validate | 202607 bcch | - |
| dc.description.oa | Accepted Manuscript | en_US |
| dc.identifier.FolderNumber | a4743 | en_US |
| dc.identifier.SubFormID | 53841 | en_US |
| dc.description.fundingSource | RGC | en_US |
| dc.description.pubStatus | Published | en_US |
| dc.description.oaCategory | Green (AAM) | en_US |
| Appears in Collections: | Conference Paper | |
Files in This Item:
| File | Description | Size | Format | |
|---|---|---|---|---|
| Gan_Grouped_Knowledge_Distillation.pdf | Pre-Published version | 1.09 MB | Adobe PDF | View/Open |
Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.



