Please use this identifier to cite or link to this item:
http://hdl.handle.net/10397/120074
| DC Field | Value | Language |
|---|---|---|
| dc.contributor | Department of Computing | en_US |
| dc.contributor | Department of Land Surveying and Geospatial Science | en_US |
| dc.creator | Chen, Y | en_US |
| dc.creator | Li, M | en_US |
| dc.creator | Rao, Z | en_US |
| dc.creator | Zeng, D | en_US |
| dc.creator | Guo, S | en_US |
| dc.creator | Guo, J | en_US |
| dc.date.accessioned | 2026-07-22T03:54:20Z | - |
| dc.date.available | 2026-07-22T03:54:20Z | - |
| dc.identifier.uri | http://hdl.handle.net/10397/120074 | - |
| dc.language.iso | en | en_US |
| dc.title | Learning by neighbor-aware semantics, deciding by open-form flows : towards robust zero-shot skeleton action recognition | en_US |
| dc.type | Conference Paper | en_US |
| dc.identifier.spage | 3374 | en_US |
| dc.identifier.epage | 3383 | en_US |
| dcterms.abstract | Recognizing unseen skeleton action categories remains highly challenging due to the absence of corresponding skeletal priors. Existing approaches generally follow an “align-then-classify” paradigm but face two fundamental issues: (i) fragile point-to-point alignment arising from imperfect semantics, and (ii) rigid classifiers restricted by static decision boundaries and coarse-grained anchors. To address these issues, we propose a novel method for zero-shot skeleton action recognition, termed Flora, which builds upon FlexibLe neighbOr-aware semantic attunement and open-form distRibution-aware flow clAssifier. Specifically, we flexibly attune textual semantics by incorporating neighboring inter-class contextual cues to form direction-aware regional semantics, coupled with a cross-modal geometric consistency objective that ensures stable and robust point-to-region alignment. Furthermore, we employ noise-free flow matching to bridge the modality distribution gap between semantic and skeleton latent embeddings, while a condition-free contrastive regularization enhances discriminability, leading to a distribution-aware classifier with fine-grained decision boundaries achieved through token-level velocity predictions. Extensive experiments on three benchmark datasets validate the effectiveness of our method, showing particularly impressive performance even when trained with only 10% of the seen data. Code is available at https://github.com/cseeyangchen/Flora. | en_US |
| dcterms.accessRights | embargoed access | en_US |
| dcterms.bibliographicCitation | The IEEE/CVF Conference on Computer Vision and Pattern Recognition 2026, June 3 - June 7, 2026, Colorado Convention Center, p. 3374-3383 | en_US |
| dcterms.issued | 2026 | - |
| dc.description.validate | 202607 bcwc | en_US |
| dc.description.oa | Not applicable | en_US |
| dc.identifier.FolderNumber | a4666 | - |
| dc.identifier.SubFormID | 53538 | - |
| dc.description.fundingSource | RGC | en_US |
| dc.description.fundingSource | Others | en_US |
| dc.description.fundingText | This research was supported by the Hong Kong RGC General Research Fund (Grant Nos. 15221123, 15216424, and 15211525) and the Hong Kong PolyU Internal Research Fund (Grant Nos. P0058468 and P0056171). | en_US |
| dc.description.pubStatus | Early release | en_US |
| dc.date.embargo | 0000-00-00 (to be updated) | en_US |
| dc.description.oaCategory | Green (AAM) | en_US |
| Appears in Collections: | Conference Paper | |
Google ScholarTM
Check
Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.


