Please use this identifier to cite or link to this item: http://hdl.handle.net/10397/119978
PIRA download icon_1.1View/Download Full Text
Title: Which model mimics human mental lexicon better? A comparative study of word embedding and generative models
Authors: Song, H 
Feng, Z 
Chersoni, E 
Huang, CR 
Issue Date: 2025
Source: In Proceedings of the 16th International Conference on Computational Semantics, p. 208-230. Kerrville : Association for Computational Linguistics, 2025
Abstract: Word associations are commonly applied in psycholinguistics to investigate the nature and structure of the human mental lexicon, and at the same time an important data source for measuring the alignment of language models with human semantic representations.Taking this view, we compare the capacities of different language models to model collective human association norms via five word association tasks (WATs), with predictions about associations driven by either word vector similarities for traditional embedding models or prompting large language models (LLMs).Our results demonstrate that neither approach could produce human-like performances in all five WATs. Hence, none of them can successfully model the human mental lexicon yet. Our detailed analysis shows that static word-type embeddings and prompted LLMs have overall better alignment with human norms compared to word-token embeddings from pretrained models like BERT. Further analysis suggests that the performance discrepancies may be due to different model architectures, especially in terms of approximating human-like associative reasoning through either semantic similarity or relatedness evaluation. Our codes and data are publicly available at: https://github.com/florethsong/word_association.
Publisher: Association for Computational Linguistics
ISBN: 979-8-89176-316-6
Description: 16th International Conference on Computational Semantics, Düsseldorf, Germany, September 22-23, 2025
Rights: ©2025 Association for Computational Linguistics
Licensed under the Creative Commons Attribution 4.0 International License (https://creativecommons.org/licenses/by/4.0/)
The following publication Huacheng Song, Zhaoxin Feng, Emmanuele Chersoni, and Chu-Ren Huang. 2025. Which Model Mimics Human Mental Lexicon Better? A Comparative Study of Word Embedding and Generative Models. In Proceedings of the 16th International Conference on Computational Semantics, pages 208–230, Düsseldorf, Germany. Association for Computational Linguistics is available at https://aclanthology.org/2025.iwcs-main.19/.
Appears in Collections:Conference Paper

Files in This Item:
File Description SizeFormat 
2025.iwcs-main.19.pdf3.92 MBAdobe PDFView/Open
Open Access Information
Status open access
File Version Version of Record
Access
View full-text via PolyU eLinks SFX Query
Show full item record

Google ScholarTM

Check


Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.