Abstract
Word associations are commonly applied in psycholinguistics to investigate the nature and structure of the human mental lexicon, and at the same time an important data source for measuring the alignment of language models with human semantic representations.Taking this view, we compare the capacities of different language models to model collective human association norms via five word association tasks (WATs), with predictions about associations driven by either word vector similarities for traditional embedding models or prompting large language models (LLMs).Our results demonstrate that neither approach could produce human-like performances in all five WATs. Hence, none of them can successfully model the human mental lexicon yet. Our detailed analysis shows that static word-type embeddings and prompted LLMs have overall better alignment with human norms compared to word-token embeddings from pretrained models like BERT. Further analysis suggests that the performance discrepancies may be due to different model architectures, especially in terms of approximating human-like associative reasoning through either semantic similarity or relatedness evaluation. Our codes and data are publicly available at: https://github.com/florethsong/word_association.
| Original language | English |
|---|---|
| Title of host publication | Proceedings of the International Conference on Computational Semantics (IWCS 2025) |
| Editors | Kilian Evang, Laura Kallmeyer, Sylvain Pogodalla |
| Publisher | Association for Computational Linguistics |
| Pages | 208-230 |
| ISBN (Electronic) | 9798891763166 |
| Publication status | Published - Sept 2025 |
| Event | International Conference on Computational Semantics (IWCS 2025) - Heinrich Heine University, Dusseldorf, Germany Duration: 22 Sept 2025 → 24 Sept 2025 https://iwcs2025.github.io/ |
Conference
| Conference | International Conference on Computational Semantics (IWCS 2025) |
|---|---|
| Country/Territory | Germany |
| City | Dusseldorf |
| Period | 22/09/25 → 24/09/25 |
| Internet address |
Fingerprint
Dive into the research topics of 'Which Model Mimics Human Mental Lexicon Better? A Comparative Study of Word Embedding and Generative Models'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver