Please use this identifier to cite or link to this item:
http://hdl.handle.net/10397/120077
| DC Field | Value | Language |
|---|---|---|
| dc.contributor | Department of Language Science and Technology | en_US |
| dc.creator | Wang, Y | en_US |
| dc.creator | Chersoni, E | en_US |
| dc.creator | Huang, CR | en_US |
| dc.date.accessioned | 2026-07-22T03:54:22Z | - |
| dc.date.available | 2026-07-22T03:54:22Z | - |
| dc.identifier.isbn | 978-2-493814-49-4 | en_US |
| dc.identifier.uri | http://hdl.handle.net/10397/120077 | - |
| dc.description | The Fifteenth Language Resources and Evaluation Conference (LREC 2026), Palma, Mallorca, Spain, 11 - 16 May 2026 | en_US |
| dc.language.iso | en | en_US |
| dc.publisher | European Language Resources Association (ELRA) | en_US |
| dc.rights | ©ELRA Language Resources Association (ELRA), 2026 | en_US |
| dc.rights | Licensed under the Creative Commons Attribution 4.0 International License (https://creativecommons.org/licenses/by/4.0/) | en_US |
| dc.rights | The following publication Wang, Y., Chersoni, E., & Huang, C. (2026). This One or That One? A Study on Accessibility via Demonstratives with Multimodal Large Language Models. In Proceedings of the Fifteenth Language Resources and Evaluation Conference (LREC 2026) (pp. 9722–9732). European Language Resources Association (ELRA) is available at https://doi.org/10.63317/29f29zththay. | en_US |
| dc.subject | Accessibility | en_US |
| dc.subject | Cognitive evaluation | en_US |
| dc.subject | Demonstratives | en_US |
| dc.subject | Large language models | en_US |
| dc.title | This one or that one? A study on accessibility via demonstratives with multimodal large language models | en_US |
| dc.type | Conference Paper | en_US |
| dc.identifier.spage | 9722 | en_US |
| dc.identifier.epage | 9732 | en_US |
| dc.identifier.doi | 10.63317/29f29zththay | en_US |
| dcterms.abstract | Accessibility refers to the ease with which a speaker can acquire an object, and it is often conveyed through demonstrative pronouns like "this" and "that", indicating proximal or distal objects. Most importantly, accessibility also involves perspective shifts, which are essential for understanding differing viewpoints. In this case study, we adopt an evaluation dataset with a pair-to-pair question structure for referent identification based on demonstratives. Our experiments show that current Multimodal Large Language Models (MLLMs) exhibit markedly low performance in accessibility tasks requiring perspective shifts, with accuracies around 2.33% (Chinese) and 1.83% (English). Moreover, models struggle with qualitative characteristics and frame-based reasoning, often failing to apply implicit contextual rules unless explicitly encoded in training data. These limitations suggest that MLLMs rely heavily on surface co-occurrence instead of truly grounded, embodied experience. Our evaluation framework provides a robust lens revealing that MLLMs lack both self-other distinction—an essential aspect of self-awareness—and the embodied cognition necessary for reliable performance in practical embodied AI applications. | en_US |
| dcterms.accessRights | open access | en_US |
| dcterms.bibliographicCitation | The Fifteenth Language Resources and Evaluation Conference, LREC 2026, Palma, Mallorca, Spain, May 11-16 2026, https://doi.org/10.63317/29f29zththay | en_US |
| dcterms.issued | 2026 | - |
| dc.relation.ispartofbook | Proceedings of the Fifteenth Language Resources and Evaluation Conference (LREC 2026) | en_US |
| dc.relation.conference | Language Resources and Evaluation Conference [LREC] | en_US |
| dc.description.validate | 202607 bcwc | en_US |
| dc.description.oa | Version of Record | en_US |
| dc.identifier.FolderNumber | a4672 | - |
| dc.identifier.SubFormID | 53569 | - |
| dc.description.fundingSource | Others | en_US |
| dc.description.fundingText | EC acknowledges the financial support from the start-up fund project “Building and Predicting Neurocognitive-Motivated Lexical Semantic Norms for Mandarin Chinese” (1-BE8G), sponsored by the Faculty of Humanities of the Hong Kong Polytechnic University. | en_US |
| dc.description.pubStatus | Published | en_US |
| dc.description.oaCategory | CC | en_US |
| Appears in Collections: | Conference Paper | |
Files in This Item:
| File | Description | Size | Format | |
|---|---|---|---|---|
| 2026.lrec2026-1.763.pdf | 5.73 MB | Adobe PDF | View/Open |
Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.



