ICGA-GPT : report generation and question answering for indocyanine green angiography images

Chen, X; Zhang, W; Zhao, Z; Xu, P; Zheng, Y; Shi, D; He, M

doi:10.1136/bjo-2023-324446

Please use this identifier to cite or link to this item: http://hdl.handle.net/10397/104987

DC Field	Value	Language
dc.contributor	School of Optometry	en_US
dc.contributor	Research Centre for SHARP Vision	en_US
dc.creator	Chen, X	en_US
dc.creator	Zhang, W	en_US
dc.creator	Zhao, Z	en_US
dc.creator	Xu, P	en_US
dc.creator	Zheng, Y	en_US
dc.creator	Shi, D	en_US
dc.creator	He, M	en_US
dc.date.accessioned	2024-03-26T06:11:44Z	-
dc.date.available	2024-03-26T06:11:44Z	-
dc.identifier.issn	0007-1161	en_US
dc.identifier.uri	http://hdl.handle.net/10397/104987	-
dc.language.iso	en	en_US
dc.publisher	BMJ Publishing Group	en_US
dc.rights	© Author(s) (or their employer(s)) 2024. No commercial re-use. See rights and permissions. Published by BMJ.	en_US
dc.rights	This article has been accepted for publication in British journal of ophthalmology, 2024 following peer review, and the Version of Record can be accessed online at https://doi.org/10.1136/bjo-2023-324446.	en_US
dc.title	ICGA-GPT : report generation and question answering for indocyanine green angiography images	en_US
dc.type	Journal/Magazine Article	en_US
dc.identifier.doi	10.1136/bjo-2023-324446	en_US
dcterms.abstract	Background: Indocyanine green angiography (ICGA) is vital for diagnosing chorioretinal diseases, but its interpretation and patient communication require extensive expertise and time-consuming efforts. We aim to develop a bilingual ICGA report generation and question-answering (QA) system.	en_US
dcterms.abstract	Methods: Our dataset comprised 213 129 ICGA images from 2919 participants. The system comprised two stages: image–text alignment for report generation by a multimodal transformer architecture, and large language model (LLM)-based QA with ICGA text reports and human-input questions. Performance was assessed using both qualitative metrics (including Bilingual Evaluation Understudy (BLEU), Consensus-based Image Description Evaluation (CIDEr), Recall-Oriented Understudy for Gisting Evaluation-Longest Common Subsequence (ROUGE-L), Semantic Propositional Image Caption Evaluation (SPICE), accuracy, sensitivity, specificity, precision and F1 score) and subjective evaluation by three experienced ophthalmologists using 5-point scales (5 refers to high quality).	en_US
dcterms.abstract	Results: We produced 8757 ICGA reports covering 39 disease-related conditions after bilingual translation (66.7% English, 33.3% Chinese). The ICGA-GPT model’s report generation performance was evaluated with BLEU scores (1–4) of 0.48, 0.44, 0.40 and 0.37; CIDEr of 0.82; ROUGE of 0.41 and SPICE of 0.18. For disease-based metrics, the average specificity, accuracy, precision, sensitivity and F1 score were 0.98, 0.94, 0.70, 0.68 and 0.64, respectively. Assessing the quality of 50 images (100 reports), three ophthalmologists achieved substantial agreement (kappa=0.723 for completeness, kappa=0.738 for accuracy), yielding scores from 3.20 to 3.55. In an interactive QA scenario involving 100 generated answers, the ophthalmologists provided scores of 4.24, 4.22 and 4.10, displaying good consistency (kappa=0.779).	en_US
dcterms.abstract	Conclusion: This pioneering study introduces the ICGA-GPT model for report generation and interactive QA for the first time, underscoring the potential of LLMs in assisting with automated ICGA image interpretation.	en_US
dcterms.accessRights	open access	en_US
dcterms.bibliographicCitation	British journal of ophthalmology, First published March 20, 2024, Online First, https://doi.org/10.1136/bjo-2023-324446	en_US
dcterms.isPartOf	British journal of ophthalmology	en_US
dcterms.issued	2024	-
dc.identifier.eissn	1468-2079	en_US
dc.description.validate	202403 bcch	en_US
dc.description.oa	Accepted Manuscript	en_US
dc.identifier.FolderNumber	a2661	-
dc.identifier.SubFormID	48031	-
dc.description.fundingSource	Others	en_US
dc.description.fundingText	Start-up Fund for RAPs under the Strategic Hiring Scheme	en_US
dc.description.fundingText	Global STEM Professorship Scheme	en_US
dc.description.pubStatus	Published	en_US
dc.description.oaCategory	Green (AAM)	en_US
Appears in Collections:	Journal/Magazine Article

Files in This Item:

File	Description	Size	Format
Chen_ICGA-GPT_Report_Generation.pdf	Pre-Published version	9.65 MB	Adobe PDF	View/Open

Open Access Information

Status	open access
File Version	Final Accepted Manuscript

Access

View full-text via PolyU eLinks

Show simple item record

Page views

38

Citations as of May 5, 2024

Downloads

17

Citations as of May 5, 2024

Google Scholar^TM

Check

Files in This Item:

Open Access Information

Access

Page views

Downloads

Google ScholarTM

Altmetric

Google Scholar^TM