Can Multimodal Large Language Models Understand OCT?
This paper creates a benchmark to evaluate how well large language models can understand images from optical coherence tomography (OCT), a tool used to diagnose retinal diseases. Practitioners in medical imaging and AI might care because it could help improve the accuracy of AI models in diagnosing and treating retinal diseases.