arxiv
PublishedJuly 21, 2026 at 4:00 AM
—neutral
Can Multimodal Large Language Models Understand OCT?
Publisher summary· verbatim
arXiv:2607.16609v1 Announce Type: cross Abstract: Optical coherence tomography (OCT) imaging is essential for the diagnosis and treatment of retinal diseases. Although multimodal large language models (MLLMs) have demonstrated considerable potential in medical image analysis, existing benchmarks lar
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivShapley Context Pruning: A Cooperative Game Perspective for Context Reranking and Pruning3harxivA Survey on the Verification of Reinforcement Learning Policies3harxivSymbolic Augmentation Closes a Canonical-Equivalence Blind Spot in Neural Fact-Checkers3harxivSelKV: Selective KV Cache Merging with Per-Token Merge-or-Drop and Attention Compensation3hThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗