ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding
By Hangjie Yuan · Paper · cs.CV
Multimodal large language models (MLLMs) hold immense potential to revolutionize clinical practice, yet deploying them in the medical domain is fundamentally a vision-centric challenge: models must absorb knowledge from heterogeneous 2D and 3D medical images, and evaluation proto