ClinFusion: A Vision-Centric Multimodal LLM System for Holistic Medical Understanding

By Hangjie Yuan · Paper · cs.CV

Multimodal large language models (MLLMs) hold immense potential to revolutionize clinical practice, yet deploying them in the medical domain is fundamentally a vision-centric challenge: models must absorb knowledge from heterogeneous 2D and 3D medical images, and evaluation proto

Cs.cv

View original

HomeResourceLoading…