Anatomy Contextualized Adaption of CT Foundation Models
By Roshan Kenia · Paper · cs.CV
CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume representations that dilute fine-grained anatomical signals. Fine-grained vision-language pre-training addresses this by aligning anat