J. Biomed. Inform. 2026 – Multimodal AI in healthcare: Review of vision-language foundation models for real-world medical applications
Taha Razzaq, Murtaza Taj, Asim Iqbal Abstract: \initial{T}\textbf{he emergence of foundation models has marked a transformative shift in AI, enabling robust generalization across diverse downstream tasks through putative zero-shot learning. Large Language Models and Vision-Language Models have demonstrated strong capabilities in tasks such as image interpretation, report generation, and question answering by effectively learning from…
