A Multimodal Biomedical Foundation Model Trained from Fifteen Million Image–Text PairsSheng Zhang, Hoifung Poon, Jaspreet Bagga et al.|NEJM AI|2024Cited by 218
A foundation model for joint segmentation, detection and recognition of biomedical objects across nine modalitiesTheodore Zhao, Sheng Wang, 裕二 池谷 et al.|Nature Methods|2024Cited by 101