中文 | English
Return

Do General-Purpose Multimodal Large Language Models Really See Radiologic Images or Rely on Text?