Search papers, labs, and topics across Lattice.
2
0
5
Vision-language models struggle to leverage visual evidence in medical VQA, with only one model surpassing human performance on a subset of questions.
LLMs can appear competent in medical contexts, but a rigorous evaluation reveals a staggering drop in performance that questions their true clinical abilities.