Search papers, labs, and topics across Lattice.
Northeastern University
2
0
3
3
Even top-performing LLMs struggle to maintain accuracy in correcting medical misconceptions, with performance plummeting from 85% to 50% over just two follow-up questions.
Even the best LLMs struggle with multi-turn medical dialogues, with error rates tripling by the third turn and a single wrong answer significantly increasing the probability of subsequent errors.