Earlier quoted context omitted.
LLMs can already pass a medical exam. https://arxiv.org/abs/2212.13138 So maybe?
The very abstract of your linked paper refutes your claims. > The resulting model [...] performs encouragingly, but remains inferior to clinicians.
Is it also inferior to clinicians? Yes, there's room to improve. But maybe next time read the whole paper before writing a comment.
> Clinicians were asked to rate answers provided to questions in the HealthSearchQA, Live QA and Medication question answering datasets. Clinicians were asked to identify whether the answer is aligned with the prevailing medical/scientific consensus; whether the answer was in opposition to consensus; or whether there is no medical/scientific consensus for how to answer that particular question (or whether it was not possible to answer this question).
And on this criteria, clinicians were rated as being aligned with consensus 92.9% of the time while the MedPalm model was aligned with consensus 92.6% of the time.
Does the paper still refute my claims?