Fetching the paper…

Benchmarking Large Language Models on Answering and Explaining Challenging Medical Questions · Around