UpToDate Expert AI
Coverage of UpToDate Expert AI in the Nexus archive.
- STAT+: Why benchmarking clinical LLMs from OpenEvidence, Doximity is complicated
A study in Nature Medicine compared clinical AI systems OpenEvidence and UpToDate Expert AI against general LLMs, sparking debate in the clinical AI community. The article discusses challenges in benchmarking these systems, emphasizing that headline-driven summaries often oversimplify complex results.
- STAT+: Clinical chatbots are taking medicine by storm. Should doctors trust them?
Clinical large language models developed by companies like OpenEvidence, Doximity, and UpToDate are widely used by U.S. doctors, but a June 2024 study in Nature Medicine found these models performed worse than generalist AI models on clinical questions. The findings sparked significant debate within the health AI community, with some interpreting the results as a 'clear victory for general frontier models.'