Abstract Large language models (LLMs) have demonstrated impressive capabilities, but the bar for clinical applications is high. Attempts to assess the clinical knowledge of models typically rely on automated evaluations based on limited benchmarks. Here, to address these limitations, we present Mult...
Research Assistant
AI chat, annotations, notes & similar papers
No comments yet
Be the first to share your thoughts!