A collective capability boundary in frontier large language models on guideline-conformant and case-specific oncology decision-making
Large language models (LLMs) achieve high scores on medical knowledge examinations, yet real-world oncology is not a knowledge test--it is a sequence …
Source:arXiv cs.AI
Read original ↗