arXiv cs.AI
· Papers
MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents
arXiv:2607.02879v1 Announce Type: new Abstract: Current benchmarks for evaluating large language models (LLMs) in medical calculation are largely based on simplified settings, where each patient case corresponds to a single calculator and the required tool is explicitly specified in the query. However, real clinical sc