Skip to content
arXiv cs.AI · Papers

MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents

arXiv:2607.02879v1 Announce Type: new Abstract: Current benchmarks for evaluating large language models (LLMs) in medical calculation are largely based on simplified settings, where each patient case corresponds to a single calculator and the required tool is explicitly specified in the query. However, real clinical sc