MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents
arXiv:2607.02879v1 Announce Type: new Abstract: Current benchmarks for evaluating large language models (LLMs) in medical calculation are largely based on simplified settings, where each patient case corresponds to a single calculator and the required tool is explicitly specified in the query. However, real clinical scenarios often require multiple calculators for joint evaluation, nested-scale calculation, and fuzzy queries that do not directly specify the target calculator. To this end, we pro...
arXiv cs.AI
·Siran Zhao, Ruihui Hou, Ziyue Huai, Chennuo Zhang, Tong Ruan
·