Improving Medical Calculation of LLMs with Embedded Coding
arXiv:2609.31908v1 Announce Type: new Abstract: Large Language Models (LLMs) perform well on medical examinations and question-answering benchmarks, but remain unreliable on medical calculation tasks that require exact numerical outputs. These calculations support high-stakes decisions such as medi…
Read the full story at arXiv cs.AI ↗
Timeline · 1 report
- 2026-09-29 04:00 · arXiv cs.AI
Improving Medical Calculation of LLMs with Embedded Coding