Earlier quoted context omitted.
I feel like you didn't understand the goal of this study > The DTN-UK stated earlier this year that generic LLMs must never be used as autonomous advisory calculators for insulin delivery. This data is the quantitative evidence base for that statement. This study is to prove that you should not rely on LLMs
The thing is it doesn't really prove LLMs can't do this, it proves no existing frontier LLMs can do this. The part where they talk about sampling multiple runs is interesting - it suggests to me that in the next few years as the reasoning process is improved the models may be able to do that autonomously. My mind really is going to using a dedicated object detection models fine-tuned with nutrition information, but I…
That has nothing to do with the question being asked, can you rely on an LLM today to help you track carbs as a diabetic?
This is very explicitly what the article is all about. Potential future LLMs are entirely irrelevant.