Earlier quoted context omitted.
I can take a bet that it haha already failed - the hype cycle has already made a promise that LLMs can’t keep. Hallucinations to the normal person are a bug. The issue is that only humans can hallucinate. We know there is a “reality”. For an LLM, everything it does is a hallucination. That’s why you have more POCs than production goods. Your “hallucination rate” is unknown. Yesterday Ars has an article that described…
Why not "simply" multigen every (important) query and take the statistical average? Hallucinations are random, the truth isn't. This is absurdly expensive with GPT4, cheaper with 3, and dirt cheap locally with LLaMA
Truth is a human thing. Statistically averaging out 4, 5, 6, N text generations from an LLM will not converge to any “truth”.
You have essentially stated that outputs from a text generator are normally distributed around “Facts”.
May I gently suggest, that and older quote about an infinite number of simians, typewriters and the works of Shakespeare, is more appropriate ?