I was so optimistic about using LLMs for "write once, read many" English language documents, but the more I've used the tools, the more pessimistic I get. More and more, I try to ask it for low prose responses because its writing just seems like such a low signal to noise ratio I'm curious about why LLM writing fails. Particularly whether LLM writing is fundamentally flawed, or if it's just distinctive and since it o…
This is ok in domains if you can train against known good answers and make sure the machine generates conforming text most of the time. It falls apart in fuzzier domains where training is much harder and intent is required (i.e. having something to say).
LLM writing is generally ok in factual domains where it can regurgitate bits of wikipedia or answers to questions, they are terrible at long form writing, in particularly in literary styles, because of a lack of intelligence and taste.
I don't think the answer lies in the data or in their training. It seems we've had a few years for this problem to be solved, but nobody seems to have worked out an answer to it.