Earlier quoted context omitted.
It's so interesting to see the wild pendulum swings of LLM sentiment here. If one likes a model then it's capable of one-shotting entire apps. Otherwise it's "only suitable for the most trivial tasks". Never in between.
I have no trivial tasks. Just last week, I was trying to map the weird and wonderful column names emitted by a NetScaler’s detailed REST log into OpenTelemetry semantics. The NetScaler is basically abandoned by its dying vendor. Hence, its new features like sending logs directly to Splunk compatible receivers are basically undocumented. I’m sure there’s like three of us masochists out there stumbling our way through…
>[1] With what prompt!? I like the terse output! Do share...
Not sure if it's complete or completely right, but the matter came up in a recent session, and when asked what gave, Jim 'n' I came up with some supposedly relevant factors, at least one of which is news to me (supposedly, steering LLMs using negative instructions isn't counterproductive anymore (not that I'd been resisting the temptation anyway)):
https://gemini.google.com/share/7af54a6861d7#:~:text=What%20...
With the caveat (or bonus) that it can go (?:too)? far when told to "be blunt" and not to "pull punches":
https://i.vgy.me/WHRZD7.png (from 2024-09) (in this case, in user-config persistent instructions in Kagi's multi-LLM thing)