I've been very happy with Luna but my approach is "many bite sized edits" for which models basically hit saturation a year ago. (I also tried the "let a massive model make massive changes" approach and am still psychologically recovering from the experience. The codebase may never recover!) Also, Luna and DSV4 Flash seem to be on par now except Luna is faster and cheaper?
FWIW I run a clinical analysis backend and directly compared Luna with DSv4 Flash 0731 - DS was a bit ahead, and cheaper even considering it used 40% more tokens. The exposure to Deepseek made me question the valuation house of cards built on SOTA providers. There are more companies producing competitive and useful models than there are companies producing jet engines for airliners, and not for the lack of trying. Ch…
Why jet engines aren't made in China
https://news.ycombinator.com/item?id=48740971
That's a very interesting point of comparison though. I wonder why that should be the case? Why is it so much harder to make a jet engine than a language model?
Maybe with LLMs the iteration times are lower? Or there's more public information about technique? Or are jet engines just intrinsically a harder problem?