Live data from Hacker News

Yi-Coder: A Small but Mighty LLM for Code

01-ai.github.io

111–117 of 117 posts

Re: Yi-Coder: A Small but Mighty LLM for Code

#111
post #85

Everytime someone tells how AI 10x his programming capabilities I'm like "tell me you're bad at coding without telling me".

Everytime someone posts a comment that is just "I'm better than other people", I'm like "what a waste of time reading that was".

Re: Yi-Coder: A Small but Mighty LLM for Code

#112

Earlier quoted context omitted.

Because code either works or it doesn't. Nobody is replacing our entire income stream with an LLM. You also need a knowledge of code to instruct an LLM to generate decent code, and even then it's not always perfect. Meanwhile plenty of people are using free/cheap image generation and going "good enough". Now they don't need to pay a graphic artist or a stock photo licence Any layperson can describe what they want a p…

At best it's a prototype and concept generator. It would have to yield assets with layers that can be exported by an illustration or bitmap tool of choice. AI generated images are almost completely useless as-is.

I agree there are plenty of images with garbled text and hands with 7 fingers, but text to image has freely available generators which create almost perfect images for some prompts. Certainly good enough to replace an actor holding a product, a stock photo, and often a stylised design.

Re: Yi-Coder: A Small but Mighty LLM for Code

#113
post #4

Claude 3.5 Sonnet still holds the LLM crown for code which I'll use when wanting to check the output of the best LLM, however my Continue Dev, Aider and Claude Dev plugins are currently configured to use DeepSeek Coder V2 236B (and local ollama DeepSeek Coder V2 for tab completions) as it offers the best value at $0.14M/$0.28M which sits just below Claude 3.5 Sonnet on Aider's leaderboard [1] whilst being 43x cheaper…

Uh, what price format is "14 million cents per 28 million cents"?

14 cents per million input/28 cents per million output tokens

Re: Yi-Coder: A Small but Mighty LLM for Code

#114

Earlier quoted context omitted.

Because code either works or it doesn't. Nobody is replacing our entire income stream with an LLM. You also need a knowledge of code to instruct an LLM to generate decent code, and even then it's not always perfect. Meanwhile plenty of people are using free/cheap image generation and going "good enough". Now they don't need to pay a graphic artist or a stock photo licence Any layperson can describe what they want a p…

> " Meanwhile plenty of people are using free/cheap image generation and going "good enough". Now they don't need to pay a graphic artist or a stock photo licence " and getting sandwich photos of ham blending into human fingers: https://www.reddit.com/r/Wellthatsucks/comments/1f8bvb8/my_l...

And yet, even knowing what I was looking for, I didn't see it long enough that I guessed I misunderstood and swiped to the second image, where it was pointed out specifically. Even if I had noticed myself--presumably because I was staring at it for way too long in the restaurant--I can't imagine I would have guessed what was going on, BUT EVEN THEN it just wouldn't have mattered... clearly, this is more than merely a "good enough" image.

Re: Yi-Coder: A Small but Mighty LLM for Code

#115
post #73

Earlier quoted context omitted.

Their services will just be blocked in the EU instead.. we've seen it in Italy early on with ChatGPT..

When will we have token-flow-aware-networking gear... Surely NVIDIA and others are already doing special traffic shaping for tokenFlows? Whats the current state of such tech/thought/standards/vendors?

When do we have defcon talks about token injection poisoning to inject hallucinations into responses while housing the actual true AGI insight that will change world view.

Re: Yi-Coder: A Small but Mighty LLM for Code

#116
post #62

Sounds very promising! I hope that Yi-Coder 9B FP16 and Q8 will be available soon for Ollama, right now i only see the 4bit quantized 9B model. I'm assuming that these models will be quite a bit better than the 4bit model.

Click on "View more" in the dropdown on their page, it has many many quantized versions to choose from.

Found it, thanks
Post reply on HN