Live data from Hacker News

Ornith-1.0: self-improving open-source models for agentic coding

github.com

31–40 of 65 posts

Re: Ornith-1.0: self-improving open-source models for agentic coding

#31
post #30
post #25

Earlier quoted context omitted.

Its not any better. Most of us at LocalLLama community dont like it except a few new people poping out and making posts.

Indeed, it performed worse than Qwen3.6-27b in my basic test. It gave a fancier looking answer, but did a worse job following the prompt.

Roughly my experience so far; it trips up on itself a bit.

However, it's much more inclined to do web search unprompted, which is fascinating in its own way.

Re: Ornith-1.0: self-improving open-source models for agentic coding

#32
post #15
post #11

These are simply benchmaxxed versions of either Qwen or Gemma 4.

Citation needed

Sure. https://deep-reinforce.com/ornith_1_0.html

>Built on top of pretrained Gemma 4 and Qwen 3.5, it achieves state-of-the-art performance among open-source models of comparable size on coding benchmarks.

>Ornith-1.0 is a self-improving training framework. Instead of relying on human-designed harnesses to drive solution generation in RL, Ornith-1.0 learns to generate both solution rollouts and the task-specific harnesses that guide those rollouts.

Re: Ornith-1.0: self-improving open-source models for agentic coding

#33
post #25

This is the first Qwen fine-tune that is not immediately rejected by the local LLM community, and in some cases even being recommended. Based on my limited usage, it is good, gives creative solutions to coding problems. I don't expect 9-35B models to one-click create full apps. Most people who were complaining did so .

Its not any better. Most of us at LocalLLama community dont like it except a few new people poping out and making posts.

> LocalLLama community

Ah, the place that shit on gpt-oss because it wasn't good at porn. That place is not what it used to be, hasn't been since that karpathy tweet, tbh. It's mostly slop and vibes nowadays.

Re: Ornith-1.0: self-improving open-source models for agentic coding

#34
post #25

Earlier quoted context omitted.

Its not any better. Most of us at LocalLLama community dont like it except a few new people poping out and making posts.

> LocalLLama community Ah, the place that shit on gpt-oss because it wasn't good at porn. That place is not what it used to be, hasn't been since that karpathy tweet, tbh. It's mostly slop and vibes nowadays.

and a lot of bots advertising a rename models like this one.

Re: Ornith-1.0: self-improving open-source models for agentic coding

#35
post #3

Previously: https://news.ycombinator.com/item?id=48709744 https://swelljoe.com/post/will-it-mythos/ : "Poor performer here, only found the one bug that almost every model found, despite its performance on other benchmarks being excellent for its size. […] It also performs poorly in a chat without tools, exhibiting an ehthusiasm for hallucination. I’m currently working on a replication of this with full tool access, i…

That benchmark ranks Kimi K2.6 and K2.7 Code near the bottom. Both are below Ornith 35B. It ranks Gemma 4 26B much higher than GLM-5.2. The results don't make much sense.

Re: Ornith-1.0: self-improving open-source models for agentic coding

#37

This is the first Qwen fine-tune that is not immediately rejected by the local LLM community, and in some cases even being recommended. Based on my limited usage, it is good, gives creative solutions to coding problems. I don't expect 9-35B models to one-click create full apps. Most people who were complaining did so .

The local LLM community is now teeming with erstwhile crypto and NFT hucksters who've brought the culture of hype from their former communities with them. There still are a few deeply technical people left, but their voices are being crowded out by the vapid marketers'.

Re: Ornith-1.0: self-improving open-source models for agentic coding

#38

This is the first Qwen fine-tune that is not immediately rejected by the local LLM community, and in some cases even being recommended. Based on my limited usage, it is good, gives creative solutions to coding problems. I don't expect 9-35B models to one-click create full apps. Most people who were complaining did so .

The local LLM community is now teeming with erstwhile crypto and NFT hucksters who've brought the culture of hype from their former communities with them. There still are a few deeply technical people left, but their voices are being crowded out by the vapid marketers'.

I've also noticed this. I wonder what causes the overlap. It can't be as simple as crypto and LLMs requiring the same hardware.

Re: Ornith-1.0: self-improving open-source models for agentic coding

#39
post #38

Earlier quoted context omitted.

The local LLM community is now teeming with erstwhile crypto and NFT hucksters who've brought the culture of hype from their former communities with them. There still are a few deeply technical people left, but their voices are being crowded out by the vapid marketers'.

I've also noticed this. I wonder what causes the overlap. It can't be as simple as crypto and LLMs requiring the same hardware.

They both feed off of hype. The people posting about crypto are not the same people gpu mining crypto so I wouldn’t chalk it up to the same hardware

Re: Ornith-1.0: self-improving open-source models for agentic coding

#40
post #25

Earlier quoted context omitted.

Its not any better. Most of us at LocalLLama community dont like it except a few new people poping out and making posts.

> LocalLLama community Ah, the place that shit on gpt-oss because it wasn't good at porn. That place is not what it used to be, hasn't been since that karpathy tweet, tbh. It's mostly slop and vibes nowadays.

Where’s a good place to go instead nowdays?
Post reply on HN