Live data from Hacker News

Using an open model feels surprisingly good

matthewsaltz.com

111–120 of 157 posts

Re: Using an open model feels surprisingly good

#111
post #96

This is an ad. I know the author has a long history here and I'm sure the post is a genuine reflection. But how is posting "I used my own product and it felt really good" anything other than self promotion?

I mean, many posts on HN are self promotion; whether they reach the front page discerns the interesting ones from the ones that are not. An interesting self promotion is still interesting.

Yes, but are self promotion posts not required to use "Show HN"/"Tell HN"? Or at least a disclosure on the comments...

Re: Using an open model feels surprisingly good

#112

Earlier quoted context omitted.

I don't have a problem with that if "truly good" is measured by something other than "feels good to me".

Karpathy doing "vibe check" on models was fine, though. Why they can't do the same thing on something they provide? We have seen self-criticism and self back-pats here. There's no need to be so critical. Others can test and have their own opinions, too.

The honest thing is to add a note that tells readers that you are also selling the product you feel good about.

Advertisement is not sharing, it's manipulation.

Re: Using an open model feels surprisingly good

#114

Earlier quoted context omitted.

Karpathy doing "vibe check" on models was fine, though. Why they can't do the same thing on something they provide? We have seen self-criticism and self back-pats here. There's no need to be so critical. Others can test and have their own opinions, too.

The honest thing is to add a note that tells readers that you are also selling the product you feel good about. Advertisement is not sharing, it's manipulation.

They already did, in my opinion. The text reads:

> ...personal account. I work at Modal, and today we just launched Kimi K3 on managed endpoints, and I know Kimi K3 is supposed to be pretty solid, so instead of upgrading my Claude plan, I wanted to give it a try. (I didn't directly contribute to this feature, so I haven't gotten to play with it yet.)

Emphasis mine.

Re: Using an open model feels surprisingly good

#116

Earlier quoted context omitted.

Sorry to be pedantic, but I believe this is a major source of confusion: Claude "Code" is not a model.

Claude code is a (pretty good) harness. Any other harness I try, I realized at some point, I’m just wishing it worked as well as CC does

only needs 66gb of ram too xD

I see a lot of people saying codex is the best cli agent harness. But I haven't used either so can't compare.

Re: Using an open model feels surprisingly good

#117

Earlier quoted context omitted.

I can see why Anthropic is freaking out right now. China has undermined their whole business. AI models will be basic commodities where hosting providers earn a tiny margin over the raw costs rather than the predicted fortunes from being the gatekeepers to the technology.

Yeah I feel like China and their OSS model companies really did the world a great service by ensuring closed source US companies don't have an monopoly on LLMs and thus capture all the value of AI development and its impact on society. It's not going to be the foundational labs that will capture all the value -- I'm sure they will do just fine. A lot of it now will flow to the companies providing the inferences and s…

yet again, china saved the world

Re: Using an open model feels surprisingly good

#119

Earlier quoted context omitted.

Yeah I feel like China and their OSS model companies really did the world a great service by ensuring closed source US companies don't have an monopoly on LLMs and thus capture all the value of AI development and its impact on society. It's not going to be the foundational labs that will capture all the value -- I'm sure they will do just fine. A lot of it now will flow to the companies providing the inferences and s…

It’s not a service, I hope nobody thinks this is being done ”for the good of humanity” or whatever. Doesn’t mean there’s not a benefit to people, but the financial politics behind it could turn out extremely good for China, and they are very well aware of just that.

Do American companies think about the "financial politics" of their decisions? I somehow really doubt it. Government takes sides and sometimes applies leverage to get companies to do what they want, but that's something separate from "it's a good business decision for Chinese labs to treat models as commodities". There's no need to think about geopolitics - American AI companies could easily take the same approach as the Chinese, but they all want to get exciting and huge VC cash rather than boring and less lucrative Compute-as-a-Service money.

Re: Using an open model feels surprisingly good

#120
post #19

The reason Claude code is so popular is because it’s really good at taking super vague human prose “Claude build me a million dollar SaaS”-type prompts and spitting out thousands of lines of code which cover tons of surface-level edge cases, build in tons of functionality, etc The smaller/open models are less good at that. But that’s not how software development is done. You don’t prompt a whole app and be done with…

I had what I thought was a completely ridiculous prompt: “Build me a replacement for Microsoft Word.” ChatGPT did what I thought was the right thing: it responded asking along the lines of “you don’t really want that do you?” and explained how Word is decades of corner cases, bug fixes, obscure features, and business processes that have been built around it.

I was stunned that Claude simply started spinning its wheels in an attempt to actually build something.

Post reply on HN