Live data from Hacker News

Using an open model feels surprisingly good

matthewsaltz.com

131–140 of 157 posts

Re: Using an open model feels surprisingly good

#131
post #119

Earlier quoted context omitted.

Do American companies think about the "financial politics" of their decisions? I somehow really doubt it. Government takes sides and sometimes applies leverage to get companies to do what they want, but that's something separate from "it's a good business decision for Chinese labs to treat models as commodities". There's no need to think about geopolitics - American AI companies could easily take the same approach as…

Chinese labs are equivalent to the Chinese state, there’s no separation. Who do you think funds the Chinese labs?

Before Trump 2.0, this was a decent argument.

Re: Using an open model feels surprisingly good

#133

I love OpenCode and the blankness of it too. Clean, light, and manual. It's a good change of pace from what we are used to on the internet (very cool but slow). Also, can you hmu with that modal plan XD

Pi Agent is even leaner feeling, it's worth a try.

My own agent is cleaner and lighter. Written in D, now only the local models are the bottleneck.

Re: Using an open model feels surprisingly good

#134

This is an ad. I know the author has a long history here and I'm sure the post is a genuine reflection. But how is posting "I used my own product and it felt really good" anything other than self promotion?

Hello! It probably doesn’t make much of a difference now that the post is flagged (I assume that it has something to do with people thinking it’s an ad?) but just for the record, I really didn’t intend this as an “ad”. No one asked me to write this and I didn’t plan on writing it. It basically happened just how I described it - I went home, had some extra energy, wanted to build something that needed both 1) an LLM to help me build the thing and 2) an LLM to use within the thing I was building. So I decided to try the Modal endpoint bc it was top of mind and I didn’t want to upgrade my Claude account. Had I not thought it was cool I would not have written about it, but what made me write the thing wasn’t “Modal is cool”, it was the feeling I describe in the post. I was a little worried it might come across as an ad and considered removing the links but figured, “oh well”. It was a nice plus that it involved the company I’m at doing something cool, but that wasn’t really the point. In any case, I’m glad some people found it interesting, and maybe this was flagged for some other reason, but thought I might as well respond, even if I may not convince anyone :) Possibly even given the above it’s not HN-appropriate somehow but I thought that the sentiment might be appreciated

Re: Using an open model feels surprisingly good

#135

Earlier quoted context omitted.

Yeah I feel like China and their OSS model companies really did the world a great service by ensuring closed source US companies don't have an monopoly on LLMs and thus capture all the value of AI development and its impact on society. It's not going to be the foundational labs that will capture all the value -- I'm sure they will do just fine. A lot of it now will flow to the companies providing the inferences and s…

It’s not a service, I hope nobody thinks this is being done ”for the good of humanity” or whatever. Doesn’t mean there’s not a benefit to people, but the financial politics behind it could turn out extremely good for China, and they are very well aware of just that.

I think this is a false dichotomy between benevolence and China secretly plotting and puppeteering everything to some grand self serving scheme. Reality is probably a lot simpler and has little to do with either. In China there's intense domestic competition in most industries, including LLMs.

Going open weights is an easy way to gain mindshare in this sort of environment, which is exactly why Meta also went open weight. The difference is that in China you had leading models going open weight which puts a lot of downward pressure on other orgs to do the exact same. I doubt China's grand vision extends beyond achieving the best system possible, and fostering a highly competitive environment is exactly how you achieve that.

The fact that open weight frontier models may also cause the investment bubble in the US to burst is likely incidental. AI funding is a bubble, and it's going to burst. The exact final cause is again mostly just incidental. The economic damage this will cause in the US will also cause substantial downstream damage to China as well, and they generally aren't so big on the whole punch yourself in the face because you don't like another country, that has become trendy in the West. So the idea of a secret economic attack doesn't even make much sense. In the status quo China wins, so they don't even have any motivation to cause chaos.

Re: Using an open model feels surprisingly good

#136

Earlier quoted context omitted.

This is what I've found and would likely be the consenses of the HN community. This has lowered the barrier of entry for many non-SWEs. I'm primarily a data scientist myself in the environmental sector with limited front end development. My partner proposed an idea to help manage her horses and over the course of several weeks we fleshed out an android app that would enable/assist her with horse care. I spent a few d…

Question if you don't mind. Would you have considered outsourcing the application you mentioned? As in, paying someone else to do it? To me it is great that LLMs are allowing more people to use computers and software the way they were meant to. But the people that are using them this way, in my opinion, wouldn't have commissioned anybody to do it anyway. They'd just live with whatever process/pain they have. So jumpi…

A practical issue is that a lot of big software is used to solve small problems. For instance with his example, in the past perhaps instead of creating a little custom app it would have been a series of excel sheets with some other third parties tool as needed. And I think this is a very common use case. At most/all companies there tend to be convoluted processes to do relatively simple things, often enriching the big generalist companies in the process. As people become capable of creating competent ad-hoc solutions to problems, the utility of big high-utility software goes down.

Like a quote I've read on here multiple times about replacing e.g. excel or whatever, people often mention that users only use 5% of a software's functionality, but it's a different 5% for each person. The argument being that to replace excel you'd then need to mimic every esoteric thing it does, but when users can now achieve that 5% in a highly customized way with no real knowledge needed, it's going to have a major impact on big software.

Re: Using an open model feels surprisingly good

#138

Earlier quoted context omitted.

To answer your question probably not unless I knew the person. I could argue the case that I would have outsourced the android UI scope if I was motivated enough to develop it and kept the backend to myself. I could then tinker with the front end and build upon it. However I wouldn't outsource now if I knew the person would just use LLM anyway. I'm the type that would rather build something myself than purchase off t…

Thank you for answering. I agree with you that, if used well, LLMs can be useful tools and increase efficiency. Unfortunately, what I am increasingly seeing is an increase in confidence but no increase in knowledge or understanding. I hope this is just a side-effect of the hype and corresponding bubble and not the future trajectory of knowledge work.

Foundational knowledge can be pretty far upstream of functional knowledge.

Whether you can't describe the analog circuits that correspond to the opcodes of the major mobile architectures running the platform of the app you're designing a UI for, or you couldn't have built any of the services you're hosting on from scratch, no one is going to be disappointed.

There are more people making more money off whatever we're currently calling making computers do things than ever before in history. The rest is largely business as usual.

Re: Using an open model feels surprisingly good

#139
post #75
post #67

you can use claude code or codex with open models using AllRouter https://github.com/Lore-Hex/AllRouter

For claude it is enough to set a few env variables. It lets you easily map each model class (opus, sonnet, haiku) to an open model with Openrouter: https://openrouter.ai/docs/cookbook/coding-agents/claude-cod...

yes but AllRouter does automatic failover and can burst to cloud as well or use local when needed

Re: Using an open model feels surprisingly good

#140
post #31
post #18

Earlier quoted context omitted.

I wrote my own OpenClaw one weekend and I am running it as my assistant through Matrix with DeepSeek v4 Flash (and Qwen). It probably costs me about 2 dollars a month and is even more useful than ChatGPT would be due to me having full control on what tools it has access to. I can do things like take a photo of a doctor's note among add the appointment to my calendar, send a PDF to my archive tagged, OCR'd etc, search…

Haha! Love to hear it. That's exactly what I have too: a claw-like[0] system which I originally used Sonnet for and now use my DeepSeek v4 Flash with. In my case, I was foolish enough to run it all on my hardware which is pretty damned fast but has a duty cycle of 5% and runs idle most of the time. It's definitely better done via API. Tell me more about the Alexa-like system! I have mine at home on a custom OpenWakeW…

I have these around the house:

https://www.home-assistant.io/voice-pe/

Then you can set the background AI to be any OpenAI compatible API. So I just created one to my local Rust Agent, and connected it to home assistant. Now I can yell from the couch to create me a new Proxmox container with the next free static IP address etc. :D

Post reply on HN