Earlier quoted context omitted.
Contract? These docs are information answering user queries. So if you use a chatbot to generate them, I'd like to be reasonably sure they aren't laden with the fabricated misinformation for which these chatbots are famous.
It's a very reasonable concern. My solution is to have the bot classify what the message is talking about as a first pass, and have a relatively strict filtering about what it responds to. For example, I have it ignore messages about code freezes, because that's a policy question that probably changes over time, and I have it ignore urgent oncall messages, because the asker there probably wants a quick response from…
A few random notes from Claude coding quite a bit last few weeks
761–770 of 870 posts
Re: A few random notes from Claude coding quite a bit last few weeks
#762Earlier quoted context omitted.
I'm almost a boomer and I agree. THis dichotomy is weird. I am retired EE and I love the ability to just have AI do whatever I want for me. I have it manage a 10 node proxmox cluster in my basement via ansible and terraform. I can finally do stuff I always wanted but had no time. I got sick of editing my kids sports videos for highlights in Davinci Resolve so just asked claude to write a simple app for me and then us…
So what is your workflow now with this app for kids sports highlights?
Re: A few random notes from Claude coding quite a bit last few weeks
#763Earlier quoted context omitted.
That’s far-fetched. It’s in the interest of the model builders to solve your problem as efficiently as possible token-wise. High value to user + lower compute costs = better pricing power and better margins overall.
> It’s in the interest of the model builders to solve your problem as efficiently as possible token-wise. High value to user + lower compute costs = better pricing power and better margins overall. It's only in the interests of the model builders to do that IFF the user can actually tell that the model is giving them the best value for a single dollar. Right now you can't tell.
Re: A few random notes from Claude coding quite a bit last few weeks
#764Earlier quoted context omitted.
>"Reasoning", however, is a feature that has been bolted on with a hacksaw and duct tape. What do you mean by this? Especially for tasks like coding where there is a deterministic correct or incorrect signal it should be possible to train.
it's meant in the literal sense but with metaphorical hacksaws and duct tape. Early on, some advanced LLM users noticed they could get better results by forcing insertion of a word like "Wait," or "Hang on," or "Actually," and then running the model for a few more paragraphs. This would increase the chance of a model noticing a mistake it made. Reasoning is basically this.
Re: A few random notes from Claude coding quite a bit last few weeks
#765Earlier quoted context omitted.
Because of the collapsing empire, mind you, not because of the LLMs.
Creation of the internet, social media, everyone on the planet getting a pocket sized supercomputer, beginning of the AI boom, Trump/beginning of the end of the US, are all reasons people will study this period of time.
But, it is really hard to escape the feeling that digital technology and AI are a huge inflection point. In some ways this couple generations might be the singularity. Trump and contemporary geopolitics in general is a footnote, a silly blip that will pale in comparison over time.
Re: A few random notes from Claude coding quite a bit last few weeks
#766Earlier quoted context omitted.
> It’s in the interest of the model builders to solve your problem as efficiently as possible token-wise. High value to user + lower compute costs = better pricing power and better margins overall. It's only in the interests of the model builders to do that IFF the user can actually tell that the model is giving them the best value for a single dollar. Right now you can't tell.
Why not? Seems like you'd just build the same app on each of the models you want to test and judge how they did.
I tried that on a few problems; even on the same model the results have too much variation.
When comparing different models, repeating the experiment gives you different results.
Re: A few random notes from Claude coding quite a bit last few weeks
#767The Slopocalypse - an unexpected variant of Gray Goo: https://en.wikipedia.org/wiki/Gray_goo
Well, it may consume the AI environment. Maybe even the internet. It's not going to consume a PC with g++, though (at least if the PC doesn't update g++ any more once g++ starts accepting AI contributions). There may come a point where having a "survivor machine" with auto-update turned off may be a really good idea.
Re: A few random notes from Claude coding quite a bit last few weeks
#768Earlier quoted context omitted.
That split has always existed. LLMs can be used on either side of the divide.
We see a ton of "AI let me code a program X faster than ever before." We see almost no "AI let me code a program X better than ever before."
Re: A few random notes from Claude coding quite a bit last few weeks
#769Earlier quoted context omitted.
I'm not a big fan of LLMs, but while using it for day to day tasks, I get the same feeling I had when I first started the internet (I was lucky to start with broadband internet). That feeling was one of empowerment: I was able to satisfy my curiosity about a lot of topics. LLMs can do the same thing and save me a lot of time. It's basically a super charged Google. For programming it's a super charged auto complete co…
When I first started using the internet, I was able to instant text message (IRC) random strangers, using a fake name, and lie about my age. My teacher had us send an email to our ex-classmate who had move to Australia, and she replied the next day, I was able to download the song I just heard on the radio and play it as many times as I wanted on my winamp. These capabilities simply didn’t exist before the Internet.…
The list of things they can provide is endless.
They're not a creator, they're an accelerator.
And time matters. My interests are myriad but my capacity to pass the entry bar manually is low because I can only invest so much time.
Re: A few random notes from Claude coding quite a bit last few weeks
#770Earlier quoted context omitted.
I don't understand why anyone finds it interesting that a machine, or chatbot, never tires or gets demoralized. You have to anthromorphize the LLM before you can even think of those possibilities. A tractor never tires or gets demoralized either, because it can't. Chatbots don't "dive into a rabbit hole ... and then keep digging" because they have superhuman tenacity, they do it because that's what software does. If…
You're a machine. You're literally a wet, analog device converting some forms of energy into other forms just like any other machine as you work, rest, type out HN comments, etc. There is nothing special about the carbon atoms in your body -- there's no metadata attached to them marking them out as belonging to a Living Person. Other living-person-machines treat "you" differently than other clusters of atoms only bec…
I might feel awe or amazement at what human-made machines can do -- the reason I got into programming. But I don't attribute human qualities to computers or software, a category error. No computer ever looked at me as interesting or tenacious.