Live data from Hacker News

Things we learned about LLMs in 2024

simonwillison.net

571–580 of 615 posts

Re: Things we learned about LLMs in 2024

#571
post #235

Earlier quoted context omitted.

I'm surprised at the description that it's "useless" as a programming / design partner. Even if it doesn't make "elegant" code (whatever that means), it's the difference between an app existing at all, or not. I built and shipped a Swift app to the App Store, currently generating $10,200 in MRR, exclusively using LLMs. I wouldn't describe myself as a programmer, and didn't plan to ever build an app, mostly because in…

May I know what is the name of app that is built using LLM? 10k MRR is highly successful app.

[deleted]

Re: Things we learned about LLMs in 2024

#572
post #366
post #356

Earlier quoted context omitted.

May you expand how you did this? I'm seeing a number of apps that claim to do just this and there are number that are becoming super popular. Not just the development of the code but the entire the thing from the code, infra, auth, cc payments, etc.

Planning to write a lengthy blog post on this. Will reply here.

[deleted]

Re: Things we learned about LLMs in 2024

#573
post #202
post #54

About "people still thinking LLMs are quite useless", I still believe that the problem is that most people are exposed to ChatGPT 4o that at this point for my use case (programming / design partner) is basically a useless toy. And I guess that in tech many folks try LLMs for the same use cases. Try Claude Sonnet 3.5 (not Haiku!) and tell me if, while still flawed, is not helpful. But there is more: a key thing with L…

> Try Claude Sonnet 3.5 (not Haiku!) and tell me if, while still flawed, is not helpful. It's not as helpful as Google was ten years ago. It's more helpful than Google today, because Google search has slowly been corrupted by garbage SEO and other LLM spam, including their own suggestions.

Google Search has been corrupted by...Google.

Re: Things we learned about LLMs in 2024

#574
post #405

Earlier quoted context omitted.

Staff plus just means staff or higher. Staff, senior staff, principal, mega ultra principal etc…

Outside of big tech, those titles aren’t common. Level X SWE vs staff vs principal doesn’t mean anything to a lot of people who aren’t in that orbit.

Sure, but my point is when someone says staff plus they mean staff or higher. They don’t mean higher than staff, or the best of the best staff engineers.

It just means anyone higher than a senior engineer.

Re: Things we learned about LLMs in 2024

#575
post #494
post #453

Earlier quoted context omitted.

They are also usually selling another AI-wrapper. I don't know the parent poster either but if your LLM product is generating $10k/month, your moat is really weak and you'll probably shut the f* up because your only moat is obscurity. Why risk that?

We shouldn’t assume the app created the customer base anew or solves a novel problem. Maybe this one does, we don’t know. But, what if the app is just an app version of a existing website store? As an example I could imagine a clothing brand wanting an app that customers can install instead of using their phone browser. $10k/month in that context isn’t as surprising or impressive.

In which case the LLM contribution to the $10K/month is equivalent to hiring a mobile developer to build such an app which (given the implied simplicity) should be a few thousands one time cost. Not the $120K/month implied by PP. And don't get me wrong, paying a few dozen dollars to get a few thousand dollars worth of software is quite the value.

Re: Things we learned about LLMs in 2024

#576

Earlier quoted context omitted.

> And a big strength LLMs have is summarizing things - I’d like to see you summarize the latest 10 arxiv papers relating to prompt engineering and produce a report geared towards non-techies. And do this every 30 mins please. Also produce social media threads with that info. Is this a task you could do yourself, better than LLMs? I don't mean to nitpick, but how good do you really think the output of this would be? P…

The quality of the summary is only as good as the effort you put into writing your workflow. If you’re simply one shotting the paper into a message and saying “plz summarise this and I’ll reward you with $1m” then of course it’s gonna be shit. But if you semantically chunked along sections and do some RAG Q&A summaries before combining into a well formatted schema then it’s probably going to be better than the first…

> I’m using the summaries as a juicier abstract. I’m not taking them as gospel.

I'm not sure of the value of this. Papers already have abstracts, rewording them using LLMs is just playing with your food. If you're seeing use out of it that's awesome though.

Re: Things we learned about LLMs in 2024

#577

Earlier quoted context omitted.

Just curious, but what AI related skills do you expect them to have?

The ability to recognize and join a hype train, I presume. It’s one way to appear proactively leading-edge to marginally-informed product managers, marketers, execs and press.

That's an extremely uncharitable presumption. Although I don't agree that routine usage of AIs should be a precondition for regular software engineering jobs, there are good reasons for using LLMs besides "joining a hype train".

Re: Things we learned about LLMs in 2024

#578
post #577

Earlier quoted context omitted.

The ability to recognize and join a hype train, I presume. It’s one way to appear proactively leading-edge to marginally-informed product managers, marketers, execs and press.

That's an extremely uncharitable presumption. Although I don't agree that routine usage of AIs should be a precondition for regular software engineering jobs, there are good reasons for using LLMs besides "joining a hype train".

Nah.

Re: Things we learned about LLMs in 2024

#579
post #478

Earlier quoted context omitted.

I'm pretty sure most people, developers especially, have had magical, life-changing experiences with LLMs. I think the problem is that they can't cant do these things reliably. I get this sentiment from a lot of AI startups, that they have a product which can do amazing things, but due to its failure modes makes it almost useless as, to use an analogy from self-driving cars, the users have to still constantly pay att…

I mean… I agree that LLMs give only superficial value, but your analogy is plain wrong. I drove 3600 km Norway to Spain in 2018 with only adaptive cruise. Then again in 2023 with autonomous highway driving (the kind where you keep a hand on the wheel for failure mode) and it was amaaaazing how big the difference was.

I get how I could be wrong on that front. I guess what I was trying to say was that there needs to be legible, predictable infrastructure for these AI systems to work well. I actually think that an LLM workflow in a constrained, well understood environment would be amazingly good too.

I've been driving a lot in Istanbul lately and I'm not holding my breath for autonomous vehicles any time soon.

Re: Things we learned about LLMs in 2024

#580
post #38

Something not mentioned is AI generated music. Suno's development this year is impressive. Unclear what this will mean for music artists over next few years.

Very clear; I like buying music produced by people who play instruments.

What you think of samples and FL Studio / DAWs?
Post reply on HN