Earlier quoted context omitted.
As someone who is AI skeptical, there's so many breathless posts like "Jizz-7 Thinking (Good) (Big Balls) can order my morning coffee!" which are a lot of words talking about one person's subjective experience of using some LLM to do one specific thing.
Could you post a selection? It would be intersting to gauge what you mean by breathless. People posting their subjective experience is precisely what a lot of these pieces should be doing, good or bad, their experience is the data they have to contribute.
GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
111–120 of 268 posts
Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#112Is this the “Web Search”, “Deep Research”, or “Agent Mode” feature of ChatGPT? Navigating their feature set is… fun.
In my experience it’s “search Reddit and combine comments”.
Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#113Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#114Your Exeter cavern quandary was not exactly sorted. https://simonwillison.net/2025/Sep/6/research-goblin/#histor...
They are quite old and very well documented, so how on earth could a LLM fuck up unless, a LLM is some sort of next token guesser ...
Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#115I do miss the earlier "heavy" models that had encyclopedic knowledge vs the new "lighter" models that rely on web search. Relying on web search surfaces a shallow layer of knowledge (thanks to SEO and all the other challenges of ranking web results) vs having ingested / memorized basically the entirety of human written knowledge beyond what's typically reachable within the first 10 results of a web search (eg: digiti…
Have you just hallucinated that?
Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#116Oh FFS, "I've Chatted and stuff" Your Exeter cavern quandary was not exactly sorted. https://simonwillison.net/2025/Sep/6/research-goblin/#histor... They are quite old and very well documented, so how on earth could a LLM fuck up unless, a LLM is some sort of next token guesser ...
I made fun of its attempt at drawing a useless scatter chart.
That example wasn't meant to illustrate that it's flawless - just that it's interesting and useful, even when it doesn't get to the ideal answer.
Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#117I've also found it to be good at digging deep on things I'm curious about, but don't care enough to spend a lot of time on. As an example, I wanted to know how much sugar by weight is in a coffee syrup so I could make my own dupe. My searches were drowned out by marketing material, but ChatGPT found a datasheet with the info I wanted. I would've eventually found it too, but that's too much effort for an unimportant t…
Don’t sleep on Gemini Deep Research feature either. I use it for my car work and it beats ChatGPT’s offering at that price point every time.
The other ones will do the thing I want: search a bunch, digest the results, and give me a quick summary table or something.
Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#118I do miss the earlier "heavy" models that had encyclopedic knowledge vs the new "lighter" models that rely on web search. Relying on web search surfaces a shallow layer of knowledge (thanks to SEO and all the other challenges of ranking web results) vs having ingested / memorized basically the entirety of human written knowledge beyond what's typically reachable within the first 10 results of a web search (eg: digiti…
I feel the opposite. Before I can use information from a model's "internal" knowledge I have to engage in independent research to verify that it's not a hallucination. Having an LLM generate search strings and then summarize the results does that research up front and automatically, I need only click the sources to verify. Kagi Assistant does this really well.
Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#119The results look reasonable? It’s a good start, given how long it takes to hear back from our doctor on questions like this.
Re: GPT-5 Thinking in ChatGPT (a.k.a. Research Goblin) is good at search
#120HN is a bit weird because it's got 99 articles about how evil LLMs are and one article that's like "oh hey I asked an LLM questions and got some answers" and people are like "wow amazing".
Not that I mind. I assume Simon just wanted to share some cool nerdy stuff and there's nothing wrong with the blog post. It's just surprising that it's posted not once but twice on HN and is on the front page when there's so much anti-AI sentiment otherwise.