Humans hallucinating about AI.
"OpenAI Researcher Hallucinates GPT-5 Math Breakthrough" could be a headline from The Onion.
OpenAI researcher announced GPT-5 math breakthrough that never happened
191–200 of 258 posts
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#192> GPT-5 is proving useful as a literature review assistant No, it does not. It only produces a highly convincing counterfeit. I am honestly happy for people who are satisfied with its output: life is way easier for them than for me. Obviously, the machine discriminates me personally. When I spend hours in the library looking for some engineering-related math made in the 70s-80s, as a last resort measure, I can try to…
I wonder whether for a lot of the search & literature review-type use-cases where people are trying to use GPT-5 and similar we'd honestly be much better off with a really powerful semantic search engine? Any time you ask a chatbot to summarize the literature for you or answer your question, there's a risk it will hallucinate and give you an unreliable answer. Using LLM-generated embeddings for documents to retrieve…
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#193Wouldn't be surprised if OpenAI employees are being asked to phrase ( market ) things this way. This is not the first time they claimed GPT-5 "solved" something [1] [1] https://x.com/SebastienBubeck/status/1970875019803910478 edit: full text It's becoming increasingly clear that gpt5 can solve MINOR open math problems, those that would require a day/few days of a good PhD student. Ofc it's not a 100% guarantee, eg be…
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#194> GPT-5 is proving useful as a literature review assistant No, it does not. It only produces a highly convincing counterfeit. I am honestly happy for people who are satisfied with its output: life is way easier for them than for me. Obviously, the machine discriminates me personally. When I spend hours in the library looking for some engineering-related math made in the 70s-80s, as a last resort measure, I can try to…
I wonder whether for a lot of the search & literature review-type use-cases where people are trying to use GPT-5 and similar we'd honestly be much better off with a really powerful semantic search engine? Any time you ask a chatbot to summarize the literature for you or answer your question, there's a risk it will hallucinate and give you an unreliable answer. Using LLM-generated embeddings for documents to retrieve…
Basically we are trying to combine the benefits of chat with normal academic search results using semantic search and keyword search. That way you get the benefit of LLMs but you’re actually engaging with sources like a normal search.
Hope it was what you were looking for!
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#195Whatever happened to "don't get high on your own supply"?
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#196You would think Open AI employees have a pretty good grasp of their model capabilities, but even if you don’t, you probably always want to be on the cautious side for every claim you see on the internet. This just seems to be the Open AI culture, which for better or worse has helped foster the AI hype environment we are currently in.
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#197> GPT-5 is proving useful as a literature review assistant No, it does not. It only produces a highly convincing counterfeit. I am honestly happy for people who are satisfied with its output: life is way easier for them than for me. Obviously, the machine discriminates me personally. When I spend hours in the library looking for some engineering-related math made in the 70s-80s, as a last resort measure, I can try to…
Saying it isn't useful is a bit of an overstatement. It can search, churn through 500k words in a few minutes, and come back with summaries, answers, and sources for each point. Should you blindly trust the summary? No. Should you verify key claims by clicking through to the source? Yes. Is it still incredibly useful as a search tool and productivity booster? Absolutely.
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#198Earlier quoted context omitted.
In my experience doing literature super-deep-dives, it hallucinates sources about 50% of the time. (For higher-level literature surveys, it's maybe 5%.) Of the other 50% that are real, it's often ~evenly split into sources I'm familiar with and sources I'm not. So it's hugely useful in surfacing papers that I may very well never have found otherwise using e.g. Google Scholar. It's particularly useful in finding relev…
So, the exact stuff Google used to be good at.
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#199Earlier quoted context omitted.
So, the exact stuff Google used to be good at.
Another win for big tech: Google has been enshittified to such a point that you can now spin up a machine that consumes 1000x the power to give you a result that has a coin toss odds of being totally made up.
Re: OpenAI researcher announced GPT-5 math breakthrough that never happened
#200Earlier quoted context omitted.
So, the exact stuff Google used to be good at.
Another win for big tech: Google has been enshittified to such a point that you can now spin up a machine that consumes 1000x the power to give you a result that has a coin toss odds of being totally made up.