Live data from Hacker News

Introducing deep research

openai.com

101–110 of 445 posts

Re: Introducing deep research

#101
From the demo: “Use bullets and tables where necessary for clarity.” It’s weird that it would be necessary to specify that. I suppose they want to showcase that you can influence the output style, but it’s strange that you’d have to explicitly specify the use of something that is “necessary for clarity”. It comes across as either a flaw in the default execution, or as a merely performative incantation.

Re: Introducing deep research

#102

So much cynicism and hate in these comments, especially as we are likely witnessing AGI come to life. Its still early, but it might be coming. Where is the excitement? This is an interesting time to be alive. HN has a huge cultural problem that makes this website almost irrelevant. All the interesting takes have moved to X/twitter

We're looking at trends that may well obliterate the economic value of a well trained human mind sitting behind a keyboard all day. That is a bit of a threat to most people on HN if the trending continues at the current rate and direction.

Re: Introducing deep research

#104
post #56

This is terrifying. Even though they acknowledge the issues with hallucinations/errors, that is going to be completely overlooked by everyone using this, and then injecting the outputs into their own powerpoints. Management Consulting was bad enough before the ability to mass produce these graphs and stats on a whim. At least there was some understanding behind the scenes of where the numbers came from, and sources w…

Either you care about being correct or you don't. If you don't care then it doesn't matter whether you made it up or the AI did. If you care then you'll fact check before publishing. I don't see why this changes.

I think a lot about how differentiating facts and quality content is like differentiating signal from noise in electronics. The signal to noise ratio on many online platforms was already quite low. Tools like this will absolutely add more noise, and arguably the nature of the tools themselves make it harder to separate the noise.

I think this is a real problem for these AI tools. If you can’t separate the signal from the noise, it doesn’t provide any real value, like an out of range FM radio station.

Re: Introducing deep research

#106
Surprised more comments aren't mentioning deepseek has this feature (for free) already. Assuming this is why OpenAI scrambled to release it.

The examples they have on the page work well on chat.deepseek.com with r1 and search options both enabled.

Do I blindly trust the accuracy of either though? Absolutely not. I'm pretty concerned about these models falling into gaming SEO and finding inaccurate facts and presenting them as fact. (How easy is it to fool / prompt inject these models?)

But has utility if held right.

Re: Introducing deep research

#107
post #41

OpenAI has a deep bench. I bet they pushed this out to change the narrative about deepseek

Also named specifically to muddle the SEO for the term "deep." Nothing that OpenAI does is unintentional.

Google publicly announced a model named "Deep Research" on December 11th: https://blog.google/products/gemini/google-gemini-deep-resea...

Re: Introducing deep research

#109
post #24

If I understood the graphs correctly, it only achieves 20% pass rate on their internal tests. So I have to wait 30min and pay a lot of money just to sift through walls of most likely incorrect text? Unless the possibility of hallucinations is negligible, this is just way too much content to review at once. The process probably needs to be a lot more iterative.

I mean you want it to grill your steak and eat it for you too? I mean I too can complain that my iPhone doesn’t automatically screen out spammers and send my mom flowers on Mother’s Day.

Why doesn't the iPhone screen spammers yet? Pixel has had this feature for a decade.

Re: Introducing deep research

#110
post #97
post #67

Earlier quoted context omitted.

This is one of the actual questions: > In Greek mythology, who was Jason's maternal great-grandfather? https://www.google.com/search?q=In+Greek+mythology%2C+who+wa...

This is a hard question for language models since it targets one of their known weaknesses.

Users don’t care about how hard something is for LLMs if they receive incorrect output.
Post reply on HN