Live data from Hacker News

Open Deep Research

github.com

21–30 of 83 posts

Re: Open Deep Research

#21
post #11

https://techcrunch.com/2025/02/04/hugging-face-researchers-a... > On GAIA, a benchmark for general AI assistants, Open Deep Research achieves a score of 54%. That’s compared with OpenAI deep research’s score of 67.36%..Worth noting is that there are a number of OpenAI deep research “reproductions” on the web, some of which rely on open models and tooling. The crucial component they — and Open Deep Research — lack is…

theres always a lot of openTHING clones of THING after THING is announced. they all usually (not always[1]!) disappoint/dont get traction. i think the causes are 1. running things in production/self hosting is more annoying than just paying like 20-200/month 2. openTHING makers often overhype their superficial repros ("I cloned Perplexity in a weekend! haha! these VCs are clowns!") and trivializing the last mile, mos…

Except that in this particular case (like in many others as far as AI goes, actually), the open version came before, by three full months: https://www.reddit.com/r/LocalLLaMA/comments/1gvlzug/i_creat...

OpenAI pretty much never acknowledge prior art in their marketing material because they want you to believe they are the true innovators, but you should not take their marketing claims for granted.

Re: Open Deep Research

#22
So basically, Altman announced Deep Research less than a month ago and open-source alternatives are already out? Investors are not going to be happy unless OpenAI outperforms them all by an order of magnitude

Re: Open Deep Research

#23
post #15
post #11

Earlier quoted context omitted.

theres always a lot of openTHING clones of THING after THING is announced. they all usually (not always[1]!) disappoint/dont get traction. i think the causes are 1. running things in production/self hosting is more annoying than just paying like 20-200/month 2. openTHING makers often overhype their superficial repros ("I cloned Perplexity in a weekend! haha! these VCs are clowns!") and trivializing the last mile, mos…

Maybe these open projects start to get more attention when we have a distribution system/App Store for AI projects. I know YC is looking to fund this https://www.ycombinator.com/rfs

Is "AI appstore" envisioned on Linux edge inference hardware, e.g. PC+NPU, PC+GPU, Nvidia Project Digits? Or only in the cloud?

Apple probably wouldn't accept a 3rd-party AI app store on MacOS and iOS, except possibly in the EU.

If antitrust regulation leads to Android becoming a standalone company, that could support AI competition.

Re: Open Deep Research

#24
post #11

Earlier quoted context omitted.

theres always a lot of openTHING clones of THING after THING is announced. they all usually (not always[1]!) disappoint/dont get traction. i think the causes are 1. running things in production/self hosting is more annoying than just paying like 20-200/month 2. openTHING makers often overhype their superficial repros ("I cloned Perplexity in a weekend! haha! these VCs are clowns!") and trivializing the last mile, mos…

Except that in this particular case (like in many others as far as AI goes, actually), the open version came before, by three full months: https://www.reddit.com/r/LocalLLaMA/comments/1gvlzug/i_creat... OpenAI pretty much never acknowledge prior art in their marketing material because they want you to believe they are the true innovators, but you should not take their marketing claims for granted.

Thanks for the pointer, https://github.com/TheBlewish/Automated-AI-Web-Researcher-Ol...

Re: Open Deep Research

#25

https://techcrunch.com/2025/02/04/hugging-face-researchers-a... > On GAIA, a benchmark for general AI assistants, Open Deep Research achieves a score of 54%. That’s compared with OpenAI deep research’s score of 67.36%..Worth noting is that there are a number of OpenAI deep research “reproductions” on the web, some of which rely on open models and tooling. The crucial component they — and Open Deep Research — lack is…

Another project, https://github.com/jina-ai/node-DeepResearch

  gemini for llm
  brave/duckduckgo for search
  jina reader for reading a webpage

Re: Open Deep Research

#26

So basically, Altman announced Deep Research less than a month ago and open-source alternatives are already out? Investors are not going to be happy unless OpenAI outperforms them all by an order of magnitude

Wasn’t this technique previewed by Google Gemini 2.0 first?

Re: Open Deep Research

#27

So basically, Altman announced Deep Research less than a month ago and open-source alternatives are already out? Investors are not going to be happy unless OpenAI outperforms them all by an order of magnitude

Wasn’t this technique previewed by Google Gemini 2.0 first?

Hilarious how few folks reference this. I’ve found it to be pretty good!

Re: Open Deep Research

#28
I signed up for Gemini Advanced to get access to Deep Research on Feb 1st and felt on top of the world. So cool, so advanced.

Then OpenAI announced theirs on the 2nd: https://openai.com/index/introducing-deep-research/

Ethan Mollick called Google's undergraduate level and OpenAI's graduate level on the 3rd: https://www.oneusefulthing.org/p/the-end-of-search-the-begin...

And now this. I can't stop thinking about The Onion Movie's "Bates 4000" clip:

https://www.youtube.com/watch?v=9JCOBPMIgAA

Re: Open Deep Research

#30
post #11

https://techcrunch.com/2025/02/04/hugging-face-researchers-a... > On GAIA, a benchmark for general AI assistants, Open Deep Research achieves a score of 54%. That’s compared with OpenAI deep research’s score of 67.36%..Worth noting is that there are a number of OpenAI deep research “reproductions” on the web, some of which rely on open models and tooling. The crucial component they — and Open Deep Research — lack is…

theres always a lot of openTHING clones of THING after THING is announced. they all usually (not always[1]!) disappoint/dont get traction. i think the causes are 1. running things in production/self hosting is more annoying than just paying like 20-200/month 2. openTHING makers often overhype their superficial repros ("I cloned Perplexity in a weekend! haha! these VCs are clowns!") and trivializing the last mile, mos…

> RL in a tight loop that is not available in the open

Completely agree that a real RL pipeline is needed here, not just some clever prompting in a loop.

That being said, it wouldn’t be impossible to create a “gym” for this task. You are essentially creating a simulated internet. And hiding a needle is a lot easier than finding it.

Post reply on HN