In about 2 weeks since OpenAI launched their $200/mo version of Deep Research, it has already been open sourced within 24 hours (Hugging Face) and now being offered for free by Perplexity. The pace of disruption is mind boggling and makes you wonder if OpenAI has any moats left.
My interest was piqued and I’ve been trying ChatGPT Pro for the last week. It’s interesting and the deep research did a pretty good job of outlining a strategy for a very niche multiplayer turn based game I’ve been playing. But this article reminded me to change next month’s subscription back to the premium $20 subscription. Luckily work just gave me access to ChatGPT Enterprise and O1 Pro absolutely smoked a really…
Perplexity Deep Research
111–120 of 180 posts
Re: Perplexity Deep Research
#112I'm super happy that these types of deep research applications are being released because it seems like such an obvious use case for LLMs. I ran Perplexity through some of my test queries for these. One query that it choked hard on was, "List the college majors of all of the Fortune 100 CEOs" OpenAI and Gemini both handle this somewhat gracefully producing a table of results (though it takes a few follow ups to get a…
Hopefully the end user of these products know something about LLMs and why asking a question such as "List the college majors of all of the Fortune 100 CEOs" is not really suited well for them.
Re: Perplexity Deep Research
#113Every week we get a new AI that according to the AI-goodness-benchmarks is 20% better than the old AI, yet the utility of these latest SOTA models is only marginally higher than the first ChatGPT version released to the public a few years back. These things have the reasoning skills of a toddler, yet we keep fine-tuning their writing style to be more and more authoritative - this one is only missing the font and colo…
Re: Perplexity Deep Research
#114Earlier quoted context omitted.
Honestly I don‘t get why everybody is saying Gemini is far behind. Like for me Gemini Flash Thinking Experimental performs far far better then o3 mini
It varies a lot for me. One day it takes scattered documents, pasted in, and produces a flawless summary I can use to organize it all. The next, it barely manages a paragraph for detailed input. It does seem like Google is quick to respond to feedback. I never seem to run into the same problem twice.
I'm puzzled as to how that would work, when people talk about quick changes in model behavior. What exactly is being adjusted? The model has already been trained. I would think it's just randomness.
Re: Perplexity Deep Research
#115Re: Perplexity Deep Research
#116This is great. I haven't tried OpenAI or Google's Deep Research, so maybe I'm not seeing the relative crapness that others in the comments are seeing. But for the query "what made the Amiga 500 sound chip special" it wrote a fantastic and detailed article: https://www.perplexity.ai/search/what-made-the-amiga-500-sou... For me personally it was a great read and I learnt a few things I didn't know before about it.
Might have just gotten lucky, but as they say "this is the worst it will ever be"^
^ this is true and false. True in the sense that the technology will keep getting better, false in the sense that users might create websites that take advantage of the tools or that the creators might start injecting organic ads into the results
Re: Perplexity Deep Research
#117Earlier quoted context omitted.
You are a bit behind. All the "deep research" tools, and paid AI search tools in general, combine LLMs with search. When I do research on you.com it routinely searches a 100 sites. Even Google searches get Gemini'd now. I had to chuckle because your very link provides a demonstration.
> You are a bit behind. Quite the opposite. I'm familiar enough with these systems to know that asking the question "List the college majors of all Fortune 100 CEOs" is not going to get you a correct answer, Gemini and you.com included. I am happy to be proven wrong. :)
It seems like you don't understand or haven't tried their deep research tools.
Re: Perplexity Deep Research
#118Every week we get a new AI that according to the AI-goodness-benchmarks is 20% better than the old AI, yet the utility of these latest SOTA models is only marginally higher than the first ChatGPT version released to the public a few years back. These things have the reasoning skills of a toddler, yet we keep fine-tuning their writing style to be more and more authoritative - this one is only missing the font and colo…
Re: Perplexity Deep Research
#119[flagged]
Re: Perplexity Deep Research
#120Every week we get a new AI that according to the AI-goodness-benchmarks is 20% better than the old AI, yet the utility of these latest SOTA models is only marginally higher than the first ChatGPT version released to the public a few years back. These things have the reasoning skills of a toddler, yet we keep fine-tuning their writing style to be more and more authoritative - this one is only missing the font and colo…
As someone who's been using OpenAI's ChatGPT every day for work, I tested Perplexity's free Deep Research feature today and I was blown away by how good it is. It's unlike anything I've seen over at OpenAI and have tested all of their models. I have canceled my OpenAI monthly subscription.
Every time I see a comment about someone getting excited about some new AI thing, I want to go try and see for myself, but I can't think of a real world use case that is the right level of difficulty that would impress me.