Earlier quoted context omitted.
Gemini was also "use us through this weird interface and also you can't if you're in the EU"; that + being far behind OpenAI and Anthropic for the past year means, they failed to reach notoriety, partly because of their own choices.
Honestly I don‘t get why everybody is saying Gemini is far behind. Like for me Gemini Flash Thinking Experimental performs far far better then o3 mini
Perplexity Deep Research
101–110 of 180 posts
Re: Perplexity Deep Research
#102That's the third product to use "Deep Research" in its name. The first was Gemini Deep Research: https://blog.google/products/gemini/google-gemini-deep-resea... - December 11th 2024 Then ChatGPT Deep Research: https://openai.com/index/introducing-deep-research/ - February 2nd 2025 Now Perplexity Deep Research: https://www.perplexity.ai/hub/blog/introducing-perplexity-de... - February 14th 2025.
It failed my first test which concerned Upside magazine. All of these deep research versions have failed to immediately surface the most famous and controversial article from that magazine, "The Pussification of Silicon Valley." When hinted, Perplexity did a fantastic job of correcting itself, the others struggled terribly. I shouldn't have to hint though, as that requires domain knowledge that the asker of a query m…
"Did you miss anything?"
"Can you fact check this?"
"Does this accurately reflect the range of opinions on the subject?"
Taking the output to another LLM with the same questions can wring out more details.
Re: Perplexity Deep Research
#103Are there good benchmarks for this type of tool? It seems not? Also, I'd compare with the output of phind (with thinking and multiple searches selected).
Re: Perplexity Deep Research
#104Unrelated question: would most people consider perplexity to have reached product market fit?
Re: Perplexity Deep Research
#105Earlier quoted context omitted.
Wild. My results are literally dozens of posts about the article. https://imgur.com/a/1hTJVkl
About the article, not any link to the article itself.
The only link I have found is a reproduction of the article[1], but I am unable to access the full text due to a paywall. I no longer have access to academic resources or library memberships that would provide access.
My Google search query was:
pussification of silicon valley inurl:upside
which returned exactly one result.I suspect the article's low visibility in standard Google searches, requiring operators like 'inurl:', might be because its PageRank is low due to insufficient backlinks.
[1] https://www.proquest.com/docview/217963807?sourcetype=Trade%...
Re: Perplexity Deep Research
#106Every week we get a new AI that according to the AI-goodness-benchmarks is 20% better than the old AI, yet the utility of these latest SOTA models is only marginally higher than the first ChatGPT version released to the public a few years back. These things have the reasoning skills of a toddler, yet we keep fine-tuning their writing style to be more and more authoritative - this one is only missing the font and colo…
Not true at all. The original ChatGPT was useless other than as a curious entertainment app. Perplexity, OTOH, has almost completely replaced Google for me now. I'm asking it dozens of questions per day, all for free because that's how cheap it is for them to run. The emergence of reliable tool use last year is what has sky-rocketed the utility of LLMs. That has made search and multi-step agents feasible, and by exte…
No, these AI companies are burning through huge amounts of cash to keep the thing running. They're competing for market share - the real question is will anyone ever pay for this? I'm not convinced they will.
Re: Perplexity Deep Research
#107can someone explain what perplexity value is ? They seem like a thin wrapper on top of big AI names, and yet i find them often mentioned as equivalent to the likes of opena ai / anthropic / etc, which build foundational models. It's very confusing.
Re: Perplexity Deep Research
#108Earlier quoted context omitted.
Not true at all. The original ChatGPT was useless other than as a curious entertainment app. Perplexity, OTOH, has almost completely replaced Google for me now. I'm asking it dozens of questions per day, all for free because that's how cheap it is for them to run. The emergence of reliable tool use last year is what has sky-rocketed the utility of LLMs. That has made search and multi-step agents feasible, and by exte…
If your goal is to replace one unreliable source of information (Google first page) with another, sure - we may be there. I'd argue the GPT 3.5 already outperformed Google for a significant number of queries. The only difference between then and now is that now the context window is large enough that we can afford to paste into the prompt what we hope are a few relevant files. Yet what's essentially "cat [62 random f…
That's not a very charitable take.
I recently quizzed Perplexity (Pro) on a niche political issue in my niche country, and it compared favorably with a special purpose-built RAG on exactly that news coverage (it was faster and more fluent, info content was the same). As I am personally familiar with these topics I was able to manually verify that both were correct.
Outside these tests I haven't used Perplexity a lot yet, but so far it does look capable of surfacing relevant and correct info.
Re: Perplexity Deep Research
#109That's the third product to use "Deep Research" in its name. The first was Gemini Deep Research: https://blog.google/products/gemini/google-gemini-deep-resea... - December 11th 2024 Then ChatGPT Deep Research: https://openai.com/index/introducing-deep-research/ - February 2nd 2025 Now Perplexity Deep Research: https://www.perplexity.ai/hub/blog/introducing-perplexity-de... - February 14th 2025.
You forgot Huggingface researchers - https://www.msn.com/en-us/news/technology/hugging-face-resea... and BTW - I post an exact same spirit comment an hour ago... So I guess Today's copycat ethics aren't solely for products- but also for comment section . LOL.
Re: Perplexity Deep Research
#110Earlier quoted context omitted.
Gemini was also "use us through this weird interface and also you can't if you're in the EU"; that + being far behind OpenAI and Anthropic for the past year means, they failed to reach notoriety, partly because of their own choices.
Honestly I don‘t get why everybody is saying Gemini is far behind. Like for me Gemini Flash Thinking Experimental performs far far better then o3 mini
Now, as we can finally access it, Google has a chance to get back into the race.