Live data from Hacker News

Two upstart search engines are teaming up to take on Google

wired.com

11–20 of 299 posts

Re: Two upstart search engines are teaming up to take on Google

#11
This is somehow not about Perplexity.

Like many, I tried many other search engines, starting with DuckDuckGo way back when. I always ended up Googling (or !g… -ing).

Perplexity is the first one that consistently works for both code questions (what’s this error message) and local questions (where’s my nearest store X and when do they close). Now they just need to speed it up a bit - Google queries are effectively instant.

Re: Two upstart search engines are teaming up to take on Google

#12
post #4

Every new search engine I've seen was a Bing wrapper with sometimes light reranking. I understand that competing with Google was borderline impossible a decade ago. But in 2024, we have cheap compute, great OSS distributed DBs, powerful new vector search tech. Amateur search engines like Marginalia even run on consumer hardware. CommonCrawl text-only is ~100TB, and can fit on my home server. Why is no company buildin…

Brave search is doing it properly. Bought tech, but still.

Re: Two upstart search engines are teaming up to take on Google

#13
post #12
post #4

Every new search engine I've seen was a Bing wrapper with sometimes light reranking. I understand that competing with Google was borderline impossible a decade ago. But in 2024, we have cheap compute, great OSS distributed DBs, powerful new vector search tech. Amateur search engines like Marginalia even run on consumer hardware. CommonCrawl text-only is ~100TB, and can fit on my home server. Why is no company buildin…

Brave search is doing it properly. Bought tech, but still.

> Bought tech

Could you please elaborate?

Re: Two upstart search engines are teaming up to take on Google

#14
post #4

Every new search engine I've seen was a Bing wrapper with sometimes light reranking. I understand that competing with Google was borderline impossible a decade ago. But in 2024, we have cheap compute, great OSS distributed DBs, powerful new vector search tech. Amateur search engines like Marginalia even run on consumer hardware. CommonCrawl text-only is ~100TB, and can fit on my home server. Why is no company buildin…

I’m my second year into Kagi and loving it. I actually just upgraded to get Kagi Assistant (basically, cloud access to every LLM out there). But the search alone is worth every penny, and it’s built/operated fully in-house as far as I know. https://kagi.com

As much as I like kagi and wish it success, it's not a search engine from scratch. Kagi uses other search engines (Google and Bing) wraps them and does a light reranking

Re: Two upstart search engines are teaming up to take on Google

#15

Google isn't even a competitor in the search space anymore. They've been completely unusable for a decade.

A cursory glance at their market share in the search space clearly says that’s not true.

For a big site I help run, we’re getting about 8.2x the impressions on Google compared to Bing.

Re: Two upstart search engines are teaming up to take on Google

#18
post #9

Earlier quoted context omitted.

Because it is super expensive and difficult to keep an index up to date. People expect to be able to get current events, and expect search results to be updated in minutes/seconds.

Some sources update faster than others, you could index news sources hourly and low velocity sites weekly. Google does that. CommonCrawl gets 7TB/month, indexing and vectorizing that is quite manageable.

I think it is more complex than that. Common crawl does not index the whole web every month. So even if you use common crawl and just index it every month, which you could do pretty cheaply admittedly, I don't think that would lead to a good search index.

Running an index is an extremely profitable business, from multiple points of view (you can literally earn money, but also run ads, you get information you can sell, you can buy mindshare). Everybody is looking for indexes beyond Google and Bing, but there are none. If it really is as easy as indexing common crawl, then I think we'd have more indexes.

Re: Two upstart search engines are teaming up to take on Google

#19
post #12

Earlier quoted context omitted.

Brave search is doing it properly. Bought tech, but still.

> Bought tech Could you please elaborate?

Brave search is a continuation of cliqz. A German company that developed a proper search engine, with an independent index. They shut down, but the tech got sold off.

Cliqz was the first time for me that a Google alternative actually worked really well - and it, or now brave search, is what parent was asking for :)

Re: Two upstart search engines are teaming up to take on Google

#20
post #14

Earlier quoted context omitted.

I’m my second year into Kagi and loving it. I actually just upgraded to get Kagi Assistant (basically, cloud access to every LLM out there). But the search alone is worth every penny, and it’s built/operated fully in-house as far as I know. https://kagi.com

As much as I like kagi and wish it success, it's not a search engine from scratch. Kagi uses other search engines (Google and Bing) wraps them and does a light reranking

No they do not, they have their own indexes.

https://help.kagi.com/kagi/search-details/search-sources.htm...

Post reply on HN