Live data from Hacker News

Ask HN: Is there room for another search engine?

news.ycombinator.com

51–60 of 200 posts

Re: Ask HN: Is there room for another search engine?

#51
Here's what I want in a search engine:

Charge me $15-25 dollars per year

Let me decide what demographic information I wish to share- make it easy for me to control and help me protect my information. Because you are charging me money you can afford it and I trust you.

Give me two search options: one, I'm only seeking information. two, I'm looking to buy. Do this for me as an advertiser: help me qualify the clicks I'm paying for

Perhaps allow me to pay per 1000 impressions (CPM) instead of per click.

By the way, I would also subscribe to a facebook that did this.

Re: Ask HN: Is there room for another search engine?

#52

Earlier quoted context omitted.

Any search crawler that ignores robots.txt is going to be blocked by site operators in a hurry.

If Google today started ignoring robots.txt, not many people would start blocking Google's crawlers, assuming they continued to do their job efficiently. robots.txt is security by obscurity at best.

People wouldn't block Google because allowing Google scraping offers a return in the form of more traffic that offsets the cost. $BRAND_NEW_ENGINE wouldn't have that advantage.

Re: Ask HN: Is there room for another search engine?

#53
If by "search engine", you mean something similar to Google/Bing then probably not.

However, if we expand the concept of "search" to something beyond text on webpages and "engine" to something beyond a linear algebra pagerank problem that weighs url links, there's room for many more competitors.

Let's say we want to search for "best restaurant":

Method #1 might be searching millions of web pages, twitter posts, newspaper archives, etc where ngram such as "best restaurant" is mentioned. That's what Google/Bing engines already do.

Method #2 might rank restaurants by collecting crowd-sourced opinions. That's what Yelp & Tripadvisor does. (Although Google also piggybacks on their data and lists yelp pages in SRP.)

Method #3 might be a company like Visa/Mastercard analyzing their billions of transactions[1] and based on actual spending amounts & frequency of a billion cardholders, they can also provide their own calculation of a "best restaurant". (I know that Visa/MC already offer limited marketing data to some entities but they don't surface that data to every day web surfers.)

The idea is that there's plenty of room for more imaginative scenarios of #2 & #3. The common theme is that Google doesn't have the data (e.g. credit-card transactions) and therefore, the new "search engines" can give fresh answers that Google algorithms can't provide. To try and boil it down to a simple question: "What interesting answers can a new engine provide that _can't_ be extracted from the text of webpages?"

Btw, I ran across some posts from a Microsoft employee (but not a Bing team member) stating his opinions on building competing search engines. https://news.ycombinator.com/item?id=7011472

[1] http://marketrealist.com/2016/10/why-visas-processing-and-in...

Re: Ask HN: Is there room for another search engine?

#54
post #40
post #24

That's a good question, and something I've spent much time on. Cuil (2008-2010) tried. I knew some of those people. It cost them about $30 million to launch a full scale search engine. They had no revenue model. In retrospect, they were hoping to be acquired by somebody. It was some ex-Google people, trying to replicate older Google technology. They had a great launch, but the system wasn't very good and traffic rapi…

7% for Bing? That's huge! (Also, how about DuckDuckGo? I've got the (admittedly gut) feeling that it should at least outperform Ask and Excite.)

Ask and Excite and all of the other small search engines are showing you Bing and Google results.

Re: Ask HN: Is there room for another search engine?

#55

Earlier quoted context omitted.

There are way more than two things that Google does wrong. Remapping my search terms into oblivion so it can pretend it's fast is the worst one. Especially when this happens to a query I've modified to quote "every" "single" "flipping" "term." I think Google is cheating, and that their usable index is much shallower than they'd have you believe. What's needed is a search engine with functional queries (as opposed to…

Another issue is spam/false matches. Why does Google return illegitimate results? Because, let me tell you, any search for "some nifty computer book pdf" returns pages upon pages of bogus links leading to ad link mazes. A crawler should be able to trivially crawl such a page, determine that no PDF is linked, and blacklist the result, but this doesn't happen. The problem is interesting, but I think you make it seem ea…

I'd argue that paywalled sites, Captcha-guarded pages, things like that are second-class content and should be treated as such. If somebody wants to dig deeper and see such results, that should be an option, but as you said, this is a tough problem to solve satisfactorily and the truth lies somewhere inbetween. An obstacle to this happening right now is that Google doesn't want to shake the advertising tree too hard. Their business model is spammy so they tolerate spamminess. That's the hardest problem of all, and one I don't think anybody's close to solving yet: How do you monetize web search without automatically wrapping results in ads? But some bright person will come up with an answer and hopefully defeat this huge conflict of interest.

Re: Ask HN: Is there room for another search engine?

#56
post #40
post #24

That's a good question, and something I've spent much time on. Cuil (2008-2010) tried. I knew some of those people. It cost them about $30 million to launch a full scale search engine. They had no revenue model. In retrospect, they were hoping to be acquired by somebody. It was some ex-Google people, trying to replicate older Google technology. They had a great launch, but the system wasn't very good and traffic rapi…

7% for Bing? That's huge! (Also, how about DuckDuckGo? I've got the (admittedly gut) feeling that it should at least outperform Ask and Excite.)

DuckDuckGo doesn't crawl/index by itself--it partners with other companies to use their search indexing. Bing is one of their primary sources of indexing, actually, although Wikipedia tells me that DDG's indexing is a compilation of "about 50 sources".

https://en.wikipedia.org/wiki/DuckDuckGo

Re: Ask HN: Is there room for another search engine?

#57

Earlier quoted context omitted.

If Google today started ignoring robots.txt, not many people would start blocking Google's crawlers, assuming they continued to do their job efficiently. robots.txt is security by obscurity at best.

People wouldn't block Google because allowing Google scraping offers a return in the form of more traffic that offsets the cost. $BRAND_NEW_ENGINE wouldn't have that advantage.

Google's crawlers don't exactly perform 2FA before they crawl a website. If impersonating Google's crawlers doesn't suit your fancy, there are all manner of ways to anonymize a crawler. So I think blocking is one of the least interesting challenges. The other side of the coin is robots.txt is not inherently adversarial, and a crawler could waste quite some time and energy crawling truly meaningless content. That, in my mind, is the interesting challenge: Obsoleting robots.txt with an intelligent (and gentle) crawler.

Re: Ask HN: Is there room for another search engine?

#58
post #40
post #24

That's a good question, and something I've spent much time on. Cuil (2008-2010) tried. I knew some of those people. It cost them about $30 million to launch a full scale search engine. They had no revenue model. In retrospect, they were hoping to be acquired by somebody. It was some ex-Google people, trying to replicate older Google technology. They had a great launch, but the system wasn't very good and traffic rapi…

7% for Bing? That's huge! (Also, how about DuckDuckGo? I've got the (admittedly gut) feeling that it should at least outperform Ask and Excite.)

DuckDuckGo didn't make the world list.[1] A US survey has them at 0.41%.[2]

[1] https://www.netmarketshare.com/search-engine-market-share.as... [2] https://www.searchenginejournal.com/august-2016-search-marke...

Re: Ask HN: Is there room for another search engine?

#59

Earlier quoted context omitted.

Those are all things Google can already do but decide not to. Which means if any startup attempts and gets a moderate success, google will of course decide to do the same. Your startup will be nothing more than an R&D department Google uses for free. Disruption happens by doing something completely different that happens to cannibalize the incumbent's business in a way that they can't change. All these little feature…

Keep in mind that it's been 20 years since Google's differantiator (PageRank) was invented. It's been about 10 years since Google's last meaningful changes to its search UI (autocomplete and "instant" search). So, while I respect Google's omniscience and omnipotence, I disagree strongly that Google/Alphabet are in a position to respond to a serious upstart competitor. Their primary response, as you hinted would be to…

There are already many "serious upstarts" that are doing just fine but will never reach the status of Google, like duckduckgo, bing, etc.

Some even say bing has better search quality than google. But that's not what matters. No matter how good the product is, if you can't get distribution it doesn't matter. Also, even if you do have distribution, it may not be the right context so it won't be effective.

If anyone is trying to really build a search engine to compete directly against Google, I can only say good luck, and prepare to be the most patient person on earth, because it won't happen fast, and probably won't even happen. Instead, while you're slowly making progress day by day with small traffic increase, you'll see some random new thing come out that had no intention of becoming a "search engine" (just like Youtube) and "disrupt" Google.

I have seen a lot of people talk about "disrupting something", and also have seen a lot of people who actually have disrupted their industry. The former are just wannabes, and the latter never said anything about "disrupting" something. They just saw an opportunity and went for it. And it ended up disrupting. Hope this makes sense.

p.s. Google doesn't operate on the original pagerank anymore, there are tons of other things going on underneath. Also they now own many innovative AND popular "search UI" both through acquisitions and their R&D. Google Maps owns a lot of the map space, Youtube owns the video search space, and so on. Look deeper and you'll find even more that you've been taking for granted.

Re: Ask HN: Is there room for another search engine?

#60
post #45
post #24

That's a good question, and something I've spent much time on. Cuil (2008-2010) tried. I knew some of those people. It cost them about $30 million to launch a full scale search engine. They had no revenue model. In retrospect, they were hoping to be acquired by somebody. It was some ex-Google people, trying to replicate older Google technology. They had a great launch, but the system wasn't very good and traffic rapi…

"Fix those two problems, and a new search engine could be better than Google. Whether anyone would notice is questionable." Yeah, I don't think many people will care about the difference between good and perfect. You might be able to find a niche in search that Google is ignoring, but you would have a hard time expanding from there into general search. In that sense Google is a bet against technology - you would inve…

The next direction in search seems to be direct question-answering. Echo (Amazon) is interesting in that, being voice only, it can't punt to a screen of search results. It has to answer the question.
Post reply on HN