Ask HN: Can we create a new internet where search engines are irrelevant?
251–260 of 395 posts
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#252Indexing information is a political problem as much as a technical one. Ultimately there will always be people who will put more effort into getting their information known than others. These people would game whatever technical solution exists.
Thats true. What if you could artificially limit the amount of effort someone can put in to getting content out there? Or even make it known to the consumer how/why the content is ranked highly?
However, creating rules transitions the contention point to who makes the rules. If you think that my algorithm will rank my sources better than your sources, you may be less interested in my algorithm regardless of its technical merits.
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#253Yes, it was called Yahoo and it did a good job of cataloging the internet when hundreds of sites were added per week: https://web.archive.org/web/19961227005023/http://www2.yahoo... I'm old enough to remember sorting sites by new to see what new URLs were being created, and getting to that bottom of that list within a few minutes. Google and search was a natural response to solving that problem as the number of sites…
I used Yahoo back in those days, and it literally proved the point that hand-cataloging the internet wasn't tractable, at least not the way Yahoo tried to do it. There was just too much volume. It was wonderful to have things so carefully organized, but it took months for them to add sites. Their backlog was enormous. Their failure to keep up is basically what pushed people to an automated approach, i.e. the search e…
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#254Indexing information is a political problem as much as a technical one. Ultimately there will always be people who will put more effort into getting their information known than others. These people would game whatever technical solution exists.
Thats true. What if you could artificially limit the amount of effort someone can put in to getting content out there? Or even make it known to the consumer how/why the content is ranked highly?
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#255That was what the early internet was like (I was there). People built indexes by hand, lists of pages on certain topics. There was the Gopher protocol that was supposed to help with finding things. But this was all top-down stuff, the first indexing/crawling search engines were bottom-up and it worked so much better. And for a while we had an ecosystem of different search engines until Google came along, was genuinel…
In the very early days, you didn't need a search engine because there weren't that many web sites and you knew most of the main ones anyway (or later on had them in your own hotlists in Mosaic). Nowadays you need a search because there is so much content. The problem is that the amount of content and the size of the potential user base are so large that is is impossible to offer search as a free service, i.e. it has…
Maybe we'll see advent of specialised paid search engines SaaSs with authentic and independent content authors like professional blogs.
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#256Earlier quoted context omitted.
The problem is that current search engines are indexing what is essentially a stack of random books thrown together by anonymous library goers. Before being able to guide readers to books, librarians have to the following non-trivial tasks over the entire collection: - identify the book's theme - measure the quality of the information - determine authenticity / malicious content - remember the position of the book in…
>I welcome any attempts to build a more organized internet. I don't think the communal book pile approach is scaling very well. Let me know if I misunderstand your comment but to me, this has already been tried. Yahoo's founders originally tried to "organize" the internet like a good librarian. Yahoo in 1994 was originally called, "Jerry and David's Guide to the World Wide Web" [0] with hierarchical directories to cu…
The underlying goal: "Get a user the information they want when they don't know where it lives" isn't really going to be helped by a non-searchable directory of millions of sites.
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#257Indexing isn't the source of problems. You can index in an objective manner. A new architecture for the web doesn't need to eliminate indexing.
Ranking is where it gets controversial. When you rank, you pick winners and losers. Hopefully based on some useful metric, but the devil is in the details on that.
The thing is, I don't think you can eliminate ranking. Whatever kind of site(s) you're seeking, you are starting with some information that identifies the set of sites that might be what you're looking for. That set might contain 10,000 sites, so you need a way to push the "best" ones to the top of the list.
Even if you go with a different model than keywords, you still need ranking. Suppose you create a browsable hierarchy of categories instead. Within each category, there are still going to be multiple sites.
So it seems to me the key issue isn't ranking and indexing, it's who controls the ranking and how it's defined. Any improved system is going to need an answer for how to do it.
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#258Earlier quoted context omitted.
I think heavy reliance on human language (and its ambiguity) is one of the main problems. Maybe personal whitelist/blacklist for domains and authors could improve things. Sort of "Web of trust" but done properly. Not completely without search engines, but for example, if every website was responsible for maintaining it's own index, we could effectively run our own search engines after initialising "base" trusted webs…
This would be ludicrously easy to game. Crowdsourcing would also be ludicrously easy to game. The problem isn't solvable without a good AI content scraper. The scraper/indexer either has to be centralised - an international resource run independently of countries, corporations, and paid interest groups - or it has be an impossible-to-game distributed resource. The former is hugely challenging politically, because the…
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#259Yes, it was called Yahoo and it did a good job of cataloging the internet when hundreds of sites were added per week: https://web.archive.org/web/19961227005023/http://www2.yahoo... I'm old enough to remember sorting sites by new to see what new URLs were being created, and getting to that bottom of that list within a few minutes. Google and search was a natural response to solving that problem as the number of sites…
Re: Ask HN: Can we create a new internet where search engines are irrelevant?
#260Earlier quoted context omitted.
The problem with git is countering nefarious forces. The blockchain is better in that regard because the consensus algorithm can be used to verify that the listings are legitimate.
content change signed by creators private key, otherwise merge is rejected? or, wiki approach...