Earlier quoted context omitted.
Which results are different than Google's?
Try searching for movie name + torrent for example
Show HN: No Trash Search
81–90 of 99 posts
Re: Show HN: No Trash Search
#82Earlier quoted context omitted.
Most. Yandex is great, especially for programming searches. It generally ranks GitHub, Stack Overflow and other content-heavy sites highly. Google has been taken over by weird clones of GitHub and SO lately, Yandex has no such trash. It completely boggles my mind that the useless GitHub and SO clones rank first page on Google. Do engineers at Google not use their own product?
> Google has been taken over by weird clones of GitHub and SO lately Do you have an example search leading to a GitHub clone?
A few weeks/months ago however, while I was trying to solve an issue whith a colleague who would search using french keywords, I noticed that some websites featured on the first page of the Google results were off.
In short, they were machine-translated versions of Stack Overflow threads. And they would appear in most of the searches using french keywords.
Those websites also appeared rarely in my searches while I was using English keywords, but most of the time I never bothered opening them. But now I notice them every time.
Some examples: When searching for "wget set http proxy" on Google, the fourth result leads me to qastack.fr, and the ninth to it-swarm-fr.com, both are websites featuring scrapped and machine-translated threads from Stack Overflow.
When searching deliberately in french for "Eclipse CDT stdout ne s'affiche pas" ("Eclipse CDT stdout not displayed [in console]"), the first result leads me to askcodez.com and the fourth one to qastack.fr (askodez is the same as the other two).
I have never stumbled upon Github clones, yet, however.
Re: Show HN: No Trash Search
#83Earlier quoted context omitted.
Which results are different than Google's?
Most. Yandex is great, especially for programming searches. It generally ranks GitHub, Stack Overflow and other content-heavy sites highly. Google has been taken over by weird clones of GitHub and SO lately, Yandex has no such trash. It completely boggles my mind that the useless GitHub and SO clones rank first page on Google. Do engineers at Google not use their own product?
Re: Show HN: No Trash Search
#84Earlier quoted context omitted.
> Google has been taken over by weird clones of GitHub and SO lately Do you have an example search leading to a GitHub clone?
French is my mother tongue, but I've quickly learned during my studies that using English keywords in my STEM-related searches would simply lead me to better (and more abundant) results. A few weeks/months ago however, while I was trying to solve an issue whith a colleague who would search using french keywords, I noticed that some websites featured on the first page of the Google results were off. In short, they wer…
Re: Show HN: No Trash Search
#85Re: Show HN: No Trash Search
#86Earlier quoted context omitted.
FYI, I think this is just the case where you should prefix the submission title with “Show HN:”. Can mods update it so it shows with the others? @dang? https://news.ycombinator.com/show https://news.ycombinator.com/showhn.html
I emailed this suggestion to the mods.
Re: Show HN: No Trash Search
#87Earlier quoted context omitted.
French is my mother tongue, but I've quickly learned during my studies that using English keywords in my STEM-related searches would simply lead me to better (and more abundant) results. A few weeks/months ago however, while I was trying to solve an issue whith a colleague who would search using french keywords, I noticed that some websites featured on the first page of the Google results were off. In short, they wer…
One huge help here is uBlacklist, which has filter lists for search engine results. Of course, the Chrome version will be crippled more as Google feels the knife in its revenue artery, so FF is advised! https://github.com/iorate/uBlacklist
Re: Show HN: No Trash Search
#88Earlier quoted context omitted.
"(BTW, any good search engines these days that aren't indirectly using Google or Bing ?)" The code for Gigablast is open-source, including the crawler. I could be wrong but I do not think search.marginalia.eu nor wiby.me use Google or Bing. The comment about "hundreds of millions" is interesting. Assume hypothetically a search engline claimed to be searching millions of sites for a given query but in truth it was act…
"How would a user verify the search engine's claim about searching millions of sites was true." Search for things specifically on those pages, by very specific phrases and such. Of course you have to find them yourself first for that verification. I can say having set up some very teeny tiny websites here and there that the googlebot is hooked up to a lot of stuff. I'm not even sure how it found a couple of them as q…
A search engine can tell users some large number of sites were searched at the time of the user's query and some large number of results exist, but what if it does not allow the user to actually view all the results.
To put it another way, the question is not what Google has discovered about the www,^1 but what Google is willing to let the user search and retrieve. If retrieving the 963rd result for a common string is not allowed, then it is impossible for the user to verify that the site containing that result was searched when the user submitted her query. Even if the search produced a 963rd result, what difference does it make if the user cannot retrieve it. What is the point of the search engine locating the 963rd result if it never has to show this result to the user querying a common string.
1. What Google has discovered about the www^2 and what Google users are able to discover about the www through Google may be two different things.^3 Google has its own interests to pursue in the name of online advertising and these may conflict with users' interests. "Censorship" is one concept that often draws negative connotations but there are many more subtle forms of filtering and manipulation that are possible here, including unintentional ones.
2. The most important focus would be what is "popular".
3. Some users might care less about what is "popular". Such users would, by and large, be less interesting to an advertising company. Individual interests might become subverted in favour of "popular" interests, to the extent they conflict. An advertising company (that runs a search engine) will favour the larger audience.
Re: Show HN: No Trash Search
#89Earlier quoted context omitted.
While I can understand the appeal, restricting your search engine to only ~120 websites out of hundreds of millions (?) is basically giving up on the Web. (BTW, any good search engines these days that aren't indirectly using Google or Bing ?)
Try Mojeek https://blog.mojeek.com/2021/03/to-track-or-not-to-track.htm... Disclosure: team member. Feedback good or bad appreciated
Re: Show HN: No Trash Search
#90Earlier quoted context omitted.
Most. Yandex is great, especially for programming searches. It generally ranks GitHub, Stack Overflow and other content-heavy sites highly. Google has been taken over by weird clones of GitHub and SO lately, Yandex has no such trash. It completely boggles my mind that the useless GitHub and SO clones rank first page on Google. Do engineers at Google not use their own product?
Don't have time to mess with it right now, but does it normally return about half results in Russian or is that something my phone/browser is doing?