They talk about using query logs to optimize their search results: >Queries performed by people, if associated to a web page, serve as even cleaner summaries than anchor text. This is because all the logic put in place by the search engine, who resolved the query with a list of web pages, and all human understanding and experience that led one to select the best page from the offered result list end up embedded in th…
A New Search Engine
11–20 of 107 posts
Re: A New Search Engine
#12> Money and Resources : We have been lucky enough to have fantastic investors, who fund and help us in our journey. Just as an FYI, this company Cliqz is owned by Hubert Burda Media, a large media conglomerate based in Germany. That doesn't necessarily inherently mean anything negative, but it's important to understand the potential underlying incentives given their marketing as such a strongly privacy oriented servi…
Re: A New Search Engine
#13They talk about using query logs to optimize their search results: >Queries performed by people, if associated to a web page, serve as even cleaner summaries than anchor text. This is because all the logic put in place by the search engine, who resolved the query with a list of web pages, and all human understanding and experience that led one to select the best page from the offered result list end up embedded in th…
Re: A New Search Engine
#14> Money and Resources : We have been lucky enough to have fantastic investors, who fund and help us in our journey. Just as an FYI, this company Cliqz is owned by Hubert Burda Media, a large media conglomerate based in Germany. That doesn't necessarily inherently mean anything negative, but it's important to understand the potential underlying incentives given their marketing as such a strongly privacy oriented servi…
Isn’t Google basically an add company (if you go by their revenue). A search engine provided by a media company or ad company. I’m curious what people think the debate is between these.
Choice is good.
Re: A New Search Engine
#15I don't know how many of you tried the engine, but there are 2 features that instantly took my attention: 1) Trackers Stats. Essentially, you can see how many and what trackers there are on the page you are about to visit. Before visiting it. 2) Page previews (I'm not sure about whether I like that)
This feature is powered by another project we run, where we measure the tracking landscape in the web (most popular domains): https://whotracks.me. Details on how that works can be found in our paper [0]. Also - we are flirting with the idea of providing a mode where the ranking is informed by the trackers in the destination site. Would love to hear your thoughts on whether you'd like smth like this.
> 2) Page previews (I'm not sure about whether I like that)
At the moment it's only a placeholder for a lengthier title and description (if available), but we are planning to use the space for rendering a short summary of the content/media in that site + similar sites in terms of content (query-relevant of course). This is more work in progress as we want to make sure content creators are on board. Again: would love to hear your thoughts on that.
Disclaimer: I work at Cliqz.
[0] WhoTracks .Me: Shedding light on the opaque world of online tracking - https://arxiv.org/abs/1804.08959
Re: A New Search Engine
#16> Money and Resources : We have been lucky enough to have fantastic investors, who fund and help us in our journey. Just as an FYI, this company Cliqz is owned by Hubert Burda Media, a large media conglomerate based in Germany. That doesn't necessarily inherently mean anything negative, but it's important to understand the potential underlying incentives given their marketing as such a strongly privacy oriented servi…
An excerpt from the 1st post of this series: "Why would a team be motivated to build another search engine? Why would Hubert Burda Media finance this over several years (they continued to back us especially in times when things got tough)?" https://0x65.dev/blog/2019-12-01/the-world-needs-cliqz-the-w...
They mostly push a narrative of privacy and censorship, when in the end the answer is probably close to "we want a piece of the pie" or "we want to be that monopoly".
Re: A New Search Engine
#17They talk about using query logs to optimize their search results: >Queries performed by people, if associated to a web page, serve as even cleaner summaries than anchor text. This is because all the logic put in place by the search engine, who resolved the query with a list of web pages, and all human understanding and experience that led one to select the best page from the offered result list end up embedded in th…
Surfacing new content in search engines is a very challenging problem. I am guessing they use a combination of social signals (twitter, facebook) popularity and domain popularity amongst other signals.
Re: A New Search Engine
#18They talk about using query logs to optimize their search results: >Queries performed by people, if associated to a web page, serve as even cleaner summaries than anchor text. This is because all the logic put in place by the search engine, who resolved the query with a list of web pages, and all human understanding and experience that led one to select the best page from the offered result list end up embedded in th…
Your point is spot on. Old pages tend to have more association to seen queries, which does not play in favor for new pages.
That said, however, there are a couple of things to consider: 1) seen queries is not the only way to create queries, we are pretty good creating synthetic queries based on the content, descriptions, etc. This queries are more noisy that the seen queries of course, but good enough. And 2)novelty, freshness and popularity are very important features on the ranking. Feel free to try out any new topic you might think of on https://beta.cliqz.com, you will see that is not only "stale" content.
Re: A New Search Engine
#19I would like to read this but I can't reach the web server. Is it just me? Had the same issue with another article from this same site a couple of days ago. Looks like everyone else is able to read it but for some reason not me. Anyone know what's going on?
Hi, Interesting, could you tell what's the error to see. Other ways you can reach the blog: If you use Tor browser can you try opening: http://cliqzdevxo33b4h6.onion/ Or if you use Beaker browser: dat://ee172d7cd9235b2cf86ea9481e8a40e48cea29c743036621edc79a4765aa0281 Disclaimer: I work for Cliqz.
This site can’t be reached0x65.dev refused to connect.
Try:
Checking the connection Checking the proxy and the firewall ERR_CONNECTION_REFUSED
This happens on Chrome, Firefox, Safari, and Opera on my Mac.
Re: A New Search Engine
#20I want to be able to open a jupyter-like notebook with the start of my search query, and the first step should be to show me the available eigencontexts, from which I can establish the gross context for my entire search. After this first click, none of the results should be about the board game or the english word--unless the relevant search results happen to include an implementation of Go the board game in Go the language.
And then when I'm done, I want to name and archive that notebook so I can return to it at a later date--whether to refresh my memory of the ultimate answer, or to continue the search.
I guess I would call this a 'research engine' instead of a search engine.