Earlier quoted context omitted.
You know, that actually makes a lot of sense. Recently, I attempted to bid on some long-tail keywords in AdWords for some targeted ads, but unfortunately they don’t have the “search volume” to qualify.
You mean you cannot bid on keywords with very low search volume ?
Google Memory Loss
511–520 of 552 posts
Re: Google Memory Loss
#512Earlier quoted context omitted.
I think the biggest irony is that the web allows for more adoption of long-tail movements than ever before, and Google has gotten significantly worse at turning these up. I assume this has something to do with the fact that information from the long tail is substantially less searched for than stuff within the normal bounds. Google wants you to be mainstream now. If everyone thinks the same and wants the same things…
I see quite the opposite inceintive for Google. If you are a very eccentric individual and they know those quirks, they have a huge competitive advantage in targeting ads to you vs. some bulk radio broadcast ad etc.
However, the observation of this article, and my observation as well, is that Google isn't currently capable of parsing very individual quirks. Rather, Google is able to place you into one of a number of highly conformist boxes. They don't have to understand you as an individual. They just have to 'box' you more effectively than their competitors.
There is nothing in the market or otherwise emergent in the nature of data and such categorization which fundamentally motivates Google to be able to parse anyone's quirks or understand the essence of a scene or artistic movement. If Google can gain a competitive advantage by creating a number of honeytrap doppelgangers which draw people away from the long tail and sequester them into un-creative, imitative, and highly conformist boxes, then so much the better for them.
https://www.youtube.com/watch?v=R9OHn5ZF4Uo
http://lesswrong.com/lw/l8/conjuring_an_evolution_to_serve_y...
In much the same way, I find that recommendation engines come up with annoying pale imitations of bands/musicians I like. I also wonder why authoritarianism seems to spread so effectively across social media, and why certain authoritarian movements seem to get such ready support from within Google and various social media companies. It's because, as a product, conformist/authoritarian screechers are more easily herded, replicated, categorized, and packaged than real individuals who think for themselves and apply principles.
Re: Google Memory Loss
#513Earlier quoted context omitted.
I can understand "firce" and "fprce" or even "f9rce" but "furce" is, on a standard QWERTY keyboard, two keys away so more unlikely to be a typo.
Do Google do spelling correction based on letter locality on the expected user keyboard? Never seen any corrections that would suggest that, often wondered why not.
Re: Google Memory Loss
#514Earlier quoted context omitted.
No, they don't but I know what you are referring to. Most of the time I get the result: No results found for [phrase] Results for [phrase] (without quotes) ..but the thing is that I know websites containing [phrase] exist. Often many of them and they aren't "dark web" either. Google used to be able to find them but no longer is. This gets more confusing because sometimes it does work. Namely if you look for phrases w…
I had also wondered whether its ability to match exact phrases has degraded. It would make sense, since keeping individual words on a distributed index is a lot easier than keeping a long phrase. But I had no way of telling, not having benchmarked it in the past.
Further, if they retain copies of the full text in their database they could do a filtered scan of the documents that hit on all subphrases to guarantee exact match. I could see that having too much of a performance impact at scale though.
In any case the dumbing down of Google search over the last few years is immensely frustrating to me.
Re: Google Memory Loss
#515Earlier quoted context omitted.
I start feeling like the web is being de-optimized for nerds & super-users
Nothing wrong with that. You generally shouldn't optimize applications for power users. Software should be usable.
Re: Google Memory Loss
#516Hmm. I just tried to reproduce this with old posts of my own and couldn't. I picked random phrases from five early 2006 blog posts that get basically no traffic and searched for them: "I had been playing the accordion Davy lent to Rosie during winter break" "The language they're using is not that different from the one I wrote PlayGUI to use" "I've been playing a decent amount of music lately, mostly guitar and piano…
In fairness to Google, that's always been a problem, and even in the good-old-days in which keyword-based searches were more effective, there were content aggregators that would copy the entire contents of phpBB-style bulletin boards (and USENET newsgroups) in order to rehost them and get clicks.
On the one hand, I want to say that it's precisely the sort of SEO/spammy practice that Google should be deprioritizing in search results. On the other hand, sometimes these copies/mirrors of content are the only extant copies of content when an original blog goes away. Although the motivations of the owners of these sorts of sites may not be as pure as that of archive.org, the result for the searcher is equivalent: the desired information is found even if it's only a rehosted copy.
Re: Google Memory Loss
#517Earlier quoted context omitted.
Startup idea: a service that will let you search your inbox. Aka google for searching. Seriously, this is egregious. You rely on your email provider to accurately search your inbox - some emails are important business, tax, and legal documents that are relevant for years, even decades. Or at least be fucking transparent about the fact that you are not really searching all emails. I know Gmail is a free service and in…
Thunderbird (and I'm guessing most of the offline, true email clients) has this built-in.
Well... this is why you might want it. Your data under your control. Your choice of tools.
If you're using GMail and Google decides to turn GMail to crap, well, bad luck.
Re: Google Memory Loss
#518Re: Google Memory Loss
#519Earlier quoted context omitted.
Which means the corpus is broken. But regular people rarely care about correct spelling, in my experience, and so I doubt corpus maintainers will care either...
You're imagining that people who make NLP corpora actually vet the text going into them? I dream of a world where people can be convinced to care that much. I'm not even talking about the scenario you suggest of filtering for proper word usage, I'm talking about filtering at all . The corpora used for popular word embeddings are full of weird nonsense text (in the case of word2vec) or autogenerated awfulness like spa…
Or don't use a machine learning model. I honestly don't care, just don't automatically turn a correct "its" into an incorrect "it's".
Re: Google Memory Loss
#520Earlier quoted context omitted.
That was just a random example of a phrase. I made it up on the fly.
Not sure what you expect then.