Live data from Hacker News

Google Cache is fully dead

seroundtable.com

191–200 of 249 posts

Re: Google Cache is fully dead

#191

Earlier quoted context omitted.

But the arena for that fight is legislation. Weed didn't become legal through lawsuits, it became legal because laws were repealed. I hope IA prevails but it's long shot, even more with the Heritage infestation of the courts.

Everyone just ignoring bad laws and contradicting them can remove laws too. But of course this is a niche topic that would never get such broad support. A lot of people smoking weed is certainly a component for the prohibition to fail at some point. Writing mails to legislative members isn't enough if you don't have any form of leverage.

Jury nullification is the real mechanism for We The People when we don't consent to be held to laws passed by They The Wealthy/Bribed Lawmakers

It requires that people refuse plea deals and demand jury trials, and that the jury is educated on what jury nullification is but when prosecutors can't get a conviction regardless of much evidence they have of guilt the laws will get changed or at least they stop being enforced.

Re: Google Cache is fully dead

#192

> Then a couple of weeks ago, added [direct] links to the Wayback Machine Hopefully they are also making substantial donations to the Internet Archive, since they will be directing a lot of traffic into it and basically using their infrastructure as a feature on their main product... EDIT: Apparently they are collaborating but there are not much details [0] [0] https://blog.archive.org/2024/09/11/new-feature-alert-ac…

>Hopefully they are also making substantial donations to the Internet Archive, since they will be directing a lot of traffic into it and basically using their infrastructure as a feature on their main product WebArchive link is hidden so deep in the "About the source" page that vast majority of Google users won't even know that it exists. There is excellent browser extension called Web Archives[0] that hooks all majo…

No kidding:

Click a result's three dots menu. Underneath all the main call to action buttons (Visit, share, save) is a Wikipedia description of the site. Underneath that is a "More about this page" button. On this separate page is a description of the company, social media links, reviews, generic results for the company, and, finally, some 1100px down, "See previous versions on Internet Archive's Wayback Machine" in a 14px font: https://imgur.com/a/IMgVDpV

What's the ETA for this being removed due to lack of use...

Re: Google Cache is fully dead

#193

Earlier quoted context omitted.

Yeah, but no person who would worry about this would have made the IA in the first place. IA itself is a massive copyright suit waiting to happen. I know, I know, if you were them and had bought bitcoin at $10 you would have sold at precisely the top at $70k per and neither before nor after.

[flagged]

Sure, but the guy who would conceive and execute on this idea was never going to be a guy who would stop there.

Folks like this don’t aim at some point and then achieve it and stay there. They aim higher, land where they do, and continue to target the higher point. It’s how it is.

You can tell because how many of the rest of the people who would have stopped and flown under the radar have duplicated the Archive and served it without the taint of the ebook lending? Precisely zero.

Re: Google Cache is fully dead

#194

I would presume Google still has all this data. They just will not let anyone else use it. Could this be an advantage that Google can use to train their models on but others won't have access? Google wants it to be more difficult to notice rewrites? Journalists to often have found valuable information with it?

As I understand it, Google does a decent amount of rendering of a page before indexing; this a) allows it to index content loaded by JS and b) prevents some ways spammers show Google different content from users. Perhaps Google's main way of storing a page no longer matches something that can be easily served as a cache page. This might be a way to remove a legacy copy of each page and reduce storage costs.

Re: Google Cache is fully dead

#195

> Then a couple of weeks ago, added [direct] links to the Wayback Machine Hopefully they are also making substantial donations to the Internet Archive, since they will be directing a lot of traffic into it and basically using their infrastructure as a feature on their main product... EDIT: Apparently they are collaborating but there are not much details [0] [0] https://blog.archive.org/2024/09/11/new-feature-alert-ac…

>Hopefully they are also making substantial donations to the Internet Archive, since they will be directing a lot of traffic into it and basically using their infrastructure as a feature on their main product WebArchive link is hidden so deep in the "About the source" page that vast majority of Google users won't even know that it exists. There is excellent browser extension called Web Archives[0] that hooks all majo…

That’s probably a good thing, people who really do research old archived stuff will dig and find it but others who casually click won’t bring archive.org to its knees

Re: Google Cache is fully dead

#196

Google Cache was useful because you could sometimes not find a term or keyword in the web site, but it would be in the cache. Or for sites that have gone offline, or no longer have the item. "It's still in the Google Cache!" you can't say that anymore. I use Google less and less these days. What's the point when you can just ask an LLM, and it gives you an answer within seconds, with no ads? You can ask for reference…

LLM… no ads…. *For now

Re: Google Cache is fully dead

#197

Earlier quoted context omitted.

I'm not so familiar with this area but my guess is that if you turned used noarchive, Google would not cache the page at all and therefore would not be able to use the text in your page as keywords for search results. So most sites therefore did not use noarchive because it improved discoverability/SEO to allow Google to cache your site. This is just a guess though and what I always assumed to be the case. This seems…

Nah, it's not a Google thing although Google honored it. Here's a reference to the Internet Archive using it: https://archive.org/post/31561/robots-archive-noarchive-meta...

I didn't mean to say that noarchive is only a google thing. My only point was that my assumption is that until now if Google didn't cache your site its contents would not be used for SEO.

Re: Google Cache is fully dead

#198
post #135

Earlier quoted context omitted.

I disagree. I am happy the Internet Archive are fighting the draconian copy right laws that exist.

They aren't really fighting it, because they never picked a winnable battle. Rather, they overextended themselves massively in a blunder akin to just throwing themselves on their enemy's sword. They decided to go all-or-nothing on uncontrolled digital lending when there wasn't a snowball's chance in hell that the current laws would give them any wiggle room. And unsurprisingly, it will give them a mortal wound.

"Pick a winnable fight" means the internet archive does not exist. Copyright in the US is very clear cut. There is no fight to "win" without changing the law.

That means advocacy. That sometimes means civil disobedience and getting society to fight for them. You want an internet archive? We need to reform copyright law.

Re: Google Cache is fully dead

#199

Google Cache was useful because you could sometimes not find a term or keyword in the web site, but it would be in the cache. Or for sites that have gone offline, or no longer have the item. "It's still in the Google Cache!" you can't say that anymore. I use Google less and less these days. What's the point when you can just ask an LLM, and it gives you an answer within seconds, with no ads? You can ask for reference…

This still happens all the time.

* I search a keyword * I see a google result * I see the keyword IN THE PREVIEW on Google * I click on the link * No keyword

And this isn't hidden SEO spam stuff, it was literally removed. The cache doesn't match the live result.

No recourse.

Re: Google Cache is fully dead

#200
post #150

Earlier quoted context omitted.

With a Real Remote Control (that nobody owns except for the coffee table that nobody owns): Anyone involved in a group viewing can just push the pause button when that is useful for one or more members of that group. Nobody's locked-down personal pocket computer needs to be involved at all for this most mundane task. Simplicity.

At home, it just pops up for everyone on the local network so anyone can do it. Works nicely for us to be honest. I quite like it. And the remote works too. Overall, I’m quite happy with this stuff. The only thing is that if you set up with all HomePods you can have your TV audio go to them too in stereo. Very cool. And Google doesn’t have that feature.

Why not both?

Can my neighbor not stop by and watch some TV with me and have the ability to pause it?

Post reply on HN