Live data from Hacker News

Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

lajili.com

211–220 of 392 posts

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#211
post #58

Imagine a world where the only content you see is from publishers that you trust, and that your friends trust, and their friends, to maybe 4 or 5 hops or so, and the feed was weighted by how much they are trusted by your particular social graph. If you start seeing spammy content, you downvote it, and your trust level from that part of your social graph drops, and they are less likely to be able to publish things tha…

just look at reddit to see how quickly self curation devolves into complete and utter garbage

[dead]

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#212

Earlier quoted context omitted.

So have there been lawsuits against Google or Bing by people led astray because they believe the search results are the final source of truth?

No, what's your point? Google doesn't purport to offer legal advice, unlike the AI companies that do but simultaneously disclaim any liability. You can be obtuse if you want, but it's completely disingenuous to compare companies purporting to offer legal advice with a google search.

I'm thinking about Chat GPT in the general sense as a search engine replacement. I have no knowledge of or interest in apps that offer legal advice using Chat GPT as a back end. Seems risky unless you show a lot of disclaimers.

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#213

Earlier quoted context omitted.

In b4 controversial users get their content banned by accidentally tripping the AI filter

The other problematic area with that detection is people who have disabilities and use an AI assistant to help them type.

Excellent point

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#214

Earlier quoted context omitted.

The worst part about this is that if there is another set of bots that tries to generate engagement, then the training data isn't coming from humans either. You have one set of actors spamming. And another set of actors upvoting stuff, predominantly their own but maybe also other random posts. So the resulting posts don't necessarily even cater to humans. It will be real online hellscape.

The web will transition strongly to verified identities, like we have with SSL certs. Along with filtering out people who use AI to post under a verified identity and get caught, It’s the only way to help ensure you’re reading actual human content.

And with that and stylometric analysis, online pseudonymity effectively dies.

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#215

Online dating is going to be a nightmare with chatbots flooding the dating sites catfishing everyone. User contrib sites like Reddit will be flooded with bots that keep the conversation going, but drop in sponsored mentions into things for revenue. I think a push for real, verifiable identities and digitally signing content may happen so people can attempt to wade through what is real and what is fake.

Welcome to online dating since its inception.

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#216
post #87
post #42

Earlier quoted context omitted.

IMO it's more vital than ever to fund projects like the Internet Archive. They're the only ones incentivized to maintain a snapshot of un-LLM-clouded training data of human knowledge, unclouded by the hubris of "who cares about the old stuff, we should focus our archiving on the web as it exists today" that inevitably will take hold (or already has) in big tech companies who will have laid off the vast majority of th…

The problem with the Internet Archive, which does an amazing job, is that they do an amazing job despite the problem being fundamentally intractable . Web content expands too quickly and too massively. I wonder if the answer is a network of topic-focused archives; like moving from a "Library of Alexandria" model to a modern nationwide system of libraries.

” Web content expands too quickly and too massively.”

If most of it is crap I would call not archiving it a feature.

There is a weird convoluted analogue to CERN particle detectors. They smash particles together and then image the resulting storm of particle contrails via detector that is basically a sandwhiched ccd detector (like you have in camera, but different) the size of a cathedral. Resulting in far too much data for any system to analyze or even store in the first place. Hence they need/needed to runtime filter the massive amount of particle trail signals and only pick out the critical ones.

If there is too much data you simply need to drop the parts you are fairly confident you don’t need.

There is no reason there should be only one internet archive, there might very well be parallel operations filtering a bit different things.

I guess it’s a bit odd Unesco does not already have a parallel effort.

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#217

Stop guys, stop. Be young in your mind. Be young. How are young people interpreting this? "Oh wow, I can get it to write or help edit essays" "I can use it as something to bounce ideas off of" "I can use it to take ideas from my head into the digital realm." Stop being old people.

Myopic view, "older" folks are also excited they can leverage this new tool, you can be excited for a thing while also being apprehensive about some of its uses.

i love how you point out "write or help edit essays" and cant see how that could have potentially negative effects on society. If someone can generate an essay for class in 2 minutes that's better than their writing produced in hours, why would they ever bother to improve their writing?

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#218

Earlier quoted context omitted.

The worst part about this is that if there is another set of bots that tries to generate engagement, then the training data isn't coming from humans either. You have one set of actors spamming. And another set of actors upvoting stuff, predominantly their own but maybe also other random posts. So the resulting posts don't necessarily even cater to humans. It will be real online hellscape.

The web will transition strongly to verified identities, like we have with SSL certs. Along with filtering out people who use AI to post under a verified identity and get caught, It’s the only way to help ensure you’re reading actual human content.

Sorry, you can’t comment because your certificate was revoked when you died. You didn’t die?

——

We will pay you 100 bucks to withhold filing your partners death certificate for one week and providing their certificates to us.

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#219

Earlier quoted context omitted.

It's already pretty bad with Github/SO threads. Guys will scrape threads on GH/SO and repost them to their sites, usually with a ton of ads but the post ranks higher than the original thread so it will come up first when you google an error.

How could it rank higher tho? SO has a huge domain ranking. How an arbitrary website can compete with that? I always thought it’s the opposite and platforms like SO and Medium incentivise posting there exactly via their crazy domain ranking.

They probably rank higher on long tail keywords. Usually for more in the weeds issues that don't get as much search volume

Re: Chat GPT is the birth of the real Web 3.0, and it's not going to be fun

#220

An endless loop of AI generated content that gets posted to the web as original human generated content, with LLMs getting re-trained on this content and spitting out more content that also gets re-posted, resulting in a cesspool of BS masquerading as organic knowledge. I'm old enough to remember when Google provided meaningful search results rather than just SEO spam, the problem is about to get an order of magnitud…

One solution is pretty simple: pedigree.

Divide up the 'net into trusted and untrusted sources. Make the trust ratings public. Use search tools and corpuses such as the Google Books dataset to source "knowledge" back to pre-Internet roots, when necessary. In short: bring academic reputation back and bring it back hard.

It will make for a more elitist web, but given that even without ChatGPT we've had a problem with wildfire misinformation spread in social media networks it might be a change that's a long time coming.

Post reply on HN