Live data from Hacker News

Knowledge Should Not Be Gated

formaly.io

1–10 of 57 posts

Re: Knowledge Should Not Be Gated

#2
Yep, knowledge should not be gated:

Imagine Google search without any links or sources named

This is the “modern” AI chatbot:

It never mentions the training data it used, in fact has no idea what it used (often FB, Reddit and partisan websites)

Update: I added the reply about after the fact Googling chatbots do - it’s different

Re: Knowledge Should Not Be Gated

#3
Sadly it has been during most of human history. I think the establishment resents the masses becoming over educated. The 1990s internet had a wealth of views and information on it. Now you can only access approved sources via search engines thanks to scaremongering, and have CloudFlare monitoring everything you do.

Re: Knowledge Should Not Be Gated

#4
post #2

Yep, knowledge should not be gated: Imagine Google search without any links or sources named This is the “modern” AI chatbot: It never mentions the training data it used, in fact has no idea what it used (often FB, Reddit and partisan websites) Update: I added the reply about after the fact Googling chatbots do - it’s different

Secifically in Google AI Overview I always see links to sites where the information is sourced from.

Or at least some of the sites, if the same info is sourced from 100 pages then it only shows 2 or 3, maybe the ones with the biggest PageRanks.

Re: Knowledge Should Not Be Gated

#6
Sdks/libs, especially open source sdks, were never about gated knowledge. They were about the providing company making it as easy as possible for you to integrate. You would not need to know the idiosyncrasies behind api retries, paging, rate limits, auth flow, and on and on. The third party developers needed a resource, they call a method and get it. Open source libraries especially are about pooling knowledge, not gating it. This is propaganda for pooling that knowledge inside a service you have to pay to use, and instead of developers all using and improving the same codebase together, they have to spend money to rewrite the same code repeatedly. This is AI companies further trying to undercut open source because it’s free.

Re: Knowledge Should Not Be Gated

#8

It seems beyond naive, rather malicious, to upload any useful private data to SaaS LLMs. Like, you are letting them data mine your business. Why are corporations not panicing over this?

Most corporations likely have zero data retention agreements with LLM providers, at least for API usage.

(Sure, you could be sceptical on whether the LLM provider is upholding that, but I personally do trust them. The trust betrayal if ZDR wasn't actually ZDR would be too great and commercially damaging for them to lie.)

Re: Knowledge Should Not Be Gated

#9

It seems beyond naive, rather malicious, to upload any useful private data to SaaS LLMs. Like, you are letting them data mine your business. Why are corporations not panicing over this?

because corporations are using providers with ZDR in the contract. If OAI or any of the cloud providers violate this they're getting sued to oblivion.
Post reply on HN