Live data from Hacker News

GPTBot – OpenAI’s Web Crawler

platform.openai.com

41–50 of 327 posts

Re: GPTBot – OpenAI’s Web Crawler

#41
post #5

If you scrape my hobby website about photography, scuba diving or let's say baking or gardening which improves your model by let's say a delta of 0.00000000001 than shouldn't I get some free credits to use that model or proportionate share in the revenue stream? EDIT: scuba diving NOT scooba diving

Why? I wouldn't pay you for marginally improving my baking skills either. It is an interesting question. I would have no qualms paying for a textbook or university course for curated learning (worth noting OpenAI has paid datasets too), but paying for (or being paid for) relatively diffuse and low quality content through hobby blogs seems at odds with my expectations as an individual, and as a society we were never (…

You might not pay, but ad revenue might.

Re: GPTBot – OpenAI’s Web Crawler

#42
post #17
post #5

If you scrape my hobby website about photography, scuba diving or let's say baking or gardening which improves your model by let's say a delta of 0.00000000001 than shouldn't I get some free credits to use that model or proportionate share in the revenue stream? EDIT: scuba diving NOT scooba diving

This doesn't appear consistent with other visitors to your website. If a cafe owner uses info on your site to improve their baking, should they also be required to share their revenue with you?

Are you considering ad revenue?

Re: GPTBot – OpenAI’s Web Crawler

#43
post #9
post #5

If you scrape my hobby website about photography, scuba diving or let's say baking or gardening which improves your model by let's say a delta of 0.00000000001 than shouldn't I get some free credits to use that model or proportionate share in the revenue stream? EDIT: scuba diving NOT scooba diving

quick back-of-envelope/googling: openai is worth $29,000,000.00 you contributed 0.00000000001 punches numbers in calculator thus the value of your free credits is 0.001 cents. minus any accounting fees.

>openai is worth $29,000,000.00

You might have missed a few zeros.

Re: GPTBot – OpenAI’s Web Crawler

#44
post #5

If you scrape my hobby website about photography, scuba diving or let's say baking or gardening which improves your model by let's say a delta of 0.00000000001 than shouldn't I get some free credits to use that model or proportionate share in the revenue stream? EDIT: scuba diving NOT scooba diving

Since that delta is clearly a transformative use of your photo -- as in, the output doesn't even remotely resemble the input -- you don't have any legal claim to it, no. I'm not sure what the plaintiffs arguing otherwise are smoking if they think they can argue it isn't transformative.

Re: GPTBot – OpenAI’s Web Crawler

#45
post #5

If you scrape my hobby website about photography, scuba diving or let's say baking or gardening which improves your model by let's say a delta of 0.00000000001 than shouldn't I get some free credits to use that model or proportionate share in the revenue stream? EDIT: scuba diving NOT scooba diving

Every response so far is no i.e. the hobby website doesn't merit any compensation. A contrarian take to support the original commenter is that if the site owner had ads, i probably got him or her some increment in site visits and helped in some small way with monetization, site ranking and boosted his / her public persona, credibility. When GPT bot visits, none of that happens. Much worse - people who might have visi…

The LinkedIn case has already established the legality of scraping, so this argument falls flat too.

Re: GPTBot – OpenAI’s Web Crawler

#46
post #25
post #19

Just a random thought: While they are already earning money by scraping data, would it not be nice if they pay the site owners a certain amount of the money they earn?

If a human read a website and profited from the knowledge obtained, I don’t think we’d expect them to pay royalties to the site owner.

What about lost ad revenue?

Re: GPTBot – OpenAI’s Web Crawler

#47
post #5

If you scrape my hobby website about photography, scuba diving or let's say baking or gardening which improves your model by let's say a delta of 0.00000000001 than shouldn't I get some free credits to use that model or proportionate share in the revenue stream? EDIT: scuba diving NOT scooba diving

If your website appears on the search results of Google and they show ads next to it, aren't you entitled to that revenue too?

Re: GPTBot – OpenAI’s Web Crawler

#48
post #25
post #19

Just a random thought: While they are already earning money by scraping data, would it not be nice if they pay the site owners a certain amount of the money they earn?

If a human read a website and profited from the knowledge obtained, I don’t think we’d expect them to pay royalties to the site owner.

A human reads a website, watches/clicks on ad, buys merch, subscribes to website, sets a bookmark to a website, shares an article, invites others etc.

What will be the point of sharing knowledge or content if it will no longer be associated to an individual or organization?

Re: GPTBot – OpenAI’s Web Crawler

#49
post #16
post #5

If you scrape my hobby website about photography, scuba diving or let's say baking or gardening which improves your model by let's say a delta of 0.00000000001 than shouldn't I get some free credits to use that model or proportionate share in the revenue stream? EDIT: scuba diving NOT scooba diving

Counterpoint (not just to be annoying — I think you pose a very interesting unanswered question): If I read your hobby website about photography and use it to take 1% better pictures, do I owe you 1% of what my clients pay me? I think that probably most people would say no, assuming you could even determine that 1% in a way that both parties agreed was fair. I think generally, we have an understanding that some stuff…

Interesting point though I'd go with another analogy.

You can go to a library to borrow a book, but you can't go to the library and copy all the books for your own use.

Re: GPTBot – OpenAI’s Web Crawler

#50
post #5

If you scrape my hobby website about photography, scuba diving or let's say baking or gardening which improves your model by let's say a delta of 0.00000000001 than shouldn't I get some free credits to use that model or proportionate share in the revenue stream? EDIT: scuba diving NOT scooba diving

If your website appears on the search results of Google and they show ads next to it, aren't you entitled to that revenue too?

Google allows me to limitlessly search their index that allows me to find other pages too and in turn, they sell my attention so it is somewhat fair proposition in contrast to a wall gardened AI model being charged by per token such as GPT 4 that includes my content as well.
Post reply on HN