Big Tech's underground race to buy AI training data
41–50 of 152 posts
Re: Big Tech's underground race to buy AI training data
#42>Rates vary by buyer and content type, but Braga said companies are generally willing to pay $1 to $2 per image, $2 to $4 per short-form video and $100 to $300 per hour of longer films. The market rate for text is $0.001 per word, she added. This is high enough that there should be a market to compensate the end users who created these
Re: Big Tech's underground race to buy AI training data
#43Re: Big Tech's underground race to buy AI training data
#44Nobody's going to mention Worldcoin?
Re: Big Tech's underground race to buy AI training data
#45>Rates vary by buyer and content type, but Braga said companies are generally willing to pay $1 to $2 per image, $2 to $4 per short-form video and $100 to $300 per hour of longer films. The market rate for text is $0.001 per word, she added. This is high enough that there should be a market to compensate the end users who created these
Re: Big Tech's underground race to buy AI training data
#46Google having so many private photos in Google Photos must be a goldmine for them.
> Google having so many private photos in Google Photos must be a goldmine for them. While true, it's META who has won that arm's race long ago in my view; hell, they just disclosed that they have private access to DMs to Netflixh [0] in a lawsuit. If you don;t think they are training their own models on this data over all their platforms you have to be a complete idiot o: Facebook, Instagram, Whatsapp. That is a muc…
Re: Big Tech's underground race to buy AI training data
#47This will be a fun reminiscence once we find out how humans are able to learn with just a tiny fraction of that data volume.
Tiny fraction... if you ignore the learning data processed by a billion years of evolution.
However, the first complex nervous systems came about in the Cambrian explosion, only about half a billion years ago. And we also don’t train LLMs by random mutation and selection, it’s a much more teleological process.
But to extend the analogy, we should be able to train a model continuously, and not have to start training from scratch for each new model. Although, maybe, that would require random mutations, and thus much more time?
Re: Big Tech's underground race to buy AI training data
#48I assume some of the more shady/no-name dashcam units with Wifi capability are uploading their video and internal microphone recordings. Distributed surveillance: The Panopitcar
Re: Big Tech's underground race to buy AI training data
#49I assume some of the more shady/no-name dashcam units with Wifi capability are uploading their video and internal microphone recordings. Distributed surveillance: The Panopitcar
Re: Big Tech's underground race to buy AI training data
#50Google having so many private photos in Google Photos must be a goldmine for them.
> Google having so many private photos in Google Photos must be a goldmine for them. While true, it's META who has won that arm's race long ago in my view; hell, they just disclosed that they have private access to DMs to Netflixh [0] in a lawsuit. If you don;t think they are training their own models on this data over all their platforms you have to be a complete idiot o: Facebook, Instagram, Whatsapp. That is a muc…
https://www.appmysite.com/blog/android-vs-ios-mobile-operati...
Random link. Can't vouch for it. But US and RoW have quite different patterns.