Live data from Hacker News

Slack AI Training with Customer Data

slack.com

161–170 of 426 posts

Re: Slack AI Training with Customer Data

#161
post #155

I wonder how many people that are really mad about these guys or SE using their professional output to train models thought commercial artists were just being whiny sore losers when Deviant Art, Adobe, OpenAI, Stability, et al did it to them.

squarely in the former camp. there's something deeply abhorrent about creating a place that encourages people to share and build and collaborate, then turning around and using their creative output to put more money in shareholder pockets. i deleted my reddit and github accounts when they decided the millions of dollars per month they're receiving from their users wasn't enough. don't have the power to move our shop…

Yeah I haven't put a new codebease on GH in years. It's kind of a PITA hosting my own gitea server for personal projects but letting MS copy my work to help make my professional skillset less valuable is far less palatable.

Companies doing this would make me much less angry if they used an opt-in model only for future data. I didn't have a crystal ball and I don't have a time machine, so I simply can't stop these companies from using my work for their gain.

Re: Slack AI Training with Customer Data

#162
post #115

The gold rush for data is wild. Private companies selling us out. - Slack - Discord - Reddit - Stackoverflow Let’s just hope this data gold rush dies out faster than the web3 craze before OpenAI reaches critical mass and gets access to government server farms. Alphabet boys have server farms of domestic and foreign surveillance and intelligence. Exabytes of data [1] [1] https://en.m.wikipedia.org/wiki/Utah_Data_Cente…

I mean slack is sold off. Founders made money. For all intents and purposes it's dead software.

It's owned by Salesforce;n if they stop growing it, it'll go the way of Heroku - and that breach didn't go well.

Re: Slack AI Training with Customer Data

#163
post #127
post #119

Earlier quoted context omitted.

Because it's default opt-in, and most people won't see this announcement.

Yep, much like just about every credit card company shares your personal information BY DEFAULT with third parties unless you explicitly opt out (this includes Chase, Amex, Capital One, but likely all others).

how do you opt out of these? I do share my data with rocket money though since there's no good alternatives :(

Re: Slack AI Training with Customer Data

#164
post #89
post #61

> We offer Customers a choice around these practices. If you want to exclude your Customer Data from helping train Slack global models, you can opt out. If you opt out, Customer Data on your workspace will only be used to improve the experience on your own workspace and you will still enjoy all of the benefits of our globally trained AI/ML models without contributing to the underlying models. Why would anyone not opt…

> Why would anyone not opt-out? Because you might actually want to have the best possible global models ? Think of "not opting out" as "helping them build a better product". You are already paying for that product, if there is anything you can do, for free and without any additional time investment on your side that makes their next release better, why not do it ? You gain a better product for the same price, they ge…

[flagged]

Re: Slack AI Training with Customer Data

#165
post #156

Earlier quoted context omitted.

> Think of "not opting out" as "helping them build a better product" I feel like someone would only have this opinion if they've never ever dealt with any in the tech industry, or capitalist, in their entire life. So like 8-19 year olds? Except even they seem to understand that the profit absolutist goals undermine everything. This idea has the same smell as "We're a family" company meetings.

[flagged]

[flagged]

Re: Slack AI Training with Customer Data

#166
post #8
post #6

I'm confused about this statement: "When developing AI/ML models or otherwise analyzing Customer Data, Slack can’t access the underlying content. We have various technical measures preventing this from occurring" "Can't" is a strong word. I'm curious how an AI model could access data, but Slack, Inc itself couldn't. I suspect they mean "doesn't" instead of "can't", unless I'm missing something.

Every company that promises "end-to-end encryption" is just pinky-swearing to you also. Like Telegram or WhatsApp

Telegram client is open source, so you can see what exactly happens there when you enable E2EE.

Re: Slack AI Training with Customer Data

#167
post #68

Earlier quoted context omitted.

This reminds me of a company called C3.ai which claims in its advertising to eliminate hallucations using any LLM. OpenAI, Mistral, and others at the forefront of this field can't manage this, but a wrapper can?? Hmm...

Ah yes, the stock everyone believed in and thought it would reach the moon during 2020.

[deleted]

Re: Slack AI Training with Customer Data

#169
post #128
post #43

> For any model that will be used broadly across all of our customers, we do not build or train these models in such a way that they could learn, memorise, or be able to reproduce some part of Customer Data This feels so full of subtle qualifiers and weasel words that it generates far more distrust than trust. It only refers to models used "broadly across all" customers - so if it's (a) not used "broadly" or (b) only…

Seems like time to start some slack workspaces and fill them with garbage. Maybe from Uncyclopedia ( https://en.uncyclopedia.co/wiki/Main_Page )

The Riders of the Lost Kek dataset is an excellent candidate https://arxiv.org/abs/2001.07487

Re: Slack AI Training with Customer Data

#170
> We offer Customers a choice around these practices. If you want to exclude your Customer Data from helping train Slack global models, you can opt out. If you opt out, Customer Data on your workspace will only be used to improve the experience on your own workspace and you will still enjoy all of the benefits of our globally trained AI/ML models without contributing to the underlying models.

Sick and tired of these default opt in explicit opt out legalese.

The default should be opt out.

Just stop using my data.

Post reply on HN