Live data from Hacker News

Slack AI Training with Customer Data

slack.com

121–130 of 426 posts

Re: Slack AI Training with Customer Data

#121
post #74

In summary, you must opt-out if you want to exclude your data from global models. Incredibly confusing language since they also vaguely state that "data will not leak across workspaces". Use tools that cannot leak data not "will not".

Most, if not all SaaS software is multi-tenant, so we've been living in the "will not" world for decades now.

Re: Slack AI Training with Customer Data

#122
post #115

The gold rush for data is wild. Private companies selling us out. - Slack - Discord - Reddit - Stackoverflow Let’s just hope this data gold rush dies out faster than the web3 craze before OpenAI reaches critical mass and gets access to government server farms. Alphabet boys have server farms of domestic and foreign surveillance and intelligence. Exabytes of data [1] [1] https://en.m.wikipedia.org/wiki/Utah_Data_Cente…

I mean slack is sold off. Founders made money. For all intents and purposes it's dead software.

Re: Slack AI Training with Customer Data

#123

The incentive for first party tool providers to do this is going to be huge, whether its Slack, Google, Microsoft, or really any other SaaS tool. Ultimately, if business want to avoid getting commoditized by their vendors, they need be in control of their data, and their AI strategy. And that probably ultimately means turning off all of these small-utility-very-expensive-and-might-ruin-your-business features, and act…

It’s definitely a moral hazard (/opportunity). As a reminder, by default on windows 11 Microsoft syncs your files to their server.

Re: Slack AI Training with Customer Data

#125
post #15
post #3

Eugh. Has anyone compiled a list of companies that do this, so I can avoid them? If anyone knows of other companies training on customer data without an easy highly visible toggle opt out, please comment them below.

Synology updated this policy back in March (Happened to be a Friday afternoon). Services Data Collection Disclosure "Synology only uses the information we obtain from technical support requests to resolve your issue. After removing your personal information, we may use some of the technical details to generate bug reports if the problem was previously unknown to implement a solution for our products." "Synology utili…

Like the other poster it would be great to have a name and shame site that lists companies training on customer data

Re: Slack AI Training with Customer Data

#126

Story time. I was at a VC conference last year and if I learned nothing else there, I learned how to spell "AI". Every single exhibitor just about had their signage proudly proclaiming their capabilities in this area, but one in particular struck me. They were touting the API integrations they could offer to train their "Enterprise AI"/LLM, and among those integrations were things like M365, Slack, etc. It struck me…

What do you think chatGPT uses as training data?

The whole world’s “sh*tposting”: Reddit, blogs, and the rest of the internet.

But also books and Wikipedia and what not.

You can “smooth” all the crap out via the training procedure.

But even more, Slack can easily filter training data to, say, only posts in high-use channels.

Further, slack has other options: eg, use their customer data only for marginal fine-tuning, for example.

Or, they don’t even know their use case yet - but want to wrap their arms around your data pronto.

Re: Slack AI Training with Customer Data

#127
post #119
post #61

> We offer Customers a choice around these practices. If you want to exclude your Customer Data from helping train Slack global models, you can opt out. If you opt out, Customer Data on your workspace will only be used to improve the experience on your own workspace and you will still enjoy all of the benefits of our globally trained AI/ML models without contributing to the underlying models. Why would anyone not opt…

Because it's default opt-in, and most people won't see this announcement.

Yep, much like just about every credit card company shares your personal information BY DEFAULT with third parties unless you explicitly opt out (this includes Chase, Amex, Capital One, but likely all others).

Re: Slack AI Training with Customer Data

#128
post #43

> For any model that will be used broadly across all of our customers, we do not build or train these models in such a way that they could learn, memorise, or be able to reproduce some part of Customer Data This feels so full of subtle qualifiers and weasel words that it generates far more distrust than trust. It only refers to models used "broadly across all" customers - so if it's (a) not used "broadly" or (b) only…

Seems like time to start some slack workspaces and fill them with garbage. Maybe from Uncyclopedia (https://en.uncyclopedia.co/wiki/Main_Page)

Re: Slack AI Training with Customer Data

#129
> To develop AI/ML models, our systems analyze Customer Data (e.g. messages, content, and files) submitted to Slack as well as Other Information (including usage information)

> We have technical controls in place to prevent access. When developing AI/ML models or otherwise analyzing Customer Data, Slack can’t access the underlying content

*> you want to exclude your Customer Data from helping train Slack global models, you can opt out.

Yeah...

Re: Slack AI Training with Customer Data

#130
post #115

The gold rush for data is wild. Private companies selling us out. - Slack - Discord - Reddit - Stackoverflow Let’s just hope this data gold rush dies out faster than the web3 craze before OpenAI reaches critical mass and gets access to government server farms. Alphabet boys have server farms of domestic and foreign surveillance and intelligence. Exabytes of data [1] [1] https://en.m.wikipedia.org/wiki/Utah_Data_Cente…

Don't forget Dropbox: https://twitter.com/Werner/status/1734890651378975007
Post reply on HN