Live data from Hacker News

Replacing my best friends with an LLM trained on 500k group chat messages

izzy.co

121–130 of 371 posts

Re: Replacing my best friends with an LLM trained on 500k group chat messages

#121

Earlier quoted context omitted.

> It is a point I often make that we don't need AGI for AI to already become very disturbing in its potential use. But we don't need AI or LLMs at all for the above scenario. Companies don't currently pry into your e-mails to make hiring decisions, but they could (ignoring laws) do it if they wanted. No LLM or AI necessary. So why would the existence of AIs or LLMs change that? If they wanted to use the content of yo…

Whenever we get to see behind the corporate veil, we often find companies don't abide by laws. How many companies failed this year hiding nefarious activities? Also, what types of behavior did we get a glimpse of from the Twitter Files? Aren't there always constant lawsuits about bad behavior of companies especially around privacy? So yes, we are talking about the same behavior existing, but the concern is that they…

> Also, what types of behavior did we get a glimpse of from the Twitter Files?

Can you actually explain the types of bad behavior? The rhetorical question about The Twitter Files somehow being a groundbreaking expose of bad behavior doesn't really match anything I've seen. Most of what was cited was essentially a social media company trying to enforce their rules.

Might want to read up on the latest developments there. Several journalists have debunked a lot of the key claims in the "Twitter Files". Taibbi's part was particularly egregious, with some key numbers he used being completely wrong (e.g. claiming millions when the actual number was in the thousands, exaggerating how Twitter was using the data, etc.).

Even Taibbi and Elon have since had a falling out and Taibbi is leaving Twitter.

If Elon Musk so famously and publicly hates journalists for lying, spinning the truth, and pushing false narratives, why would he enlist journalists for "The Twitter Files"? The answer is in plain view: He wanted to take a nothingburger and use journalists to put a spin on it, then push a narrative.

Elon spent years saying that journalists can't be trusted because they're pushing narratives, so when Elon enlists a select set of journalists to push a narrative, why would you believe it's accurate?

> So yes, we are talking about the same behavior existing, but the concern is that they now get orders of magnitude more power to extend such bad behavior.

No they don't. The ultimate power is being able to read the e-mails directly. LLMs abstract that with a lower confidence model that is known to hallucinate answers when the underlying content doesn't have a satisfactory set of content.

Re: Replacing my best friends with an LLM trained on 500k group chat messages

#122

Earlier quoted context omitted.

> Adding an LLM abstraction layer doesn't make the existing laws (or social/moral pressure) go away. Isn't the "abstraction" of "the model" exactly the reason we have open court filings against stable diffusion and other models for possibly stealing artist's work in the open source domain and claiming it's legal while also being financially backed by major corporations who are then using said models for profit? Whose…

> for possibly stealing artist's work in the open source domain The provenance of the training set is key. Every LLM company so far has been extremely careful to avoid using people's private data for LLM training, and for good reason. If a company were to train an LLM exclusively on a single person's private data and then use that LLM to make decisions about that person, the intention is very clearly to access that p…

I've spoken with a lawyer about data collection in the past and I think there might be a case if you were to:

- collect thousands of people's data

- anonymize it

- then shadow correlate the data in a web

- then trace a trail through said web for each "individual"

- then train several individuals as models

- then abstract that with a model on top of those models

Now you have a legal case that it's merely an academic research into independent behaviors affecting a larger model. Even though you may have collected private data, the anonymization of it might fall under ethical data collection purposes (Meta uses this loophole for their shadow profiling).

Unfortunately, I don't think it is as cut and dry as you explained. As far as I know, these laws are already being side-stepped.

For the record, I don't like it. I think this is a bad thing. Unfortunately, it's still arguably "legal".

Re: Replacing my best friends with an LLM trained on 500k group chat messages

#123

Earlier quoted context omitted.

> Adding an LLM abstraction layer doesn't make the existing laws (or social/moral pressure) go away. Isn't the "abstraction" of "the model" exactly the reason we have open court filings against stable diffusion and other models for possibly stealing artist's work in the open source domain and claiming it's legal while also being financially backed by major corporations who are then using said models for profit? Whose…

> for possibly stealing artist's work in the open source domain The provenance of the training set is key. Every LLM company so far has been extremely careful to avoid using people's private data for LLM training, and for good reason. If a company were to train an LLM exclusively on a single person's private data and then use that LLM to make decisions about that person, the intention is very clearly to access that p…

> Every LLM company so far has been extremely careful to avoid using private people's data for LLM training

No, they haven’t. (Now, if you said “people's private data” instead of “private people's data”, you’d be, at least, less wrong.)

Re: Replacing my best friends with an LLM trained on 500k group chat messages

#125
post #51

Earlier quoted context omitted.

My interest is training a model on myself and everything I’ve learned through life so that if I die, my kids might be able to extract some value from my experience. Learning life on your own without help can be tiring and costly (both emotionally and financially), and bad advice can be worse than no advice. A guide would be helpful imho. Step 1: survive. Step 2: enable yourself to thrive. I already have boxes of pape…

This all assumes the model would give them good advice, which is sort of based on the assumption you would give them good advice, right?

If the AI could extract advice to give people from his life experience, wouldn’t that be an advanced enough AGI not to need his personal experience to begin with? It’d just analyze the inputs and dispense personalized wisdom.

Re: Replacing my best friends with an LLM trained on 500k group chat messages

#127

Earlier quoted context omitted.

Not sure if this is a joke but this in case it isn’t, no, that’s not how it works, it will just make up random plausible-ish guesses

Isn't that what humans do?

I think the closest analogy to what the above poster wants to do would be to talk to a fortune teller when your wife won't tell you something, give the fortune teller information about your wife, and then "present" the fortune teller's fortune reading stating what your wife was up to in a divorce case.

It's true that humans sometimes do things like this! It doesn't go well for them.

Re: Replacing my best friends with an LLM trained on 500k group chat messages

#129
post #49

While I love all these stories of turning your friends and loved ones into chat bots so you can talk to them forever, my brain immediately took a much darker turn because of course it did. How many emails, text messages, hangouts/gchat messages, etc, does Google have of you right now? And as part of their agreement, they can do pretty much whatever they like with those, can't they? Could Google, or any other company…

The first instance of this would most likely alienate a lot of users. What is more likely to happen is the development of new products that basically cater to social needs through mimicking real world interactions. Subscribe for 15$ a month to feel like you have an unending flow of conversations with interesting bots that mimic your friends! I am sure there is a market for this.

This product could be advertised as a way for people who are not that socially inclined to practice their social skills. Or learn other languages through fake immersion. The use cases to make this seem like a benefit are pretty limitless.

Re: Replacing my best friends with an LLM trained on 500k group chat messages

#130

I was talking to a non-tech friend about all the AI advancements lately and when she asked me what I thought the biggest risk was I said it's exactly what we all just experienced the past 3 years and realized is awful for human - prolonged social isolation. My biggest worry is that AI generated art (be it photos, music, code, etc.) and AI assistants will become so good we won't need other humans to get our social fix…

LLM bro. Just Learn (to) Love (the) Machine.
Post reply on HN