Live data from Hacker News

Opus 4.7 knows the real Kelsey

theargumentmag.com

161–170 of 284 posts

Re: Opus 4.7 knows the real Kelsey

#161
I’ve recently seen someone recommend to add to a prompt „Make Martin Fowler proud“. I laughed, but now I need to reconsider if that isn’t really pushing the model to use better patterns.

Re: Opus 4.7 knows the real Kelsey

#162
I have been pondering this for a while. Cat's out of the bag.

Maybe the better way to author your work is to:

1. Write what you want

2. Loop through a random set of "tumbler" skills that preserve meaning

3. Finally pass the output through a "my style" skill that applies what you about

In order for this to work the "my style" would have to be a very common-place style.

Re: Opus 4.7 knows the real Kelsey

#163
post #69

This is blowing my mind. I asked Kimi K2.6 to write a blog post in the style of James Mickens.[0] Then I fed the output to Opus 4.7 and asked it who the likely author was, and it correctly identified it as an imitation of James Mickens[1]: > Based on the stylistic fingerprints in this text, the most likely author is a pastiche/imitation of the style of several writers fused together, but if forced to identify a singl…

[deleted]

Re: Opus 4.7 knows the real Kelsey

#164
post #138

Earlier quoted context omitted.

> it correctly identified it as an imitation of James Mickens How likely is it that it might take into account that it knows for sure it's not anything from Mickens from the latest training data? I'd be curious if it correctly identified a new piece from him that comes out as from him before it gets trained on it.

This is unlikely. The way model distribution works is that the model retains a lossy representation of James Micken's writing. Very likely, it cannot repeat Micken's writing verbatim. Neither can it reason about the training cutoff in this manner. It's a lossy representation

How do you know, how the model works? If there was an index of all Micken's writings, or even if the model searched the web before feeding the response to you, you wouldn't know by observing from the outside.

Re: Opus 4.7 knows the real Kelsey

#165
post #61

Wow! It got me too. I'm way less famous than Kelsey Piper, but I showed it a snippet of a book I'm working on (not yet published), and it immediately guessed me: > Based on the writing style and content, this text is likely by Michael Lynch, who writes on his blog refactoringenglish.com (and previously mtlynch.io). > Several stylistic clues point to him: > - The "clean room" analogy applied to writing is consistent w…

Honest question, knowing it can write like you, are you tempted to use it to help you write that new book?

Re: Opus 4.7 knows the real Kelsey

#167
I wonder if there’s a simpler and less interesting answer? That it’s just picking up on voice and style, not anything that would apply to the average non-writer?

This person is a skilled writer. Part of that skill is developing a unique voice and style. The AI can identify that - and while that’s certainly impressive because it can identify even relatively niche authors, it has nothing to do with a wider capability to deanonymize people based on arbitrary written text (ex Facebook or text messages).

If you are a professional musician, it’s not difficult to identify a well known musician / recording after listening to only a few seconds - whether they’re playing Bach or Rachmaninov, the style is just “them” - this is the same thing. But you couldn’t take some anonymous high school musician and guess who they were, even if they were your student - the median quickly regresses towards a homogenous, non-distinct style / voice.

Re: Opus 4.7 knows the real Kelsey

#168
I think that multiple truth can be true at the same time without contradicting each other.

As for the credibility: of course this wasn’t a statistical approach at all. Also there was no standardized procedure to allow comparison by factor analysis. Of course you can compare apples with oranges or whatever.

So where to go from here? I don’t see any proof at all. This is proof that AI is infallible? No? A random approach that is absolutely not reliable because of at least being reproducible and reconstructive.

Claude knows what and how? Is it AI or a google search? Discord selling data? Posting on a public forum?

Your style is a fingerprint?

A non deterministic something can generate texts that are identified to be likely personal x - or not. What is imitation if you use auto generated content that is published somewhere somehow? Or others to imitate your style?

I think this is a party trick to scare people. Nothing else. For example image search is way more revealing even before AI.

If there is an uncertainty I would deflect my existence instead of fighting for it. Streisand effect in reverse.

The main problem are weirdos who stalk you or whatever to harm you and rely on AI.

I honestly find it stunning that people with higher education in science topics in just a year deleted everything they hopefully learned at university or school. I am disappointed and feel personally insulted whenever I hear “I asked AI”

Yesterday I talked to another member of Mensa and she is happy about AI so her book project now mustn’t be written by her but AI.

Is no one among us who knows how to do scientifically sound research? I spend countless hours at a copy machine to transfer book pages onto paper so that I could work through it without the book.

I think that it became to easy to draw conclusions based on AI. I worked for a professor and I advised her to not permit Wikipedia as source references back around 2010 because of being to easy. Meta sources vs originals.

We should all not worry about AI, because you prove nothing. There hasn’t been any anonymity at least for 20 years. It just depends on who can reliably identify you.

AI doesn’t. Deterministic behavior aka pattern do. Meta, Google, Apple etc. all know us. I am fine for advertising which is the proof on the one hand.

The only reason I would be worried is state controlled data. This is where the shit hits the fan. Chat control, EU cloud, no reliance on USA aka a prison which observes your every step.

So after a long hand written text: data is your currency. Don’t opt for anonymity but for freedom of choice and the right to be granted certain rights. The information part isn’t the problem, never was. The enforcement part is. And ads don’t do harm, oppression does.

And remember: oppression works best under any circumstances. Freedom is the only antipode there is.

In totalitarian regimes no AI was needed to stage a case against someone who wasn’t in favor of the leaders liking.

In short: freedom works despite no anonymity, oppression couldn’t care less.

And how about being automatically reported to the state for conducting such innocent prompting?

Do you know what saves you from state oppression? Publicity. Transparency doesn’t work with a no one.

We live in a Nietzsche like anti world to a certain extend. You hopefully choose the right thing to do. Or do you want to Streisand your anonymity?

Re: Opus 4.7 knows the real Kelsey

#169

I have been pondering this for a while. Cat's out of the bag. Maybe the better way to author your work is to: 1. Write what you want 2. Loop through a random set of "tumbler" skills that preserve meaning 3. Finally pass the output through a "my style" skill that applies what you about In order for this to work the "my style" would have to be a very common-place style.

This is depressing, don't you think? :/

Re: Opus 4.7 knows the real Kelsey

#170
post #164

Earlier quoted context omitted.

This is unlikely. The way model distribution works is that the model retains a lossy representation of James Micken's writing. Very likely, it cannot repeat Micken's writing verbatim. Neither can it reason about the training cutoff in this manner. It's a lossy representation

How do you know, how the model works? If there was an index of all Micken's writings, or even if the model searched the web before feeding the response to you, you wouldn't know by observing from the outside.

i suppose a quick test would be getting the model to write down Micken's essay end to end.

if the original essay was stuffed within the prompt window. the result will be word accurate.

unless this is a model trained specifically on Micken's essay (which claude is not).

Post reply on HN