Untitled topic
1–8 of 8 posts
Re: undefined
#2Re: undefined
#3Is the model "thinking of the last word for the next line already" or is "the word that came last in the second line was already showing high p from considering only the first line"? Shouldn't we care about such a distinction?
A more cool question is whether the model is carrying a latent representation of the destination that isn’t yet reflected in the immediate token probabilities?
I don’t know. Anthropic is investigating: https://www.anthropic.com/research/tracing-thoughts-language...
Re: undefined
#4Tested a passage on Pangram and 100% AI generated btw https://www.pangram.com/history/f6b20d8e-4eb4-4bc2-b325-ad39...
Re: undefined
#5Not that I don't try, but I still get bored halfway through because it sounds like I've read the same thing before.
Re: undefined
#6It’s never useful information. It’s like asking a class of highschoolers to explain in a report why they made certain choices in an art project. The real reason is “i liked it like that” but if you ask them to produce 10 pages of fluff justifying the choices based on nothing they’ll be happy to. (except of course for that one uber diligent kid who actually thought about concepts etc before starting, sorry if that was you, my point is that the AI isn’t like that)
Re: undefined
#7All the signs of slop, short one line paragraphs. Tested a passage on Pangram and 100% AI generated btw https://www.pangram.com/history/f6b20d8e-4eb4-4bc2-b325-ad39...