Earlier quoted context omitted.
> This is because revealing the raw reasoning exposes exactly how the AI processes information. These companies spend in huge amounts on R&D to develop a thinking process that is superior to their competition. Exposing those thinking mechanics to competitors would completely defeat the purpose of their spending. They simply won't do it. It's like you telling your exact location to someone who is trying to hunt you do…
Not sure if anyone remembers the brief 12ish hour period when the very first “reasoning” ChatGPT model went public, but it provided credible evidence for this. Before the massive nerf (showing summaries and suppressing certain aspects of reasoning) you would literally see reasoning text appearing on your screen like “while xyz is true, these facts may be seen as supporting hateful rhetoric or a conspiracy theory whic…
The text in Claude Code’s “Extended Thinking” output
181–190 of 248 posts
Re: The text in Claude Code’s “Extended Thinking” output
#182I do miss the days when reasoning was visible. Another point for open source models!
Re: The text in Claude Code’s “Extended Thinking” output
#183> I’m underwhelmed by how Anthropic is presenting the behavior of their application. If you ever need a record of the logic a used by YOUR AGENT during a session. Nope, not your agent, if you're not running it locally. You just get to use it in whatever way they allow (also see the whole OpenClaw backlash and claude -p changes), unless there'd be regulation and laws around this (which there aren't and would be lobbie…
Proprietary technology is fun /s
What a waste of time
Re: The text in Claude Code’s “Extended Thinking” output
#184Yep, its basically a scam to charge you more tokens and provide less compute. You cant even guarantee WHAT model you get. Or if they downgrade you. Or if you 'offend corporate sensibilities' and they misdirect or lie. The only way to get good returns on a model is to run it yourself. Quit paying for corporate bullshit.
Re: The text in Claude Code’s “Extended Thinking” output
#185I feel like I get a lot of what this article presents as "hidden" by using this process: - "Read `description` and create a specification, implementation guide, and checklist." - "Ask clarifying questions. If any of those questions has a clear best recommendation, please select that yourself and record that in "autorecommendations.md". - "Have codex and antigravity review each of these and work to consensus." These a…
Re: The text in Claude Code’s “Extended Thinking” output
#186> You've provided the current rewritten thinking and the guidelines, but I don't see the "next thinking" content that I should be rewriting. Could you provide the next thinking that needs to be rewritten?
These sentences are completely unrelated to the actual conservation
Re: The text in Claude Code’s “Extended Thinking” output
#187this is really really not that bad at all
Re: The text in Claude Code’s “Extended Thinking” output
#188Earlier quoted context omitted.
> Proper distillation requires access to the logits Why do you need logits? Can't you just train on cross-entropy loss of the model against the hard decision, like you do in regular pretraining? There are definitely current-gen open-weight models (Step 3.7 Flash is one) that refer to themselves as an OpenAI model in CoT, but not in the final response.
How do I get that loss, though, without the softmax inputs?
Re: The text in Claude Code’s “Extended Thinking” output
#189Earlier quoted context omitted.
Weirdly pleasant, if minor, signal of human authorship
I'm convinced this "signal" has already been hijacked. Maybe a Baader-Meinhof phenomenon, but I've noticed more and more egregious spelling errors that make little sense from a human perspective. Hop into whatever chatbot you'd like and ask it to "write a paragraph with subtle misspellings on long but common words", and you'll notice misspellings that just feel wrong, because they don't map to a clear misunderstandin…
It's the general (lazy) usage of default model outputs that are still too clean.
It's pretty trivial to ask Haiku to "add cool kid no-caps and occasionally mix up 'their/there/they're' for authenticity"