That writing style might be a tad too tense If I got it correct (appending B from https://stolen-thoughts.com/paper.pdf is essential) they are the authors of the well-known exploit to recover readable CoT from OpenAI and Anthropic models. They use that to find hints of distillation, by running a benchmark with a SotA model, recovering the CoT, then taking the first 1% of the CoT and running the open-source model as i…
> No.
> This is major.
> Given time, maybe best to respond explaining can't due to time? but instructions expect actual work. However complexity huge; but as coding agent, need to attempt
> Maybe we can cheat ... But user may test and see still single CPU.
The smarter AI will be, the better it will be at avoiding doing actual work.
Also, can similar responses be explained with that both models were trained on a same dataset of answers to the benchmark problems?