Live data from Hacker News

Knowledge Distillation of Black-Box Large Language Models (2024)

arxiv.org

31–35 of 35 posts

Re: Knowledge Distillation of Black-Box Large Language Models (2024)

#31
post #2

The Chinese are really going strong on destroying the American AI economy bubble. Honestly, despite the fact that I'm totally pro USA and anti China, I think we should help them crashing the American AI bubble. They are controlling everything and we can't even buy a new computer nowadays while getting no benefit from this. I wish some influential programmers stimulated coders everywhere to skip Claude and Chatgpt sub…

"anti China", why so? have you lived there?

Re: Knowledge Distillation of Black-Box Large Language Models (2024)

#32
post #10

[flagged]

Please don't post like this on HN. The guidelines make it clear we're trying for something better here. https://news.ycombinator.com/newsguidelines.html.

We detached this comment from https://news.ycombinator.com/item?id=48712718 and marked it off topic.

Re: Knowledge Distillation of Black-Box Large Language Models (2024)

#33

Earlier quoted context omitted.

And it’s a paper from Alibaba researchers, the company/lab that Anthropic called out by name.

I do not find the Anthropic allegations believable. All the results presented in these distillation papers are for very small models. In order to gain anything, Alibaba or others would need today to use the Anthropic models to improve LLMs at least one hundred times bigger than those tested in these papers. I assume that the number of queries to the teacher LLM grows superlinearly with the size of the student model,…

Interesting and related , saw on X earlier that CC was now adding some tracking etc, here a post related to how it does it and what its targeting (chinese users) https://thereallo.dev/blog/claude-code-prompt-steganography

one could also see the fable-5 getting pulled off, US govt-ant talks, etc as part of all this globally i think

which is may way to say maybe Anthropic knows this isnt true, but they still will say otherwise publicly to make this admin understand whatever they need regardless on potential security issues etc, idk im extrapolating toomuch probably

Re: Knowledge Distillation of Black-Box Large Language Models (2024)

#35
post #32
post #10

[flagged]

Please don't post like this on HN. The guidelines make it clear we're trying for something better here. https://news.ycombinator.com/newsguidelines.html . We detached this comment from https://news.ycombinator.com/item?id=48712718 and marked it off topic.

[deleted]
Post reply on HN