Smaller, faster, safer: running Kimi and GLM at scale
blog.cloudflare.com
Smaller, faster, safer: running Kimi and GLM at scale
1–10 of 74 posts
Re: Smaller, faster, safer: running Kimi and GLM at scale
#2I love AI, but I really hate reading it.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#3I was interested in reading this until my slop detector went off at the paragraph starting with “It's worth being precise about where the benefit comes from, because it isn't raw speed.” I love AI, but I really hate reading it.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#4However, I wish their testing were more detailed. Firstly, some model families are more sensitive to KV quantisation than others (only Kimi K2.6 was tested). Secondly, the evaluation suite they use to claim that FP8 KV quantisation is indistinguishable is noticeably lacking coding benchmarks; in long-running tasks, minor tool call errors compound over time.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#5I was interested in reading this until my slop detector went off at the paragraph starting with “It's worth being precise about where the benefit comes from, because it isn't raw speed.” I love AI, but I really hate reading it.
I have had to stop commenting this because it would end up on 50% of the posts here. I really wish we could flag prose as ai-generated on here and just filter it out.
How well it would work on this site, I'm not sure.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#6I was interested in reading this until my slop detector went off at the paragraph starting with “It's worth being precise about where the benefit comes from, because it isn't raw speed.” I love AI, but I really hate reading it.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#7Earlier quoted context omitted.
I have had to stop commenting this because it would end up on 50% of the posts here. I really wish we could flag prose as ai-generated on here and just filter it out.
LinkedIn (of all places!) announced a button for flagging this recently: https://www.linkedin.com/posts/hsrinivasan1_ai-slop-is-a-top... How well it would work on this site, I'm not sure.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#8I was interested in reading this until my slop detector went off at the paragraph starting with “It's worth being precise about where the benefit comes from, because it isn't raw speed.” I love AI, but I really hate reading it.
I have had to stop commenting this because it would end up on 50% of the posts here. I really wish we could flag prose as ai-generated on here and just filter it out.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#9Earlier quoted context omitted.
I have had to stop commenting this because it would end up on 50% of the posts here. I really wish we could flag prose as ai-generated on here and just filter it out.
LinkedIn (of all places!) announced a button for flagging this recently: https://www.linkedin.com/posts/hsrinivasan1_ai-slop-is-a-top... How well it would work on this site, I'm not sure.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#10Earlier quoted context omitted.
I have had to stop commenting this because it would end up on 50% of the posts here. I really wish we could flag prose as ai-generated on here and just filter it out.
LinkedIn (of all places!) announced a button for flagging this recently: https://www.linkedin.com/posts/hsrinivasan1_ai-slop-is-a-top... How well it would work on this site, I'm not sure.