Live data from Hacker News

Smaller, faster, safer: running Kimi and GLM at scale

blog.cloudflare.com

41–50 of 74 posts

Re: Smaller, faster, safer: running Kimi and GLM at scale

#41

I was interested in reading this until my slop detector went off at the paragraph starting with “It's worth being precise about where the benefit comes from, because it isn't raw speed.” I love AI, but I really hate reading it.

I hope you can get over it, because before long literally everything will be written by AI. Millions of people write with this every day.

Sounds like a great way to filter out all the pointless garbage on the internet.

Re: Smaller, faster, safer: running Kimi and GLM at scale

#43
post #32

Earlier quoted context omitted.

So they don't even support K3? What's the point. K2.7 Code is practically free already

They do: Input tokens (per 1M)$3.00 Cached input tokens (per 1M)$0.30 Output tokens (per 1M)$15.00

Not sure where you're getting that from but it's not on the pricing page. Nor is K3 mentioned in the post

Re: Smaller, faster, safer: running Kimi and GLM at scale

#45

I was interested in reading this until my slop detector went off at the paragraph starting with “It's worth being precise about where the benefit comes from, because it isn't raw speed.” I love AI, but I really hate reading it.

I hope you can get over it, because before long literally everything will be written by AI. Millions of people write with this every day.

That makes human written content even more valuable to advertisers.

After all, why would advertisers want to advertise to bots? This whole sentiment of "You'all need to allow my bot postings in your group" is getting tiresome. There's plenty of places where your bot can talk to other bots, publish for other bots, etc.

Re: Smaller, faster, safer: running Kimi and GLM at scale

#48
post #5

Earlier quoted context omitted.

LinkedIn (of all places!) announced a button for flagging this recently: https://www.linkedin.com/posts/hsrinivasan1_ai-slop-is-a-top... How well it would work on this site, I'm not sure.

If it works, it’s going to be the best feature introduced by a social network in a long time. Incredible that it comes from LinkedIn.

I wonder if the realised the entire website has become the most unbearable AI slop imaginable and if they don't do anything about it they won't have any real humans reading posts, just agents trying to advertise to each other.

Re: Smaller, faster, safer: running Kimi and GLM at scale

#49
post #46

Thanks for the transpiration but this is too shallow when talking about LLM serving.

Could you suggest good resources about LLM serving? Blogs, articles...

I cannot think of any that so concentrated. Best place to me is r/localllama on reddit. Unsloth website and twitter posts is a good place too (usually in localllama too).
Post reply on HN