I was interested in reading this until my slop detector went off at the paragraph starting with “It's worth being precise about where the benefit comes from, because it isn't raw speed.” I love AI, but I really hate reading it.
I hope you can get over it, because before long literally everything will be written by AI. Millions of people write with this every day.
Smaller, faster, safer: running Kimi and GLM at scale
41–50 of 74 posts
Re: Smaller, faster, safer: running Kimi and GLM at scale
#42We let all traffic get MITM'd, now we're letting our AI conversation get tracked. Cloudflare reeks like a US Honeypot.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#43Earlier quoted context omitted.
So they don't even support K3? What's the point. K2.7 Code is practically free already
They do: Input tokens (per 1M)$3.00 Cached input tokens (per 1M)$0.30 Output tokens (per 1M)$15.00
Re: Smaller, faster, safer: running Kimi and GLM at scale
#44I think Cloudflare not providing ZDR on their inference is the biggest public indicator that Cloudlare glows. We let all traffic get MITM'd, now we're letting our AI conversation get tracked. Cloudflare reeks like a US Honeypot.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#45I was interested in reading this until my slop detector went off at the paragraph starting with “It's worth being precise about where the benefit comes from, because it isn't raw speed.” I love AI, but I really hate reading it.
I hope you can get over it, because before long literally everything will be written by AI. Millions of people write with this every day.
After all, why would advertisers want to advertise to bots? This whole sentiment of "You'all need to allow my bot postings in your group" is getting tiresome. There's plenty of places where your bot can talk to other bots, publish for other bots, etc.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#46Re: Smaller, faster, safer: running Kimi and GLM at scale
#47Thanks for the transpiration but this is too shallow when talking about LLM serving.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#48Earlier quoted context omitted.
LinkedIn (of all places!) announced a button for flagging this recently: https://www.linkedin.com/posts/hsrinivasan1_ai-slop-is-a-top... How well it would work on this site, I'm not sure.
If it works, it’s going to be the best feature introduced by a social network in a long time. Incredible that it comes from LinkedIn.
Re: Smaller, faster, safer: running Kimi and GLM at scale
#49Thanks for the transpiration but this is too shallow when talking about LLM serving.
Could you suggest good resources about LLM serving? Blogs, articles...