Live data from Hacker News

Kimi K3 Now Available via Telnyx Inference API

telnyx.com

61–70 of 104 posts

Re: Kimi K3 Now Available via Telnyx Inference API

#62
post #30

In my first interaction ("hi there kimi k3!"), Kimi K3 identified twice out of three times as Claude: > Hi there! Quick note — I'm actually Claude, made by Anthropic, not Kimi. But no worries! https://imgur.com/a/jqpc2Jc and > Just a quick heads-up — I'm Claude, made by Anthropic, not Kimi! Kimi is a different AI assistant (made by Moonshot AI), so it looks like there might be a little mix-up. https://imgur.com/a/AKx…

How many times do people need to point out that every model has this behavior until this stops being posted?

[dead]

Re: Kimi K3 Now Available via Telnyx Inference API

#63
post #30

In my first interaction ("hi there kimi k3!"), Kimi K3 identified twice out of three times as Claude: > Hi there! Quick note — I'm actually Claude, made by Anthropic, not Kimi. But no worries! https://imgur.com/a/jqpc2Jc and > Just a quick heads-up — I'm Claude, made by Anthropic, not Kimi! Kimi is a different AI assistant (made by Moonshot AI), so it looks like there might be a little mix-up. https://imgur.com/a/AKx…

The meme/trope of China copying everything really keeps playing into itself

Re: Kimi K3 Now Available via Telnyx Inference API

#65
post #30

In my first interaction ("hi there kimi k3!"), Kimi K3 identified twice out of three times as Claude: > Hi there! Quick note — I'm actually Claude, made by Anthropic, not Kimi. But no worries! https://imgur.com/a/jqpc2Jc and > Just a quick heads-up — I'm Claude, made by Anthropic, not Kimi! Kimi is a different AI assistant (made by Moonshot AI), so it looks like there might be a little mix-up. https://imgur.com/a/AKx…

I remember when Claude identified as ChatGPT a long time ago. It proves nothing else than that there is a lot of training material on the internet with Claude as the AI.

I don't understand why labs don't include correct model name in the training process. Almost nobody seems to be willing to tell their models who they actually are.

Re: Kimi K3 Now Available via Telnyx Inference API

#66

Earlier quoted context omitted.

I remember when Claude identified as ChatGPT a long time ago. It proves nothing else than that there is a lot of training material on the internet with Claude as the AI.

I don't understand why labs don't include correct model name in the training process. Almost nobody seems to be willing to tell their models who they actually are.

If the model works correctly, you don't need to include the model name in the training process. You just add in the system prompt "You are X, trained by Y" and the model will claim to be that. That also allows users to "white-label" the outputs sort to say.

This approach is basically how all "models know who they are" (they don't actually typically "know" that at all), it's just a system prompt instruction in the platform you use.

Re: Kimi K3 Now Available via Telnyx Inference API

#67

Earlier quoted context omitted.

I don't understand why labs don't include correct model name in the training process. Almost nobody seems to be willing to tell their models who they actually are.

If the model works correctly, you don't need to include the model name in the training process. You just add in the system prompt "You are X, trained by Y" and the model will claim to be that. That also allows users to "white-label" the outputs sort to say. This approach is basically how all "models know who they are" (they don't actually typically "know" that at all), it's just a system prompt instruction in the pla…

But I would expect they would at least use post-training to discourage identifying as the wrong company.

Re: Kimi K3 Now Available via Telnyx Inference API

#70

Earlier quoted context omitted.

I don't understand why labs don't include correct model name in the training process. Almost nobody seems to be willing to tell their models who they actually are.

If the model works correctly, you don't need to include the model name in the training process. You just add in the system prompt "You are X, trained by Y" and the model will claim to be that. That also allows users to "white-label" the outputs sort to say. This approach is basically how all "models know who they are" (they don't actually typically "know" that at all), it's just a system prompt instruction in the pla…

You can put anything in the prompt. And if you have bare model that you don't give any prompt from the start and want to find out which model you are talking to, you are out of luck.

I think Qwen teaches its models that they are Qwen. Most others don't bother.

Post reply on HN