Live data from Hacker News

ChatGPT for Teams

openai.com

111–120 of 453 posts

Re: ChatGPT for Teams

#111

Earlier quoted context omitted.

I have 20 years of software development experience, and I couldn’t understand anything you said. Is there a dictionary for this new lingo, or am I just too mid?

He speaks very unclearly, instead of saying GPT-4-turbo he says 4.5 preview. 4.5 is invention of his. Also mixtral medium - no idea of what he means by that. Not to mention a claim that mixtral is as good as gpt-4. It’s on the quality of gpt3.5 at best, which is still amazing for an open source model, but a year behind openai

Mistral-medium is a model that mistral serves only via API since it's a prototype model. It hasn't been released yet and it's bigger than the mixtral-8x7b model

Re: ChatGPT for Teams

#112
post #13

Ok, so there are now 2 tiers where they don't use our data to train the model? The higher bandwidth is to clearly entice new customers, but the question remains, what happens to the old ChatGPT Plus users? Do their quotas get eaten up by these new teams?

Looks like the $20/month PLUS plan DOES use your data to train the model now... (they seem to have removed that "feature" from the list in the side-by-side comparison)

AFAIK, Plus has always trained on your conversation data. Enterprise and the API do not.

Re: ChatGPT for Teams

#114
post #98

Earlier quoted context omitted.

> Mistral Medium destroys the 4.5 preview. On what metrics? LMSys shows it does well but 4-Turbo is still leading the field by a wide margin. I am using 8x-7b internally for a lot of things and Mistral-7b fine-tunes for other specific applications. They're both excellent. But neither can touch GPT-4-turbo (preview) for wide-ranging needs or the strongest reasoning requirements. https://huggingface.co/spaces/lmsys/cha…

It just feels like “what LLM is better” becomes new “what GPU is better” type of talk. It’s great to find a clear winner, but at the end the gap between the leaders isn’t an order of magnitude.

These days the question is more about which LLM is second best. It’s very tight while ChatGPT 4 is in its own league.

Re: ChatGPT for Teams

#116
post #98

Earlier quoted context omitted.

> Mistral Medium destroys the 4.5 preview. On what metrics? LMSys shows it does well but 4-Turbo is still leading the field by a wide margin. I am using 8x-7b internally for a lot of things and Mistral-7b fine-tunes for other specific applications. They're both excellent. But neither can touch GPT-4-turbo (preview) for wide-ranging needs or the strongest reasoning requirements. https://huggingface.co/spaces/lmsys/cha…

It just feels like “what LLM is better” becomes new “what GPU is better” type of talk. It’s great to find a clear winner, but at the end the gap between the leaders isn’t an order of magnitude.

I think people are missing the context that the prices of even the largest LLMs trend towards $0 in the medium term. Mistral-medium is almost open source, and we are still early days

Re: ChatGPT for Teams

#117

I’ve got my stuff rigged to hit mixtral-8x7, and dolphin locally, and 3.5-turbo, and the 4-series preview all with easy comparison in emacs and stuff, and in fairness the 4.5-preview is starting to show some edge on 8x7 that had been a toss-up even two weeks ago. I’m still on the mistral-medium waiting list. Until I realized Perplexity will give you a decent amount of Mistral Medium for free through their partnership…

Was this generated by some AI? It it a parody?

Re: ChatGPT for Teams

#118

Earlier quoted context omitted.

> It's easy and cheap to just try both these days, don't take my word for which one is better. I literally use 8x-7b on my on-prem GPU cluster and have several fine tunes of 7b (which I said in the previous post). I've used mistral-medium. GPT-4-turbo is better than them all on all benchmarks, human preference, and anything that isn't biased vibes. My opinion - such that it is - is that GPT-4-turbo is by far the best…

Mistral Medium beats 4.5 on the censorship benchmark. It doesn't refuse to help with anything that could be vaguely non-PC or could potentially be used to hurt anyone in the wrong hands, including dangerously hot salsa recipes.

That's not a metric.

That's a use case.

Certainly, no one here is arguing that there are things openai refuses to allow, and given that the effectiveness of using GPT4 on them is literally zero, a sweet potato connected to a spring and keyboard will "beat" GPT-4, if that's your scoring metric.

If you want a meaningful comparison you need tasks that both tools are capable of doing, and then see how effective they are.

Claiming that mistral medium beats it is like me claiming the RenderMan beats DALLE2 at rendering 3d models; yes, technically they both generate images, but since it's not possible to use DALLE2 to render a 3d model, it's not really a meaningful comparison is it?

Re: ChatGPT for Teams

#120

Earlier quoted context omitted.

> It's easy and cheap to just try both these days, don't take my word for which one is better. I literally use 8x-7b on my on-prem GPU cluster and have several fine tunes of 7b (which I said in the previous post). I've used mistral-medium. GPT-4-turbo is better than them all on all benchmarks, human preference, and anything that isn't biased vibes. My opinion - such that it is - is that GPT-4-turbo is by far the best…

Mistral Medium beats 4.5 on the censorship benchmark. It doesn't refuse to help with anything that could be vaguely non-PC or could potentially be used to hurt anyone in the wrong hands, including dangerously hot salsa recipes.

I'm a big proponent of freedom in this space (and remain one), but Dolphin is fucking scary.

I don't have any use cases for crime in my life at the moment beyond wanting to pirate like Adobe Illustrator before signing up for an uncancelable subscription, but it will do arbitrary things within it's abilities and it's google with a grudge in terms of how to do anything you ask. I stopped wanting to know when it convinced me it could explain how to stage a coup d'etat. I'm back on mixtral-8x7b.

Post reply on HN