Live data from Hacker News

Mistral releases Pixtral 12B, its first multimodal model

techcrunch.com

1–10 of 44 posts

Re: Mistral releases Pixtral 12B, its first multimodal model

#4
post #2

I’d love to know how much money Mistral is taking in versus spending. I’m very happy for all these open weights models, but they don’t have Instagram to help pay for it. These models are expensive to build.

No license with this one yet, though you can probably assume it's Apache like the others.

Re: Mistral releases Pixtral 12B, its first multimodal model

#5
post #2

I’d love to know how much money Mistral is taking in versus spending. I’m very happy for all these open weights models, but they don’t have Instagram to help pay for it. These models are expensive to build.

No license with this one yet, though you can probably assume it's Apache like the others.

The article says they confirmed it's Apache via email

Re: Mistral releases Pixtral 12B, its first multimodal model

#7
The "Mistral Pixtral multimodal model" really rolls off the tongue.

> It’s unclear which image data Mistral might have used to develop Pixtral 12B.

The days of free web scraping especially for the richer sources of material are almost gone, with anything between technical (API restrictions) and legal (copyright) measures building deep moats. I also wonder what they trained it on. They're not Meta or Google with endless supplies of user content, or exclusive contracts with the Reddits of the internet.

Re: Mistral releases Pixtral 12B, its first multimodal model

#8
post #3

12B is pretty small, so I’m doubting it’ll be anywhere close to internvl2 however mistral does great work and likely this model is still useful for on device tasks

It appears to be slightly worse than Qwen2VL 7B, a model almost half it's size, if you look at the Qwen's official benchmarks instead of Mistral's.

https://xcancel.com/_philschmid/status/1833954941624615151

Re: Mistral releases Pixtral 12B, its first multimodal model

#9
post #7

The "Mistral Pixtral multimodal model" really rolls off the tongue. > It’s unclear which image data Mistral might have used to develop Pixtral 12B. The days of free web scraping especially for the richer sources of material are almost gone, with anything between technical (API restrictions) and legal (copyright) measures building deep moats. I also wonder what they trained it on. They're not Meta or Google with endle…

What do you mean by copyright measures? Has anything changed on that front in the last two years?

My hunch is that most AI labs are already sitting on a pretty sizable collection of scraped image data - and that data from two years ago will be almost as effective as data scraped today, at least as far as image training goes.

Post reply on HN