Mistral releases Pixtral 12B, its first multimodal model
techcrunch.com
Mistral releases Pixtral 12B, its first multimodal model
1–10 of 44 posts
Re: Mistral releases Pixtral 12B, its first multimodal model
#2Re: Mistral releases Pixtral 12B, its first multimodal model
#3Re: Mistral releases Pixtral 12B, its first multimodal model
#4I’d love to know how much money Mistral is taking in versus spending. I’m very happy for all these open weights models, but they don’t have Instagram to help pay for it. These models are expensive to build.
Re: Mistral releases Pixtral 12B, its first multimodal model
#5I’d love to know how much money Mistral is taking in versus spending. I’m very happy for all these open weights models, but they don’t have Instagram to help pay for it. These models are expensive to build.
No license with this one yet, though you can probably assume it's Apache like the others.
Re: Mistral releases Pixtral 12B, its first multimodal model
#6New Mistral AI Weights
Re: Mistral releases Pixtral 12B, its first multimodal model
#7> It’s unclear which image data Mistral might have used to develop Pixtral 12B.
The days of free web scraping especially for the richer sources of material are almost gone, with anything between technical (API restrictions) and legal (copyright) measures building deep moats. I also wonder what they trained it on. They're not Meta or Google with endless supplies of user content, or exclusive contracts with the Reddits of the internet.
Re: Mistral releases Pixtral 12B, its first multimodal model
#812B is pretty small, so I’m doubting it’ll be anywhere close to internvl2 however mistral does great work and likely this model is still useful for on device tasks
Re: Mistral releases Pixtral 12B, its first multimodal model
#9The "Mistral Pixtral multimodal model" really rolls off the tongue. > It’s unclear which image data Mistral might have used to develop Pixtral 12B. The days of free web scraping especially for the richer sources of material are almost gone, with anything between technical (API restrictions) and legal (copyright) measures building deep moats. I also wonder what they trained it on. They're not Meta or Google with endle…
My hunch is that most AI labs are already sitting on a pretty sizable collection of scraped image data - and that data from two years ago will be almost as effective as data scraped today, at least as far as image training goes.
Re: Mistral releases Pixtral 12B, its first multimodal model
#10Like writing on an ePaper tablet, exporting the PDF and feed this into this model to extract todos from notes for example.
Or what would be the SotA for this application?