Mistral's Shieldstral: 3B open-weights model for multimodal moderation
61–70 of 154 posts
Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#62I would be curious if this can do moderation with an arbitrary ruleset, or if it's just "that one moderation style" we already know from current big tech platforms. The kind where malicious intent is okay if the words are nice. ___ Or, rephrased: How big is the space in which you can tune this model without retraining. Is it just "we hate sex"/"we don't hate sex" "We hate violence"/"we don't hate violence" or is it _…
They got a lot of hate for not keeping up with frontier model releases, but have managed to carve out a nice business that isn't even really niche.
Before the datacenter deals their revenue was higher than xAI's
There is a whole world out there of purpose built and hosted task specific vertical llms - especially with an emphasis on cost.
Mistral, Microsoft model releases and Thinking Machines are all over this, and it's smart. Scoop up all the tasks that don't require large and expensive frontier general-purpose llms.
Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#63Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#64I would be curious if this can do moderation with an arbitrary ruleset, or if it's just "that one moderation style" we already know from current big tech platforms. The kind where malicious intent is okay if the words are nice. ___ Or, rephrased: How big is the space in which you can tune this model without retraining. Is it just "we hate sex"/"we don't hate sex" "We hate violence"/"we don't hate violence" or is it _…
Isn't Mistral a French company? Not that the French can't do cultural imperialism either, but they (the French) don't strike me as very SV.
Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#65Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#66I would be curious if this can do moderation with an arbitrary ruleset, or if it's just "that one moderation style" we already know from current big tech platforms. The kind where malicious intent is okay if the words are nice. ___ Or, rephrased: How big is the space in which you can tune this model without retraining. Is it just "we hate sex"/"we don't hate sex" "We hate violence"/"we don't hate violence" or is it _…
Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#67Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#68I would be curious if this can do moderation with an arbitrary ruleset, or if it's just "that one moderation style" we already know from current big tech platforms. The kind where malicious intent is okay if the words are nice. ___ Or, rephrased: How big is the space in which you can tune this model without retraining. Is it just "we hate sex"/"we don't hate sex" "We hate violence"/"we don't hate violence" or is it _…
Having grown a large healthcare review platform, I can attest to the success we had mapping specific policy violations to natural language is incredibly useful. At scale, patients having terrible situations and/days can write about in ways that can be deeply unhealthy for the community or the doctors reading/receiving the feedback and sometimes very threatening beyond that purposes for the community. We built a custo…
Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#69Earlier quoted context omitted.
The problem is that their performance is too far away from the latest generation of Asian models. They had kept up in the mid-range a few years ago. But this standing is sadly long gone. If you need a fast Opensource'ed LLMs you can go for EU-hosted DeepSeek or Qwen.
By this logic the Chinese should have just given up and let the American AI companies have the market because they were so far behind. I'm sure Europe has the capability to distill other people's frontier models to catch up if they wish to do so.
There definitely have been Chinese companies with models that fell behind, which is the more direct comparison.
As for the EU in general, there are not a lot of known options. There are some working on things.
The US, the EU, and China all have frontier labs that have yet to release anything.
Re: Mistral's Shieldstral: 3B open-weights model for multimodal moderation
#70I'm a bit doubtful that a black box approach like this to moderation will ever catch on.