> Mistral OCR 3 is ideal for both high-volume enterprise pipelines and interactive document workflows. I don’t know how they can make this statement with 79% accuracy rate. For any serious use case, this is an unacceptable number. I work with scientific journals and issues like 2.9+0.5 and 29+0.5 is something we regularly run into that has us never being able to fully trust automated processes and require human verif…
Mistral OCR 3
71–80 of 137 posts
Re: Mistral OCR 3
#72From a tweet: https://x.com/i/status/2001821298109120856 > can someone help folks at Mistral find more weak baselines to add here? since they can't stomach comparing with SoTA.... > (in case y'all wanna fix it: Chandra, dots.ocr, olmOCR, MinerU, Monkey OCR, and PaddleOCR are a good start)
after clicking on your link I browsed twitter for a minute and damn that place has become weird (or maybe it always was?)
Re: Mistral OCR 3
#73Earlier quoted context omitted.
I've worked on document extraction a lot and while the tweet is too flippant for my taste, it's not wrong. Mistral is comparing itself to non-VLM computer vision services. While not necessarily what everyone needs, they are a very different beasts compared to VLM based extraction because it gives you precise bounding boxes, usually at the cost of larger "document understanding". Its failure mode are also vastly diffe…
Why not use both? I just built a pipeline for document data extraction that uses PaddleOCR, then Gemini 3 to check + fix errors. It gets close to 99.9% on extraction from financial statements finally on par with humans.
Re: Mistral OCR 3
#74Gave it a birth registry from a Portuguese locality from 1755 which my dad and I often decipher to figure out geneology and it did a terrible job. Regular Gemini Thinking can actually get 70-80% of the documents correct except lots of mistakes on given names. Chatgpt maybe understands like 50-60%. This Mistral model butchered the whole text, literally not a word was usable. To the point I think I'm doing something wr…
Re: Mistral OCR 3
#75Earlier quoted context omitted.
Where are you seeing 79% accuracy? 79% only occurs on the page as a win rate, not an accuracy
And I believe the number is 74%, compared to OCR 2. What matters is whether this is better than competition/alternatives. Of course nobody is just going to take the output as is. If you do that, that's your problem.
Re: Mistral OCR 3
#76Earlier quoted context omitted.
Where are you seeing 79% accuracy? 79% only occurs on the page as a win rate, not an accuracy
Right! I didn’t know the difference. Does it mean for 79 out of 100 documents they produce 100% accurate OCR, I doubt it. The win rate sounds like a practical approximation of accuracy here to me. If I am wildly off, I am happy to learn.
The previous version already achieved up to 99% accuracy in multiple benchmarks, already better than most OCR software.
Re: Mistral OCR 3
#77I'm reading worse performance than many OSS offerings like Paddle, MinerU, MonkeyOCR, etc: https://www.codesota.com/ocr
Re: Mistral OCR 3
#78Does it handle math expressions (those rendered from LaTeX) well? I've been looking for a good OCR model to transcribe my math textbooks into markdown (obviously ignoring the images and figures) with LaTeX as math expressions, and none of the current OCR models work reliably enough. EDIT: you can try it yourself for free at https://console.mistral.ai/build/document-ai/ocr-playground once you create a developer accoun…
I've just finished processing thousands of documents using the Gemini Pro 3 vision model and it outperformed every OCR and image model I've tested by a long shot, perfect markdown with latex for the math every time.
Re: Mistral OCR 3
#79Earlier quoted context omitted.
I'm assuming you're interested in studying Ayahuasca traditions? I recently learned that traditionally in Shipibo culture, ayahuasca was never meant to be given to "the normal mind". Instead the maestras would be the ones taking the ayahuasca in order to help guide them into diagnosing people dealing with various sicknesses. These maestras were also ranked by how many different plants they'd done a dieta on. A dieta…
Yes essentially. I've got a few resources cobbled together over the last few years but it'd be really nice to have this reference (my Spanish isn't the best, and running to the translator for a definition can be a little annoying). Also to share with fellow learners/apprentices I know. There are a couple of classes out there (which are actually geared more toward the ceremonial/icaro language, not purely conversation…
Re: Mistral OCR 3
#80Earlier quoted context omitted.
I'm assuming you're interested in studying Ayahuasca traditions? I recently learned that traditionally in Shipibo culture, ayahuasca was never meant to be given to "the normal mind". Instead the maestras would be the ones taking the ayahuasca in order to help guide them into diagnosing people dealing with various sicknesses. These maestras were also ranked by how many different plants they'd done a dieta on. A dieta…
Yes essentially. I've got a few resources cobbled together over the last few years but it'd be really nice to have this reference (my Spanish isn't the best, and running to the translator for a definition can be a little annoying). Also to share with fellow learners/apprentices I know. There are a couple of classes out there (which are actually geared more toward the ceremonial/icaro language, not purely conversation…
FYI - Lens on Android does in-place language translation including attempting to use the same/similar font that the original language is written/printed.
Unfortunately, I don't think Lens can be used in an automated batch translation mode to convert an entire book/multiple pages