Live data from Hacker News

Mistral OCR 3

mistral.ai

111–120 of 137 posts

Re: Mistral OCR 3

#111
post #94

My current holy grail is my attempt to convert a Shipibo (an indigenous Peruvian language)-to-Spanish dictionary into a Shipibo-to-English dictionary. The pdf I have (available freely on archive.org) isn't a great scan (though I think it'd be a heck of a lot easier than some of the handwritten examples they show). Layout (2-columns) along with header/footers can cause some headaches, but it is all Latin script. This…

I applaud your efforts, but that seems difficult to me. There's so much nuance in language, and the original spanish translation would even be dependent upon locale-destination of the original dictionary. Which would also be time based, as language changes over time. And that translation is likely only a rough approximation, as words don't often translate directly. To add in an extra layer (spanish -> english) seems…

Thanks for the suggestions, I do appreciate it. I was being pretty brief with my post but I really have spent a lot of time and tried this from a number of angles. I've had good luck with non-LLM tools to do the initial OCR, but it's not context aware especially about column/page breaks (like I mentioned it's kind of a dirty scan, and if the breaks happen on a Shipibo part it barfs a bit. Good for a rough search at least).

I would love to create a json version of it that would essentially have a bunch of fields for each word (Shipibo/Spanish/English word/definition/example, type of word, etc). It's further complicated by how words can be modified in Shipibo (it's actually a very technical language- words can have any number of prefixes and suffixes tagged on to change their meaning and their precision. In their "icaros", the healing songs they sing in ceremony, the most technical use of the language is considered to be the most beautiful. Essentially poetry from their "medical" jargon).

I've done some human-in-the-loop attempts but still come up short in one way or another (I end up getting frustrated and throwing my hands up after seeing how much time I dump on it). So I figure this will remain a good test as the tools (and my prompting abilities) get better. It's definitely not urgent for me.

Re: Mistral OCR 3

#112
post #91

My current holy grail is my attempt to convert a Shipibo (an indigenous Peruvian language)-to-Spanish dictionary into a Shipibo-to-English dictionary. The pdf I have (available freely on archive.org) isn't a great scan (though I think it'd be a heck of a lot easier than some of the handwritten examples they show). Layout (2-columns) along with header/footers can cause some headaches, but it is all Latin script. This…

Once you have managed to get the data out and structured, you may want to check out dict.press. It's a dictionary publishing and management tool (which I maintain). Multiple widely used Indian dictionary projects run on it.

Will take a look, I assumed there'd be some tools along those lines, thanks for the suggestion

Re: Mistral OCR 3

#113
post #92
post #59

Earlier quoted context omitted.

I'm assuming you're interested in studying Ayahuasca traditions? I recently learned that traditionally in Shipibo culture, ayahuasca was never meant to be given to "the normal mind". Instead the maestras would be the ones taking the ayahuasca in order to help guide them into diagnosing people dealing with various sicknesses. These maestras were also ranked by how many different plants they'd done a dieta on. A dieta…

I suppose both of us watched the same Youtube video by Metta Beshay (i think that is his name?)

I actually did too lol. I was pleasantly surprised because it was actually decent and realistic about the situation (a lot of people get this romantic idea about going to the jungle to live and learn with the indigenous and have an "authentic" experience, and this does a pretty good job if dispelling that).

Re: Mistral OCR 3

#114
post #4

It seems like Mistral is just chasing around sort of "the fringes" of what could be useful AI features. Are they just getting out-classed by OAI, Google, Anthropic? It seems like EU in general should be heavily invested in Mistral's development, but it doesn't seem like they are.

Mistral is pursuing pursuing B2B use cases. Thats because they're releasing open models and the big thing about B2B is they HATE sending their data off-prem. OCR'ing and organizing old docs is a huge feature in B2B. Mistral's strategy seems smart to me.

Re: Mistral OCR 3

#115

I appreciate having an OCR interface rather than having to chat with a bot, but unfortunately chatting with Gemini 3 gives far better results than this. I gave it the document Gemini 3 got a surprisingly good result on: https://urn.digitalarkivet.no/URN:NBN:no-a1450-rk10101508282... and the output wasn't even recognizably Danish. Just out of pity I gave it a birthday card from my sister written in very readable moder…

[deleted]

Re: Mistral OCR 3

#116
Sadly, only available through a hosted API. I don't see how this is useful for OCR, unless you are OK with uploading your confidential documents to "the cloud"?

I'm still hoping for improved locally hosted models: qwen3-vl:30b-a3b-thinking-q4_K_M is already really good.

Re: Mistral OCR 3

#117
post #116

Sadly, only available through a hosted API. I don't see how this is useful for OCR, unless you are OK with uploading your confidential documents to "the cloud"? I'm still hoping for improved locally hosted models: qwen3-vl:30b-a3b-thinking-q4_K_M is already really good.

Businesses sign contracts about what happens when the data is uploaded. Ultimately your purpose is to make money more than maximally locking down your IP.

Re: Mistral OCR 3

#118

Earlier quoted context omitted.

Following the leaders too closely seems like a bad move, at least until a profitable business model for an AI model training company is discovered. Mistral’s models are pretty good, right? I mean they don’t have all the scaffolding around them that something like chatGPT does, but building all that scaffolding could be wasted effort until a profitable business model is shown. Until then, they seem to be able to keep…

They can't hire the best talent because the most experienced people will not leave their homes to chase a high-risk role with questionable remuneration by relocating their whole life to Paris or London. This goes to show how leaders in Mistral don't quite get that they are not special as they seem to think they are. Anthropic or OpenAI also require their talent to relocate but with stakes that are at least a high rew…

If somebody is in the EU already that calculation completely flips. We have a strong software startup industry in the US, would it really be that surprising if there was more unallocated talent in the EU, at this point?

Re: Mistral OCR 3

#119

I appreciate having an OCR interface rather than having to chat with a bot, but unfortunately chatting with Gemini 3 gives far better results than this. I gave it the document Gemini 3 got a surprisingly good result on: https://urn.digitalarkivet.no/URN:NBN:no-a1450-rk10101508282... and the output wasn't even recognizably Danish. Just out of pity I gave it a birthday card from my sister written in very readable moder…

> got a surprisingly good result

> the output wasn't even recognizably Danish

How would you know that it's good then?

Re: Mistral OCR 3

#120

I appreciate having an OCR interface rather than having to chat with a bot, but unfortunately chatting with Gemini 3 gives far better results than this. I gave it the document Gemini 3 got a surprisingly good result on: https://urn.digitalarkivet.no/URN:NBN:no-a1450-rk10101508282... and the output wasn't even recognizably Danish. Just out of pity I gave it a birthday card from my sister written in very readable moder…

> got a surprisingly good result > the output wasn't even recognizably Danish How would you know that it's good then?

I believe you misread. My reading is that Gemini 3 gave a good result on a certain input, so they gave the same input to this model and the result was poor.
Post reply on HN