GPT‑NL: a sovereign language model for the Netherlands
321–325 of 325 posts
Re: GPT‑NL: a sovereign language model for the Netherlands
#322Earlier quoted context omitted.
Seems like you don’t understand. You take current version and build on top of it. You have the weights. You might not get some n+1 version at some point but the n version you will have will be still most likely much better than whatever you come up with burning good will money of people believing in „sovereignty”. You are not getting ahead in this game by being „true to your local values” capital expenditure is insan…
It seems like you don't understand. For fine tuning, it's cheaper to fine tune an existing model. For massive changes, it's better to retrain from scratch. Otherwise, model will UNLEARN a lot first, and then you will train about twice longer to the same result. https://en.wikipedia.org/wiki/Catastrophic_interference
I am speaking from very practical point of view. English and whatever frontier models are trained on is lingua Franca of software/tech/science currently - you don’t want to make massive changes because just exactly like you wrote model will unlearn a lot.
Current models translate easily between the languages.
So from my point of view even if we as a smaller country will have N-2 model that we slightly fine tune or just give it a harness with national RAG it will be better than wasting money on training a model from scratch only on „texts in country own language „ because that is loosing proposition. It’s usefulness is going to be really limited compared to model trained on body of knowledge in English.
It is a lot like company CEO thinking they have to „train AI model” for their company on their company materials - well no, you just make RAG and eventually fine tune some models, give access to dat, give access to MCP.
Because even if you are F1000 company you don’t have resources to train your own generic model, as a nation there is no that have resources to train „national model” on par with frontier labs.
Re: GPT‑NL: a sovereign language model for the Netherlands
#323Earlier quoted context omitted.
Idk which models you refer to, but I tested a bunch recently, and they performed well on Dutch. Only the smallest, such as qwen 3.6 27B, made up words and switched languages.
There's a large gap between making up words and an actually native text distribution. LLMs have a clear pattern, clear tells, a "feel" in English, and it's normally even more pronounced in non-English languages. Lots of bias towards English sentence structure, idioms, etiquette, etc.
Re: GPT‑NL: a sovereign language model for the Netherlands
#324Earlier quoted context omitted.
I don't know, I don't really buy that this all went downhill with Trump. It's easy to say that about someone you don't like and fail to overlook all the signs with someone you like.
I didn’t have a problem with Trump during his first term. It was clear he was unfit to be a president, but he was mostly golfing.
We call this confirmation bias.
I doubt any president would really be able to golf for much of his time. Though you can believe what you'd like. I would love to find evidence for this belief though, where do I Look?
Re: GPT‑NL: a sovereign language model for the Netherlands
#325Earlier quoted context omitted.
Simplistic reply without substance. The EU economic growth is influenced by much more than just the integration. It can be stagnant not because but despite the integration. It could also be the case that EU integration was an attempt to improve economy but it didn't work out.
So no evidence of economic importance.