Earlier quoted context omitted.
It's also a bit funny because Norway definitely has enough money to hire a team of Anthropic's best to go out there and train them a model that does whatever they want. They probably have enough money to fund their own Anthropic competitor.
>They probably have enough money to fund their own Anthropic competitor. Which is bizarre to me Norway doesn't have a booming tech sector with all hat wealth fund acting as the biggest VC. They instead use their wealth fund to invest in US's tech sector. Baffling.
Norway's 2 petabytes of Huawei flash storage and LLM training
151–160 of 227 posts
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#152How true is this statement: "He asserted that any country with its own language that did not have a sovereign LLM trained in that language was at a disadvantage as a globally trained, English-speaking LLM would not know about that country’s history, news and culture that was described in the local language." I thought all big players already train on basically everything remotely available to them no matter the langu…
If you want LLMs to have knowledge of the Norwegian language, wouldn't the most obvious thing to do be to build a good training dataset and make the dataset widely available? Why go to the expense of training your own model, especially when it will be inferior to state of the art models.
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#153Earlier quoted context omitted.
If you want LLMs to have knowledge of the Norwegian language, wouldn't the most obvious thing to do be to build a good training dataset and make the dataset widely available? Why go to the expense of training your own model, especially when it will be inferior to state of the art models.
I task GPT/Claude with researching stuff that pertains to very specific cultural or legal aspects in French politics, on a daily basis. Even though French is a way more common language globally than Norwegian, these models still haven't figured out that, no matter the language I myself speak to them (German or English depending on my mood) their web searches need to be done in French to return reasonable results. I h…
I was trying to work out how and when to use swear words, and the relative power index of them. it translated english swear words into the target language then lectured me on not using them.
It took a bunch of prodding for it to actually think as the target language to then get the (mostly) correct response.
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#154Earlier quoted context omitted.
>Current-best models are pretty fluent at major languages and cultures strong disagree on that one. As a German interacting with ChatGPT, even in German it gives me the feeling of talking to the Pluribus people, which reminds me of an anecdote of Walmart failing in Germany because people were freaked out by the constantly upbeat, smiling employees. Understanding a culture is a very different task than translating the…
I'm Finnish and dear god I hate the default overtly friendly tones of LLMs. Always the first thing to tune in system prompt. You're a machine, stop anthropomorphizing yourself and pretending to be my best friend, and just give me the damn answer and nothing else. :D
I do understand where proponents of language equivalency are coming from. LLMs seem to be extremely good at answering simple, one-shot type questions and mechanical 'low-level' translations for most languages. I feel like as soon as you introduce complex chains of thought or multi-step cross-linguistic tasks, minor imperfections stack and become magnified, just as with coding tasks or context rot.
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#155Earlier quoted context omitted.
Why would the gap grow? There is no more training data to acquire, frontier model are training on the entire internet. Everything from now on is just fine-tuning.
Your statement assumes training data is the only thing that matters for the big players, while not considering it limiting for the small Norwegian model. That’s a fallacy.
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#156> The Olivia system is an HPE Cray Supercomputing EX system, with 448 GPUs and 64,512 CPU cores. Training a sovereign LLM with this meager hardware as opposed to a LORA on some open source model seems like a huge mistake and a potential red flag. There is no way these people have the resources to train a fully fledged LLM, so claiming that is their goal makes me think they don't intend for the LLM to be useful. Which…
Depends on what they are doing and why. but at most big labs, only the final model training happens on the big clusters. a lot of experimentation happens on So for fast iteration, this seems fine.
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#157> Marius Husnes, the Head of IT Platform at the library (Nasjonlbiblioteket) discussed the project at Huawei’s ID Forum 2026 in Paris, saying that no commercial LLM provider was developing a local (Norwegian) language LLM. He asserted that any country with its own language that did not have a sovereign LLM trained in that language was at a disadvantage as a globally trained, English-speaking LLM would not know about…
It seems like you've made an assertion but not provided evidence. Why is it not a disadvantage to only have english LLMs?
Can you get the nuance of Norwegian history/culture with present models?
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#158Earlier quoted context omitted.
It's also a bit funny because Norway definitely has enough money to hire a team of Anthropic's best to go out there and train them a model that does whatever they want. They probably have enough money to fund their own Anthropic competitor.
>They probably have enough money to fund their own Anthropic competitor. Which is bizarre to me Norway doesn't have a booming tech sector with all hat wealth fund acting as the biggest VC. They instead use their wealth fund to invest in US's tech sector. Baffling.
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#159How true is this statement: "He asserted that any country with its own language that did not have a sovereign LLM trained in that language was at a disadvantage as a globally trained, English-speaking LLM would not know about that country’s history, news and culture that was described in the local language." I thought all big players already train on basically everything remotely available to them no matter the langu…
Not remotely true in my estimation. I don't really speak Norwegian, but I do speak Swedish(which means I mostly understand Norwegian as they're very similar). Every model I've tried speaking Swedish to does it perfectly. I'd be surprised if the same isn't true for Norwegian already
Re: Norway's 2 petabytes of Huawei flash storage and LLM training
#160Earlier quoted context omitted.
>They probably have enough money to fund their own Anthropic competitor. Which is bizarre to me Norway doesn't have a booming tech sector with all hat wealth fund acting as the biggest VC. They instead use their wealth fund to invest in US's tech sector. Baffling.
Considering the fact that the US is complaining about Norway putting too much money into the US market, imagine what would happen if all that money was spent in Norway. It would be chaos.
It would create jobs, sovereignty, intellectual property and soft power?
Instead it goes to strengthening the tech monopoly of a country that threatens to invade your neighbour.