Apertus – Open Foundation Model for Sovereign AI
151–160 of 198 posts
Re: Apertus – Open Foundation Model for Sovereign AI
#152The previous version of this model has been pretty bad, but claimed to adhere to copyright laws. However, based on my testing, that's not true either. So in my view this is completely useless.
It uses fineweb, which is derived from Common Crawl, which is an unlicensed scrape of web pages.
Re: Apertus – Open Foundation Model for Sovereign AI
#153Apertus V1 performance were sub-par. The Team is working on v2 ATM. Looking forward to testing it.
Re: Apertus – Open Foundation Model for Sovereign AI
#154Earlier quoted context omitted.
As long as the following remains true, this release ends up a bigger contribution to science at large than most other models trained "behind closed doors": > Fully open model: open weights + open data + full training details including all data and training recipes
Is a recipe useful if no one likes it? There are equally open, much more useful models out there: https://artificialanalysis.ai/?models=nvidia-nemotron-3-ultr...
Re: Apertus – Open Foundation Model for Sovereign AI
#155As an opesource AI researcher with a lot of models and datasets on huggingface I am very appreciative of these types of project but we are ignoring the elephant in the room here ( or lack of ) the swiss have no gpus
Re: Apertus – Open Foundation Model for Sovereign AI
#156I'm mildly surprised that more people aren't using Nemo models for this reason. We've moved most of our processing to a combination of Nemo Ultra and Super, with some support for multi-model-specific tasks on Omni. The setup is working REALLY well for us, and I'm comfortable with the more measured pace of improvements. We work with many long-context problems, and the ecosystem is great. There were a number of use cas…
Re: Apertus – Open Foundation Model for Sovereign AI
#157Earlier quoted context omitted.
Allen AI do not get enough love. They are doing GenAI how it should have always been done. In fact, if the frontier companies had taken their approach, it would have started much slower, but I think we would be far more advanced by 2035. Instead we have a majority of society that wants to see AI fail.
> Instead we have a majority of society that wants to see AI fail. Do you talk to regular people? I work out of coffee shops routinely and literally like 90% of laptops have ChatGPT or Claude open. I was shocked at how many of my friends love the silliest of AI features (like Slack bot summarizing your day or your upcoming meetings), and a lot of decks, proposals, SOW's, etc. are (at least in part) generated with AI…
I hate cars but I still drive to the office 1x / week because I have to.
Re: Apertus – Open Foundation Model for Sovereign AI
#158What's the community's take on Sovereign AI being funded by states around the world? Why the emphasis on sovereign? Open is good enough. No?
Re: Apertus – Open Foundation Model for Sovereign AI
#159Earlier quoted context omitted.
Allen AI do not get enough love. They are doing GenAI how it should have always been done. In fact, if the frontier companies had taken their approach, it would have started much slower, but I think we would be far more advanced by 2035. Instead we have a majority of society that wants to see AI fail.
> Instead we have a majority of society that wants to see AI fail. Do you talk to regular people? I work out of coffee shops routinely and literally like 90% of laptops have ChatGPT or Claude open. I was shocked at how many of my friends love the silliest of AI features (like Slack bot summarizing your day or your upcoming meetings), and a lot of decks, proposals, SOW's, etc. are (at least in part) generated with AI…
Re: Apertus – Open Foundation Model for Sovereign AI
#160Other fully open LLMs include Allen AI's OLMo 3.1 and MBZUAI's K2 Think V2, both of which have released their full training pipelines and datasets. Nvidia Nemotron is also an open training source model, though a portion of its dataset remains proprietary. Quoting lambda's comment: > Note that the Nemotron models are generally stronger than Olmo and K2 Think V2 (according to Artificial Analysis benchmarks), and there…
Allen AI do not get enough love. They are doing GenAI how it should have always been done. In fact, if the frontier companies had taken their approach, it would have started much slower, but I think we would be far more advanced by 2035. Instead we have a majority of society that wants to see AI fail.