Live data from Hacker News

Apertus – Open Foundation Model for Sovereign AI

apertvs.ai

151–160 of 198 posts

Re: Apertus – Open Foundation Model for Sovereign AI

#152
post #11
post #3

The previous version of this model has been pretty bad, but claimed to adhere to copyright laws. However, based on my testing, that's not true either. So in my view this is completely useless.

It uses fineweb, which is derived from Common Crawl, which is an unlicensed scrape of web pages.

You don't need a license to scrape the public web and analyze it, turn it into tokens and other transformations. Let's not expand copyright beyond the horrible monster it already is.

Re: Apertus – Open Foundation Model for Sovereign AI

#153
post #151

Apertus V1 performance were sub-par. The Team is working on v2 ATM. Looking forward to testing it.

I don't know, I'm implementing a translation system right now, and Apertus is very good for the model size. I wished they added some chain of thought training to increase precision and context understanding.

Re: Apertus – Open Foundation Model for Sovereign AI

#154

Earlier quoted context omitted.

As long as the following remains true, this release ends up a bigger contribution to science at large than most other models trained "behind closed doors": > Fully open model: open weights + open data + full training details including all data and training recipes

Is a recipe useful if no one likes it? There are equally open, much more useful models out there: https://artificialanalysis.ai/?models=nvidia-nemotron-3-ultr...

Nemotron still has partial closed data. Having multiple models to chose from is a good thing

Re: Apertus – Open Foundation Model for Sovereign AI

#155
post #56

As an opesource AI researcher with a lot of models and datasets on huggingface I am very appreciative of these types of project but we are ignoring the elephant in the room here ( or lack of ) the swiss have no gpus

Do some research before posting that kind of stuff

Re: Apertus – Open Foundation Model for Sovereign AI

#156

I'm mildly surprised that more people aren't using Nemo models for this reason. We've moved most of our processing to a combination of Nemo Ultra and Super, with some support for multi-model-specific tasks on Omni. The setup is working REALLY well for us, and I'm comfortable with the more measured pace of improvements. We work with many long-context problems, and the ecosystem is great. There were a number of use cas…

[dead]

Re: Apertus – Open Foundation Model for Sovereign AI

#157
post #96

Earlier quoted context omitted.

Allen AI do not get enough love. They are doing GenAI how it should have always been done. In fact, if the frontier companies had taken their approach, it would have started much slower, but I think we would be far more advanced by 2035. Instead we have a majority of society that wants to see AI fail.

> Instead we have a majority of society that wants to see AI fail. Do you talk to regular people? I work out of coffee shops routinely and literally like 90% of laptops have ChatGPT or Claude open. I was shocked at how many of my friends love the silliest of AI features (like Slack bot summarizing your day or your upcoming meetings), and a lot of decks, proposals, SOW's, etc. are (at least in part) generated with AI…

They use it, but do they love it or do they feel like they need it to do their best work and stay ahead?

I hate cars but I still drive to the office 1x / week because I have to.

Re: Apertus – Open Foundation Model for Sovereign AI

#158

What's the community's take on Sovereign AI being funded by states around the world? Why the emphasis on sovereign? Open is good enough. No?

It was in reaction to the possible threat of main actors restricting use. The latest US gov stunt with Fable just made it concrete and pressing.

Re: Apertus – Open Foundation Model for Sovereign AI

#159
post #96

Earlier quoted context omitted.

Allen AI do not get enough love. They are doing GenAI how it should have always been done. In fact, if the frontier companies had taken their approach, it would have started much slower, but I think we would be far more advanced by 2035. Instead we have a majority of society that wants to see AI fail.

> Instead we have a majority of society that wants to see AI fail. Do you talk to regular people? I work out of coffee shops routinely and literally like 90% of laptops have ChatGPT or Claude open. I was shocked at how many of my friends love the silliest of AI features (like Slack bot summarizing your day or your upcoming meetings), and a lot of decks, proposals, SOW's, etc. are (at least in part) generated with AI…

Ironic that you should question if the commenter talks to regular people and then cite people who work on laptops from the coffee shop, use Slack etc.

Re: Apertus – Open Foundation Model for Sovereign AI

#160
post #8

Other fully open LLMs include Allen AI's OLMo 3.1 and MBZUAI's K2 Think V2, both of which have released their full training pipelines and datasets. Nvidia Nemotron is also an open training source model, though a portion of its dataset remains proprietary. Quoting lambda's comment: > Note that the Nemotron models are generally stronger than Olmo and K2 Think V2 (according to Artificial Analysis benchmarks), and there…

Allen AI do not get enough love. They are doing GenAI how it should have always been done. In fact, if the frontier companies had taken their approach, it would have started much slower, but I think we would be far more advanced by 2035. Instead we have a majority of society that wants to see AI fail.

LLMs were invented by AI2, before Transformers were a thing - with RNN-based ELMO.
Post reply on HN