Live data from Hacker News

Apertus – Open Foundation Model for Sovereign AI

apertvs.ai

121–130 of 198 posts

Re: Apertus – Open Foundation Model for Sovereign AI

#121
post #56

As an opesource AI researcher with a lot of models and datasets on huggingface I am very appreciative of these types of project but we are ignoring the elephant in the room here ( or lack of ) the swiss have no gpus

the Apertus model was trained on the Alps supercomputer, operational at CSCS since September 2024, a data center of over 10'000 top-of-the-line NVIDIA Grace-Hopper chips

https://log.alets.ch/110/

Re: Apertus – Open Foundation Model for Sovereign AI

#122
post #56

As an opesource AI researcher with a lot of models and datasets on huggingface I am very appreciative of these types of project but we are ignoring the elephant in the room here ( or lack of ) the swiss have no gpus

How is this a real problem? Genuine question, because i don’t really understand the urgency of everyone buying up ram and gpus as prices for those skyrocket. I can run the 8B version of this swiss-ai model on a ten year old GPU. For the larger one, $2000 consumer hardware can run it fine. Beyond that, there are plenty of places where time on a GPU can be rented, and if the model is good, there will be hardware to run…

You can run it, but you can't train it. While this type of toy model could actually be trained in Swiss equipment, a state-of-the-art LLM probably could not.

My charitable reading of GP's point is that the bottleneck for true compute sovereignty is the chips, not the models.

Re: Apertus – Open Foundation Model for Sovereign AI

#123
post #96

Earlier quoted context omitted.

Allen AI do not get enough love. They are doing GenAI how it should have always been done. In fact, if the frontier companies had taken their approach, it would have started much slower, but I think we would be far more advanced by 2035. Instead we have a majority of society that wants to see AI fail.

> Instead we have a majority of society that wants to see AI fail. Do you talk to regular people? I work out of coffee shops routinely and literally like 90% of laptops have ChatGPT or Claude open. I was shocked at how many of my friends love the silliest of AI features (like Slack bot summarizing your day or your upcoming meetings), and a lot of decks, proposals, SOW's, etc. are (at least in part) generated with AI…

I'm calling yesterday as Peak Height of AI...

I was at my daughter's football game, and another father from the club came up to me and asked if I were in IT and knew how AI worked. He then asked if I could help him setup an AI agent to generate passive income.

We're at the equivalent of December 2017 for crypto. Hang on to your hats!

Re: Apertus – Open Foundation Model for Sovereign AI

#125
post #8

Other fully open LLMs include Allen AI's OLMo 3.1 and MBZUAI's K2 Think V2, both of which have released their full training pipelines and datasets. Nvidia Nemotron is also an open training source model, though a portion of its dataset remains proprietary. Quoting lambda's comment: > Note that the Nemotron models are generally stronger than Olmo and K2 Think V2 (according to Artificial Analysis benchmarks), and there…

> an open training source model

It's always funny to see people tempted to call open-blobs/open-weights, which are literally shareware like WinRAR or Adobe PDF Viewer, open source, and then need to invent a new term for what is actually open source.

Re: Apertus – Open Foundation Model for Sovereign AI

#126
post #92

Earlier quoted context omitted.

Is there any evidence that "a majority of society wants AI to fail" Or is it just vibes?

There’s a few polls that have shown most people use AI, but they also dislike it. I’m in that boat, where my company pays for my subscription, and I use it to be productive. But I don’t really feel good about it. https://gizmodo.com/people-hate-ai-even-more-than-they-hate-...

Does it count when I hate the Dario, Altman and their weird cult more and more, every time they open their mouths? I think that I would not hate the tech in isolation, but considering who tech elite became, their rhetoric and how they behave , I want them fail just because of that.

Re: Apertus – Open Foundation Model for Sovereign AI

#128

Earlier quoted context omitted.

Can you give examples?

Is this a good faith question? It would take several hundred pages to document even a fraction of the violations. How about deporting people without a hearing or opportunity to present evidence about their charges. And then violating the judges order to turn the planes around. How about systematically ignoring judicial rulings. How about detaining people based on the color of their skin and spoken language/accent. Ho…

> How about deporting people without a hearing or opportunity to present evidence about their charges.

Not to detract from your general point about the US, your first point is something that's happened recently in Switzerland:

https://truthout.org/articles/swiss-police-arrest-deport-pal...

Re: Apertus – Open Foundation Model for Sovereign AI

#129
post #31
post #10

For a model that claims to focus on many languages, it's quite unreliable when it comes to simple questions like "how to say X in language Y" or "how to conjugate verb X in language Y". It keeps hallucinating words that do not exist, and when corrected, it only hallucinates a new lie.

it probably doesnt know what language each set of words is referencing. i doubt they are including a lot of training data labeled with the language. "how to say X in language Y" is a different task from saying X in language Y

Actually, it isn't all that different. There are only two words separating "how to say X in language Y" from "say X in language Y". And this "vulgar" metric is actually quite relevant for an LLM, which answers based on conversational context.

Re: Apertus – Open Foundation Model for Sovereign AI

#130

A chat interface where you can try Apertus: https://chat.publicai.co

You will need to register with an email and password though, i.e. your sessions will be recorded and identified.

Also even after you do that, and start a chat, you currently get:

  "JSON.parse: unexpected character at line 1 column 1 of the JSON data"
so it's not quite there yet.
Post reply on HN