This has nothing to do with superintelligence, it's just the people that were working on the paper prior to the re-org happened to publish after the name change. Though it is notable that contrary to many (on HN and Twitter) that Meta would stop publishing papers and be like other AI labs (e.g. OpenAI). They're continued their rapid pace of releasing papers AND open source models.
Open weights models, not open source. And even their weights are under a specific license not as permissive as apache 2.
Meta Superintelligence Labs' first paper is about RAG
201–210 of 283 posts
Re: Meta Superintelligence Labs' first paper is about RAG
#202Earlier quoted context omitted.
This is the right terminology. Model weights are literally compiled binary data; they are the output of an algorithm run on a bunch of source data. That training dataset is the "source" of the model. Training data (or the scripts used to generate it) is human-readable and modifiable, like source code. Binary weights are not.
Just to note though, source copyright extends to its compiled form. There is probably an analogue there for model weights.
Re: Meta Superintelligence Labs' first paper is about RAG
#203Earlier quoted context omitted.
Yes but it doesn't generalize very well. Even on simple features like gender. If you go look at embeddings you'll find that man and woman are neighbors, just as king and queen are[0]. This is a better explanation for the result as you're just taking very small steps in the latent space. Here, play around[1] mother - parent + man = woman father - parent + woman = man father - parent + man = woman mother - parent + wom…
so addition is not associative?
Re: Meta Superintelligence Labs' first paper is about RAG
#204Can we have a more informative, less clickbaity, title?
Re: Meta Superintelligence Labs' first paper is about RAG
#205This was inevitable. You can't keep training LLMs and expect that's the answer to the evolution of AI. Yes it'll happen and we'll keep creating new more refined and bigger models but it's like DNA or something like the cortex of the brain. After that you need these systems that essentially "live" for years digesting information and develop a more refined way to process, store and retrieve the information. Compression…
Re: Meta Superintelligence Labs' first paper is about RAG
#206Re: Meta Superintelligence Labs' first paper is about RAG
#207Re: Meta Superintelligence Labs' first paper is about RAG
#208Earlier quoted context omitted.
Alexandr Wang is not interesting and a few steps short of a fraud that Mark had to bail out because he was so co invested. Shareholders should be livid if they knew a single thing about what was going on.
Tell me more
Re: Meta Superintelligence Labs' first paper is about RAG
#209I came to believe the LLMs work with token embeddings. Is then the REFRAG only "something" in front of the LLM, and the decoder is the RL policy which expands only some token chunk embeddings into token embeddings feedable to LLM? Or the REFRAG needs you to 'tune' the LLM to be able to work with both token embeddings and token chunk embeddings?
Re: Meta Superintelligence Labs' first paper is about RAG
#210Earlier quoted context omitted.
That's why I still use an abacus.
The abacus skills are safely obsolete, the skills of general thinking and creativity must not become that. This couldn't be more specious. Meme thinking like this, repeating something you've heard as reflex without regard to whether it fits a situation, is the exact kind of unoriginality we can't allow to become the default mode of thinking.