Earlier quoted context omitted.
I've heard of one study that said AI slows developers down, even when they think it's helping. https://www.infoworld.com/article/4061078/the-productivity-p...
AI may slow coding a bit but dramatically reduces cognitive load. The real value of AI isn't in helping coding. It's in having a human-like intelligence to automate processes. I can't get into details but my team is doing things that I couldn't dream of three years ago.
Meta Superintelligence Labs' first paper is about RAG
81–90 of 283 posts
Re: Meta Superintelligence Labs' first paper is about RAG
#82Earlier quoted context omitted.
I thought Alex Wang was a very curious choice. There are so many foundational AI labs with interesting CEOs... I get that Wang is remarkable in his own right, but he basically just built MTurk and timed the bubble. Doesn't really scream CEO of AGI to me.
A lot of people also don't know that many of the well known papers are just variations on small time papers with a fuck ton more compute thrown at the problem. Probably the strongest feature that correlates to successful researcher is compute. Many have taken this to claim that the GPU poor can't contribute but that ignores so many other valid explanations... and we wonder why innovation has slowed... It's also weird…
I worked for a small research heavy AI startup for a bit and it was heart breaking how many people I would interact with in that general space with research they worked hard and passionately on only to have been beaten to the punch by a famous lab that could rush the paper out quicker and at a larger scale.
There were also more than a few instances of high-probability plagiarism. My team had a paper that had been existing for years basically re-written without citation by a major lab. After some complaining they added a footnote. But it doesn't really matter because no big lab is going to have to defend themselves publicly against some small startup, and their job at the big labs is to churn out papers.
Re: Meta Superintelligence Labs' first paper is about RAG
#83Earlier quoted context omitted.
This was actually shown to not really work in practice.
I have seen this particular work example to work. You don't get the exact match but the closest one is indeed Queen.
Having set of "king - male + female = queen" like relations, including more complex phrases to align embeddings.
It seems like terse, lightweight, information dense way to address essence of knowldge.
Re: Meta Superintelligence Labs' first paper is about RAG
#84Earlier quoted context omitted.
Could you elaborate or link something here? I think about this pretty frequently, so would love to read something!
Metric: time to run 100m Context: track athlete Does it cease to be a good metric? No. After this you can likely come up with many examples of target metrics which never turn bad.
You're misunderstanding the root cause. Your example works as the the metric is well aligned. I'm sure you can also think of many examples where the metric is not well aligned and maximizing it becomes harmful. How do you think we ended up with clickbait titles? Why was everyone so focused on clicks? Let's think about engagement metrics. Is that what we really want to measure? Do we have no preference over users being happy vs users being angry or sad? Or are those things much harder to measure, if not impossible to, and thus we focus on our proxies instead? So what happens when someone doesn't realize it is a proxy and becomes hyper fixated on it? What happens if someone does realize it is a proxy but is rewarded via the metric so they don't really care?
Your example works in the simple case, but a lot of things look trivial when you only approach them from a first order approximation. You left out all the hard stuff. It's kinda like...
Edit: Looks like some people are bringing up metric limits that I couldn't come up with. Thanks!
Re: Meta Superintelligence Labs' first paper is about RAG
#85Seems very incremental and very far from the pompous 'superintelligence' goal.
A 30 fold improvement seems a tad more than incremental.
Re: Meta Superintelligence Labs' first paper is about RAG
#86Earlier quoted context omitted.
Could you elaborate or link something here? I think about this pretty frequently, so would love to read something!
Metric: time to run 100m Context: track athlete Does it cease to be a good metric? No. After this you can likely come up with many examples of target metrics which never turn bad.
Re: Meta Superintelligence Labs' first paper is about RAG
#87Seems very incremental and very far from the pompous 'superintelligence' goal.
"Send this through the math coprocessor." "Validate against the checklist." "Call out to an agent for X." "Recheck against input stream Y." And so on.
Retrieval augmentation is only one of many uses for this. If this winds up with better integration with agents, it is very possible that the whole is more than the sum of its parts.
Re: Meta Superintelligence Labs' first paper is about RAG
#88Earlier quoted context omitted.
Metric: time to run 100m Context: track athlete Does it cease to be a good metric? No. After this you can likely come up with many examples of target metrics which never turn bad.
So what is your argument, that it doesn't apply everywhere therefore it applies nowhere? You're misunderstanding the root cause. Your example works as the the metric is well aligned. I'm sure you can also think of many examples where the metric is not well aligned and maximizing it becomes harmful. How do you think we ended up with clickbait titles? Why was everyone so focused on clicks? Let's think about engagement…
I never said that. Someone said the law collapses, someone asked for a link, I gave an example to prove it does break down in some cases at least, but many cases once you think more about it. I never said all cases.
If it works sometimes and not others, it's not a law. It's just an observation of something that can happen or not.
Re: Meta Superintelligence Labs' first paper is about RAG
#89Earlier quoted context omitted.
Personal experience here in a FAANG, there has been a considerable increase in: 1. Teams exploring how to leverage LLMs for coding. 2. Teams/orgs that already standardized some of the processes to work with LLMs (MCP servers, standardized the creation of the agents.md files, etc) 3. Teams actively using it for coding new features, documenting code, increasing test coverage, using it for code reviews etc. Again, perso…
Im sure the MBA folks love stats like that - theres plenty that have infested big tech. I mean Pichai is an MBA+Mckinsey Alumni. Ready for the impending lay off fella?
Re: Meta Superintelligence Labs' first paper is about RAG
#90Earlier quoted context omitted.
Could you elaborate or link something here? I think about this pretty frequently, so would love to read something!
Metric: time to run 100m Context: track athlete Does it cease to be a good metric? No. After this you can likely come up with many examples of target metrics which never turn bad.
Yes if you run anything other than the 100m