Earlier quoted context omitted.
This already exists: https://www.youtube.com/watch?v=UAG_FBZJVJ8
Could you point me to the timestamp where it empties a dishwasher?
Q-Transformer
41–50 of 74 posts
Re: Q-Transformer
#42It's clear we're searching for the god algorithm of AI, just like physicists are searching for theory of everything. Are transformers the answer though?
Re: Q-Transformer
#43https://diffusion-policy.cs.columbia.edu/
Looking at the figures and videos it seems... worse? Sort of surprised they didn't compare it but I guess they're trying to limit the discussion to purely reinforcement learning methods.
EDIT: Ah, I see this is an older result and was published concurrently to the diffusion policy paper, so it's likely the authors didn't know about it in time to add the extra comparisons.
Re: Q-Transformer
#44Being implemented as we speak, by the always impressive LucidRains [1] [1]: https://github.com/lucidrains/q-transformer
It's kinda funny how Lucidrains came back to make numerous commits to this repo following Q*.
Re: Q-Transformer
#45Re: Q-Transformer
#46Earlier quoted context omitted.
That's a great start. Opening a drawer is pretty hard problem.
Correct- as a reference for OC this is known as Moravec's Paradox ( https://en.wikipedia.org/wiki/Moravec%27s_paradox )
>Encoded in the large, highly evolved sensory and motor portions of the human brain is a billion years of experience about the nature of the world and how to survive in it. The deliberate process we call reasoning is, I believe, the thinnest veneer of human thought, effective only because it is supported by this much older and much more powerful, though usually unconscious, sensorimotor knowledge. We are all prodigious olympians in perceptual and motor areas, so good that we make the difficult look easy. Abstract thought, though, is a new trick, perhaps less than 100 thousand years old. We have not yet mastered it. It is not all that intrinsically difficult; it just seems so when we do it.
Re: Q-Transformer
#47Earlier quoted context omitted.
It seems more likely that there will be multiple avenues to AGI, all with their strengths and weaknesses. But perhaps the "God AI" will be a multifaceted model composed of many different models acting in unison.
Exactly, just as the true god has seven aspects [0]. [0] https://en.wikipedia.org/wiki/Themes_in_A_Song_of_Ice_and_Fi...
Re: Q-Transformer
#48It's clear we're searching for the god algorithm of AI, just like physicists are searching for theory of everything. Are transformers the answer though?
We're searching for an efficient algorithm that leads to AGI. Given sufficient time and compute, I'm sure that we could get there with existing stuff, by accident, and we wouldn't realize it before moving on to the next thing... and there'd be a poor orphan AGI, lost in a Git repo, waiting for runtime.
Re: Q-Transformer
#49Earlier quoted context omitted.
That's a great start. Opening a drawer is pretty hard problem.
The other robot project posted here yesterday opened drawers at around 80% though https://news.ycombinator.com/item?id=38453047 And they did it in many different homes.
Still working on closing the drawers afterwards, though...
Re: Q-Transformer
#50It's clear we're searching for the god algorithm of AI, just like physicists are searching for theory of everything. Are transformers the answer though?
It seems more likely that there will be multiple avenues to AGI, all with their strengths and weaknesses. But perhaps the "God AI" will be a multifaceted model composed of many different models acting in unison.
My human brain doesn't use the same algorithm for learning to play a song on a piano as learning to play a new board game. I'm not an AI person, but it seems reasonable to imagine we'd have different "modules" to apply as needed.
AlphaGo probably sucks at conversation. ChatGPT can't play Go. The part of my brain writing this couldn't throw a baseball. The physics engine that lets me throw a baseball couldn't write this. Is there a reason we'd want or need one specific AI approach to be universally applicable?