Earlier quoted context omitted.
Seems to me that the way to architect this is to have multiple quasi independent embedded controllers at different levels. For example, you might have 3 independent finger controllers, managed by a hand controller, and so on, on up to the highest level LLM that drives everything. So you have an LLM that just says, pick up the green can, issues whatever structured data needs to go to the arm controller and on down the…
That is indeed what a lot of Machine Learning turned Robotics researchers/enthusiasts are banking on. A counter argument to that is what Dhruv Batra responded to the question "lol why not use LLM for everything" [1]. >As a human, I don't understand the detailed micro-second by microsecond movements my fingers have to do to pick something up, let alone how I'm touch-typing this sentence. It just sort of "happens" when…
Aside from some basic life support systems, don't almost all movements start with conscious effort? Whether you are deliberate about the exercise or not, you practice and practice until you develop 'muscle memory' where it becomes unconscious: walking, dribbling a basketball, holding a G chord on a guitar, etc.