Earlier quoted context omitted.
Even more fundamental than science, there is missing philosophy, both in us regarding these systems, and in the systems themselves. An AGI implemented by an LLM needs to, at the minimum, be able to self-learn by updating its weights, self-finetune, otherwise it quickly hits a wall between its baked-in weights and finite context window. What is the optimal "attention" mechanism for choosing what to self-finetune with,…
A system that self-updates its weights is so obvious the only question is who will be the first to get there?
Self-updating weights could be more like epigenetics.