I like to start with memory engineering then reach full system then reducing costs. This allows unlocking full potential of agents.
Interesting. Where can I read more about this?
For me, if I put costs in my architecture, I'm limited heavily and that can easily change how memory is shaped dramatically. The opposite, putting memory in architecture, is not true. Memory engineering first, then full scale in the system then costs considerations.
In addition, I believe this is future friendly. Because AI is advancing and getting smarter and cheaper everyday.
I couldn't find a guide on this so I share my basic thoughts.