I was thinking about this recently. I tend to run my AI at low context because the documentation states that they degrade with higher context usage. However I see tons of people on LinkedIn with ways of backing up context, not wanting to lose context, etc. This seems like another way the system is being misused. Higher context usage also uses more tokens. I suspect you get worse (and slower) output too than a dense d…
If every exchange is treated as an independent query/response then it's much easier to see how cutting out the fluff using a combination of its summaries and your own helps stay focused.