Earlier quoted context omitted.
No, it really matters because of the impact it has on context tokens. Reading on GH issue with MCP burns 54k tokens just to load the spec. If you use several MCPs it adds up really fast.
The impact on context tokens would be more of a 'you're holding it wrong' problem, no? The GH MCP burning tokens is an issue on the GH MCP server, not the protocol itself. (I would say that since the gh CLI would be strongly represented in the training dataset, it would be more beneficial to just use the CLI in this case though.) I do think that we should adopt Amp's MCPs-on-skills model that I've mentioned in my ori…
Even if the model doesn’t already know the cli commands it can interrogate them at a much lower token cost for just the commands needed.