Late, but it depends. Sometimes I ask for all of them to do the same thing and have a different model judge it, sometimes I ask for an orchestration model to run the smaller ones each on one task. I have a 'model cohort' where I run the same request across 5 very inexpensive models. Then have my driving LLM judge or synthesize.

You just come up with what you want in your head, and tell it to do it in that way and it does it. I trust 5 independent smartest programs ever over the SOTA smartest program's only.