Thank you for your opinion & recommendations. Something I saw today related to "sub-agents" is in Kimi 2.6's model card it says
Elevated Agent Swarm: Scaling horizontally to 300 sub-agents executing 4,000 coordinated steps, K2.6 can dynamically decompose tasks into parallel, domain-specialized subtasks, delivering end-to-end outputs from documents to websites to spreadsheets in a single autonomous run.
So maybe Kimi 2.6 is doing the "type of thing" I am looking for, but I don't have the means to run it practically. Maybe at 1 token per second which would be brutal.
I tried out Qwen 3.6 27B but not yet in an agentic setting, so I can't really judge yet. Maybe it's just me but the small model size seems limiting. I thought gpt-oss-120b was good.
Sure, if you have a micro swarm architecture laid out, I would love to hear what it is.