Just have the harness able to choose which model its sub-agents use, then tell it how to split up tasks and which models to use when doing so.
Amazing! Really brilliant idea, thank you for sharing this project. There is so much ground to cover in the LLM gateway / routing / reporting world, and this is a great start. The Tinker implementation is my favorite part, fine tuning is much better than a sea of context files.
Look at the Intelligence features in the Enterprise plan:
* Per-prompt model optimization
* Caching
* A model you own, trained on your traffic