Line of inquiry
Inquiring lines›What drives capability improvement…›How should computational architect…›this line of inquiry
Does intelligent routing among smaller models outperform training larger models?
A broader line of inquiry — a family of 34 specific questions the research asks around this. Follow one into its inquiring-line page, or move sideways to a related line below.
Questions in this line of inquiry 34
Specific inquiring lines the field asks around this — ordered from the most general framing down to the most specific angle.
- Can multiple small models outperform a single large model with good routing?
- What makes routing a better investment than training larger models?
- Should model routing decisions account for prompt-tier dependencies?
- How do routing and test-time compute scaling work together as optimization axes?
- Does model selection matter more than model improvement for query routing?
- Can embedding-cluster routing outperform a single frontier model?
- How does routing decide between models before generation happens?
- Can routing policies remain meaningful over behaviorally homogeneous model pools?
- Can compute allocation and model routing be combined for better results?
- Can routing systems prevent expert models from failing outside their specialty?
- Why might diverse smaller models with routing beat one giant model?
- Can routing enable heterogeneous SLM-first architectures at scale?
- What makes query complexity a better routing signal than response quality?
- Can model routing and compute allocation work together as independent optimizations?
- How do pre-training and distillation enable minimal routing signals to work?
- Can a router predict query complexity well but still fail as governance?
- Why does single-model routing beat ensemble and cascade approaches on latency?
- Can semantic routing couple similarity matching with resource constraints?
- What makes routing a governance mechanism rather than just a predictor?
- Can routing signals organize training data into a meaningful curriculum automatically?
- How do KNN and prompted routers differ in the accuracy-stability tradeoff?
- Does surface-form query rewriting allow attackers to steer model routing decisions?
- What makes mixture-of-experts routing learn token-level specialization effectively?
- How do routers decide when to escalate from small to large models?
- Does the improved model actually return to the routing pool and shape future decisions?
- Can hierarchical vector routing reduce context overhead while maintaining tool coverage?
- How do cost-aware cascades compare to single-turn routing in the component framework?
- What learning signals best supervise router training across benchmark tasks?
- How do hierarchical architectures improve multi-hop query performance?
- Why do KNN routers collapse when queries are paraphrased?
- How should topology routing adapt to different task types?
- Why does Branch-Train-Merge fail without learned routing between experts?
- Who serves as the teacher model in the routing-guided distillation process?
- How does semantic clustering help decide which model handles each query?