When is Routing Meaningful? Diversity and Robustness in Language Model Societies
4.20T1 sourcearXiv cs.MA
Source record
Published by arXiv cs.MA (T1 source). The original is at https://arxiv.org/abs/2607.09197.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryArgues that multi-model LLM routing is evaluated too narrowly on accuracy and cost. Proposes two additional criteria: behavioral differentiation among actors and routing stability under query perturbations. Adapts Hierarchic Social Entropy to LLM societies and introduces a perturbation-based robustness metric, finding that a curated subset of under ten agents captures most diversity and that KNN and prompted routers diverge in robustness.
Why it mattersReroutes how multi-model routing should be evaluated. The coreset heuristic and the accuracy-vs-meaningfulness gap are directly usable for anyone building agent routing pipelines.
Cited by
No citations on record.
