SeLMRoute: Probabilistic Semantic Evidence for Large Language Model Routing
Abstract
Large language model (LLM) routing aims to select the most suitable model for each incoming query. Most existing routers learn this decision directly from query embeddings, model representations, preference data, or clusters of similar examples. Such approaches can be effective, yet the representation used for routing rarely states what a query actually requires. We introduce SeLMRoute, a routing framework that separates the extraction of candidate-independent semantic evidence from the learning of candidate performance and the application of deployment objectives. A decision model first evaluates a set of interpretable questions about the query, such as its reasoning requirements and use of external knowledge, with each judgment retained as a probability distribution. The resulting probabilistic semantic state is used by a lightweight supervised router to estimate candidate model performance. Routing objectives are applied after performance estimation, which allows the same semantic state to support performance-oriented and cost-aware decisions. On the LLMRouterBench (15 datasets, 20 candidate models, 11,481 queries), SeLMRoute achieves an average accuracy of 72.08% pm 0.45, while grouped five-fold out-of-fold evaluation reaches 72.64%, compared with 69.23% for the strongest fixed candidate. The representation achieves the highest mean performance among the evaluated semantic, dense, lexical, and domain-level representations. In a separate 13-model performance-cost setting, SeLMRoute improves performance in all five grouped splits, with a mean PerfGain of 2.66%. Our code is available at https://github.com/Indigma-Innovations/SeLMRoute.
Community
What if LLM routing focused on what a query requires, rather than what it resembles?
With SeLMRoute, we represent each query using 16 interpretable semantic probes, covering reasoning, coding, knowledge, exactness, ambiguity, and more, while preserving uncertainty through a compact probabilistic representation.
This gives us a simple pipeline:
Query → Semantic Evidence → Model Performance → Routing
Across 11,481 queries, 15 datasets, and 20 models, SeLMRoute achieves 72.08% average accuracy, outperforming the best known model (69.23%).
We’d love to hear your thoughts!
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- Pretrain Once, Route Anywhere: Towards a Foundation Model for LLM Routing (2026)
- Routing Should Pay for Itself: Sparse Supervision for Economical LLM Routing (2026)
- RLCascadeRouter: Quality-Estimator-Free Cascade Routing via Reinforcement Learning (2026)
- LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers (2026)
- Opportunity Is Not Realizability: Selection-Valid Diagnostics for Multi-LLM Routing (2026)
- SCX Router: Streaming Zero-Shot Model Selection with a Decoder-KV Classifier and a Real-World Task Ontology (2026)
- AgentWeave: Routing Before Reasoning for Efficient Function Calling in Tool-Rich Language Models (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2609.34736 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 2
Indigma/SeLMRoute-Laya
Datasets citing this paper 1
Indigma/SeLMRoute-Semantic-Features
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper