Explainable Model Routing for Agentic Workflows

descriptionPublicationkeyboard_double_arrow_right Article , Conference object , Preprint 01 Jan 2026Embargo end date: 01 Jan 2026Publisher:ZenodoJournal:CoRR, volume abs/2604.03527

Authors: Mika Okamoto; Ansel Kaplan Erol; Mark Riedl;

doi: 10.5281/zenodo.19700291 , 10.48550/arxiv.2604.03527 , 10.5281/zenodo.19700290

arXiv: 2604.03527

Explainable Model Routing for Agentic Workflows

- Summary
- Subjects
- Metrics

Abstract

Modern agentic workflows decompose complex tasks into specialized subtasks and route them to diverse models to minimize cost without sacrificing quality. However, current routing architectures focus exclusively on performance optimization, leaving underlying trade-offs between model capability and cost unrecorded. Without clear rationale, developers cannot distinguish between intelligent efficiency—using specialized models for appropriate tasks—and latent failures caused by budget-driven model selection. We present Topaz, a framework that introduces formal auditability to agentic routing. Topaz replaces silent model assignments with an inherently interpretable router that incorporates three components: (i) skill-based profiling that synthesizes performance across diverse benchmarks into granular capability profiles (ii) fully traceable routing algorithms that utilize budget-based and multi-objective optimization to produce clear traces of how skill-match scores were weighed against costs, and (iii) developer-facing explanations that translate these traces into natural language, allowing users to audit system logic and iteratively tune the cost-quality tradeoff. By making routing decisions interpretable, Topaz enables users to understand, trust, and meaningfully steer routed agentic systems.

Proceedings of the CHI 2026 Workshop on Human-Centered Explainable AI (HCXAI); April 13–17, 2026; Barcelona, Spain.

Keywords

Human-Computer Interaction, FOS: Computer and information sciences, Artificial Intelligence (cs.AI), Artificial Intelligence, Explainable AI, Human-computer Interaction, Agentic AI, LLM Routing, Human-centered AI, Language Model, Human-Computer Interaction (cs.HC), Agentic AI Routing

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	0
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Average
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

0

Average