One model deployment to route them all

Microsoft Developer uses a “Sparkle Cupcakes” agent scenario to show why mixed traffic (simple FAQs vs complex catering requests) benefits from routing rather than forcing every request through one model.

Overview

Meet the Azure AI Foundry Model Router

Supported models and deployments

Configure routing priorities

Test the “Cupcake Agent” with a single-line change

Route complex reasoning tasks vs general knowledge questions

Monitor usage and model selection