YouTube Video
https://aka.ms/foundry-portal
https://aka.ms/InsideMicrosoftFoundryPlaylist
Model Router routes to the best model for the task. But, before you switch, you'll want to know: will it save me money on my workload without hurting quality? In this episode we use the open-source Model Router Auto Evaluation toolkit. It compares Model Router with your current baseline model (for example GPT-5) on: quality, cost, latency, value and model distribution. At the end you get a HTML dahsboard with 8 charts which answers the question: do I need a giant model or route a selection of the models using the Model Router?
0:00 - Save Costs with Model Router
0:35 - Model Router Auto Evaluation
1:08 - Configure Models and Pricing
1:39 - Prepare the Evaluation Dataset
2:10 - Run and View the Evaluation
2:40 - Compare Cost, Latency, and Quality
3:12 - Analyze the Quality Breakdown
3:43 - Choose the Right Tradeoffs
https://aka.ms/InsideMicrosoftFoundryPlaylist
Model Router routes to the best model for the task. But, before you switch, you'll want to know: will it save me money on my workload without hurting quality? In this episode we use the open-source Model Router Auto Evaluation toolkit. It compares Model Router with your current baseline model (for example GPT-5) on: quality, cost, latency, value and model distribution. At the end you get a HTML dahsboard with 8 charts which answers the question: do I need a giant model or route a selection of the models using the Model Router?
0:00 - Save Costs with Model Router
0:35 - Model Router Auto Evaluation
1:08 - Configure Models and Pricing
1:39 - Prepare the Evaluation Dataset
2:10 - Run and View the Evaluation
2:40 - Compare Cost, Latency, and Quality
3:12 - Analyze the Quality Breakdown
3:43 - Choose the Right Tradeoffs