Back to evaluations
Public evaluation
Model Router Test
Evaluation of orchestration efficiency and routing accuracy.
Evaluation type
task based
Challenge
Multi-Model Orchestration Platform with LangGraph and GPT-5.4 Pro
Difficulty
Advanced
Rigor
Unspecified
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
Model Router Test
Verifies that the correct model is selected based on complexity.
Input format
Task description text
Output format
JSON indicating model used