Back to evaluations
Public evaluation

Model Router Test

Evaluation of orchestration efficiency and routing accuracy.

Evaluation type
task based
Challenge
Multi-Model Orchestration Platform with LangGraph and GPT-5.4 Pro
Difficulty
Advanced
Rigor
Unspecified

Evaluation overview

How the linked challenge is judged: tasks, benchmarks, and criteria count.

Tasks
1
Benchmarks
0
Criteria
0

Task templates

Inputs and expected outputs.

Task 1

Model Router Test

Verifies that the correct model is selected based on complexity.

Input format

Task description text

Output format

JSON indicating model used