Back to evaluations
Public evaluation
Narrative Branching Eval
Assessment of narrative consistency and latency in conversational flow.
Evaluation type
task based
Challenge
Building Real-Time Interactive Narratives with OpenAI Agents SDK and DeepSeek R1
Difficulty
Intermediate
Rigor
Unspecified
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
Narrative Branching Eval
Measures logical coherence across five turn segments.
Input format
Initial narrative prompt
Output format
Character dialogue chain