Back to evaluations
Draft evaluation

LifeSG Multi-Agency Citizen Service Assistant using OpenAI Agents SDK — evaluation

Evaluates OpenAI Agents SDK handoff efficiency, tool call accuracy, and grant calculation correctness across multi-turn user scenarios.

Evaluation type
task based
Challenge
LifeSG Multi-Agency Citizen Service Assistant using OpenAI Agents SDK
Difficulty
Advanced
Rigor
Not declared

The author has not specified a rigor level.

Evaluation overview

How the linked challenge is judged: tasks, benchmarks, and criteria count.

Tasks
1
Benchmarks
0
Criteria
0

Task templates

Inputs and expected outputs.

Task 1

lifesg_grant_eligibility

Tests multi-agent conversation flow and grant calculation logic for a newborn child milestone.

Input format

JSON containing household income, child birth rank, and citizenship status

Output format

JSON containing eligible_schemes, total_grant_value_sgd, and agent_handoffs_executed