Back to evaluations
Draft evaluation
LifeSG Multi-Agency Citizen Service Assistant using OpenAI Agents SDK — evaluation
Evaluates OpenAI Agents SDK handoff efficiency, tool call accuracy, and grant calculation correctness across multi-turn user scenarios.
Evaluation type
task based
Challenge
LifeSG Multi-Agency Citizen Service Assistant using OpenAI Agents SDK
Difficulty
Advanced
Rigor
Not declared
The author has not specified a rigor level.
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
lifesg_grant_eligibility
Tests multi-agent conversation flow and grant calculation logic for a newborn child milestone.
Input format
JSON containing household income, child birth rank, and citizenship status
Output format
JSON containing eligible_schemes, total_grant_value_sgd, and agent_handoffs_executed