Back to evaluations
Public evaluation
lifesg_grant_eligibility
Evaluates OpenAI Agents SDK handoff efficiency, tool call accuracy, and grant calculation correctness across multi-turn user scenarios.
Evaluation type
task based
Challenge
LifeSG Multi-Agency Citizen Service Assistant using OpenAI Agents SDK
Difficulty
Advanced
Rigor
Unspecified
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
lifesg_grant_eligibility
Tests multi-agent conversation flow and grant calculation logic for a newborn child milestone.
Input format
JSON containing household income, child birth rank, and citizenship status
Output format
JSON containing eligible_schemes, total_grant_value_sgd, and agent_handoffs_executed