Back to evaluations
Public evaluation
actuarial_risk_assessment
Evaluates Pydantic AI type safety, validator enforcement, and numerical precision when calculating climate risk premiums.
Evaluation type
task based
Challenge
Actuarial Climate Risk Scoring Agent with Pydantic AI and Type Safety
Difficulty
Intermediate
Rigor
Unspecified
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
actuarial_risk_assessment
Evaluates calculation of annualized expected loss (AEL) and type-safe schema validation.
Input format
JSON containing asset coordinates, flood probability, and asset value SGD
Output format
JSON matching Pydantic model with asset_id, annualized_expected_loss, and risk_tier