Back to evaluations
Public evaluation
casualty_ils_structuring
Evaluates multi-agent conversation termination and output yield pricing error against ground truth actuary benchmarks.
Evaluation type
task based
Challenge
AutoGen Casualty ILS: Multi-Agent Tail Risk Structuring System
Difficulty
Advanced
Rigor
Unspecified
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
casualty_ils_structuring
Agents negotiate ILS bond parameters and output agreed coupon spread.
Input format
JSON containing expected_loss_pct, attachment_point, tail_factor
Output format
JSON containing agreed_coupon_bps, rating_grade, market_clearing_status