Back to evaluations
Draft evaluation
Orchestrate Multi-Agent Actuarial Risk Rating using AutoGen & Hydra — evaluation
Evaluates the consistency, accuracy, and auditability of premium recommendations across stochastic multi-agent deliberation runs.
Evaluation type
task based
Challenge
Orchestrate Multi-Agent Actuarial Risk Rating using AutoGen & Hydra
Difficulty
Advanced
Rigor
Not declared
The author has not specified a rigor level.
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
actuarial_underwriting_eval
Evaluates catastrophe risk premium calculation and consensus report generation.
Input format
Hydra dynamic YAML config + property portfolio details
Output format
JSON actuarial opinion summary with calculated technical premium