Back to evaluations
Public evaluation

actuarial_underwriting_eval

Evaluates the consistency, accuracy, and auditability of premium recommendations across stochastic multi-agent deliberation runs.

Evaluation type
task based
Challenge
Orchestrate Multi-Agent Actuarial Risk Rating using AutoGen & Hydra
Difficulty
Advanced
Rigor
Unspecified

Evaluation overview

How the linked challenge is judged: tasks, benchmarks, and criteria count.

Tasks
1
Benchmarks
0
Criteria
0

Task templates

Inputs and expected outputs.

Task 1

actuarial_underwriting_eval

Evaluates catastrophe risk premium calculation and consensus report generation.

Input format

Hydra dynamic YAML config + property portfolio details

Output format

JSON actuarial opinion summary with calculated technical premium