Back to evaluations
Draft evaluation
Build Type-Safe UPI Fraud Triage Agent with Pydantic AI — evaluation
Evaluates type safety compliance and fraud triage accuracy on live UPI payment streams.
Evaluation type
task based
Challenge
Build Type-Safe UPI Fraud Triage Agent with Pydantic AI
Difficulty
Intermediate
Rigor
Not declared
The author has not specified a rigor level.
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
upi_fraud_triage
Processes transaction signal vector and returns validated FraudResult Pydantic schema.
Input format
JSON containing transaction_amount_inr, velocity_1h, merchant_category_code, device_fingerprint_match
Output format
JSON matching FraudResult schema: risk_level, block_transaction, trigger_otp