Back to evaluations
Public evaluation

catalog_authenticity_audit

Evaluates multi-agent catalog audit accuracy against gold luxury brand verification datasets.

Evaluation type
task based
Challenge
Build Luxury E-Commerce Verification Crew with CrewAI
Difficulty
Intermediate
Rigor
Unspecified

Evaluation overview

How the linked challenge is judged: tasks, benchmarks, and criteria count.

Tasks
1
Benchmarks
0
Criteria
0

Task templates

Inputs and expected outputs.

Task 1

catalog_authenticity_audit

Evaluates product catalog details and returns brand risk and listing clearance status.

Input format

JSON containing product_title, seller_name, price_inr, MSRP_inr, image_url

Output format

JSON containing authenticity_status, risk_score, and flagged_reasons