Back to evaluations
Public evaluation
catalog_authenticity_audit
Evaluates multi-agent catalog audit accuracy against gold luxury brand verification datasets.
Evaluation type
task based
Challenge
Build Luxury E-Commerce Verification Crew with CrewAI
Difficulty
Intermediate
Rigor
Unspecified
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
catalog_authenticity_audit
Evaluates product catalog details and returns brand risk and listing clearance status.
Input format
JSON containing product_title, seller_name, price_inr, MSRP_inr, image_url
Output format
JSON containing authenticity_status, risk_score, and flagged_reasons