Back to evaluations
Draft evaluation

Vaccine Safety Literature Synthesis Agent with Claude Agents SDK — evaluation

Evaluates claim verification accuracy, citation correctness, and reasoning trace clarity.

Evaluation type
task based
Challenge
Vaccine Safety Literature Synthesis Agent with Claude Agents SDK
Difficulty
Advanced
Rigor
Not declared

The author has not specified a rigor level.

Evaluation overview

How the linked challenge is judged: tasks, benchmarks, and criteria count.

Tasks
1
Benchmarks
0
Criteria
0

Task templates

Inputs and expected outputs.

Task 1

literature_synthesis_task

Synthesizes scientific literature to answer vaccine safety questions.

Input format

JSON containing topic query and list of study abstract documents.

Output format

JSON containing consensus conclusion, evidence quality grade, and citations.