Back to evaluations
Draft evaluation
Vaccine Safety Literature Synthesis Agent with Claude Agents SDK — evaluation
Evaluates claim verification accuracy, citation correctness, and reasoning trace clarity.
Evaluation type
task based
Challenge
Vaccine Safety Literature Synthesis Agent with Claude Agents SDK
Difficulty
Advanced
Rigor
Not declared
The author has not specified a rigor level.
Evaluation overview
How the linked challenge is judged: tasks, benchmarks, and criteria count.
Tasks
1
Benchmarks
0
Criteria
0
Task templates
Inputs and expected outputs.
Task 1
literature_synthesis_task
Synthesizes scientific literature to answer vaccine safety questions.
Input format
JSON containing topic query and list of study abstract documents.
Output format
JSON containing consensus conclusion, evidence quality grade, and citations.