Test and Evaluate Comprehensive Vulnerability Assessment

testingChallenge

Prompt Content

Provide your Claude agent with a known vulnerable code snippet and trigger a full vulnerability assessment using your implemented tools. Evaluate the agent's generated report against ground truth vulnerabilities, focusing on accuracy, clarity of explanations, and quality of remediation suggestions. Use the `VulnerabilityAssessment` evaluation task template for structured testing.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations