Integrate Giskard for Evaluation

testingChallenge

Prompt Content

Integrate Giskard into your agent's development loop. Specifically, create a Giskard test suite to evaluate the 'policy_evaluation_tool's accuracy. For instance, provide scenarios where a specific action (e.g., 'transfer data to non-EU server without SCCs') should 'FAIL' under GDPR, and ensure your agent's tool correctly identifies this. Report the Giskard evaluation results.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations