Evaluate Multi-Agent Performance and Reporting Accuracy

testingChallenge

Prompt Content

Simulate a scenario where a policy directive on 'Supply Chain Security' is provided to the 'Intelligence Analyst'. The analyst should then task the 'Tool Scout' to find relevant tools. After the report is generated, evaluate its accuracy and completeness against known cybersecurity standards. Use a prompt to trigger this full workflow and collect the final report and a simulated Cartesia output.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations