Testing and Complex Scene Analysis

testingChallenge

Prompt Content

Run the agent with both 'Object_Compliance_Check' and 'Complex_Scene_Analysis' evaluation tasks. Create varied 'scene_description' inputs with deliberate errors or ambiguities. Analyze Gemini 2.5 Pro's reasoning traces and the agent's output. Refine the agent's prompts and plugin usage to improve its accuracy in detecting subtle errors and its ability to provide clear, actionable recommendations for complex scenarios.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations