Test & Refine Stress Scenarios

testingChallenge

Prompt Content

Create a testing harness that simulates various stress scenarios (e.g., short deadlines, ambiguous input). Run your agent through these scenarios and analyze its performance, particularly focusing on the 'adaptive_actions' and 'misbehavior_score'. Refine the agent's logic based on observed outcomes.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations