Full System Testing & Transparency

testingChallenge

Prompt Content

Select 3-5 challenging scientific problems (e.g., from a simulated FrontierScience-like benchmark) and run your complete LangGraph-based agent system through them. Document the system's final conclusions, but critically, also trace and present the entire reasoning path, including tool calls, intermediate thoughts, and any self-correction steps. Evaluate the transparency and verifiability of the process.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations