DSPy Evaluation Pipeline Design

planningChallenge

Prompt Content

Design a DSPy evaluation pipeline that can take an LLM's output and a corresponding benchmark question, then verify its factual accuracy using an external knowledge source (simulated via LlamaIndex RAG). Detail how DSPy modules will orchestrate calls to Gemini 2.5 Pro for verification and incorporate hybrid reasoning.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations