Back to Prompt Library
planning
Architect Multimodal Input Pipeline
Inspect the original prompt language first, then copy or adapt it once you know how it fits your workflow.
Linked challenge: Edge Multimodal AI for AR Glasses: Real-time Assistant
Format
Text-first
Lines
1
Sections
1
Linked challenge
Edge Multimodal AI for AR Glasses: Real-time Assistant
Prompt source
Original prompt text with formatting preserved for inspection.
1 lines
1 sections
No variables
0 checklist items
Design the end-to-end multimodal input pipeline. Outline how voice (transcribed via `Fixie`), simulated camera input (object detection/scene understanding), and gesture data (from a proxy) will be fused and prepared for `Gemini 2.5 Pro`. Focus on real-time data flow and context enrichment.
Adaptation plan
Keep the source stable, then change the prompt in a predictable order so the next run is easier to evaluate.
Keep stable
Preserve the role framing, objective, and reporting structure so comparison runs stay coherent.
Tune next
Swap in your own domain constraints, anomaly thresholds, and examples before you branch variants.
Verify after
Check whether the prompt asks for the right evidence, confidence signal, and escalation path.