Back to Prompt Library
planning

Architect Multimodal Input Pipeline

Inspect the original prompt language first, then copy or adapt it once you know how it fits your workflow.

Linked challenge: Edge Multimodal AI for AR Glasses: Real-time Assistant

Format
Text-first
Lines
1
Sections
1
Linked challenge
Edge Multimodal AI for AR Glasses: Real-time Assistant

Prompt source

Original prompt text with formatting preserved for inspection.

1 lines
1 sections
No variables
0 checklist items
Design the end-to-end multimodal input pipeline. Outline how voice (transcribed via `Fixie`), simulated camera input (object detection/scene understanding), and gesture data (from a proxy) will be fused and prepared for `Gemini 2.5 Pro`. Focus on real-time data flow and context enrichment.

Adaptation plan

Keep the source stable, then change the prompt in a predictable order so the next run is easier to evaluate.

Keep stable

Preserve the role framing, objective, and reporting structure so comparison runs stay coherent.

Tune next

Swap in your own domain constraints, anomaly thresholds, and examples before you branch variants.

Verify after

Check whether the prompt asks for the right evidence, confidence signal, and escalation path.