Implement Gemini Multimodal Interaction

implementationChallenge

Prompt Content

Extend your Google ADK agent to use Gemini 1.5 Pro for processing user utterances and generating responses. Your agent should be able to interpret natural language commands related to navigation and safety. Show how you would configure the ADK to use Gemini as the primary LLM for reasoning and how to pass a user's voice input to Gemini and then use Gemini's response to drive the `speak` tool. Also, outline how you would pass environmental context (like current location and activity) to Gemini.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations