Simulate Inference Optimization with TensorRT-LLM

testingChallenge

Prompt Content

Implement a simulated tool for the 'Performance Engineer' agent that represents the functionality of TensorRT-LLM. The agent should be able to 'call' this tool to optimize a given model and report on simulated performance gains (e.g., reduced latency, increased throughput). Demonstrate this with a sample agent interaction.

Try this prompt

Open the workspace to execute this prompt with free credits, or use your own API keys for unlimited usage.

Usage Tips

Copy the prompt and paste it into your preferred AI tool (Claude, ChatGPT, Gemini)

Customize placeholder values with your specific requirements and context

For best results, provide clear examples and test different variations