Create a demo for Model inference that measures response time and prints a simple latency report.
Run your code to see output here...