Advanced
Pro
Cost and Latency Engineering for LLM Apps Quiz
Test your grasp of token accounting, prompt-prefix caching, semantic caching, model routing and cascades, streaming vs. batching, the TTFT-plus-decode latency model, and cost-per-1,000-conversations arithmetic.
10 questions
10 min
Pass: 70%