10 questions · need 7/10 to pass.
Q1.When applying "How the Model Picks Its Next Word (Sampling)" in practice, which of these holds?
single
Q2.Which statement about how "Where the Compute Actually Goes (FLOPs)" actually works is correct?
single
Q3.For "Reading Your Prompt vs. Writing the Answer (Prefill vs. Decode)", which detail or constraint from the module is accurate?
single
Q4.For "Your First Local Inference Run", which detail or constraint from the module is accurate?
single
Q5.When applying "Reading the Model's Confidence (Logprobs)" in practice, which of these holds?
single
Q6.Which definition of "Reading Your Prompt vs. Writing the Answer (Prefill vs. Decode)" matches what the module established?
single
Q7.Which fact about "Tokens, Not Characters (Tokenization at Serve Time)" matches the mechanism the module covered?
single
Q8.Which statement about how "What Happens When You Call an LLM" actually works is correct?
single
Q9."How Generation Knows When to Stop" — which of these claims is supported by the module?
single
Q10.Which of these correctly identifies the role of "What Happens When You Call an LLM" in the broader system?
single