Zenaique
Topics
Practice
Study
Browse
Reference
Pricing
Search…
⌘K
Topics
/
Inference Optimization
/
Decode Phase
Decode Phase
Subtopic
7 questions
Questions tagged with Decode Phase — part of Inference Optimization.
Premium questions for this topic
Format
Difficulty
Role
Experience
Companies
Sort
Newest
Quality
Difficulty ↑
Difficulty ↓
Questions
Find the wrong move in 'decode is slow, so let's switch to a smaller FLOP model'
Spot the Error
Medium
Walk through what actually happens during the prefill phase of an LLM forward pass.
Flashcard
Easy
What unit measures GPU memory bandwidth and why does that number cap decode speed?
Flashcard
Easy
max_tokens vs stop: pair each parameter with the cap it enforces on the completion.
Match Pairs
Easy
Define LLM inference and explain how it differs from training.
Flashcard
Easy
Predict the bandwidth vs compute latency of a single Llama-70B decode step on H100
Predict Output
Hard
Spot the errors in this 'API is slow because of network and tokenizer' explanation
Spot the Error
Medium
Cred
Jasper
Ey
NVIDIA
Canva
Induced Ai
Fireworks Ai
Jpmorgan
Datarobot
Evenup
Inflection Ai
Infosys
Airbnb
Graphcore