Accelerate LLM inference by up to 3x using speculative decoding. Learn how draft models, Medusa, and EAGLE reduce latency without compromising output quality.