2026-06-02•4 min readThe GPU Cycles You Already Paid For Are Training Your Next ModelMIT's TLT uses idle RL training compute to train adaptive drafters and accelerate long-tail rollout generation.#llm-training#efficiency#mit#speculative-decoding#reinforcement-learning