· 7 min read
TAPS: Why Your Draft Model's Training Data Matters More Than Its Architecture
TAPS shows draft training data drives speculative decoding: merged-tree verification of math and chat specialists hits 5.11 acceptance length, weight averaging 2.59.