Speculative decoding
DDD
Superseded baseline#20 of 151 most-superseded
Superseded — cited as a baseline and beaten by newer methods
0 papers critique it · 4 beat it on benchmarks
Beaten on benchmarks
Head-to-head results where a newer method reports beating DDD. Values are copied from the source paper's tables — verify against the cited paper.
ECHO beats DDD
5.25 vs 3.97
Avg. Speedup · [Vicuna-13B]
ECHO: Elastic Speculative Decoding with Sparse Gating for High-Concurrency ScenariosEVICT beats DDD
147.69 vs 112.25
Average · [Temperature = 0]
Making Every Verified Token Count: Adaptive Verification for MoE Speculative Decodingmethod beats DDD
5.41 vs 4.97
Avg. Speedup · [Vicuna-13B]
Draft Less, Retrieve More: Hybrid Tree Construction for Speculative DecodingLDLP-DDD beats DDD
71.20 vs 69.39
What to use instead
Recent methods in the same sub-problem, not yet superseded in the knowledge base — arXiv benchmark leaders, not vetted production recommendations.
- Jun 3, 2026
- Jun 2, 2026
- Hybrid Verified DecodingHybrid Verified Decoding: Learning to Allocate Verification in Speculative DecodingMay 31, 2026
- May 28, 2026
- May 28, 2026
- May 28, 2026
- May 19, 2026
- May 19, 2026
- May 9, 2026
- May 8, 2026
- May 1, 2026
- Apr 21, 2026