Method DriftSpeculative decoding

Tracked

sd.npu

Accelerating Mobile Language Model via Speculative Decoding and NPU-Coordinated Execution

Speculative decoding · first seen Oct 17, 2025

current frontier — recent, not yet superseded in the knowledge base

0 papers critique it · 0 beat it on benchmarks

Newer alternatives

Recent methods in the same sub-problem, not yet superseded in the knowledge base.