DSpark
Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation
DSpark is DeepSeek-AI's speculative decoding framework/drafter approach for accelerating LLM inference. It combines semi-autoregressive parallel draft generation with confidence-scheduled, load-aware verification, and DeepSeek released DSpark checkpoints and implementations through the DeepSpec repository.

Recent stories
0 linked stories
No linked stories yet.