DSpark
Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation
DSpark is DeepSeek's speculative-decoding framework/draft-model algorithm for large language model inference. It combines semi-autoregressive parallel block drafting with confidence-scheduled verification to speed token generation without changing the target model, and is released via DeepSpec configs/checkpoints plus DSpark-attached DeepSeek and Qwen/Gemma checkpoints.

Recent stories
1 linked story