DFlash

Speculative inference engine developed by Luce and DeepSeek to accelerate local llm inference by combining token prediction with advanced compression techniques.

Core Features

Recent Developments & Benchmarks

References