#
dflash2
Here are 4 public repositories matching this topic...
Evidence-first lab for speculative decoding, DFlash/DFlash2, block diffusion, exactness, acceptance, and inference performance
-
Updated
Aug 21, 2026 - Python
Qwen 3.8 DFlash2: 27B AI Running at 236 TOK/S Locally! - High-speed local prose and code autocomplete powered by speculative decoding with lightweight draft models.
-
Updated
Aug 19, 2026 - Python
Improve this page
Add a description, image, and links to the dflash2 topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the dflash2 topic, visit your repo's landing page and select "manage topics."