DFlash
Qwen3-4B-DFlash-b16 is a diffusion draft model trained against a frozen Qwen3 target for speculative decoding.
Example Model and Recipe
See the dLLM Fine-Tuning Guide.
Qwen3-4B-DFlash-b16 is a diffusion draft model trained against a frozen Qwen3 target for speculative decoding.
See the dLLM Fine-Tuning Guide.