WeDLM-8B Instruct does not use autoregression like regular LLMs,
but a diffusion method for text generation.
What does this provide?
🚀 In mathematical reasoning tasks, the model works 3–6 times faster
than Qwen3-8B even with vLLM optimizations - while maintaining quality.
This release breaks the old myth that "diffusion models are not suitable for precise text tasks".
In practice, WeDLM shows that such an approach can compete
and even outperform transformers in inference speed.
The model is open and available under the Apache 2.0 license:
GitHub, HuggingFace
••••••••••••••••••••••••••••••••••••••••••••••••••••
🤖 Data Science, ML & Big Data with @DataXplore
