Hugging Face (Twitter)
RT @UnslothAI: You can now train MoE models 12× faster with 35% less VRAM via our new Triton kernels (no accuracy loss).
Train gpt-oss locally on 12.8GB VRAM.
In collab with @HuggingFace, Unsloth trains DeepSeek, Qwen3, GLM faster.
Repo: github.com/unslothai/unsloth
Blog: https://unsloth.ai/docs/new/faster-moe
Post #2291
150
