Best models for your hardware this week.
8-12GB
- https://huggingface.co/LiquidAI/LFM2.5-8B-A1B incredible model, so fast, so small
16-32GB
- latest Google model, Gemma 12B: https://huggingface.co/google/gemma-4-12B really solid performance up neck and neck with a model 2x its size from a month ago.
Jetbrains new model, best in class on livecode bench
32-96gb
- Nex-N2-Mini GPT style postrain of Qwen-35B it seems to be its class leader caveman style reasoning https://huggingface.co/Nexdata/Nex-N2-Mini
- Jackrong’s Qwopus is the #1 overall Q4 of Qwen3.6-27B on our benchmark suite of 5 agent + coding benchmarks (1200 samples total) https://huggingface.co/jackrong/QwQ-32B-Preview-Qwopus
192gb
- Step-3.7-Flash is hard to beat, high scores, really fast inference, vision capable, later cutoff dates https://huggingface.co/stepfun-ai/Step-3.7-Flash
384gb
- Nex-N2-Pro GPT style post train of Qwen-3.5-397B incredibly strong and #1 on deepswe if their claims are right https://huggingface.co/Nexdata/Nex-N2-Pro
768gb
- very promising post-train of GLM-5.1 that wins out on 8 benchmarks
This post (sticker, poll or similar) has no web preview. Open in Telegram







