Current ranking
Local deploy
For the 7-day window ending 2026-09-13 21:12:03 UTC, user-experience leaders in Local deploy: closed models — No model meets the sample threshold; open models — Qwen3.8 27B (70.7/100, n=28).
No model currently meets the sample threshold.
#1 Qwen3.8 27Bn=28 70.7▾
“dense 27B fitting entirely in vram is really the move here”
via Reddit · 2026-09-11
Open model detail →“you can easily go for (good!) 5-6 bit quants with 3.8 27b imo.”
via Reddit · 2026-09-11
#2 Qwen3.6 35B A3Bn=5 65.1▾
Open model detail →“qwen3.6 35b a3b is the best for this setup, you can get it running quite well”
via Reddit · 2026-09-11
#3 GLM 5.3 Flashn=8 63.0▾
“GLM 5.3 Flash is a clear winner for 128gb vram, for me.”
via Reddit · 2026-09-10
Open model detail →“GLM 5.3 Flash at Q4”
via Reddit · 2026-09-10