← All categories

Current ranking

Local deploy

For the 7-day window ending 2026-09-13 21:12:03 UTC, user-experience leaders in Local deploy: closed models — No model meets the sample threshold; open models — Qwen3.8 27B (70.7/100, n=28).

Top 3 specific models in each category. A model appears when its category sample size is at least 5. Rank movement compares the previous published snapshot.

No model currently meets the sample threshold.

#1 Qwen3.8 27Bn=28 70.7

“dense 27B fitting entirely in vram is really the move here”

via Reddit · 2026-09-11

“you can easily go for (good!) 5-6 bit quants with 3.8 27b imo.”

via Reddit · 2026-09-11
Open model detail →
#2 Qwen3.6 35B A3Bn=5 65.1

“qwen3.6 35b a3b is the best for this setup, you can get it running quite well”

via Reddit · 2026-09-11
Open model detail →
#3 GLM 5.3 Flashn=8 63.0

“GLM 5.3 Flash is a clear winner for 128gb vram, for me.”

via Reddit · 2026-09-10

“GLM 5.3 Flash at Q4”

via Reddit · 2026-09-10
Open model detail →