← All models

Ling · Last 7 days

Ling 3.0 Tiny

0–100 · higher means more positive community experience, not benchmark performance.

Updated 2026-09-05 16:01 UTC · 2026-08-29 – 2026-09-05 UTC
Community score69.2/10019 comments in 7 days

Community reviews

Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.

2 selected excerpts

NegativeSpeed & latency

“Why I find it surprising is I can run Gemma 4 12b and Gemma 4 26b both QAT versions with some offloading at the 20-25 tok/sec range. Ling 3 Tiny should surely be much faster than them.”

NegativeSpeed & latency

“Ling-3.0-Tiny gives me only 30 t/s while Ling-mini-2.0 gives me 50-60 t/s on CPU-only inference itself. Same with GPU-CUDA. 70-80 t/s vs 150+ t/s”