← All models

Ling · Last 7 days

Ling 3.0 Tiny

0–100 · higher means more positive community experience, not benchmark performance.

Updated 2026-09-05 14:21 UTC · 2026-08-29 – 2026-09-05 UTC
Community score69.2/10019 comments in 7 days

Community reviews

Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.

3 selected excerpts

NegativeSpeed & latency

“Why I find it surprising is I can run Gemma 4 12b and Gemma 4 26b both QAT versions with some offloading at the 20-25 tok/sec range. Ling 3 Tiny should surely be much faster than them.”

NegativeSpeed & latency

“Ling-3.0-Tiny gives me only 30 t/s while Ling-mini-2.0 gives me 50-60 t/s on CPU-only inference itself. Same with GPU-CUDA. 70-80 t/s vs 150+ t/s”

PositiveSpeed & latency

“I have had good luck using ling-3.0-tiny for the OpenZim MCP server for example, that model will work its way through the damn encyclopedia and pull out the answers (and I can verify where they came from). Sure Qwen3.8 does it just as well,”