← All models

Ling · Last 7 days

Ling 3.0 Tiny

0–100 · higher means more positive community experience, not benchmark performance.

Updated 2026-09-05 14:21 UTC · 2026-08-29 – 2026-09-05 UTC
Community score69.2/10019 comments in 7 days

Community reviews

Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.

8 selected excerpts

PositiveLocal deploy

“6gb here. Put Ling tiny in a harness and go wild”

NegativeSpeed & latency

“Why I find it surprising is I can run Gemma 4 12b and Gemma 4 26b both QAT versions with some offloading at the 20-25 tok/sec range. Ling 3 Tiny should surely be much faster than them.”

NegativeSpeed & latency

“Ling-3.0-Tiny gives me only 30 t/s while Ling-mini-2.0 gives me 50-60 t/s on CPU-only inference itself. Same with GPU-CUDA. 70-80 t/s vs 150+ t/s”

NegativeGeneral text

“LFM2.5-2.6B - It turned out significantly better for me. I tested it a bit, though, specifically in terms of extracting and finding the desired text note in obsidian and ling 3.0 tiny shows more tool calls issues”

PositiveGeneral text

“Ling-tiny is really cool”

PositiveSpeed & latency

“I have had good luck using ling-3.0-tiny for the OpenZim MCP server for example, that model will work its way through the damn encyclopedia and pull out the answers (and I can verify where they came from). Sure Qwen3.8 does it just as well,”

PositiveCoding

“I find Ling 3.0 tiny particularly interesting as it looks really nice for a tiny model with 7.9B total parameters, with only 1.3B parameters activated per token”

PositiveReasoning

“smarter than many 30b parameter models”