← 全部模型

Ling · 近 7 天

Ling 3.0 Tiny

0–100 分,越高代表社区使用体验越正面,不代表能力测试成绩。

更新于 2026-09-05 14:21 UTC · 2026-08-29 – 2026-09-05 UTC
社区口碑分69.2/10019 条评论 · 近 7 天

社区评论

精选近 7 天的不同意见,摘录条数不代表真实好评比例。

3 条精选摘录

负面速度与延迟

“令我惊讶的是,我可以用一些卸载的方式以20-25 tok/sec的范围运行Gemma 4 12b和Gemma 4 26b两个QAT版本。Ling 3 Tiny应该肯定比它们快得多。”

Redditen机器翻译
展开原文

Why I find it surprising is I can run Gemma 4 12b and Gemma 4 26b both QAT versions with some offloading at the 20-25 tok/sec range. Ling 3 Tiny should surely be much faster than them.

负面速度与延迟

“Ling-3.0-Tiny只给我30 t/s,而Ling-mini-2.0在纯CPU推理上给我50-60 t/s。GPU-CUDA上也是如此,70-80 t/s对比150+ t/s”

Redditen机器翻译
展开原文

Ling-3.0-Tiny gives me only 30 t/s while Ling-mini-2.0 gives me 50-60 t/s on CPU-only inference itself. Same with GPU-CUDA. 70-80 t/s vs 150+ t/s

正面速度与延迟

“我用 ling-3.0-tiny 来处理 OpenZim MCP 服务器运气不错,比如这个模型会翻遍该死的百科全书找出答案(而且我能验证答案的来源)。当然 Qwen3.8 也能做到一样好,”

Redditen机器翻译
展开原文

I have had good luck using ling-3.0-tiny for the OpenZim MCP server for example, that model will work its way through the damn encyclopedia and pull out the answers (and I can verify where they came from). Sure Qwen3.8 does it just as well,