← 全部模型

Gemini · 近 7 天

Gemini 3.8 Flash

0–100 分,越高代表社区使用体验越正面,不代表能力测试成绩。

更新于 2026-09-05 14:21 UTC · 2026-08-29 – 2026-09-05 UTC
社区口碑分43.5/100140 条评论 · 近 7 天

社区评论

精选近 7 天的不同意见,摘录条数不代表真实好评比例。

6 条精选摘录

正面编程

“DeepSWE v1.1,3.8 Flash 得分 73.7%,较 3.7 Flash 的 65.3% 提升超过 8 个百分点,Ars Technica 报道称它已登顶该榜单。专注智能体终端编码的 Terminal-bench 2.1 上,它拿到 89.4%,超过了一众旗舰模型。”

正面编程

“编程肯定3.8flash”

褒贬兼有编程

“gemini 3.8 flash 编程确实非常出色。问题在于指令遵循和做出错误假设,还有就是你给它一个计划它却返回一个更差的计划。但我注意到如果”

Redditen机器翻译
展开原文

gemini 3.8 flash is genuinely excellent at coding. It's the instruction following and making wrong assumptions that suck and the annoying thing where you give it a plan and it comes back with an inferior plan. But one thing i've noticed if

负面编程

“3.8 flash 跑分刷到极致,垃圾模型”

Redditen机器翻译
展开原文

3.8 flash is benchmaxxed to hell, garbage model

正面编程

“基于我的使用体验,它在 coding 方面的提升是非常不错的。对比 3.6 -> 3.7 从不可用到可用,其实我个人觉得 3.7 -> 3.8 可以算是 可用到好用了”

中性编程

“Deep SWE 是单任务短编程,而 Terminal Bench 4 是平均五六个小时的长任务编程。Flash 模型,Terminal Bench 4 基本不会很高”