← All models

DeepSeek · Last 7 days

DeepSeek V4 Flash Vision Exp

0–100 · higher means more positive community experience, not benchmark performance.

Updated 2026-09-05 14:21 UTC · 2026-08-29 – 2026-09-05 UTC
Community score59.2/10033 comments in 7 days

Community reviews

Selected comments from the last 7 days. The balance of excerpts does not represent the share of positive reviews.

10 selected excerpts

NegativeImage / vision

“I tried it (full precision, DeepSeek-V4-Flash-Vision-Exp) and it turned out DeepSeek's vision capability is very limited. It can't read normal letter text. It's a single-grid design with no tiling, and everything gets squashed into a 384-to”

PositiveSpeed & latency

“you CAN NOW RUN IT ON 6K's! I'm testing it right now and it's awesome. ~230tps during reasoning, ~350tps on code generation, LMCache for instant prefix cache and ~7K prefill”

NegativeSpeed & latency

“currently on 2x Spark you lose -20% raw perf”

PositiveLocal deploy

“just switched to deepseek v4 flash vision exp since it was released earlier this week”

PositiveSpeed & latency

“In actual operation, it only took about 1 minute to complete the task, the efficiency is truly amazing”

ZhihuzhMachine translated
Show original

实操过程中仅仅用了1分钟左右就完成了任务,效率之高令人惊叹

NegativeImage / vision

“In my anecdotal testing this model seems less capable than for example Qwen 3.8 27B, which for me was a big step up in vision e.g. in its precision to come up with coordinates for bounding boxes on an image using a normalized 0-100 scale.”

PositiveReasoning

“DeepSeek v4 Flash Vision Exp looks to have an incremental improvement in addition to supporting vision and for users with dual DGX Sparks it's still the best model available and really the only large model that can be run at the native quan”

PositiveImage / vision

“Works very well so far, including vision”

PositiveGeneral text

“better at talking”

PositiveGeneral text

“another gift !!! Best summer ever for local LLM”