DeepSeek V4 Flash · Community source · Personal experience
An OpenCode user subjectively finds Flash cheap and fast enough to replace parts of a Claude/GPT workflow, while describing Pro as slower and costlier.
The author previously used Claude heavily and GPT/Codex at work (via OpenCode). After their Claude subscription expired, they switched to domestic models to replace it across the board:
Their main model is MiniMax, with Kimi used to understand screenshots (by copying and pasting them), but it sometimes gets stuck
Xiaomi's model is good but slow and expensive; deepseek v4 pro is good but slow, inefficient, and expensive
After using deepseek-v4-flash for about two days, they can recommend it:
"Insanely fast, 100 to 150 TPS, super good. It's a bit outdated and needs reminders (for example, context7) about what w… This is a necessary excerpt; read the original source for full context.
If you haven't tried it yet: API usage may cost less than $1 a day, so it's worth a try
"This is good stuff—cheap, insanely fast, and good. Fuck Claude and Anthropic."
Tested speed: 100–150 tokens/second
Estimated cost: < $1/day (API usage)
The community rates its price-performance ratio extremely highly, with "monster" becoming the keyword of the post
The figures, task set, reasoning tier, and client conditions apply only to the listed source and collection snapshot. Different versions, harnesses, or providers must not be compared directly; undisclosed parameters remain unknown.
For a reproduction, fix the model version, provider or client, reasoning tier, tools, task-set version, sample count, and collection date, and record failures, retries, and human corrections. Full steps are in the source notes below.
Reddit r/opencodeCLI · u/Maleficent-Movie-625 · Original publication date Unknown · Site edit date 2026-09-20
Open original sourceDeepSeek V4 Flash
Download the Tabbit client to check model access