The author previously used Claude heavily and GPT/Codex at work (via OpenCode). After their Claude subscription expired, they switched to domestic models to replace it across the board:
Their main model is MiniMax, with Kimi used to understand screenshots (by copying and pasting them), but it sometimes gets stuck
Xiaomi's model is good but slow and expensive; deepseek v4 pro is good but slow, inefficient, and expensive
After using deepseek-v4-flash for about two days, they can recommend it:
"Insanely fast, 100 to 150 TPS, super good. It's a bit outdated and needs reminders (for example, context7) about what w… This is a necessary excerpt; read the original source for full context.
If you haven't tried it yet: API usage may cost less than $1 a day, so it's worth a try
"This is good stuff—cheap, insanely fast, and good. Fuck Claude and Anthropic."
Tested speed: 100–150 tokens/second
Estimated cost: < $1/day (API usage)
The community rates its price-performance ratio extremely highly, with "monster" becoming the keyword of the post
DeepSeek V4 Flash