After DeepSeek's new pricing took effect, InferX shared an alternative pricing option for V4 Flash:
"After DeepSeek's new pricing took effect, I wanted to share an alternative for those running V4 Flash at scale. We offe… This is a necessary excerpt; read the original source for full context.
u/V5489: "Over the past few hours today, I used 27,453,272 tokens across 181 API requests and spent $0.28 in total. 🤷♂️" (showing Flash's actual low cost)
u/FidgetsAndFish: "CommandCode looks like a better deal. Why should I choose InferX?" (price competition among third-party aggregation platforms)
u/Important-Fly-2105: "Don't use inferx; the cache hit rate is 60–70%, and the TTL is only 5 minutes."
After DeepSeek's official price increase, third-party aggregation platforms (InferX, CommandCode, and others) resell V4 Flash at lower prices, creating price competition.
User testing still shows an extremely low cost (about $0.28 for 27 million tokens).
After the price increase, "finding alternative channels" has become a mainstream topic of discussion in the community.
DeepSeek V4 Flash