"Kill line" sounds like clickbait, but the blue dot on the chart really earns it:
"DeepSeek V4 Flash 0731 scores about 50 on the Artificial Analysis index, at about 3 cents per weighted task. On… This is a necessary excerpt; read the original source for full context.
Another awkward detail: The current Pro preview is worse than Flash 0731 on this chart; Pro has not beaten anything on this chart so far.
u/StupidScaredSquirrel: If models on the Pareto frontier "kill" all offline models, then most of this chart would have already been killed by other models—in reality, that's not how it works; every use case has other requirements
u/Subject-18: OpenAI deliberately left Mimo 2.5 out of its Luna price-cut post; hilarious
u/FilterJoe: Asked what "per task" actually means (whether retry tokens, task types, etc., are included)
The "kill line" concept is similarly popular in English-speaking communities: V4 Flash 0731 achieves an AA score of about 50 at roughly 3 cents per task, occupying an advantageous position at the far lower-left of the price-performance curve
The post is a highly upvoted r/LocalLLM post (812 points), showing that awareness of its price-performance value has spread beyond the niche
DeepSeek V4 Flash