Source: DeepSeekSeptember 11, 2026

DeepSeek Releases V4.1-Flash Model at $0.15 Per Million Tokens

DeepSeek released V4.1-Flash, a lightweight open-weight model priced around $0.15 per million tokens, positioned for high-throughput, cost-sensitive workloads rather than benchmark-topping reasoning.

Key Points:

• V4.1-Flash is an open-weight model with native vision support, built for speed and throughput rather than topping benchmark leaderboards.

• The pricing is dramatically below most Western frontier models, continuing a pattern of Chinese open-weight labs pricing aggressively even as the performance gap between the two narrows.

• The release is aimed at developers and small businesses running high-volume AI tasks — like summarizing large batches of documents or classifying customer messages — who don't need top-tier reasoning for every task.

Open access and aggressive pricing win adoption faster than proprietary control does. That trend quietly erodes the pricing power frontier labs have relied on to justify premium subscription tiers.

Why It Matters: At roughly $0.15 per million tokens, the cost gap alone may justify switching for high-volume, low-complexity workloads — benchmark DeepSeek's V4.1-Flash against your current model this week.

DeepSeek Releases V4.1-Flash Model at $0.15 Per Million Tokens | AI Onboarded