DeepSeek's New Flash Model Costs $0.14/Mtok and Outranks MiniMax M3
DeepSeek-V4-Flash-0731 is a 304 billion parameter open-weight model with enhanced agentic capabilities, priced at $0.14 per million input tokens and $0.27 per million output tokens.
What it is
DeepSeek-V4-Flash-0731 is the newest release in DeepSeek's V4 model family, a 304B parameter, 167GB open-weight model available on Hugging Face.
What it does
It delivers what DeepSeek describes as substantially enhanced agentic capabilities, and Artificial Analysis ranks it ahead of MiniMax M3, a 428B model, despite being significantly smaller.
Why it matters
At $0.14/$0.27 per million tokens it may currently offer the best value-per-intelligence of any available model, an alternative worth checking before defaulting to a larger, pricier model for agentic workloads.
How to use it
Pull the weights from Hugging Face or use it via API where offered, and benchmark it against your current agentic model before switching.