DeepSeek V4-Flash
BudgetOpen weightsDeepSeek
Open-weight price-performance pick: MIT weights, 1M context, and API rates far below closed budget tiers.
- text
Best for: Cost-sensitive hosted agents · High-volume coding assist · Open-weight deployments with cluster VRAM
Watch out: API list prices move to peak/off-peak from 2026-08-16 16:00 UTC (see pricing note). Self-hosting the ~284B MoE still needs roughly 90–170GB+ class VRAM depending on quant — not a laptop budget build. DeepSWE snapshot also uses the shared mini-swe-agent harness — verify cost, effort, and serving conditions before treating the result as a forecast.
- Released
- July 31, 2026
- Input / 1M
- $0.14
- Output / 1M
- $0.28
- Context
- 1M
- Max output
- 384K