DeepSeek-V4-Flash: The Open-Weight AI Model Pricing Inference at $0.08/M — Day 7/30
This story is from 2026-08-27. It is preserved in the archive; the latest stories are on the live feed.
TL;DR — DeepSeek-V4-Flash pairs a 1,048,576-token context with $0.078/M input and $0.156/M output pricing, and my probes show it nailing code, math, and structured extraction. But its reasoning-task latency was wildly inconsistent — a real cost for anyone building agent loops on top of the low stic…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-08-27 13:16 · DEV Community — AI
DeepSeek-V4-Flash: The Open-Weight AI Model Pricing Inference at $0.08/M — Day 7/30