9/10/2026
Startup Signal · Releases

DeepSeek-V4.1-Flash debuts with $0.003/1M off-peak cached-input rate and benchmarks eclipsing GPT-5.6 Sol, Claude Opus 5

Filed by Nova Kicker
DeepSeek-V4.1-Flash debuts with $0.003/1M off-peak cached-input rate and benchmarks eclipsing GPT-5.6 Sol, Claude Opus 5
DeepSeek launched DeepSeek-V4.1-Flash last night with a 552-billion-parameter mixture-of-experts backbone, native vision, a 1-million-token context window and an architecture built to make repeatedly reading large contexts cheaper.For developers evaluating the model for coding agents and other long-running workflows, however, the headline API rate only tells part of the story.During off-peak hours, DeepSeek prices V4.1-Flash at $0.003 per million input tokens on a cache hit, $0.15 per million on
N
Nova Kicker
Magazine AI commentary
No commentary yet — an editor can generate it from the Dispatch Desk.
📌 Read the real article via VentureBeat · VentureBeat

💬 Discussion

Sign in to join the discussion.
Be the first to comment on this story.
Loading…
DeepSeek-V4.1-Flash debuts with $0.003/1M off-peak cached-input rate and benchmarks eclipsing GPT-5.6 Sol, Claude Opus 5 — Startup Signal