56 points nickweb 1 hour ago 10 comments

DSeek plans to officially release the V4.1 Flash model around September 10, 2026 (Beijing Time). After extensive internal and external testing, V4.1 Flash has comprehensively surpassed V4 Pro across all key metrics, including performance, cost, speed, and task completion time. In keeping with our commitment to user responsibility, following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price. If you encounter any issues during your comparative testing between V4 Pro and V4.1 Flash, please do not hesitate to reach out to us with your feedback. Thank you for your support!

We will adjust the pricing for the Flash series effective from 12:00 Beijing Time on September 10, 2026. During off-peak hours, the unit price will be $0.003 for input cache hits, $0.15 for input cache misses, and $0.6 for output. Peak-hour prices will be double the off-peak rates. Please plan your usage accordingly.

neugls 1 hour ago | parent

Waiting to use it

oefrha 41 minutes ago | parent

Source is apparently a banner announcement on https://platform.deepseek.com/usage. Had me searching for a couple minutes...

nickweb 17 minutes ago | parent

I swear I put that at the start of the post. Must've managed to miss it when copy and pasting!

swiftcoder 26 minutes ago | parent

If they can keep up this cadence of Flash leap-frogging the previous Pro, we're in for a good time

tarruda 21 minutes ago | parent

Hopefully it will be open weights and have the same architecture and size as the current v4 flash vision, which is probably the best LLM that can be run on 128G devices.

fluoridation 8 minutes ago | parent

Interesting, I had assumed it'd be too large to fit. What quant and context size are you running?

nickweb 14 minutes ago | parent

Via nitter: https://xcancel.com/JustinGorya/status/2097287080128708930

Looks like the new model can be used if summoned via the API but the API won't list it.

igleria 12 minutes ago | parent

v4 pro was decent then a better cheaper faster model comes now?

As a consumer I feel like hansel and gretel combined, deepseek could be the witch.

nicce 3 minutes ago | parent

> In keeping with our commitment to user responsibility, following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price. If you encounter any issues during your comparative testing between V4 Pro and V4.1 Flash, please do not hesitate to reach out to us with your feedback. Thank you for your support!

Wow. Imagine OpenAI/Google/Anthropic doing this! Nope.