parsimo2010
today at 5:16 PM
The timing looks like they are trying to take the wind out of Qwen's sails by releasing this on the same day that Qwen released the weights of Qwen3.8-max. Or maybe it's coincidence...
For comparison I looked at Qwen's claimed benchmarks for Qwen3.8-max (https://qwen.ai/blog?id=qwen3.8). Assuming each published set of benchmarks is believable, it looks like v4 Pro 0813 is better on average but overall performance is comparable. Pro 0813 is much cheaper. If you don't need vision capabilities then you don't have much reason to use Qwen3.8-max.
- 43.6 on HLE (Presumably without tools). Pro 0813 is a little worse.
- 86.6 on Terminal Bench 2.1. Pro 0813 is better.
- 55.9 on NL2Repo. Pro 0813 is better.
- 27 on Agent's Last Exam. Pro 0813 is a little worse.
- 72.5 on Toolathon-Verified. Pro 0813 is better.
- 56.6 on DeepSWE 1.1. If the DeepSWE listed for Pro 0813 is the same version, then Pro is better.
- 27.3 on AutomationBench. If the AutomationBench (Public) listed for Pro 0813 is the same, then Pro is better.
I guess we do need to wait to see if the upcoming DS pricing increase is enough to change the value proposition. As it is now, they could double or triple prices and it still would be a better value to use DS. I bet they know that.
trollbridge
today at 5:23 PM
By that standard, the release of Grok 4.6 was also timed on the same day.
Given how I think DeepSeek operates... I think they just release it when they feel it's ready, and don't even seem that concerned with what other people are doing.
somenameforme
today at 5:28 PM
Their leaks would confirm this sort of attitude. They're not trying to become the top player or anything like that - just working to play their part in pushing LLM tech forward and going from there. It was quite refreshing from the 'here's how we're going to dominate the world' nonsense. It's undoubtedly the same attitude that just lets them shrug and cancel the fund raising round after the leaks came from said funding round.
trollbridge
today at 5:36 PM
The founder of DS's stated goal is to get to AGI. He thinks this is the path to get there.
Kind of interesting, when compared to the hubris from American frontier labs.
Benefits of having a well performing hedge fund funding DeepSeek.
IIRC, Demis attempted to start a fund inside DeepMind but it was killed off. In an alternative world where he manages to pull that off, perhaps DeepMind would still be independent with Demis at the helm.
surgical_fire
today at 5:33 PM
Their stance on LLM development is why they earned my respect in a time when OpenAI and Anthropic only earn my mistrust.
That, and the fact that DS is an insanely capable model.
parsimo2010
today at 5:32 PM
Actually, yes. I just didn't know about Grok's release because they aren't on the front page of HN.
Official pricing only kinda matters for an open weight model, no?
parsimo2010
today at 5:31 PM
It still matters as a point of comparison until other providers come online. If the consensus price from other providers is much different that can be compared then. But for now we have $0.435 / $0.87 for v4 Pro 0813 (with increase announced but we don't know the new pricing), and $2 / $6 for Qwen3.8-max. So until we get other data points that is what we have to look at.
I wondered if the promised change in pricing is actually going to be deepseek bringing up their cached costs. They're extremely inexpensive.
I mean at the rate of model releases happening, I think a lot of these will collide more often than expected!