Top
Best
New

Posted by jonotime 2 days ago

Why isn't the industry freaking out about DeepSeek 4.1 Flash?(www.dgt.is)
1097 points | 964 commentspage 12
gsky 1 day ago|
America bans Chinese models sooner or later just the China banned American big tech
aussieguy1234 1 day ago||
What blows me away about this model is it's speed.

It's way faster than Opus or any of the GPT models.

I have a coding harness which is opencode plus a few skills relevant to my workflow. Deepseek 4.1 Flash does very well in this environment. I haven't noticed much difference quality wise compared to Opus 5, which I use in my day job as my employer pays for it (although I'm considering using DeepSeek here too given how cheap it is).

hypfer 1 day ago||
Is it known why unsloth seems to not have touched DeepSeek 4.1 Flash?
lanesun 21 hours ago||
Because the version officially released by DS is an extremely quantized version, and there are many new things in the architecture, Unsloth needs time to handle this, just as was the case with the previous DS V4 Flash.
zozbot234 1 day ago|||
It's still lacking llama.cpp support, and the work on that isn't moving very fast either. Looks more like a general community issue, where this model isn't drawing much interest.
jacquesm 1 day ago||
You can ask them directly, Daniel Han-Chen is pretty responsive.
anguralbanish2 1 day ago||
I would love to get them more better, it's good not a bad thing.
Frannky 1 day ago||
I mean, they kinda tried to regulatory-capture the market after trying to scare the public, possibly because those models will be a cheap option that gets the job done?

For now, I think everyone is still using Anthropic and OpenAI because if you use a subscription you pay 1/40–1/50 of the API prices, and the models are good when they don’t nerf them, and they are also way cheaper than open models’ API prices.

The interesting thing will happen when they pull the plug and become economically smarter to stop using them. I regularly try alternatives to avoid being locked in and found GLM-5.3 as an orchestrator and GLM5.3 Flash + OMP and DeepSeek Flash as advisor to be able to get jobs done just fine. Space Bunny too was pretty great, which was probably MiniMax’s new model.

I think they are using an Uber like strategy but without the network effects that justify losing money for so long

lmeyerov 1 day ago||
GLM 5.3 Flash even more so... But yes :)
robertlane0 1 day ago||
Honestly for me the intelligence gap between DS 4.1 Flash and Muse Spark 1.3 makes Muse more worth it for me, especially on a $10 OpenCode Go sub, with the caveat that everything I use it on is open source which makes the fact that I'm sharing it with Meta a little moot because it's already published permissively on GitHub anyways.
vietvu 1 day ago||
This artcile is like 2 months too late?
james221 15 hours ago||
this was beautiful to read thanks
athrael-soju 1 day ago|
Because it will be replaced within weeks?
More comments...