Top
Best
New

Posted by jonotime 2 days ago

Why isn't the industry freaking out about DeepSeek 4.1 Flash?(www.dgt.is)
1099 points | 966 commentspage 15
cactusplant7374 1 day ago|
Because engineers are lusting for 1000 tokens per second. You can only achieve something like that with OpenAI.
pessimizer 1 day ago||
I'm no expert, but it think that it's the pricing on GPT-6 Luna. I'm also guessing that it's been underpriced just for this reason. I also don't think it's all that great, but it's definitely very cheap.

If it's underpriced, it's a loss leader to sell the other models, so it actually can't be too good.

I really put these things through their paces because I use them to review and work with new abstract game rules and models, so they're always flying blind. Luna misses the obvious (and more importantly, the clearly explained) consistently. My second prompt is listing all of the points in its first response, and saying "No, it doesn't work like that." The third prompt is picking out the two or three suggestions it made after correcting itself on all of the original points and saying "That's how it already works." The fourth prompt is "Now that we're done going over the rules, can we start?"

I actually feel like 5.6 Luna seemed better.

sergiotapia 1 day ago||
In my experience it just takes so much longer to arrive at "done" state for me. It thinks for soooooo long. I guess if you're running 12 sessions at once you don't really notice.
AIblemblio 1 day ago||
No they can't.

And as long as I pay as little for claude opus 5.5 i do right now, i'm using it.

But yes i'm glad that we have alternatives.

joshrw 15 hours ago||
I am rooting for the Chinese models to win so they can save us from our capitalist overlords.

However, the quality of the open-source Chinese models is terrible. They game or fake their benchmarks because in real-world usage they suck.

verdverm 2 days ago||
Why would we freak out? The systems we use have always gotten better, faster, cheaper with time
samyar 1 day ago||
it's good but not good enough
tonyhart7 1 day ago||
it literally hallucinating a lot

I dont get why people says D4.1 flash is good

m3kw9 1 day ago||
i thought 6.1sol copied the caching architecture so this isn't such a big deal no more
doctorpangloss 1 day ago|
because it doesn't work very well?

if you have a legitimate coding application, it isn't very good. if you have some kind of inauthentic activity, which could be what it is trained for for all sorts of reasons...

computerex 1 day ago|
What is your evidence? Deepseek v4.1 Flash is by far the most popular coding model on openrouter, having processed 38.7T tokens in just the last 7 days, over 3x the usage of the 2nd rank model.

So I ask again, what are you basing your assertion on?

doctorpangloss 1 day ago||
my own usage of deekseep v4.1 flash, and that among the dozens of great programmers i know, not a single person is using it

BUT. they are employed to do / deciding-to-do authentic (if often meaningless) stuff.

here's a short list of inauthentic activity that claude and openai refuse to do:

- chat services that, when you ask them, say they are not chatbots when they are

- code to work around software licenses or DRM

- code to scrape or download copyrighted material

- directly cheating on homework

- adopting a persona in social media that spreads misinformation or propaganda

this is but a short list. but ask me, "are there enough inauthentic activity demands such that someone who CANNOT USE claude or gpt as the LLM would use dsv4.1 on openrouter instead?" yes. i mean there are whole countries right now where the culture can be summarized as, "bottom to top, inauthentic activity." i am surprised it is not more usage!

computerex 1 day ago||
So your argument is that the 38T of tokens used in the last 7 days is by moron programmers or people doing "inauthentic" tasks? You think the person who made this post is also an idiot?

Do you realize how incredibly delusional/self-centered you sound?

doctorpangloss 1 day ago||
do YOU know anyone gainfully employed in programming who is using dsv4.1 to do work? what kind of work is it? why don't you ask them if it is good?

in the market, where you cannot fake or hide stuff very easily: the outsource customer services and cheating sectors have been the most disrupted. Cheating company Chegg lost 99% of its market value. CS it remains to be seen - https://www.reuters.com/technology/teleperformance-shares-pl... - certainly perceived to be disrupted, but they are not dead yet.

in my personal usage: dsv4 is generally pretty buggy. for example, if you give it a needle-in-the-haystack simple copying problem, it catastrophically fails to find needles if they happen to be positioned at index 250k tokens out of 1m. it can also be triggered to spew all sorts of garbage when DSpark is enabled during ordinary long-context coding, such as spewing weird DSML tool call errors after a normally parsed tool call error.

i don't know why you have to attack me personally, i think you're a bright and otherwise nice person and you understand the thrust of my POV.

pimeys 1 day ago|||
Yes. Hi. From our team 3/5 of us use 4.1 to do our daily tasks. For paid work for a company who pays us salary. From people around me I hear a lot of my friends being really happy with it especially for the price.

I don't know man, maybe this is not super serious what I'm doing. Some systems stuff with rust, implementing my own desktop apps with iced, porting old DOS games to Linux...

It is a very good model.

doctorpangloss 1 day ago||
so what you're saying is though, if they could afford it they would just use claude or codex?
pimeys 1 day ago||
Of course. It is a message to both: drop your prices. Opus gous down to 0.3/0.007/1.2 and we will definitely take another look.
computerex 1 day ago|||
You are speaking out of your ass, that’s what I take issue with. Falsifiability is something I hold sacred and you are taking a dump on it.

Fwiw I work in a company producing software for many fortune 500’s you have heard about and many people from our team use deepseek.

I am literally using it right now. Your entire line of reasoning rubs me the wrong way.

Btw check your provider and harness… improperly configured deepseek can emit dsml. If you are not passing thinking tokens back to the model it tends to do that.

Use a proper harness and good provider.

doctorpangloss 1 day ago||
"Hey Mr. Fortune 500 Client, would you prefer us to use something called DeepSeek V4.1 Flash, made by the Chinese, sending your Fortune 500 code to some random service provider on something called OpenRouter, where they promise according to something called Zero Data Retention that--"

Mr. Client: "I'm going to stop you right there. Why aren't you using Claude, or Codex, or Claude on Bedrock? Don't we deserve the best?"

You: ...

Look I don't know. I can tell from the hyperbole of your language, talking out of asses and such, that there is more to the story than you are letting on. Like Chinese users are banned from officially using Claude and Codex, for example. So many reasons that you cannot use Claude, not so much reasons to not choose to use Claude. All I am really saying is, I know DSV4 is kind of bad, that there is a lot of inauthentic activity, and that Claude and Codex refuse to do many kinds of inauthentic activity, and that a lot of coding done by outsourced shops has always been of questionable quality and purpose. I mean in my personal life, I know more people who have been scammed by Bulgarian code body shops than I know people who have used DSV4.1.

computerex 1 day ago||
You have NO IDEA what you're talking about. You are clueless.

Deepseek v4.1 flash is an open weights model. You can run it on your own hardware. You have no idea how my companies gets access to it. A very cursory Google search would reveal to you that there are many enterprise grade LLM providers that host this model on US soil with SOC2 protections.

Like: https://fireworks.ai/

Try not to talk about subjects you have no knowledge about because you are making yourself look like an idiot.

Edit: It's also clear to me that you don't deploy any LLM based system on scale because if you had you'd know why open weights models are so compelling.

Hint: it's the cost.

doctorpangloss 1 day ago||
okay, but are you US based? and can you specifically describe one of the pieces of software you are developing? it's okay if not. i am just wondering. i certainly believe that crappier stuff is cheaper!
computerex 1 day ago||
Yes the company I work for is based in San Diego, I work remotely from Alabama.

I bet you voted for trump. With brains like that.

More comments...