Posted by jacquesm 11 hours ago
Edit: Does anyone else notice the switching between dashes and em-dashes between paragraphs? Tells you a lot about the man, doesn't it.
That he doesn’t know how to use basic word processing, or even agent SKILLS.md or whatever it’s called now. Or maybe it’s just the next generation of vagueposting.
The CEO of Anthropic not using AI extensively would be news.
> "I know that there’s a sort of Silicon Valley shorthand where regulation = regulatory capture = concentration of power, but I’ve always found this to be an overly simplified picture of the world."
> wall of text follows
> doesnt proceed to clearly tell us what the actual picture of the world is, then
am i correct in summarizing that the line of reasoning is- frontier llm access means you are at an economic advantage
- a big risk of this is ongoing wealth concentration
- the "open weights" approach can't solve the problem of wealth and llm access being linked; you need compute too, and compute is expensive, thus "open weights" still favors the wealthy
- instead we need "objective and fair institutional processes", as this will allow small labs cook up their stuff while frontier labs get regulated
why would i care about what these smaller players do, if economic advantage = frontier model access? also, the reasoning around "why bother with open weights because compute is expensive too" seems like the kind of logic a motivated 12 year old could work their way around in 30 seconds. things are not either/or dario, you said so yourself.
seems like a 400 word corpo misdirection essay. par for the course.
The tech execs waking up on a random monday:
> I think that ordinary people don’t trust companies, governments, or the tech industry and always suspect that we are cooking up some new way to screw them over.
No shit Dario, no shit...
To paraphrase Sinclair "It's difficult to make a man realize the issues with regulatory capture, if they're to be the benefactor of regulatory capture".
> A crude analogy is that the formal court system can sometimes feel stuffy and elitist, but it does a much better job of defending the rights of vulnerable individuals than the alternative, mob justice.
Not so sure. After all lack of the latter did help establish the current tech overlord rule.
Especially when said individuals are rich.
> ...
> completely exempt any company below a certain amount of revenue or model training costs from being covered at all
One could argue that "Frontier AI" company know they have nothing to fear from company with less than XM$ revenue, and so their support for this type of regulation is still a way to force regulatory capture. In any case, whatever regulation you support, it's a regulation that you didn't have to handle when you were growing, but that incumbent will have to deal with.
Whatever it is Dario's intention or not does not matter. Capitalism push to consolidation and the eventual regulation that will need to be applied to mitigate the externality from a new industry, will mean that their will only be a handful of "very big" winner. Same story since the beginning of the industrial era. Even if that's not what Dario personally want, it's in the best interest of the shareholders, which will force Anthropic to do everything it can to be one of the big one.
> Overall my view is that AI is structurally a technology that tends to concentrate power, for reasons that have nothing to do with regulation (more to do with the extreme implications of the scaling laws).
The same can be said about a lot of other industry (aviation, oil, chip manufacturing, ...). Every country / union big enough will finance their own champion to try to keep a foot in the industry even if they are not the best.
Nothing beats the hubris of the Silicon Valley CEOs and self proclaimed “elite”. I’m missing people calling out “we are better than that”… NOT
The distillations of Qwen 3.8 down to a 27B model are good, but they're not on a par with frontier models.
When Dario talks about open weights not being a solution this is what he means - if you don't have 400GB of VRAM lying around the fact that there's an open model like Qwen3.8-2.4T-A95B doesn't really help much. If we're not regulating how models are available, or making sure access is open, then RAM prices will mean everything concentrates on a few very rich companies.
DeepSeek V4 Flash 0731 is 167 gigabytes from the developer and as a GGUF with no additional quantization. It limps along on my 192GB M2 Mac from several years ago [0]. This model tests better[1] than Claude Opus 4.6 released in February. That's six months ago - what will be available 6 months from now?
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731/tr...
https://huggingface.co/unsloth/DeepSeek-V4-Flash-0731-GGUF
So yeah, enthusiasts aren't going to run frontier models on their gaming machines, but a small office could easily justify the $30k - $100k cost to run something like this at high speed. The small company I worked for routinely spent that kind of money on Dec Alphas twenty five years ago, and that's not accounting for inflation adjustment.
And this is completely discounting the advances smaller models are making. You're right that Qwen 3.8 comes in different sizes. However, Qwen 3.8 27B and Qwen 3.6 27B do run on gaming cards, and they're better than the frontier models from twelve months ago.
I have no idea what will happen in the future, but I wouldn't base my guesses solely on the largest open weight models.
[0] Yes, it's unpleasantly slow (5-8 tok/sec)
[1] Yes, benchmarks should be taken with a lot of salt.
Does it have to be? There are plenty of coding tasks, where it's good enough.
However, if you're having a discussion about access to frontier models, and using Qwen 3.8 as an example of how open weights is a solution, then you should be honest and accurate about what you're talking about. Making an argument like "People can run Qwen 3.8 at home. That shows open weights are great." is a bit disingenuous if you're not also making it clear that you're not talking about Qwen 3.8 Max (or that you have a beast of a PC at home :D ).
I suppose the road to technical hell is paved with marketers and grifters. :)
[1]: https://ollama.com/library/deepseek-r1:1.5b [2]: https://huggingface.co/deepseek-ai/DeepSeek-R1-Distill-Qwen-...
Exactly. There are common coding tasks that these models can adequately do. They are absolute trash for anything that isn't coding. And even with coding, they are so, so far behind frontier models.