Top
Best
New

Posted by louiereederson 11 hours ago

Nvidia, Microsoft, Meta warn against overregulating open-weight models(www.cnbc.com)
Letter: https://images.nvidia.com/pdf/Open-Weights-and-American-AI-L... [pdf]

https://x.com/JensenHuang/status/2080643682408321103, https://xcancel.com/JensenHuang/status/2080643682408321103

https://www.wired.com/story/silicon-valley-is-completely-div..., https://archive.ph/LAkVI

452 points | 216 commentspage 5
throwaway323929 8 hours ago|
[flagged]
pegasus 7 hours ago||
Regardless of the motivations behind the paper, the fact is that they are right. It's a topsy-turvy world where America closes up and China opens, and I hope it gets corrected soon.

And besides, it seems to me like the ethical choice. Given how these models are, in some very real sense, mechanical plagiators, built on the generosity of creators past and present, some of them now in danger of being replaced by the machine. I think the least the labs can do is open these models up. These and other such considerations were the reason OpenAI started with that name. Of course, it was questionable that those ideals would survive the encounter with generational wealth. Just look up what the founders of Google were saying about advertising when they were two students tinkering at an as of yet unproven tech. Same thing for OpenAI, self-interest speaks that much louder when there's real money on the table.

It just boggles the mind that people now make excuses for their all-too-predictable about-turn.

throw0101d 5 hours ago|||
> I think the least the labs […]

Perhaps worth not calling them "labs". Are they not (for-profit) companies?

astro1234 5 hours ago||
Of course but I think labs is a good term. Like a lot of terms like this it comes from history. The people working on AI at these places come predominantly from academia and predominantly do research. Of course now these companies are far more than just a “lab” and I can imagine their headcount may be more non-research folks at this point but most of the results are indeed coming from the “lab”-y areas
swingandamiss 5 hours ago||||
> It's a topsy-turvy world where America closes up and China opens

There's nothing "open" about China. Google, meta, openai, etc all blocked. Go visit and see how it goes when you try to access your gmail or open facebook. Try to use chatgpt. Try to get citizenship and see how that goes. China blocks many western companies with their great firewall and force internal similar products. This is smart, China wants to prioritize their own.

smallmancontrov 5 hours ago|||
Your examples are the typical case, which is why OpenAI being closed and Kimi/DeepSeek/Qwen being open is topsy-turvy.
plufz 5 hours ago||||
Im guessing parent meant more open in this particular case, i.e open weight models vs closed cloud hosted. Obviously china is not an open country.
klik99 5 hours ago|||
He means specifically open vs closed AI model development, and how that’s topsy-turvy for the reasons you mention.
chasd00 7 hours ago|||
well let's be clear, there's nothing open about an open-weight model.

To me, "open these models up. " must mean provide all the data and supporting documentation required to reproduce the model. That would be "open". Postgres is open because you can download all the data and supporting documentation and reproduce the binary yourself. However, being able to only download a postgres binary would make it no longer open.

Additional training on top of an open-weight model sounds analogous to writing mods for minecraft. You may change some behavior but that doesn't make minecraft "open".

true_religion 3 hours ago|||
It's open, with a small 'o', and no moral judgements attached to it.

1. You have the weights, so you can run the model yourself on your own hardware. 2. You have the weights, so you can do post-training and shift those weights for your own purposes. It's not the same as training the model, but for many people its fine as what we want is a quantization, or a fine tune, or to create hybrid models.

What they don't give you are the training data, and reproduction instructions but... the toolchain to create the software has never been a part of 'Open Source'.

Even though it feels like a huge loophole, it's technically open source if you deliver the source code, without having a compiler that's available so you force them to recreate the toolchain from scratch.

To me, though, a better analogy is to research science where you'll be happy when they give you the full result set they compiled even if you don't get the raw data which may have IP or privacy concerns, or their often poorly documented lab notes so you can actually reproduce.

What you want to do is run your own experiment, and get your own results... not duplicate theirs directly. Even if reproduction is your aim in science, being unable to reproduce without copious notes sometimes points out that the original experimental process must have been flawed.

LLMs have the same issue. The creation process isn't entirely well documented, and the raw data can't be released since although the company have the right to use certain sources, they can't transfer those rights to others.

pegasus 4 hours ago||||
Sure, I'm all for opening the whole enchilada. But at least Chinese AI development is more open than the US one. And not just on the weights front, but AFAIK also more open when it comes to the techniques used and lessons learned in the process (DeepSeek at least is). Which is in a sense even more valuable and laudable.
jfrbfbreudh 5 hours ago||||
Weights are the cake and source is the recipe. Except you aren’t going to spend millions working through the recipe to arrive at the same exact cake anyway, so why do you care?
cyberax 6 hours ago||||
The problem with that analogy is that compiling Postgres takes a few minutes. Training a model takes hundreds of thousands of dollars (at least) and specialized hardware.

Having a standardized training set is valuable, though.

fragmede 6 hours ago|||
Here here! Open weights is the term because it's close to open source but everyone thinks everyone else is an idiot and saying the model is downloadable but you can't recreate it is too difficult for our feeble minds to understand. There are actually open source models out there though, with datasets and training code.
StableAlkyne 6 hours ago||
I think it has more to do with the fact that there has to be some term to describe the concept of "model which has weights that are openly accessible"

Like, that's just a logical thing to call it. I don't believe anyone is making a judgement on the intelligence of the reader to call it "open weight" when it refers to weights that are openly available.

"Open source" would be a more appropriate term to describe a model which also includes the training source.

pegasus 4 hours ago||
Indeed.
ImprobableTruth 6 hours ago|||
This is some bizarre victim inversion. The providers of closed models are the ones who are trying to use regulation to stop their open model competition, not the other way around.
sigbottle 5 hours ago||
This is how it is with all of these guys, their only principle is "what's good for me", and will twist all narratives to fit it
vatsachak 5 hours ago||
That's why competition is good though
dofm 5 hours ago|||
I don't think that is what is happening.

Instead a bunch of tech companies are gathering to try to stop OpenAI and Anthropic fear-bouncing the White House and the Republican Congress into giving them regulatory capture and repeating the mistakes they are making around RISC-V.

Those mistakes won't just entrench two companies, they will entrench the bigger-better-faster-more model (closed companies making ever bigger cloud-bound models) when it is abundantly clear that enormous progress can still be made on smaller, even desktop-bound models (where, due to distribution, open weights are essentially inevitable).

Regulatory capture that stops open weights work will also have impacts on local and on-device AI work, as well as on academic research.

furyofantares 8 hours ago|||
I'm very confused by this comment, I don't know who you're referring to.
gizmodo59 8 hours ago||
None of the people signed this have ever produced a frontier model at a given date (Which is to say its neither Google/OAI/Ant). The ones that sign are meta, musk (who is in the shovel selling business as well), hugging face obviously and few others

Note: Just pointing out the comment intent and nothing else

furyofantares 8 hours ago|||
OK, I guess my confusion was that the effort to ban open weights models is what I would categorize as "groveling at the white house to try to stop their competition".
segmondy 7 hours ago||||
Google has produced Gemini Pro, Meta produced llama3-70B/llama3-405B and now has Muse, Musk has Grok. These are very capable models and to claim that they are not frontier is to stretch the meaning unless you want to say "#1 model" Gemini Pro at one point even if for a week, was #1 model.
gizmodo59 1 hour ago||
I included Google along with Ant and OAI. OSS models have never been the highest performing model unless you refer to benchmaxxing.
mirekrusin 5 hours ago|||
Musk didn’t sign it.

nVidia did and they released good models – same with Meta, Microsoft, IBM, Mistral – all are signatories.

segmondy 7 hours ago|||
ha! who is groveling to the white house to stop competition? openai, anthropic and google, they are the losers. if they want to compete, compete.
PeterStuer 5 hours ago|||
You can call out regulatory capture evwn when you would do the same if the shoe were on the other foot.
lenerdenator 8 hours ago||
I thought the whole point of open software development was anyone could take your thing and improve it.

It would seem as if the community either isn't doing that or is relying on the Chinese to do that.

daveguy 7 hours ago||
I thought this would be a push for a nationwide open weights effort / initiative, but I was surprised to see this:

> Distillation ... reflects a long tradition of learning from, building upon, and improving existing technologies, a tradition that has helped drive innovation since the rise of the open-source software movement. By contrast, unlawful efforts to extract value from closed models raise legitimate concerns. Those concerns should be addressed through targeted legal and commercial frameworks rather than sweeping restrictions on techniques that play an important role in AI innovation.

Sounds like they are saying "please protect our IP theft" that created closed weight frontier models in case we arbitrarily decide to close our models. But don't get rid of distillations in general so that we can all also keep benefiting from open models. We don't want to lose the ability to benefit from the work of others as we launder IP into closed models.

Surprised Linux Foundation kept their name on it with that.

chasd00 7 hours ago||
"..a tradition that has helped drive innovation since the rise of the open-source software"

i've said this in other comments but the fact that these companies are trying to align with open-source when there's no "source" included with their models is pretty damning. I think it lays bare the absence of any kind of noble or righteous motive with respect to distillation.

paxys 9 hours ago|
I agree with the message, but I'm not going to accept it from this particular crowd.

Microsoft, NVIDIA, Meta, Palantir, IBM...They have all been actively hostile to open source for decades, and have a history of embracing it only when convenient and profitable.

Microsoft benefited from its close partnership with OpenAI for years right up until it went sour. Where was this enthusiasm for open weights then?

Meta was developing open models and then abandoned that effort in search for profits. Muse Spark is now fully closed.

All these companies have the resouces to train and release frontier open weights models today, but choose not to. So spare me the marketing and virtue signaling.

ryandrake 9 hours ago||
I’ll never forgive Microsoft for their hostile actions against open source and open standards in the 90s. I’ll never believe any endorsement from them of open-anything.
Arubis 8 hours ago|||
Yeah, the endorsement logos at the bottom are majority bad actors with regards to openness in code. Many are bad actors with regards to openness in _society_.
balls187 8 hours ago||
Yeah, never ask a fox to count your chickens.