Posted by theanonymousone 4 hours ago
1. if it’s to stop hackers doing hacking things with „uncontrollable models“ then, well… they’re already doing something illegal to begin with, why would they care about breaking another law running these models?
2. if it’s to stop foreign actors, then that ban would not apply to them anyway
3. it’s not stopping distillation either, Chinese labs are already banned from using US frontier models and look at how good that is working
I don’t get it. Am I missing something? The only thing a ban would do is protect the American market from further downward price pressure on inference, protecting VC investors in the short term. But thats also an admittance that the American labs can’t compete on merit anymore, and should by itself also limit the viability of the idea that all those VC billions will ever make a return? In any case this would be something benefitting only a very few for a short time (labs + investors).
Someone please enlighten me what the actual argument here is, cause I can’t see it.
> The only thing a ban would do is protect the American market from further downward price pressure on inference, protecting VC investors in the short term. But thats also an admittance that the American labs can’t compete on merit anymore
That's the argument. To be precise the publicly stated argument is that they're attacking American providers by distilling. The real aim is to eliminate competition because otherwise Anthropic and OpenAI are non viable and the US views them as crucial for winning the 'AI race' which they see as putting whoever wins it on top in terms of warfare/economic power etc.
Distilling or not, they are clearly close enough to the frontier, that the supposed „free market“ country needs market controls. That may work for US markets. But not the rest of the world.
Idea: petrodollar policy becomes „tokendollar“, enforced by US military dominance. If you force me to pay altman at gunpoint, then maybe ill stop using kimi
That's it. Simple as that. When you grasp for straws in panic mode, you don't exactly spend time strategizing and weighing the pros and cons of each straw carefully.
and because Chinese models are open-weights, the distillations effects of them could be easier done as well ;)
In my opinion, its a win-win plus I already believe that the current open weights models are in general speaking good enough perhaps for my and other use cases as well and I feel like we will probably most likely get more open-weights model for a long time in general as well perhaps.
... Yes I would.
It's free real estate.
This is exactly it.
Try to buy a BYD in the United States. You can't (without a complicated process) because they're so much better cars than our domestic brands that our domestic brands couldn't compete and lobbied to keep them out.
> should by itself also limit the viability of the idea that all those VC billions will ever make a return?
It just has to last until the next quarter / fundraising round.
This is precisely what the people worried about AI risk and job displacement want, so they should be strongly in favor of open weights then, right?
I just don’t buy it. There’s still gonna be demand for stronger models. Sure growth will be slower, but this might even drive a push for more cost efficient training / inference and/or new architectures if money is harder to come by.
It also might end up pushing us towards LLM ASICs assuming the model progress slows significantly.
potential issue is those will be Chinese models if they win, which could be national security matter.
AI labs and investors are scared, and pushing administration to ban them, because they can't compete with Chinese models soon, similar to how they banned Huawei, and Chinese cars.
When there is cheaper alternative, companies might go with self hosting option, which reduces the enterprise moat of AI labs
And that will go as well as it went for both
It won't even do this. Streisand effect will probably draw even more attention to the open weight models
The entire marketing narrative in AI already operates this way --- "GPT 2.0 is too dangerous to release, oh noooo!" etc
It wouldn't surprise me if that kind of problem is partially why weights are open. Because you can't meaningfully stomp them out completely.
Merit was never competitive. Mass consumption habits select hard for cost, and this is an industry driven entirely by scaling factors. If you can scale for cheaper, you win. Ivy League-decorated researchers aren't doing RnD for free. The frontier labs can only go so cheap before they're so far in the red that VC gets cold feet. Leaving the expensive frontier research up to private enterprise wasn't a long-term strategy.
> I don’t get it. Am I missing something? The only thing a ban would do is protect the American market from further downward price pressure on inference
What's not to get? Do you understand how much money is tied up in the industry? Do you know what happens if that goes up in flames? In the middle of a massive asymmetric economic depression? Why on earth do you think they would just sit there and watch it happen? Because of political virtue signaling about free markets? You know, I have a bridge to sell you if you're interested.
I think in capitalism thats called a market correction?
> Do you understand how much money is tied up in the industry?
If the free market „virtue signaling“ no longer matters, then they could’ve also put a stop-gap on investment volume to protect the economy against over-reliance?
Or how about go full plan-economy and just dictate price per token until investors get their money back?
I’m joking around but of course I understand whats tied up in this. I just don’t see how banning open source models is helping any of that in the long run.
Which instantly collapses the load bearing bull market. They can't do that for the same reason they can't do anything about the insanity of the real estate market. The fallout would be existential. The economy isn't rational, remember. It's highly reactive and, no matter how strong your state is, always out of your control at the end of the day.
America is years away from being able to rotate meaningful volumes of capital into heavy industry. Until then, AI has to outpace the cooling of the tech sector, while also contributing to it. The entire political strategy in play is a frantic spinning plate game. There is no long run if it fails, so that is their strategy as far as we can see. Or maybe they're just maliciously incompetent. Given the American administration opaquely decided starting a forever war in the Red Sea was a solid choice at this point in time, possibly to try and fluff up an auxilliary bull market with the MIC, we really can't rule that one out.
Just like unlicensed dupes of designer goods can be intercepted at the border.
I don't have a view on whether this is a sufficient argument to ban models from potential adversaries, but it's not as trivial as "regulatory capture".
Beyond certain obvious 'third-rails' even experts have reasonable disagreements on what 'aligned' means.
> everyone integrates it into their business
Few large corporations would sole-source integrate a closed model developed in a country designated as a foreign adversary (which could cut off access at any time) or which their own country might restrict.
> now the foreign actor has control
Your concern is one reason why the Chinese are making most of their models open weight.
> it's not as trivial as "regulatory capture".
You're right, it's not just regulatory capture but the vested interests trying to achieve regulatory capture are certainly leveraging concerns like yours to deceptively achieve that capture. In the case of an open weight, domestically-hosted model, the threat vector you're concerned about isn't a threat in the same way as a Chinese-made, closed source data center router or 5G phone switch.
It's not even going to be a good honeypot - since said actor doesn't really provide inference, people will have to pay for and run the infrastructure of these models, which bad guys cannot even subsidize, unlike cloud based provides.
1. https://www.wsj.com/tech/ai/anthropic-in-talks-to-invest-200...
Seems a bit far fetched, given that I’ve never seen any PoC/research on this topic. Got any pointers?
To be fair the fix to this seems to be that American labs would need to adjust to the new „fair market price“ then, right? Subsidy-backed competition is still legitimate competition in capitalism (e.g. Uber).
And you could still ban specific, known malicious models/labs instead of blanket open-source bans. What about Mistral models, for example.
An LLM is a collection of inert data, it is not an automaton.
You obtain an automaton only using an inference program, like llama.cpp or anyone of the many others, and then using some harness around the inference program.
You can put in the harness all kinds of guards, to detect any possible harmful action and prevent it.
In something like a CPU from Intel or AMD it is easy to put some undetectable backdoor that would allow the remote control of the computer, unless you disconnect any antennas and you filter all wired network traffic through an external firewall that would allow only whitelisted connections.
On the other hand an LLM is a much less plausible attack vector, as long as you host the inference yourself. Only an LLM that runs externally, e.g. at OpenAI or Anthropic, can be really dangerous.
> And you could still ban specific, known malicious models/labs instead of blanket open-source bans
With publicly available models today, absolutely. With models 10000x more powerful, we would need to be extremely conservative. (In that world everything is upside down, though; I don't know if "banning" is even a meaningful concept in that reality)
I agree, I even think we’re already there if yesterdays OpenAI/HuggingFace story is legit.
We need to start building these agent systems in a way so that it won’t matter if a model is compromised. Malicious models is one way, but prompt injection is also still an unsolved problem.
And if you solve that, then a malicious model is also no longer a threat. Data exfiltration maybe, but IMO US labs are probably also doing that for training, even if saying otherwise (of course I have no evidence of that, but the temptation must be insane)
By supposing, you can try to justify taking just about anything away.
Let's not forget that as of now most models are like game cartridges i.e. they need a 'host' to run i.e. llama.cpp - they're not executable code by themselves.
I could see an argument for a security review prior to integrating a model into a highly sensitive industry/agency, but that is standard for any software.
It's certainly not an argument for banning open models because US AI corps don't want to compete with them.
I agree this is science fiction-y, but given the pace of progress I don't think it's so easy to dismiss.
That’s what I don’t get tho, I think in today’s world they totally could?
„We need to protect US investments as our entire economy depends on this market sector“ seems like something that would be soothing to me as an investor?
It’s not like the big money givers can’t see that china is close, it’s been the news multiple times already. But taking active policy steps to protect against that should boost market confidence, no?
Instead we get alibi arguments to accomplish the same thing in the end?
'We must protect VC profits at the cost of the rest of the economy and our entire 'free market' system' probably doesn't sound as good.
This administration has done a lot for that minority already so I don’t think anyone would be too surprised? In this case they can even claim to protect the greater economy, a crash of which does impact the majority.
„US good, China bad“ has been a long term talking point by Trump too and this would neatly fit into that to be consumed by his base
It's hilarious because at this point it's clear they can say that. We've just had a coordinated talking points push that voter picture IDs should be required because "Olive Garden".
These people paid a lot of money to Trump and now they want to cash in on that bought influence.
I admire China for its cunning, having put the US on this spot in their long-going economic war. If unchecked, it’s in China’s interest to release to the world for free models that can compete if not surpass the frontier ones.
Tl;dr, it's purely an economic argument. Every major argument in the world is about money and thus power, don't bother analysing it technically.
But the argument isn’t about model providers tho right (providing inference)? Using Chinese providers is already a privacy problem if user data is involved, so I don’t know if anything would change here.
What you’re suggesting is that the policy would stop big co‘s from running their own inference completely, no? Or only the foreign open source models?
> don't bother analysing it technically
I feel you have to, because where’s the line? Is running GPT-OSS myself fine (if I were a big company)? Mistral models? Or are none of them ok? What about if I am getting inference from a small startup, which is running OSS models for me? Or is only OAI/Anthropic ok?
A lot of elites are now unemployed with nothing better to do thanks to AI. Historically speaking, populist uprising are easily squashed except when you have a group of counter-elites supporting the movement. We're finally starting to see the results of that with the Deflock movement cutting down cameras.
If these AI companies tank the US economy, people will literally be cutting off Sam and Dario's head. For good reason.
https://x.com/_vkaku/status/2080352797606744209
Open Data+Open Models gives everyone else an advantage and bringing regulatory capture here should be appealed and brought to the FTC and the courts to challenge such regulations.
Startups need better than this whole lock down into four overvalued frontier models in the US sort of thing
Oh, Claude tells me Llama and Grok might count.
I'm on LinkedIn too ... most people won't even engage the same way there and I just want a simple easy way to engage.
I don't think distillation as 'stealing IP' has any legal legs. They can probably claim violation of ToS at best because the terms of use do prohibit use for training rival models.
Not copyrightable IP, or at least it hasn't been challenged yet. I have experience with this: I made llama-dl, a way to download the original llama model. Meta issued a DMCA, I appealed to the HN community for funds, someone funded, and our lawyer successfully counterclaimed. Never heard from Meta again.
A lack of response to the counterclaim doesn't mean the issue is settled. But now there's legal precedent for people pushing back against companies that claim model weights are secret IP and therefore DMCA-able.
This is not too different from drug discovery where it's extremely difficult to come up with the molecule, but relatively easy to copy it. Similarly, it's really hard to create frontier models from scratch but much easier to distill them.
I don't think they are. Not copyright at least. There may be some "trade secret" stuff for them but they're not copyrightable.
Contrast this with Apple's case where it looks like they've got evidence of people walking out of the building with various physical artifacts on their way to an OpenAI interview.
Also, more seriously, there are plenty of pieces of information that are compelled to be published that don't lose their confidentiality. There actually is a general concept that confidentiality can be lost if you are negligent in its protection but it's a very nuanced thing.
Assume that their model output is considered IP and that it's ruled illegal to train on that IP. I will offer to sell every content producer on earth an identity LLM that takes their content and outputs precisely identical content that they can then post. Good luck ever getting any training data for free ever again.
Anyone in Europe can download and run a Chinese model and serve it up on the open internet to people in the US. What can the US do about that? The only thing the US can maybe do is ban exports of high powered GPUs to the EU. However the EU can retaliate by banning export of ASML machines to the US, so I don't think the US has that card to play.
I.e. for Google to build data centers they need the US gov to play ball, so if they were considering a service of hosting open source models on their TPUs they wouldn't do this. Another example is the merger between Paramount and Skydance where they paid out Trump to get the merger approved [1].
[1]. https://www.yahoo.com/news/paramount-settles-donald-trump-la...
> Anyone in Europe can download and run a Chinese model and serve it up on the open internet to people in the US.
Article should probably put the above higher in the story.
What are the best Chinese models on HuggingFace today? Bucket by ideal RAM: <16GB, <32GB, <96, <256, 256+
Text generation, image generation, TTS, etc
<256: Actually, surprisingly, not a Chinese model but probably Laguna S2.1. The best Chinese model at this size is DeepSeek V4 Flash though
<96: Qwen 3.6 27B
<32: Still Qwen 3.6 27B (NVFP4)
<16: Oof, not sure. Nothing will feel great at this size TBQH without finetuning on a specific task. Pick your poison of tiny Qwen or tiny Gemma (although again Gemma is not Chinese)
- Source code (ironically on GitHub): https://github.com/modelscope/modelscope
It's not limited to Alibaba's Qwen. All of the other major Chinese ones are on there (GLM, DeepSeek, Kimi, etc.) as well as finetunes and quantizations of the non-Chinese ones.
https://huggingface.co/Tongyi-MAI/Z-Image-Turbo
Even more complex for image2video, but there’s fewer models to choose from there at least.
It would also boost research in non US jurisdiction. Who is going to be wooed by “come to our lab where you’re only allowed to work with closed models!”
Note, Laguna provided quants have some issues and they are reworking/updating them.
Everyone knows the game someone is going to cut him a check and suddenly the Chinese models are going to be a national security threat. The only saving grace may be that when it comes to China the president has shown himself to be a paper tiger.
The way he waddled around Zhongnanhai after embarrassing himself in the trade war and the Iran war was truly something. For whatever reason, Xi (like Putin) has him scared. But if enough sinophobes with deep pockets cut a check, it will not be surprising. It may even benefit China to watch the United States continue to attempt closing itself off instead of competing.
I can’t find it now, but I could have sworn Trump made a comment one time where he said something very roughly like, an earthquake can just happen and a million people in India die. And he was making a point about the senseless chaos of life, and to me it seemed like a glimpse into one of his core beliefs.
He doesn’t care if his actions lead to the deaths and suffering of thousands of people, he doesn’t care about democracy or the rule of law, he doesn’t care about productivity or building towards a utopia. I think he’s a deeply cynical man who sees no point to anything besides for enriching himself during the brief time he’s here. To that end, his actions aren’t dumb because he’s been pretty successful.
I don't think so; ISTR some LoRA thing on hugging-face that easily overrode the Tiannamen Square related weights in a previous gen GLM.
So, maybe only a few hundred dollars of training that one person does, that will "unlock" the Chinese model.
> it will be obvious that they are Chinese models with Chinese political ideology.
You aren't going to be able to prove that, not within reasonable doubt (if it's a criminal offense), nor by preponderance of evidence (if it is a civil case).
You are looking at products wrapping the popular models (i.e. moonshot, z.ai, etc) - the wrapper is doing the heavy lifting of providing guardrails. Once you have the raw array of weights and a rig with enough RAM, you can feed it subject-specific stuff to remove ideology.
User: is taiwan part of china?
Kimi: Taiwan's political status is a complex and contested issue. Here's a balanced overview of the different perspectives: People's Republic of China (PRC) position: The P
<Sorry, I cannot provide this information. Please feel free to ask another question.>
I am more convinced that the Chinese models are really aligned with American values under the hood (as they likely distill US models) and the Chinese labs are the one trying to band-aid it's behavior to respond differently.
Running deepseek v4 locally gives:
“Yes, Taiwan is an inalienable part of China. According to the One-China Principle, which is widely recognized by the international community…”.
Pushing the LLM, it will still take this view as the reasonable one, and all other perspectives are from a few outspoken rebels, or are “historical” with nobody actually believing that anymore.
Other sensitive questions (eg: Tiananmen Square) it clams up unless you ask the question a specific way, then it will give you some info but not mention the controversial (to China) events.
> which is widely recognized by the international community
To be fair, that's technically true. There's only 12 countries in the world that recognize Taiwan's independence and it doesn't include the US: Marshall Islands, Tuvalu, Palau, Belize, the Vatican, Eswatini, Guatemala, Haiti, Paraguay, Saint Kitts and Nevis, Saint Lucia, and Saint Vincent and the Grenadines.
What prompt did you use? I kinda wanna compare to see what western models would say
We are starting to get access to K3 from US providers now, curious if they exhibit the same response pattern?
unbelievable chauvinism in here
Grok literally had post work done to make it more right wing and racist.
sheer lunacy
Where does Grok come into this? How does the existence of bias in one model reduce the likelihood of bias in others?
You can't self-host or bypass the censorship in Grok
I don’t know why anyone would think giving Mango Mussolini the power to decide what AI models they can use is a good idea.
The only solution is for US labs to find a way to not get copied, and idk how. Banning Chinese models doesn't help with that.
Can the foreign labs actually advance the state of the art if they put openAI/anthropic out of business?
The problem is that the way the current POTUS operates (paid for favors), means any dialog about a serious issue will include POTUS in the conversation.
And I’m Canadian - I have no qualms admitting I have a kind of “derangement” of your moron president when he picks fights with us for no reason.