Top
Best
New

Posted by wsxiaoys 17 hours ago

Qwen 3.8 follows GPT-5.5 Pro reasoning prefills(gist.github.com)
215 points | 84 commentspage 2
CamperBob2 16 hours ago|
News flash: people who scraped the Internet without permission to build their product complain when something vaguely similar is done to them. Water still wet, sky still blue. Film at 11.

(slibhb: Don't get me wrong, I agree with you 99%. But the frontier labs have zero moral authority here.)

codedokode 9 hours ago||
As I understand, they got paid for the traces unlike owners of scraped websites. They sell text generation tool, so what's the problem if someone generates texts using it?
noir_lord 16 hours ago|||
I'd send them the worlds smallest violin but Rufus is getting in the way of me finding it.
UberFly 15 hours ago||
I had to install an add-on in Waterfox to stop Rufus from following me around and interjecting every 2 minutes.
noir_lord 13 hours ago||
I'd already stopped using Amazon for geopolitical reasons but I needed to get something in an emergency the last week (family member in the hospital, so I bent the rule) first time I'd seen Rufus, even if I wasn't boycotting Amazon for other reasons that monstrosity would have made me consider it.
verdverm 16 hours ago|||
I would not be surprised in the slightest if we later find out they are running those same open models to find useful traces or bits to incorporate into their own training. Lots of rules for thee but not for me from Big Ai

I look forward to a day when open models are so dominant that we stop considering traces to be some form of intellectual property that must be hidden from / manipulated for paying users.

It's that manipulation of inputs and outputs that really rubs me the wrong way

vezycash 15 hours ago||
They are copying useful parts of open models 100% especially from deepseek.
throw10920 9 hours ago||
Evidence?
vlyan 15 hours ago|||
not just the internet, but every commercially published written work in existence, and I doubt their highly publicized destructive scanning thing had managed to legitimize even a fraction of a percent.

this what is permissible for Jupiter is not permissible for a cow bullshit alone should tell people all they need to know about what kind of greasy sociopaths run "open"ai and (mis)anthropic, and how seriously you should take their purported stances on "safety" and other self-serving shit.

slibhb 15 hours ago|||
I understand people just get off posting stuff like this. But creating LLMs from the entire corpus of human text was a huge achievement. Distilling those models is much less of an achievement. It means China is further behind than we thought.
vanviegen 14 hours ago|||
They're using competitor output as additional training input. Shrug.. It's not like they have access to the weights.
ezekiel68 14 hours ago||||
First movers rarely take the prize, though, do they? Les Paul and Mary Ford pioneered overdubbing voices back in the 1940s and 50s. The Beatles stole it from Buddy Holly. And Elton John from the Beatles.
Zambyte 13 hours ago||||
Distilling is a massive achievement. I can run Qwen. I can't run GPT (TM). It's not a matter of X is better than Y. It's a matter of Y exists, X does not.
davyAdewoyin 14 hours ago|||
How is China further behind if distillation cannot stop? I think it's a reasonable strategy to follow, even if they could train from scratch.
undeveloper 13 hours ago||
distillation results in a worse product than the actual teacher model iirc
SXX 11 hours ago||
Nobody care if Chinese models are only 99%, 95% or 90% as good as SotA US models.

Because we only have weights and able to self-host Chinese ones. Gemma 4 and GPT OSS are nice to have, but nowhere close to that.

RivieraKid 15 hours ago||
For some reason, what China is doing seems worse. Part of it is that I want the US to stay ahead of China.
tizerluo 8 hours ago||
[flagged]
wip0 4 hours ago||
[dead]
brcmthrowaway 15 hours ago|
This makes me very sad

If Qwen and other Chinese labs are just copying reasoning traces, then those labs are more than a year behind the frontier.

levocardia 15 hours ago||
It's been pretty obvious to me that the Chinese labs are operating mostly on a fast-follow strategy. The distillation attacks are well-documented, and there is good reason to believe they are able to copy architectural innovations as well. If US labs stagnate I would expect Chinese labs to stagnate as well. Their engineering is great, but in terms of frontier innovation (which requires heavy compute to search for new strategies that work at frontier scale) they are very far behind.
jarym 13 hours ago||
Japanese electronics started off the same way post-war.
Roland303 11 hours ago||
well then theyll be do better post war, just in time for fallout timeline
Daishiman 15 hours ago|||
The way I see it the Chinese labs are optimizing for other things, including effective compact models that don't need to run on top-of-the-line nVidia hardware.
atomicnumber3 15 hours ago|||
It makes me happy because it means that these misanthropic technofascists have no moat. They can spend trillions of dollars only for it to be largely copied in short order.

Even if they weren't political adversaries of freedom, I would still feel 0% bad given all their training is already on data they got for free.

Information continues to want to be free. To the benefit of us all.

vipa123 15 hours ago|||
Preach brother, they stole everything on the internet, and beyond, to train their models. They thought all that information was free, and everyone a few months beyond them is just following their example.
polotics 14 hours ago||
They didn't actually steal in the sense that the information is still there on the internet.... About these shredded rare books, now we're talking.

If I may propose instead of "steal" I think we could agree to write they "Aaron-Swartz'ed" the information from the internet, what do you think, is this too harsh on Sam Altman or Carmen Ortiz ?

vipa123 14 hours ago||
Is anything too harsh for these new robber barons?
9864325789976 13 hours ago|||
Oh, we got a real rebel among us.
nater5000 14 hours ago||
Well, yeah? I didn't think anybody seriously thought otherwise?