Top
Best
New

Posted by volf_ 1 day ago

Xiaomi MiMo v2.6(mimo.xiaomi.com)
282 points | 122 commentspage 2
eriquesito 1 day ago|
Funny that all but one video has audio, the house 3D model one, where you can hear (what I assume are) Xiaomi's engineers talking about who knows what.
ddxv 1 day ago||
This looks great in terms of cost and capabilities, truly pushing the frontier forward in terms of open weight light weight models.
thrownawaysz 1 day ago||
>Night 0.8x Usage, 00:00-08:00 -UTC+8

It's because offpeak electricity is cheaper?

Funnily it's perfect if you are in the Pacific Time Zone because you can use it daytime 9am to 5pm

SSLy 1 day ago||
I believe it's a function of their primary user base being in china
eunos 1 day ago||
Same as DeepSeek, non busy time for UTC+8, maybe also cheaper electricity during night
MisterMunchkin 1 day ago||
I really liked MiMo 2.5, it was really affordable and actually had vision, unlike DeepSeek. (DeepSeek has only recently added it)

Just tried 2.6 flash on a really niche topic I specialise in and it has done a really good job. They’ve definitely polluted their training data with claudeslop, but looking past the slop there is a decent model.

perrygeo 1 day ago||
Can we afford to look past it? If/when claudeslop starts infecting every new model to such an extent, that model will produce its own slop, infecting new models... At what point do we lose all reliable methods for establishing "truth"? This is epistemic collapse waiting to happen. I honestly thought it would take longer... holding out for a coherent shared reality in 2030 seems optimistic.
omani 1 day ago||
how do you recognize "claudeslop"?
Bluestein 1 day ago||
It's an honest, load-bearing, simple thing.-
SSLy 1 day ago|||
that's belt and suspenders too
pimeys 1 day ago||
a smoking gun
Bluestein 16 hours ago||
... and a caveat worth flagging.-
nullc 18 hours ago|||
You're right to push back. That's on me. There is one thing I must flag which will move the needle. My honest take: It's all about the shape of your priors. That's the lever, and that's not nothing. It's worth your attention before you land your next remark. Here's why that matters: It re-contextualizes everything.
DanMcInerney 1 day ago||
This is a big week. Probably getting next OpenAI and Anthro models, Grok 4.7, Mimo, etc. These open source model releases are why I can't take the "slow down" crowd seriously. I pitted older Mimo, qwen, step, gpt-oss, and other models against each other playing games like Werewolf and Sketch.io-like games where I let them talk shit while they played against each other. Mimo was by far pareto frontier of game-playing for the models that were <$0.15/m input tokens on OpenRouter. Qwen was pareto frontier in the shit talking game though. Qwen's hilarious. https://www.tiktok.com/@clankerfights/video/7642862917582425...
algoth1 1 day ago||
Finally a lab that doesn't cheat on the charts
bertili 1 day ago||
They mixed up DeepSeek 4.1 Flash with something else on this page, possibly DeepSeek 4.1 Flash means Gemini 3.8 Flash.
varispeed 1 day ago||
These benchmark are useless as they don't say whether they were done before or after Fable and Astra got nerfed.
system2 22 hours ago|
That's not chinese models' fault. Fable and Astra deserve to be punished for their scammy bait-and-switch.
gigatexal 1 day ago||
Leaning into what it cost to train is hilarious and an obvious shot at US frontier labs spending tens to hundreds of millions or more to train their models.
alfalfasprout 1 day ago|
The moat for OAI and anthropic seems to be very quickly shrinking. Chinese labs are now using RSI-like approaches and even without resorting to heavy distillation they're catching up in a couple of months vs. what would have been 6-12 months a year prior.

And as these models get better the pace of training is quickly speeding up too.

This doesn't bode particularly well for anthropic/OAI after they go public.

mdale 9 hours ago||
I don't know if we will look at OpenAI and Anthropic as moating on frontier models & selling tokens.

They are banking on the application layer and accumulated business and end user context. They have to quickly make that systems integrated value out weigh the model choice price value in the broader market

Unknown if they will be able to pull that off.

verdverm 1 day ago||
token vendors are headed to the same place mobile data vendors went, this is good for everyone but those who thought they could maintain exorbitant prices
pvab3 22 hours ago||
Explain about mobile data vendors?
verdverm 22 hours ago||
when mobile data first arrived, it was expensive and people bragged about their bills (token counts today)

with time, it became commoditized, people now have unlimited plans, and the money is made by the applications that sit on top (token generation is increasingly undifferentiated low-level infra)

This is not to say there has not been significant innovation in the time since, but it's a low margin business (tokens look to be headed this way)

More comments...