Top
Best
New

Posted by bvaldivielso 11 hours ago

Muse Spark 1.3(developer.meta.com)
https://research.meta.ai/blog/introducing-muse-spark-1-3
502 points | 340 commentspage 3
dcl 7 hours ago|
Very keen to try this after using Claude Code over the last few months. Should I just point Claude Code to Muse Spark endpoint (because I'm familiar with Code)? What do people think of Muse Code or other coding agent harnesses?
alexboehm 7 hours ago|
Just try opencode, it comes with 1.3 contributor free.
dcl 7 hours ago||
Well thats very interesting. Thank you. Will be interesting to see how hard/easy it is to translate my Claude skills, loop design, etc to the new harness.

This kind of raises another question to me regarding the coding benchmarks, how much of it is model versus harness?

jonahhorowitz 2 hours ago||
Coming from Claude Code, I initially went with opencode but switched to pi.dev after a while and I think I like it more. It's lighter weight. It's worth trying both.
ydna404 4 hours ago||
For folks who are impressed with costs, why does it matter to you? Is subscriptions not a thing? I may be missing something but only companies should really care about this I would think?
MitziMoto 4 hours ago|
Some of us own and run companies? Cost per performance is a huge deal.
fibonacci112358 10 hours ago||
Is everyone rushing to launch something before Astra tomorrow?
ryanschaefer 8 hours ago||
For all of the comments about training: I thought that subscription plans for other models allow the same. Am I mistaken?
Aurornis 8 hours ago|
It's a toggle. Some will automatically enable it and you have to turn it off. People who rapidly click through setup flows can miss it and leave it enabled.
maciejgryka 10 hours ago||
Does anyone know what the license for this model is? Specifically any word on restrictions about what it can be used for?
gehsty 9 hours ago||
As a product, would developers switch to a meta model/harness? I don’t think so.

Only way I see is if it becomes the new SOTA / frontier, does anyone think Meta will surpass Anthropic or OpenAI?

I still can’t get my head around why language models are an existential threat to Meta - they own the platforms people watch adds on?

phyrex 9 hours ago|
Meta also has 50k engineers. Not to mention that tons of meta infrastructure - including ads! - use AI. Would you want that sort of business be this dependent on someone else?
unsupp0rted 1 hour ago||
I'm annoyed my (US-bought) Meta glasses still block me from using the AI features, months after moving back to a country where it's generally enabled.
souvlakee 11 hours ago||
Why they didn't use LLM to create html table instead of https://lookaside.fbsbx.com/elementpath/media/?media_id=1048...?
LZ_Khan 10 hours ago||
Ha, even with monitoring engineers keystrokes and mouse movements not SotA on OSWorld.
scotty79 11 hours ago|
Is the fact that everybody almost catches up with the frontier a sign that we are entering a new region of sigmoid curve?
schopra909 11 hours ago||
Progress is iterative. Everyone is always riffing on other’s ideas and can execute on them given enough support (eg $$). The person to get to an idea first is just 5% away, so it’s possible to catch up.

Moreover,I think it’s impossible to know if you’re hitting a portion of the sigmoid, because there will often be an idea that changes the trajectory altogether.

In 2024, there was a ton of talk about the plateau. Reasoning was an iteration on chain of thought, but it didn’t really work. Deepseek proposes RLVR as a way to get around the lack of $ they have to produce human reasoning trace data. That small iteration catches the eye of OpenAI and Anthropic, turns out to be way more important than even DeepSeek could have ever expected when it comes to improving LLMs for coding, and last 18 months have been an exercise on riding that insight to the nth degree.

That one small iteration brought us a lot of progress. Now we’re seemingly exhausting the impact of that one insight, but there may be another soon enough.

stymaar 11 hours ago|||
> Deepseek proposes RLVR as a way to get around the lack of $ they have to produce human reasoning trace data.

What was the difference between what deepseek did for R1 and what OpenAI did for o1?

npn 4 hours ago||
openai did human crafted chain of thought dataset training. deepseek didn't have the resources so they attempted RL. doing RL correctly is hard because of the risk of model collapsing.
refulgentis 10 hours ago||||
I don’t know why people think DeepSeek did reasoning models / RLVR before OpenAI, there was a gap of months.
Philpax 6 hours ago|||
o1 was first, and Anthropic were doing a bit of it; DeepSeek brought it to the masses, but did not invent it.
schopra909 4 hours ago||
Totally, RLVR as a concept predates DeepSeek; but they proposed a version that was simple and scalable. Popularizing a specific version of a technique is exactly what I mean by iterations on a theme. It’s only 5% different from what others tried before, but that 5% difference showed a lot more potential than other versions of the same idea.

Since DeepSeeks GRPO, they’ve been improvements as well like AliBabas GSPO that have gotten wide adoption. Again iterations

danielmarkbruce 9 hours ago|||
Even if all the big ideas are gone and we are entering a new part of the curve, there is still an enormous amount of improvement possible. Just iterating on data mix/quality etc, training pipelines, reward functions, specific ways of reasoning (which i guess is mostly just data still) for the next 20 years will yield a looooot. And that's just the models. The harnesses/application layers/whateveritgetscallednext space has 20 years of progress to make.
samuelknight 11 hours ago|||
Meta has an enormous amount of compute. They are either going use it making and inferencing models or they are going to sell their excess capacity to model providers. Zuck had to completely rebuild his AI team after the Llama 4 launch mess.
ipsum2 9 hours ago|||
Yes. It's really up to OpenAI/Anthropic to release a new paradigm to shift the curve now, before everyone catches up entirely.
gdiamos 7 hours ago|||
I think it means that we should be aiming further ahead
redox99 11 hours ago|||
No because the frontier keeps advancing very fast.
dominotw 11 hours ago||
meta fails at everything yet is frontier on this one
More comments...