Top
Best
New

Posted by gmays 6 hours ago

Ember-1(fireworks.ai)
277 points | 156 commentspage 3
blissofbeing 3 hours ago|
Would be nice to include in fire pass.
tdhz77 5 hours ago||
Does anybody know if this would be a good model for creative writing?
combobyte 4 hours ago|
> model

> creative

Choose one.

tdhz77 1 hour ago||
Are you a bot?
combobyte 1 hour ago||
You really have no sense of irony, do you?
dbuxton 4 hours ago||
Do they mean Opus 5.5 or Opus 5?
themgt 4 hours ago||
The result? Ember-1 set a new Pareto frontier for Bedside Bench across both open and closed models including GPT-5.6 Sol, GPT-6 Astra, and Claude Opus 5 on cost/task.

"Pareto": 8 hits

"Opus 5.5": zero hits

wmf 4 hours ago|
Obviously this research was done before 6.0 Sol and Opus 5.5 came out. Your point stands that the frontier moves quickly and small gains can be eclipsed quickly.
ls612 5 hours ago||
On the smaller end, Quen 3.8, while being extraordinarily capable for a small local model, also suffers from extreme thinking. I wonder if the techniques described here generalize to other models too.
spijdar 5 hours ago||
I suspect it might generalize to other large models, but I don't think Qwen3.8 27B is one of them. Kimi K3 is a 2.8 trillion parameter model, and I suspect that is playing a big role in being able to reduce the length of CoT without taking a hit in quality.

That's just vibes, though.

KaoruAoiShiho 4 hours ago||
https://www.reddit.com/r/LocalLLaMA/comments/1wj3s31/thank_y...
monkey_monkey 5 hours ago||
I don't think the article mentions Pareto frontier enough.

Also, did I miss a memo? Suddenly every article on AI seems to be talking about the Pareto frontier - or have I just not been paying attention?

AnodicElegy 5 hours ago||
I guess they figure "best bang for your buck" comes off a little too colloquial.
swiftcoder 3 hours ago||
I would really love if we brought back some colloquialisms in this field. Not that long ago most folks in tech would have had pretty blank looks on their faces when someone started talking about the "Pareto frontier"
user43928 4 hours ago|||
Pareto frontier on some benchmark that I am hearing of for the first time.

Kimi K3 with less reasoning tokens isn't exactly exciting either, and particularly so if the license is less open than original Kimi K3.

DonsDiscountGas 4 hours ago|||
They want it to be the best at something. And it's obviously not the absolute smartest. So here we are.
intothemild 2 hours ago|||
Is there a Pareto frontier for the number of times articles mention or don't mention a Pareto frontier.
alienbaby 3 hours ago||
when everyones fighting to be 'somewhere in the pile' they need some way to advertise they have made progress while not being the best.
logicallee 4 hours ago||
This is really interesting. I think the Fireworks Serverless Training infrastructure they used to develop it is also unique and needed. Except if someone works at one of a handful of the largest labs, it is very difficult to set up or try any sort of training pipeline. The managed training infrastructure makes it available to more people.
nostrebored 4 hours ago|
I can’t help but think it’s more expensive tinker.
esafak 5 hours ago||
It looks like it would be similar to GLM 5.3 Flash, had they tested it...
justmeeew 4 hours ago||
[dead]
huflungdung 5 hours ago|
[dead]