Top
Best
New

Posted by OfficialTurkey 2 days ago

GPT-6 Sol and Luna(openai.com)
1732 points | 822 commentspage 4
Sinidir 1 day ago|
Wow. Luna 6 is an insane value bargain at this point. If Anthropic doesn't finally come out with their own small low cost model instead of still having haiku 4.5 they'll be history soon.
badatnames 2 days ago||
It's asking a lot to trust they can or will maintain this new pricing. In any case it's exciting to think this might lead to further price cuts in the highly competent and competitive Chinese clones. I'm still using ChatGPT for interactive queries, but at this point pretty much only because of its familiar UI
chaos_emergent 2 days ago|
Curious why you think it's unsustainable?
badatnames 2 days ago|||
Because at some point keeping it up involves filing an S-1 that doesn't look like a garbage fire
wyre 2 days ago||
Didn't SpaceX already set a precedence for garbage fire S-1s? I don't think OpenAI has to worry about that?
badatnames 2 days ago||
SpaceX is a different beast with extremely high friction to enter its market, a massive technology lead, and well developed preferential high level relationships with just about every country worth worrying about.

OpenAI/Anthropic meanwhile feel a bit like they're hoping to sell iPhones in a market about to be flooded by $20 flip phones, with almost no channel of their own to do it. And for whatever mad reason OpenAI are now signalling they will attempt to compete on price with flip phones despite their cost of labour, energy, and just about everything else being far higher

wyre 2 days ago||
Wasn't SpaceX's insane valuation largely based off of Grok, because their rocket and satellite businesses could never be valued at over a trillion $$?

I don't see your metaphor to iphones and flip phones. This new Luna model is cheaper than deepseek 4.1 flash, except for cache reads. OpenAI having to compete with China is a much larger economic-political issue that is far larger than just our AI labs.

cmrdporcupine 2 days ago|||
Well, they do rug pull constantly. This week and last leading up to this the cost to use Codex was overwhelmingly perceived as terrible. People running out of usage all over the place. Reddit full of people crying. I noticed it myself.

Then they do a new model launch, issue quota resets all around, and it's a party for 2-3 weeks before things return to normal.

FergusArgyll 2 days ago||
Oh, I'm happy I'm not the only one. Astra was feasting on tokens!
cmrdporcupine 2 days ago||
It wasn't just Astra. Sol 5.6 was a hog, too. They futzed with the formula and it pissed people off royal.
GodelNumbering 2 days ago||
Gpt 6 Luna is cheaper than Deepseek 4.1 flash! Today is wild in terms of intelligence/price across the board!
gizmodo59 2 days ago|
it was expected no? if cost is the only reason to use oss models, they can do much better than small providers who don't have much compute.
jdprgm 2 days ago||
I wish there was more transparency on the plus plans usage limits showing actual token usage and prices per model that eats away at remaining usage.

Does anyone know how exactly these price differences for example between sol6 and sol5.6 translate to codex percentages? In theory it seems like for "high" on both it should result in ~3x more usage. If that is actually the case it would be huge! But all we see is % left and % changes while using and we really have no idea when or how those numbers are being calculated or when they change. So there is a 50% price reduction on API but who knows how the hell that translates to whatever price calculation is used on codex.

hehimself 2 days ago||
Love the price reductions across major players
madduci 2 days ago|
Because Qwen4 has been announced!
eloisant 2 days ago|||
And GLM 5.3 works great
system2 2 days ago||
Except for the censorship. We use it for massive data crunching, and roughly 5-8% (depending on the day) gets censored and doesn't get a response. We switched to Mimo 2.6, which is relatively better. For censored stuff, we use Sonnet and OpenAI Nano models.

Also Mimo 2.6 is roughly 30% cheaper. Without batch.

Havoc 2 days ago||
What sort of content is it censoring? Politics I assume?
system2 2 days ago||
News mostly. Anything China-related gets censored without hesitation. Some random stuff got censored too. It is borderline unusable, to be honest, unless only numbers are crunched.
Havoc 1 day ago||
Interesting. Was planning to use it for a news related thing too. I guess one can throw Jev at it first to ask whether it relates to China and then decide?

Or use the failure to get a response like you say

system2 1 day ago||
We are using it with OpenAI Luna. We send any failed query to Luna, and the operation is complete.
blovescoffee 2 days ago|||
and to squeeze anthropic, and other research innovations, not just chinese models but those help bring price down
Readerium 2 days ago||
Opus 5.5 seems better? Can someone attach both scores
hehimself 2 days ago|
Not the direct competitor to Opus 5.5, cuz 6 Sol is 50% cheaper.
Readerium 2 days ago||
Same price on Cache Reads 0.2/M So won't be 50 percent cheaper, more like 25% cheaper assuming half cost is cache read.
blovescoffee 2 days ago||
cost is dominated by non cached reads
pinkgolem 2 days ago||
that might depend on usecase, half of my cost is cache reads usally
jacobgold 2 days ago||
These counter-launches are starting to seem kind of tacky and boring. Just launch on your own schedule guys.
magarnicle 2 days ago|
Maybe this is what they meant by "pacing"?
Alifatisk 2 days ago||
So with GPT-6 Astra, Codex introduced an experimental feature for context management that’s supported to be beneficial for long conversations. Will that experimental feature now also apply to Sol and Luna?

https://community.openai.com/t/experimental-context-manageme...

I would also like to point out that it was quite predictable that Terra got discontinued, it didn’t make sense to have it when both Sol and Luna overlapped it.

Lunas insane discount is a game changer, OpenAI knows what they are doing here. Luna at max reasoning effort, even though its not optimal for long conversations, its incredibly intelligent while dirty cheap. Its not even competition anymore.

Whats even crazier is that I’ve underestimated how good Luna actually is. I’ve seen colleges create fantastic things with just Luna medium. This basically means you never have to think about your Codex usage anymore. You can run all day and not

have to worry about your 5h or weekly usage limit. To me, the discounts OpenAI is offering with Sol and Luna is truly a new milestone.

kreitter 2 days ago||
are you using that feature? i turned it on and then had second thoughts (not sure why) and deactivated it before ever using it lol
Alifatisk 1 day ago||
Yes, I have it turned on. But I’ve been using Luna model so I don’t think this feature have been applied yet.
iamthe0ne23 2 days ago||
[dead]
XCSme 1 day ago||
A good improvement overall.

GPT-6 Luna now is 50% cheaper, which makes it have one of the best intelligence per cost ratios.

GPT-6 Sol is smarter, but seems to reason 2x more than GPT-5.6, which makes it 2x slow3r and 25% more expensive in practice.

[0]: https://aibenchy.com/compare/openai-gpt-6-sol-high/openai-gp...

jumploops 2 days ago|
I’m still finding context is king, even with the best models.

For example, I had Fable review Astra’s output yesterday, and it found some issues and fixed them. Passing the fixes back, Astra then uncovered additional issues with Fable’s fixes (and yes, this will go on ad infinitum if you let it, but these were “real” issues).

It seems the big story here is the reduced Luna pricing. It’s a fantastic model that can handle most automation needs (though I still use the big models for day-to-day development).

More comments...