Top
Best
New

Posted by meetpateltech 6 hours ago

Grok 4.7(x.ai)
437 points | 362 commentspage 2
maz1b 5 hours ago|
Either way, the fact that xAI or SpaceXAI or whatever the name is, I can commend the team behind it on their rapid ascent and progress by being close and or on the frontier in several respects.
avazhi 5 hours ago|
Your comment is like 6 months to a year late.

There for awhile it seemed like we’d have 3 big competitors but then Grok 4.2 or 4.4 was just diabolical while OAI and Claude continued their significant improvements. Grok was/is so bad that I was convinced musk was gonna shut it down and just fund Anthropic compute once they reached their compute agreement.

sarjann 1 hour ago||
Why are they comparing grok 4.7 xhigh to grok 4.6 high?

Unless they produce the same token output on the face of it, it looks like they're trying to cover for 4.7 not having good model perf?

WarmWash 5 hours ago||
Good thing they used 5.6 sol instead of Astra for benchmarks, the EEbench one is crazy[1]

[1]https://eebench.org/

johnfahey 5 hours ago||
No doubt xAI has seen rapid progress, but it's been several months of them being "just behind" OpenAI and Anthropic. It seems the gap between just behind the frontier and pushing it is a lot wider than most people thought it was a year ago, and that's why a clear third contender in the frontier model space has yet to materialize.
stiltzkin 2 hours ago|
[dead]
notduckrabbit 5 hours ago||
Significant regression in token efficiency compared to Grok 4.6 suggested by artificialanalysis.ai Intelligence Index Comparisons.
sourcecodeplz 3 hours ago||
Output tokens from Intelligence Index:

- grok 4.6 (xhigh): 97M (for 44 score)

- grok 4.7 (xhigh): 240M (for 46 score)

everfrustrated 4 hours ago||
That is comparing Grok 4.6 high to Grok 4.7 xhigh tho.
notduckrabbit 4 hours ago||
No, you can add Grox 4.7 high to the chart. 36k vs 66k
GodelNumbering 4 hours ago||
Every Grok release obscures their cache pricing while highlighting their input/output pricing

From their headline comparison:

  Grok: $2/$6 per million
  
  Fable: $10/$50 per million


  What this doesn't say: Grok costs 0.50/M cache read, Fable $0.25/M cache read
Long running agentic workflows are dominated by cache reads.

Just makes Grok sound deceptive, and more importantly, reliant on user's lack of understanding of costs aka predatory (which in turn is more infuriating)

sourcecodeplz 3 hours ago|
muse spark 1.3 contribs cache read is $0.002 btw (~220x diff).
bastawhiz 1 hour ago||
Is anyone treating Meta's offerings as a serious contender in any real use case? Zuck and co are burning cash hard to try to get people using their models after falling off the wagon for a couple years. It would be wild if those prices aren't total loss leaders.
c0rruptbytes 4 hours ago||
as someone who is limited by amazon bedrock support at work (no idea why we got stuck with the worst one) - grok is literally the only budget-ish model option, so nice to see it updated, Sol and Opus are just too rich for my blood. Luna is good but so slow at getting things done (tps wise it's fast)
mh- 2 hours ago|
Are the prices on Bedrock substantially different to the rate cards of the direct APIs? Just trying to understand whether this is Opus-through-Bedrock is too expensive, or Opus is too expensive.
gslepak 5 hours ago||
Does anyone have any experience with Grok's subscription? How does it compare price-wise to the API?
daquisu 3 hours ago||
There are some users reporting it improved a lot in the last few weeks. The max sub usage for Grok is around $12,000 of API pricing now, so a 40x multiplier for the $300 plan.

It is the same multiplier for Sol with subscription. For Astra though the multiplier is ≈20x, so half of Sol usage.

For Claude it seems to be ≈40x too for Opus, but less for Fable (similar to Astra in GPT).

All on the most expensive plan. Previously, Grok usage escalated linearly from the $100 plan to $300 plan. That would be a really good $100 plan if it is still true.

Some sources:

1. https://x.com/kunchenguid/status/2098256018836963382

2. https://x.com/stevenzhang/status/2092110386569089311

3. https://github.com/openai/codex/issues/43731

4. https://redd.it/1wciwc1

5. https://x.com/SemiAnalysis_/status/2064815044085318040

6. https://redd.it/1vx0k69

everfrustrated 4 hours ago|||
I find I can just about get by with coding every day on a Cursor $60/mth sub with Grok fast mode disabled. Doing pretty heavy coding work/requirements etc, but not much sub agents and no loops.

For me and what I’m doing that’s insanely good value.

I find grok build chews through my SuperGrok sub very quick - but I think that is due to it having the 500k context window which uses more credits. Cursor limits it to 256K (tho I see in today’s update for Grok 4.7 there’s now a toggle for context size).

thefourthchime 3 hours ago|||
There are two ways to subscribe, and it’s very confusing, but the best value is to get cursor ultra for $200 a month. I basically have infinite tokens with that plan, plus grok bot, which I really like
andreyvit 4 hours ago|||
Well when I ran out of Grok SuperHeavy subscription ($300) once and tried to use extra credits to cover half a day remaining till reset, $50 in extra credits went in two hours. Based on that, subscription definitely lasts longer; Grok subscription just about covers a week of my work (sometimes a bit extra remains unused, sometimes it runs out half a day to a day early). And as a point of comparison, it lasts for doing same tasks as 2.5-3 weekly limits of Codex on 5.6 Sol did (using xhigh on both Sol and Grok); I needed 3x$200 Codex subscriptions to cover my weekly usage.
nwienert 4 hours ago||
By far the worst value subscription of any. I tried Superheavy and got about 5-10% the usage of CC/Codex.
ls1911 6 hours ago||
after using cursor grok & trae.ai for several months , grok curor is highly superior results to trae.ai
shdtabasum 5 hours ago|
Why Chinese models from Kimi, Deepseek are not added in comparison benchmarks?
xquce 4 hours ago|
Same reason Coca-Cola only mention Pepsi and Pepsi only mention Coca-Cola. It's an proven way to capture the market. You would rather split the pie in two rather than in 4,12 or 50 right?
More comments...