Top
Best
New

Posted by meetpateltech 8 hours ago

Grok 4.7(x.ai)
466 points | 382 commentspage 4
andsoitis 7 hours ago|
Congratulations to the team!
oh_no 5 hours ago||
the AA numbers are generationally bad. double token use (the one thing Grok was good at was low reasoning usage!) to gain 5% in the benchmark score. with reportedly a larger model. maybe it shows gains IRL but wow, I've never seen a new generation model look so underwhelming compared to the last.
jascha_eng 5 hours ago||
32 on the omniscience index. Not terrible but far from Astra and fable: https://artificialanalysis.ai/evaluations/omniscience
MuffinFlavored 7 hours ago||
If the CursorBench 4.0 score diagram is the headline, I read it as "Grok 4.7 xHigh is almost the same as Fable5.1 on low".

Is there a metric for like... time taken when comparing these two? I see score and cost.

If Fable5.1 can knock it out more quickly on low but Grok4.7 might take twice as long to stumble through a problem (and leave behind a bunch of yucky comments or un-needed extra unit tests), are they really comparable?

Or like... the "quality" of the solution? "It works" versus "it's unmaintainable/very messy/hacky".

Invictus0 4 hours ago||
SpaceX AI releasing "Grok" has to be some of the worst branding I've seen in my lifetime
DonHopkins 4 hours ago|
xAI should offer persecuted White South Africans deep reverse-discrimination-victim discounts on Grok tokens.
inshard 4 hours ago||
Any real world experience with Grok Ultra $300 monthly subscription vs Claude Code Max in terms of overall built work mileage, or general token limits?
mpalczewski 3 hours ago|
Yeah I switched and the 300 plan is basically introductory 100/ month and I never hit the limit. While constantly hammering on it
gaigalas 5 hours ago||
Pacing the frontier, with an aggressive release cadence. Gotta love the US tech industry.
brcmthrowaway 5 hours ago||
Dumb question. Are these products really winner-take-all? Why is there such a furious rate of development?
dgellow 4 hours ago||
It’s not at all winner takes all, it’s a race to the bottom. Models are becoming a commodity
hdhdjdif 5 hours ago||
because boomers will give you free money + tip

musk can fund the space stuff with this

thih9 6 hours ago|
I refuse to use Grok. Mostly because of the usual reasons - somehow this high profile AI model seems more disgusting than others and it is in a way impressive.

But also Xai doesn’t seem to care about user experience and long term support.

eknkc 6 hours ago||
I am subscribed to ChatGPT, Claude, Kimi and GLM coding plans. 200$ one on GPT and the 20$ ish ones on all others. Recently added Grok and it has somehow bacome my second most used model.

For daily one off questions I prefer it because it is fast enough and I like the way it responds. I also use it for basic research like “find me a battery drill for this and that”.

Kimi and GLM feel extremely coding oriented. I use them for code reviews basically. I hate the way Anthropic models talk. GPT takes too much time and effort for that kind of stuff for some reason.

Grok happened to be a nice middle ground.

brandonagr2 6 hours ago|||
You should try it, it is less sycophantic than other models and is faster and better at most reasoning levels, don't confuse the twitter bots and services also named Grok with the frontier model itself
venzaspa 2 hours ago||
Perhaps he doesn't want to use it because it's owned by human being who many people view as vile.
swozey 6 hours ago||
I can't take anyone seriously who uses grok seriously. I like to look at the cybertruck owners forum every so often because it's just... hilarious. And the amount of superfluous grok use over there is just insane. Half the posts I click in there will have a bunch of people dumping entire grok takes "why do people hate cybertruck owners?" "Because they're jealous and poor," sort of stuff that they just LOVE to post.

As a technical point of reference to compare against other llm stuff, sure, I'll glance at a report or benchmark but I really couldn't care less about anything to do with the project and it could blow other options away and I wouldn't touch it.

ElectronCharge 5 hours ago||
Possibly interestingly, I can't take you seriously for having such a superficial approach.

You probably shouldn't cut off your nose to spite your face.

totallymike 7 minutes ago|||
Having ethics is not generally considered superficial. Musk is a deplorable person, and choosing not to use his CSAM generator seems like a pretty good idea.
mempko 4 hours ago|||
I don't know man, Musk doing Nazi salutes doesn't seem that superficial. He did help get Trump in power and also killed a lot of aid to children that need it.

What's superficial about refusing to use a product from someone like that? Or are you one of those 'technology isn't about politics' people? That's a superficial take if you ask me.

All technology is political, and understanding that is a deep, not superficial take. It requires systems thinking which unfortunately many people building technology seem to lack, despite software being a sophisticated complex system.

More comments...